Summary
✨ AI‑Generated
A staff-level software engineering role focused on secure, high-performance runtime infrastructure and large-scale API systems. You will build fault-tolerant distributed components in Rust, debug complex production issues, guide architectural decisions, support enterprise users, and mentor other engineers.
Highlights
High-impact staff-level ownership of secure, reliable runtime infrastructure, with opportunities for architectural leadership, mentorship, production problem solving, and work on large-scale distributed systems.
Description
About the Role
This opportunity is for a Staff Software Engineer focused on building, operating, and evolving secure, high-performance, and highly reliable runtime infrastructure.
The role centers on large-scale GraphQL systems, fault-tolerant architecture, public APIs, production debugging, and distributed systems engineering.
You will take significant technical ownership of critical runtime infrastructure, solve complex production problems, work directly with users and enterprise customers, and help guide architectural decisions across engineering teams.
This is a hands-on role that combines software development, systems engineering, operational ownership, mentorship, and cross-team technical leadership.
What You’ll Do
Build, test, and maintain fault-tolerant infrastructure for GraphQL runtime platforms, primarily using idiomatic RustDesign systems with strong security, performance, reliability, and maintainability requirementsTriage, debug, and resolve escalations from enterprise customers operating large-scale GraphQL deployments with hundreds of subgraphs and trillions of monthly requestsOperate and improve durable, stable public APIs used by demanding GraphQL workloadsEngage directly with community members and enterprise customers to understand technical needs, investigate issues, and use real-world feedback to improve the platformDesign scalable and observable systems that integrate effectively with a wide range of customer infrastructure environmentsUse independent technical research and production feedback to guide system improvementsCollaborate with engineers across teams through clear communication and constructive code reviewsMentor and guide engineers in designing systems and writing idiomatic, maintainable Rust codeEvaluate the end-to-end impact of technical changes and ensure alignment with cross-domain concernsLead architectural discussions and cross-team technical initiativesDrive significant technical improvements while helping other engineers develop their own leadership capabilitiesCreate comprehensive technical designs and documentation addressing cost efficiency, security, reliability, and observabilityParticipate in on-call rotations as a core responsibility to maintain the reliability of mission-critical systems
Qualifications
Professional experience with Rust and a strong ability to write performant, maintainable codeStrong systems engineering expertiseKnowledge of stateless and fault-tolerant system designUnderstanding of event-driven architectures and distributed systems paradigmsStrong debugging skills and comfort performing hands-on operational work in production environmentsAbility to diagnose complex technical problems rather than focusing exclusively on feature developmentStrong cross-team collaboration skills and the ability to positively influence engineers across an organizationInterest in GraphQL, modern developer tooling, and advanced infrastructure engineeringGrowth mindset with a commitment to continuous learning and staying current with evolving engineering practices
Nice to Have
Experience with GraphQL or large-scale runtime systemsExperience working in environments with frequent high-priority escalations or production incidents
Pay: $192,000 - $230,000 per year.
Benefits
Equity participationChoice of multiple medical plans for U.S.
employeesAdditional medical plan options available to California residentsDental benefitsVision benefits