Senior Backend Software Engineer - Agent Platform

Jobgether — United States · Posted ~1 day ago

Senior Visa History ✓

Skills

Backend software development High availability High-throughput systems Distributed systems Production troubleshooting Systems engineering Incident response Scalability Operational excellence Backend systems High-availability systems

🔓 Log in to save this job, tailor your resume & track your apply process — 7 days free, no card needed.

Log in to add to target list

Summary ✨ AI‑Generated

A senior backend engineer is needed to build and operate highly available, high-throughput systems at global scale. The role blends hands-on development with production troubleshooting, incident resolution, service-boundary problem solving, and turning operational insights into durable platform improvements.

Highlights

Senior-level opportunity focused on globally scaled critical systems, deep production problem-solving, reliability engineering, and direct technical impact on essential platform services.

Description

This position is listed on behalf of a partner company, who manages all applications and next steps. Our partner is looking for a Senior Backend Software Engineer - Agent Platform based in the United States. This is a senior engineering opportunity focused on building and operating highly available, high-throughput backend systems that protect millions of devices worldwide. You’ll join an R&D environment where reliability, scalability, and operational excellence are central to engineering decisions. The role combines hands-on software development with deep production troubleshooting and systems-level problem solving. You’ll work across service boundaries, helping resolve complex incidents and turning operational learnings into durable platform improvements. Your work will directly influence critical services responsible for policy, configuration, and command distribution at global scale. You’ll collaborate with engineers and cross-functional partners while helping shape technical direction and platform architecture. The environment rewards ownership, curiosity, continuous learning, and engineers who enjoy solving difficult distributed-systems challenges. Accountabilities Lead rapid response to customer-critical production incidents, diagnosing, triaging, and resolving complex issues spanning multiple services across the agent platform.Conduct systematic root-cause analysis and translate incident findings into lasting improvements across reliability, scalability, observability, operability, and system resilience.Quickly understand unfamiliar services, architectures, and large codebases while partnering with engineering teams to troubleshoot and resolve cross-service failures.Design, develop, test, document, deploy, and operate large-scale distributed systems capable of processing millions of events per second with high availability and low latency.Build and evolve backend services responsible for policy, configuration, and command distribution to millions of endpoints worldwide.Maintain and improve existing services through refactoring, feature development, architectural enhancements, and ongoing operational improvements.Monitor application health, system metrics, and data integrity while strengthening operational visibility and platform stability.Translate business and functional requirements into robust, scalable, maintainable, and operationally sound technical solutions.Collaborate across engineering and other stakeholder groups to solve complex technical challenges, influence architecture and technical direction, and deliver scalable solutions.Evaluate and adopt technologies, engineering practices, and tooling that improve platform reliability, scalability, performance, and developer productivity. Requirements 5+ years of professional backend software engineering experience, with strong expertise in at least one of Java, Go, or Python and the ability to work across the broader technology stack.Demonstrated reliability-first mindset, including experience handling complex production incidents, performing root-cause analysis, and implementing preventative engineering improvements.Strong experience designing, building, and operating large-scale distributed systems, with a solid understanding of failure modes, performance trade-offs, resilience patterns, and operational excellence.Ability to quickly navigate unfamiliar systems and codebases and troubleshoot complex technical problems across multiple service boundaries.Hands-on experience with cloud platforms such as AWS or GCP and technologies including Docker, Helm, and Kubernetes.Experience with distributed messaging and data technologies such as Kafka, PostgreSQL, Redis, Cassandra, ClickHouse, or comparable platforms.Familiarity with service communication technologies such as gRPC, REST, GraphQL, or similar APIs.Strong software engineering fundamentals, including testing, documentation, deployment, monitoring, refactoring, and production operations.Excellent communication and collaboration skills, with the ability to work effectively with engineering teams, product partners, technical stakeholders, and customer-facing groups.Demonstrated ownership, autonomy, curiosity, and sound judgment when driving ambiguous and technically complex problems toward successful outcomes.Ability to influence technical direction and mentor other engineers while working effectively across organizational boundaries.Experience in enterprise SaaS, cybersecurity, endpoint security, or another high-scale technology environment is highly desirable. Benefits Base salary range of $132,000–$182,000 USD, with actual compensation varying based on candidate location and other relevant factors.Restricted Stock Units (RSUs) and Employee Stock Purchase Plan (ESPP).Flexible time off, paid company holidays, and paid sick time.Gender-neutral parental leave and grandparent leave.Medical, dental, and vision insurance coverage.401(k) retirement plan with company match.Life and disability insurance.Health and dependent-care flexible spending accounts.Voluntary benefits, including hospital, accident, and critical illness coverage.Employee Assistance Program and prepaid legal services.Pet insurance and specialized cancer-care support.Global business travel medical insurance.Home office allowance and mobile phone reimbursement.Wellness coaching and gym/wellness reimbursement.Fertility coverage and adoption and surrogacy reimbursement.Fully remote work opportunity within the United States. How Jobgether Works We use an AI-powered matching process to ensure your application is reviewed quickly, objectively, and fairly against the role's core requirements. Our system identifies the top-fitting candidates, and this shortlist is then shared directly with the hiring company. The final decision and next steps (interviews, assessments) are managed by their internal team. We appreciate your interest and wish you the best! Why Apply Through Jobgether? Data Privacy Notice: By submitting your application, you acknowledge that Jobgether will process your personal data to evaluate your candidacy and share relevant information with the hiring employer. This processing is based on legitimate interest and pre-contractual measures under applicable data protection laws (including GDPR). You may exercise your rights (access, rectification, erasure, objection) at any time. We may use artificial intelligence (AI) tools to support parts of the hiring process, such as reviewing applications, analyzing resumes, or assessing responses and identifying potential inconsistencies or verification signals in application materials based on available information. These tools assist our recruitment team but do not replace human judgment. Final hiring decisions are ultimately made by humans. If you would like more information about how your data is processed, please contact us.