Summary
✨ AI‑Generated
Seeking a full stack engineer who can own production health, troubleshoot complex issues, and improve applications across frontend, backend, databases, and infrastructure.
Highlights
High-impact engineering role focused on solving complex production issues, improving reliability, and working across the full technology stack.
Description
Full Stack Production Software Engineer
We are looking for a highly capable and curious Full Stack Production Software Engineer who thrives on solving complex production problems.
You will serve as the first engineering responder for production incidents, owning technical triage across our entire application stack.
Your responsibility is to rapidly diagnose issues, restore service, identify root causes, and partner with engineering teams to ensure problems are permanently resolved.
Your core responsibility is to own the health of our production applications.
You will investigate issues spanning React frontends, Laravel APIs, databases, cloud infrastructure, and third-party integrations.
You’ll determine whether a problem is caused by application code, configuration, infrastructure, data, or external dependencies, then either resolve it directly or coordinate with the appropriate engineering teams.
You bring strong full stack engineering experience, excellent debugging skills, and the ability to remain calm under pressure.
You enjoy understanding how systems work, communicating clearly during incidents, and eliminating recurring problems rather than repeatedly fighting the same fires.
Beyond incident response, you will contribute to improving our operational maturity by enhancing observability, refining on-call processes, building internal tooling, documenting runbooks, and driving postmortem action items that improve long-term reliability.
This is a high-trust, high-impact role for someone who enjoys working across the entire technology stack and believes production excellence is a core engineering discipline.
If this role speaks to your strengths, we’d love to meet you.
Key Responsibilities:
Primary: Production Incident Response & Technical Triage
Serve as the primary engineering on-call responder for production incidents.Diagnose issues across React applications, Laravel services, APIs, databases, cloud infrastructure, and third-party integrations.Restore service quickly while balancing immediate mitigation with long-term reliability.Determine root cause and coordinate resolution across engineering teams when necessary.Clearly communicate incident status, customer impact, and resolution progress during production events.
Supporting: Application Reliability
Perform root cause analysis for production incidents and recurring issues.Implement permanent fixes when appropriate or partner with feature teams to drive resolution.Improve application performance, stability, and resilience through engineering improvements.Identify recurring operational issues and eliminate them through automation or architectural improvements.
Supporting: Observability & Operational Excellence
Build and maintain dashboards, alerts, logging, and tracing to improve system visibility.Improve monitoring strategies to reduce alert fatigue while increasing issue detection.Develop internal tools that improve engineering efficiency and operational workflows.Create and maintain runbooks, troubleshooting guides, and operational documentation.
Supporting: Engineering Collaboration
Partner with Product Engineering, Infrastructure, QA, and Support to resolve customer-impacting issues.Participate in incident reviews and drive corrective actions to completion.Contribute code changes to Laravel and React applications when needed.Advocate for engineering practices that improve reliability, scalability, and maintainability.
Qualifications:
Required
3+ years of professional software engineering experience.Strong experience developing applications with PHP (Laravel).Strong experience building modern web applications using JavaScript and React.Experience troubleshooting production systems in a cloud environment.Solid understanding of REST APIs, relational databases, and distributed application architecture.Experience working with Git and modern software development workflows.Comfortable working in Linux environments.
Key Competencies
Excellent debugging and problem-solving skills across frontend and backend systems.Ability to rapidly isolate issues and determine root cause under pressure.Strong systems thinking and curiosity about how complex applications behave in production.Excellent communication skills during incident response and cross-functional collaboration.Ownership mindset with the ability to independently drive issues to resolution.Ability to balance urgency with thoughtful engineering decisions.Continuous improvement mindset with a passion for automation and operational excellence.
Nice to Have
Experience participating in an engineering on-call rotation.Experience with AWS or other cloud platforms.Familiarity with Docker, Kubernetes, Redis, queues, or event-driven architectures.Familiarity with microservice architectures and developing services in Go (Golang).Experience using observability platforms such as Datadog, Grafana, New Relic, Sentry, or similar.Experience with CI/CD pipelines and deployment automation.Experience conducting postmortems and driving long-term reliability improvements.