Senior Site Reliability Engineer

Neuroflow — United States · Posted ~13 hours ago

Senior

Skills

AWS Azure Terraform

🔓 Log in to save this job, tailor your resume & track your apply process — 7 days free, no card needed.

Log in to add to target list

Summary ✨ AI‑Generated

Join a healthcare technology organization working to improve access to behavioral health support. Contribute to reliable, scalable technology that helps integrate behavioral healthcare into the broader healthcare ecosystem.

Highlights

Opportunity to improve healthcare outcomes through technology and contribute to a more integrated behavioral healthcare ecosystem.

Description

Who We Are NeuroFlow CEO and West Point graduate Christopher Molaro served in the army for five years, including a tour in Iraq as a platoon leader. Coming back home, he experienced firsthand the gaps in the behavioral health system and how veterans and civilians alike face too many barriers when it comes to receiving appropriate, timely care. While pursuing his MBA at Wharton, Chris met his future co-founder Adam Pardes, and the two agreed – even the most engaging digital mental health apps in the world wouldn’t truly change the problem; only a solution that systematically integrated behavioral health into the full healthcare ecosystem could create meaningful change. And so they created NeuroFlow. What We Do We pride ourselves on partnering with healthcare leaders to assist in driving better outcomes, lowering total cost of care, and making behavioral health risk more predictable and transparent. NeuroFlow exists to make sure no one who needs behavioral health support falls through the cracks. We build more than just engaging digital health tools for self-care: we create platforms that identify population behavioral health risk early, engage individuals with acuity-specific resources, and enable care teams to make smarter and more efficient decisions. Together, NeuroFlow’s solutions arm healthcare organizations with the insights they need to overcome the systemic challenges in today’s healthcare ecosystem. How We Do It The award-winning culture at NeuroFlow is one built around encouragement and daring to be great. Our core values have been displayed in our office since day one, and each team member is responsible for carrying out these values and keeping each other accountable to them. We succeed through our flexibility and agility, navigating and transforming an industry ripe for change where “no” or “can’t” is too often the default. NeuroFlow offers unique opportunities to work in a fun and challenging fast-paced environment with direct, meaningful impact on helping to close the divide between mental and physical health. NeuroFlow is backed by investors including HLM Venture Partners, SEMCAP Health, Concord Health Partners, Builders VC, and Dreamit, alongside grant funding from the National Science Foundation and the U.S. Department of Defense. Our work has been recognized by TIME Magazine as one of the World's Top HealthTech Companies of 2025, by Fierce Healthcare's 2026 Fierce15, and by MedTech Breakthrough's Best Overall Mental Health Solution award. About The Role As our Senior Site Reliability Engineer, you'll own how reliable our production systems are and how we measure it. You'll set our SLOs, lead incident response, build the tooling that lets teams ship and experiment safely, and keep our infrastructure secure and compliant across AWS, Azure, and our data centers. You'll work with engineering, security, and product partners to set the standard for how we build and ship. What You'll Do Reliability & Incident Response Define SLOs and SLIs with stakeholders inside and outside engineering, back them with error budgets, and use them to decide when to ship and when to slow down.Run and improve our SaaS monitoring, alerting, and logging stack (Dynatrace, New Relic, Datadog, or similar) so we meet compliance requirements and catch problems before customers do.Take part in on-call for production incidents and help engineers work through customer issues.Lead blameless postmortems. When a failure keeps showing up, find the root cause and fix it for good.Find the manual, repetitive work that eats engineering time, measure what it costs, and automate it away. Track and report the reduction.Leave behind automation and documentation that survives a change of owner. Platform & Infrastructure Own the design, build, and upkeep of the core infrastructure that lets NeuroFlow engineers test and ship product changes quickly and safely.Set the priorities for reliability and platform work, bring them to your manager for approval, and adjust when business needs shift.Weigh risk against impact on high-visibility systems, and roll out improvements in steps you can measure and reverse.Make the call on build vs. buy. Keep engineering time focused on technology that sets NeuroFlow apart, and recommend proven tools for the rest. Delivery Partner with engineering teams to define and evolve our delivery standards, including CI/CD, testing, and safe deployment practices like canaries and automated rollback. Security & Compliance Work with the security team to keep our infrastructure and operations compliant, and raise risks before an audit or incident surfaces them. Influence & Mentorship Be the person engineers go to for reliability, Azure, AWS, Terraform, and operational practice, and help them grow their own skills in these areas.Write documentation that someone new to a system can pick up and follow.Lead with solutions. When you spot a gap, come with a plan and the tradeoffs already laid out. Qualifications 8+ years in site reliability, DevOps, or infrastructure engineering, including ownership of production systems from end to end.Built and managed AWS and Azure accounts and resources in line with the Well-Architected Framework, using services such as ECS, Fargate, RDS, Lambda, SNS, SQS, S3, EventBridge, and Step Functions.Deployed Docker-based software to production, with a solid grasp of what it takes to run containers reliably.Worked hands-on across Azure, AWS, and traditional data centers, including managing Windows VMs and IIS.Built deep expertise in SQL and relational database administration, including query tuning, index optimization, and resolving high-load production incidents (SQL Server preferred).Held yourself and your team to an infrastructure-as-code standard, with Terraform behind everything you create.Put DevOps principles, the 12-Factor App, least privilege access, and zero-trust architecture into practice in real systems.Led cross-team work to pin down requirements and define SLAs, adjusting the technical detail for each audience.Defined and operated against SLOs, and led incident response and postmortems for production outages.Worked within ITIL-aligned service management, including incident, problem, and change management.Mentored other engineers on reliability and delivery practices. Preferred Qualifications Azure or AWS certifications, such as AZ-104, AZ-305, or AWS Certified Solutions Architect.Experience with Azure Government or other federal cloud environments.Experience supporting the VA, DoD, or other federal health programs, including FedRAMP or ATO work.PowerShell scripting and automation.Experience in a HIPAA-compliant environment.Security work beyond day-to-day operations, such as red teaming or penetration testing. Security Requirements United States Citizenship.Applicants selected will be subject to a security investigation and eligibility requirements for access to classified (Public Trust) information. Company Benefits Applicable for full time employees Flexible work schedule, unlimited PTO, physical and mental wellness benefits, medical coverage, parental leave, 401K, company-sponsored events, referral program, onsite gym, dog friendly office, snacks in the office, commuter benefits, onsite massages. What We Believe NeuroFlow prohibits unlawful discrimination against any applicant or employee on the basis of race, color, religion, gender, gender identity, gender expression, sexual orientation, national origin, family or parental status, disability, age, veteran status, or any other status protected by applicable law. All employment decisions are based on qualifications, merit, and business needs. Applicants with disabilities may be entitled to reasonable accommodation under the terms of the Americans with Disabilities Act and certain state or local laws. A reasonable accommodation is a change in the way things are typically done which will ensure an equal employment opportunity without imposing undue hardship on NeuroFlow. Please inform our Talent team if you need any assistance completing any forms or to otherwise participate in the application process. As a HIPAA compliant organization, NeuroFlow expects all team members to: Act in accordance with NeuroFlow’s Information Security Policies.Protect organizational assets from unauthorized access, disclosure, modification, destruction or interference.Report security events or other risks to the organization.Execute organizational security processes or activities.Perform security responsibilities that defined and communicated for their role.Be responsible for their actions regarding the security of the organization.