Summary
An experienced Site Reliability Engineer is needed to own and maintain cloud-based infrastructure deployed to customer environments. You will collaborate with application, IT, and software teams to ensure reliable deployments and operational performance while working on sophisticated AI and robotics systems.
Highlights
Senior infrastructure role supporting the deployment of cloud-based systems to customer environments. Offers the chance to work across applications, IT, software, cloud infrastructure, and advanced robotics and AI systems.
Description
Apptronik is a human-centered robotics company developing AI-powered robots to support humanity in every facet of life.
Our flagship humanoid robot, Apollo, is built to collaborate thoughtfully with people, starting with critical industries such as manufacturing and logistics, with future applications in healthcare, the home, and beyond.
We operate at the cutting edge of Applied AI, applying our expertise across the full robotics stack to solve some of society's most important problems.
You will join a team dedicated to bringing Apollo to market at scale, tackling the complex challenges like safety, commercialization, and mass production to change the world for the better.
JOB SUMMARY
We are seeking an experienced Site Reliability Engineer to own and maintain the deployment of our cloud-based infrastructure to customer sites.
In this role, you will work closely with our Applications Engineers, IT, and software teams to ensure the smooth deployment of our solution which collects training data and deploys models to real-time robotic systems.
Providing a reliable framework will accelerate progress integrating Google DeepMind's Gemini Robotics Model to humanoid robot hardware.
ESSENTIAL DUTIES AND RESPONSIBILITIES or KEY ACCOUNTABILITIES
Partnering with customers and Applications Engineers to remove roadblocks to deployment successWriting and fixing Infrastructure as Code (Terraform, Helm, Ansible)Developing and maintaining code in Python / TypescriptTroubleshooting networking, performance, and security challengesCollaborating with engineering and product teams to shape improvementsResponding to outages, participating in on-call rotations and traveling occasionally to customer sites to support deployment and integration
SKILLS AND REQUIREMENTS
Strong communication skills, customer empathy and flexibility to adaptHands-on Linux systems engineering and networking experienceProficiency with Infrastructure as Code tools (Terraform, Helm, Ansible)Development experience in Kubernetes / Python / Typescript / C++Experience creating intuitive and high-utility dashboards and vizualizations (Grafana)Integrating real-time monitoring and alerting (eg PagerDuty)Takes initiative and seeks ownership of the end-to-end infrastructure solutionWillingness to travel as needed for client support
EDUCATION and/or EXPERIENCE
Bachelor's or Master's degree in Computer Science, Engineering, or a related technical field.Minimum of 5 years of professional, full-time experience building and maintaining reliable, scalable systems.
PHYSICAL REQUIREMENTS
Prolonged periods of sitting at a desk and working on a computerMust be able to lift 15 pounds at timesVision to read printed materials and a computer screenHearing and speech to communicateThis is a direct hire.
Please, no outside Agency solicitations.
Apptronik provides equal employment opportunities to all employees and applicants for employment and prohibits discrimination and harassment of any type without regard to race, color, religion, age, sex, national origin, disability status, genetics, protected veteran status, sexual orientation, gender identity or expression, or any other characteristic protected by federal, state or local laws.