Description
SRE Engineer – Cloud Environment
About the Job
The SRE Engineer – Cloud Environment will be responsible for building, operating, and improving Bambu Lab’s U.S.-based cloud infrastructure and production reliability environment.
This role will work closely with engineering, security, and infrastructure teams and will play a key role in maintaining cloud stability, Kubernetes reliability, observability, incident response, and secure infrastructure operations.
Location: Santa Clara, CA
Work Model: Onsite
Employment Type: Full-Time
Client Details
Happy Global’s client, Bambu Lab, is a technology company specializing in 3D printing products and solutions for consumer and professional markets.
The company develops an integrated ecosystem of 3D printing hardware, software, and related technologies and continues to expand its operations in the United States.
As Bambu Lab’s U.S.
technology organization continues to grow, the company is strengthening its local cloud infrastructure, reliability engineering, and security capabilities to support scalable and resilient systems.
Description
The SRE Engineer – Cloud Environment will be responsible for:
▪ Build and maintain cloud infrastructure across compute, networking, storage, and related platform services within AWS and GCP environments.
▪ Operate and maintain Kubernetes environments, ensuring platform stability, availability, scalability, and reliable production performance.
▪ Lead and support Tier 2 incident response, root-cause analysis, post-incident reviews, and remediation initiatives while participating in an on-call rotation.
▪ Develop and optimize observability systems, including monitoring, alerting, dashboards, and operational metrics, while defining and tracking SLIs and SLOs.
▪ Implement DevSecOps practices and U.S.-based infrastructure security policies, supporting the engineering implementation and operation of security governance requirements.
▪ Manage and automate cloud infrastructure through Infrastructure as Code (IaC), particularly Terraform, to improve consistency, scalability, and operational efficiency.
▪ Automate infrastructure and operational workflows using Python, Go, Shell, or related scripting and engineering tools.
▪ Partner with development, security, and infrastructure teams to improve system reliability, troubleshoot production issues, and continuously strengthen platform resilience.
Profile
A successful SRE Engineer – Cloud Environment should have:
▪ 3+ years of experience in Site Reliability Engineering, DevOps, systems operations, cloud infrastructure, or a related engineering function.
▪ Strong hands-on experience with Linux, TCP/IP networking, Kubernetes, and cloud platforms such as AWS or GCP.
▪ Demonstrated experience operating and troubleshooting production Kubernetes and cloud infrastructure environments.
▪ Experience with Infrastructure as Code, particularly Terraform, for provisioning and managing cloud infrastructure.
▪ Proficiency in at least one scripting or programming language such as Python, Go, or Shell.
▪ Hands-on experience with monitoring, alerting, observability, incident response, and root-cause analysis.
▪ Strong ownership of system reliability, availability, and operational stability.
▪ Ability to work onsite in Santa Clara, CA and participate in an on-call rotation.
▪ Legal authorization to work in the United States or eligibility for employer-sponsored work authorization.
Bambu Lab is willing to support qualified candidates requiring H-1B sponsorship or H-1B transfer.
Preferred Qualifications
▪ Bachelor's degree in Computer Science, Software Engineering, Communications Engineering, Information Security, or a related field; equivalent professional experience will also be considered.
▪ Experience defining and managing SLIs, SLOs, and reliability metrics for production systems.
▪ Experience implementing DevSecOps practices, cloud security controls, and infrastructure security policies.
▪ Experience supporting large-scale, highly available, or rapidly growing cloud environments.
▪ Experience working closely with software engineering, security, DevOps, and infrastructure teams.
▪ Familiarity with incident management, postmortem processes, and reliability improvement initiatives.
Compensation & Benefits
Base Salary: $150,000–$200,000 USD annually.
Additional compensation may include:
▪ Annual performance bonus
▪ Other performance-based incentive compensation, if applicable
Benefits include:
▪ Medical insurance
▪ Dental insurance
▪ Vision insurance
▪ 401(k) retirement plan
▪ Paid time off and other company-sponsored employee benefits, as applicable
▪ Professional development and career growth opportunities
Compensation may vary depending on experience, qualifications, technical expertise, location, and other job-related factors.
How to Apply
If your background aligns with this opportunity, we encourage you to apply.
Qualified candidates will be contacted by a member of the Happy Global recruiting team for an initial discussion regarding the opportunity with Bambu Lab.
This position includes participation in an on-call rotation.
All on-call arrangements, including standby requirements and applicable compensation, will be administered in compliance with FLSA and applicable California labor laws.
By applying for this position, candidates acknowledge that their application may be subject to Bambu Lab’s applicable candidate privacy policies.
Equal Employment Opportunity
Happy Global and its client, Bambu Lab, are committed to providing equal employment opportunities to all qualified applicants without regard to race, color, religion, religious creed, sex, sexual orientation, gender, gender identity or expression, national origin, ancestry, age, physical or mental disability, medical condition, genetic information, marital status, pregnancy, childbirth or related medical conditions, protected veteran status, or any other characteristic protected by applicable federal, California state, or local law.
All qualified applicants are encouraged to apply.