Summary
A technology services organization is seeking a machine learning infrastructure engineer to build reliable platforms for serving large AI models. The role focuses on distributed systems, GPU utilization, scalability, cloud infrastructure, and production reliability.
Highlights
Fully remote opportunity focused on cutting-edge AI infrastructure, large-scale systems, performance optimization, and advanced machine learning workloads.
Description
Machine Learning Infrastructure Engineer – Remote
Bright Vision Technologies is a technology consulting and software development company delivering cloud, AI, data, and enterprise solutions across the United States.
This is a fantastic opportunity to join an established and well-respected organization offering tremendous career growth potential.
Job Title: Machine Learning Infrastructure Engineer
Location: 100% Remote (U.S.)
Position Type: Full-time, Direct W2
Salary Range: $100,000–$150,000 Annually
Experience Required: 6+ years
Sponsorship: U.S.
Citizens, Green Card Holders, EAD Holders, and H-1B transfer candidates are encouraged to apply.
We are unable to sponsor new H-1B visa petitions for this position.
Job Summary
We are seeking a Machine Learning Infrastructure Engineer to design, build, and operate high-performance, highly reliable inference platforms for serving large machine learning models in production.
The role focuses on the systems engineering side of AI deployment, including request routing, batching, caching, autoscaling, GPU utilization, and end-to-end observability across diverse model workloads.
The ideal candidate brings strong distributed systems and performance engineering expertise, has shipped serving systems at scale, and understands the trade-offs between latency, throughput, cost, and quality in ML serving.
Required Qualifications
Bachelor’s or Master’s degree in Computer Science or a related fieldSix or more years of experience in distributed systems, infrastructure, or ML platform engineeringStrong proficiency in Python and a systems language such as Go, Rust, or C++Deep experience operating high-throughput, low-latency services in productionHands-on experience with LLM or large model inference frameworks such as vLLM or TensorRT-LLMStrong understanding of GPU architecture, memory hierarchies, and accelerator utilizationFamiliarity with Kubernetes, autoscaling, and modern cloud platformsExperience with observability stacks including metrics, tracing, and structured loggingSolid grounding in performance engineering and capacity planningStrong communication and incident response skills
Preferred Qualifications
Open-source contributions to model serving infrastructureExperience with multi-region or globally distributed AI servingFamiliarity with model quantization, distillation, and compression techniquesExposure to FinOps for AI workloads and cost-efficient serving designExperience supporting external-facing AI APIs at scale
How to Apply
Would you like to know more about this opportunity? For immediate consideration, please send your resume to hilda@bvteck.com.
Bright Vision Technologies is an Equal Opportunity Employer.
Equal Employment Opportunity (EEO) Statement
Bright Vision Technologies (BV Teck) is committed to equal employment opportunity (EEO) for all employees and applicants without regard to race, color, religion, sex, sexual orientation, gender identity or expression, national origin, age, genetic information, disability, veteran status, or any other protected status as defined by applicable federal, state, or local laws.
This commitment extends to all aspects of employment, including recruitment, hiring, training, compensation, promotion, transfer, leaves of absence, termination, layoffs, and recall.
BV Teck expressly prohibits any form of workplace harassment or discrimination.
Any improper interference with employees' ability to perform their job duties may result in disciplinary action up to and including termination of employment.
Powered by JazzHR
floBpwUgCF