Description
Bright Vision Technologies is a technology consulting and software development company delivering cloud, AI, data, and enterprise solutions across the United States.
This is a fantastic opportunity to join an established and well-respected organization offering tremendous career growth potential.
Job Title: AI Platform Engineer
Location: 100% Remote (Continental United States)
Position Type: Full-time, Direct W2
Salary Range: $100,000 – $150,000 per annum
Experience: 6+ years
Sponsorship: U.S.
Citizens, Green Card Holders, EAD Holders, and H-1B transfer candidates are encouraged to apply.
We are unable to sponsor new H-1B visa petitions for this position.
Job Summary:
We are seeking an AI Platform Engineer to design, build, and operate scalable AI inference platforms for production ML workloads.
The ideal candidate will have expertise in distributed systems, LLM serving, GPU optimization, autoscaling, and cloud-native infrastructure, with a strong focus on performance, reliability, and observability.
Key Responsibilities:
Design and maintain scalable AI model serving platformsOptimize inference performance, GPU utilization, and request routingBuild autoscaling, deployment, and monitoring solutionsImplement caching, security, and high-availability strategiesCollaborate with ML teams to deploy and support production AI models
Required Qualifications:
6+ years of experience in distributed systems, infrastructure, or ML platform engineeringStrong proficiency in Python and Go, Rust, or C++Experience with LLM inference frameworks (vLLM, TensorRT-LLM), Kubernetes, cloud platforms, and GPU optimization
Preferred Qualifications:
Bright Vision Technologies is a technology consulting and software development company delivering cloud, AI, data, and enterprise solutions across the United States.
This is a fantastic opportunity to join an established and well-respected organization offering tremendous career growth potential.
Job Title
AI Platform Engineer
Location: 100% Remote (Continental United States)
Position Type: Full-time, Direct W2
Salary Range: $130,000–$180,000 Annually (based on experience)
Experience Required: 10+ Years
Sponsorship: U.S.
Citizens, Green Card Holders, EAD Holders, and H-1B transfer candidates are encouraged to apply.
We are unable to sponsor new H-1B visa petitions for this position.
Job Summary
Bright Vision Technologies is seeking a highly experienced AI Platform Engineer with 10+ years of experience in distributed systems, cloud-native infrastructure, and AI platform engineering to design, build, and operate enterprise-scale AI inference and machine learning platforms.
The ideal candidate will possess deep expertise in LLM serving, GPU optimization, Kubernetes, cloud infrastructure, distributed systems, and MLOps, with a proven ability to deliver highly scalable, reliable, secure, and cost-efficient AI platforms supporting production machine learning workloads.
Key Responsibilities
Design, build, and maintain scalable AI inference and model-serving platforms for enterprise production environmentsArchitect highly available, cloud-native infrastructure supporting Large Language Models (LLMs), foundation models, and machine learning servicesOptimize inference latency, throughput, GPU utilization, memory management, and request scheduling across distributed AI workloadsDesign autoscaling, workload orchestration, traffic management, and intelligent request routing strategies for AI servicesImplement model deployment, versioning, rollback, and lifecycle management using modern MLOps practicesDevelop monitoring, observability, logging, distributed tracing, and alerting solutions to ensure platform reliability and performanceImplement caching strategies, API gateways, security controls, authentication, authorization, and high-availability architecturesCollaborate with AI researchers, ML engineers, DevOps teams, and software engineers to deploy and support production AI modelsDrive cloud infrastructure optimization, resource utilization, FinOps initiatives, and operational excellenceMentor engineering teams, conduct architecture reviews, and establish best practices for AI platform engineering and cloud-native developmentEvaluate emerging AI infrastructure technologies, model-serving frameworks, and GPU acceleration techniques to drive continuous innovation
Required Qualifications
Bachelor's or Master's degree in Computer Science, Computer Engineering, Artificial Intelligence, or a related technical discipline10+ years of professional experience in distributed systems, infrastructure engineering, cloud platforms, or machine learning platform engineeringStrong programming skills in Python and at least one systems programming language such as Go, Rust, or C++Extensive experience with Large Language Model (LLM) serving, model inference optimization, and production AI infrastructureHands-on experience with vLLM, TensorRT-LLM, Triton Inference Server, Ray Serve, or similar AI serving frameworksStrong expertise in Kubernetes, container orchestration, Docker, and cloud-native application architecturesExperience optimizing GPU workloads using CUDA, NVIDIA GPU technologies, distributed inference, and high-performance AI infrastructureExperience with cloud platforms including AWS, Microsoft Azure, or Google Cloud Platform (GCP)Strong understanding of distributed systems, networking, scalability, observability, and security best practicesExcellent analytical, communication, collaboration, and technical leadership skills
Preferred Qualifications
Experience designing and operating multi-region AI platforms and globally distributed inference servicesKnowledge of model optimization techniques such as quantization, pruning, compression, speculative decoding, KV cache optimization, and mixed-precision inferenceExperience with MLOps, GitOps, Infrastructure as Code (Terraform, Bicep, CloudFormation), and CI/CD automationFamiliarity with service mesh technologies such as Istio or Linkerd, API gateways, and event-driven architecturesContributions to open-source AI infrastructure projects, technical publications, patents, or conference presentationsExperience implementing FinOps strategies, cloud cost optimization, and enterprise AI governanceExperience with multi-region AI deployments and AI infrastructureFamiliarity with model optimization techniques such as quantization or compressionOpen-source contributions or experience supporting large-scale AI APIs
Interested in this opportunity? Apply today for immediate consideration!
Email your updated resume to jaya@bvteck.com
Call or Text (908) 505-3545 if you have any questions.
Learn more about Bright Vision Technologies at www.bvteck.com.
We look forward to connecting with talented professionals and helping you take the next step in your career.
Bright Vision Technologies is an Equal Opportunity Employer.
Equal Employment Opportunity (EEO) Statement
Bright Vision Technologies (BV Teck) is committed to equal employment opportunity (EEO) for all employees and applicants without regard to race, color, religion, sex, sexual orientation, gender identity or expression, national origin, age, genetic information, disability, veteran status, or any other protected status as defined by applicable federal, state, or local laws.
This commitment extends to all aspects of employment, including recruitment, hiring, training, compensation, promotion, transfer, leaves of absence, termination, layoffs, and recall.
BV Teck expressly prohibits any form of workplace harassment or discrimination.
Any improper interference with employees' ability to perform their job duties may result in disciplinary action up to and including termination of employment.
Powered by JazzHR
9blt5tsDAm