Backend Engineer - AI

Cygnify — Singapore · Posted ~8 hours ago

Mid Full-time

Skills

Backend engineering API development AI inference systems Production systems Monitoring Performance optimization AI LLMs APIs Backend systems Cloud

🔓 Log in to save this job, tailor your resume & track your apply process — 7 days free, no card needed.

Log in to add to target list

Summary ✨ AI‑Generated

A technology company is hiring a Backend Engineer to build AI-powered backend services. The role involves designing inference systems, scalable APIs, observability solutions, and performance improvements for production workloads.

Highlights

Build production AI infrastructure focused on reliability, performance, and scalable user-facing services.

Description

Backend Engineer, AI Role As a Backend Engineer, AI, you own the inference and orchestration layer that powers every AI interaction in the product. Your work sits between models and users, where latency, correctness, reliability, and cost directly impact real-world experience. You will build and operate production systems that turn model capability into fast, stable, observable APIs used across mobile and desktop clients. Focus Build and operate backend systems that serve AI-powered features in production.Design inference pipelines, orchestration layers, and service boundaries around models.Own production concerns: monitoring, logging, alerting, and incident response.Optimize latency and throughput across inference, caching, batching, and streaming. Ideal Experiences Strong backend engineering fundamentals in production environments.Experience running high-throughput, low-latency services.Familiarity with AI inference patterns (LLMs, embeddings, multimodal).Comfortable debugging distributed systems under load.Bias toward shipping and learning from production behavior. Outcomes Backend systems run reliably at scale, handling production AI traffic with low latency and high throughput.APIs are stable, clear, and support seamless integration with frontend and ML systems.Production incidents are quickly detected, diagnosed, and resolved, minimizing user impact.Iterative improvements based on real usage continuously increase system performance and reliability. Tech Stack PythonNodeJsPytorchOpenAI / Anthropic / open-source LLMsSQl & noSQLKubernetesDocker