Senior Generative AI Engineer

Look4It — Poland · Posted ~1 hour ago

Senior Full-time

Skills

LLM RAG LangGraph FastAPI Python AI agents Azure OpenAI

🔓 Log in to save this job, tailor your resume & track your apply process — 7 days free, no card needed.

Log in to add to target list

Summary ✨ AI‑Generated

A technology team is seeking a senior AI engineer to build and operate production LLM applications. The role covers agent workflows, retrieval systems, APIs, evaluation processes, and performance optimization.

Highlights

Work on production-grade AI systems with ownership across development, deployment, evaluation, and optimization of advanced language model solutions.

Description

We are looking for a Senior AI Engineer to join a team building and operating a production-grade LLM system used by real users. This is not a proof-of-concept project. You will be responsible for taking LLM-powered features from idea through development and deployment to continuous improvement, with a strong focus on answer quality, performance, observability and cost efficiency. You will work with modern LLM technologies, including LangGraph, RAG, tool calling, Azure OpenAI, Gemini and Claude, and have a real impact on how AI-powered products are built and operated in production. Responsibilities: Build LLM-powered features end to end. Design and implement agentic flows, retrieval, and tool calling using LangGraph - then ship them as FastAPI services with streaming, persistence and proper tests.Own answer quality. Build evaluation datasets, regression suites and LLM-as-judge checks so we know whether a prompt or model change made things better before it reaches users.Get the right context to the model. Turn user questions into effective queries against our search platform, orchestrate multi-step research loops, and shape the context the model reasons over. When an answer is wrong, work out whether retrieval, the query or the prompt is at fault - and fix the right one.Debug production. Instrument flows with tracing (Langfuse), investigate bad answers from real traces, and manage latency, token and cost budgets - including routing across model sizes and families behind an AI gateway. Requirements: Experience in building and operating backend services - APIs, async, testingHands-on experience taking LLM features to production and keeping them running - not only prototypesAgent / orchestration frameworks - LangGraph ideallyPractical RAG experienceExperience debugging LLM systems in production - tracing, evaluation, cost and latencyExperience running services in the cloud (we're on Azure)Strong problem-solving skills, analytical thinking, and technical decision-makingFluent in English, proactive communicator, and a collaborative team playerOpen-minded, creative, and motivated to push boundaries in AI and automation Nice to have: Azure OpenAI, AI Search, App ServiceInfrastructure-as-code (Bicep)Mentoring or tech-lead experienceTech stack:PythonFastAPILangGraphAzure OpenAI, Google Gemini, Anthropic ClaudeFAISSPostgreSQLLangfuseAzure App InsightsBicepDocker/Podman Benefits: B2B contract100% remote workLong-term engagementWork on a live LLM productModern AI/LLM technology stackFlexible working environment