Forward Deployed Engineer, AI Inference Intern

Lyceum Technology — Germany · Posted ~4 hours ago

Junior Contract

Skills

AI inference Customer technical support Technical requirements analysis AI model selection GPU configuration Technical presentations Customer communication GPU Open-source AI models

🔓 Log in to save this job, tailor your resume & track your apply process — 7 days free, no card needed.

Log in to add to target list

Summary ✨ AI‑Generated

A fast-growing European AI infrastructure organization is looking for an intern to work at the intersection of engineering and customer engagement. You will analyze customer requirements, recommend suitable AI models and GPU configurations, lead technical sessions, support commercial discussions and help turn manual technical workflows into scalable products. This is an ideal opportunity for someone who enjoys both engineering and direct customer interaction.

Highlights

Gain hands-on experience at the intersection of AI infrastructure, engineering and customer-facing work. You will help match models and GPU configurations to customer needs, run technical sessions, support commercial opportunities and help automate currently manual processes into scalable products.

Description

About Lyceum Lyceum is a sovereign European AI inference provider. We run open-source models on our own GPU infrastructure, powered by 100% renewable energy, so teams can build with AI on their own terms – without giving up their data or getting locked into a single vendor. Backed by tier-1 investors, we're growing fast and our inference business is about to scale strongly. The Role As a Forward Deployed Engineer Intern, you own the technical side of our AI inference deals. You help customers figure out which models, GPUs and configurations fit their needs, run technical sessions with them, and work hand in hand with our commercial team to get deals closed. This is not a pure engineering role: you'll spend a lot of time with customers, and we're looking for someone who enjoys exactly that. Much of this work is still manual today – you'll help us turn it into product. What You'll Do Match customer requirements to the right models, GPUs and configurations for dedicated inferenceRun technical sessions with customers and help them make confident decisionsWork in tandem with our commercial team to move deals forwardSupport serverless and API customizations, and help turn recurring ones into productTranslate customer needs into clear technical specs for our engineering teamCollect benchmarks and learnings that help us automate matching and customizations What We're Looking For Studies in computer science, data science or a closely related fieldInterest in or first exposure to AI inference: LLMs, inference engines, GPUsReal excitement about working with customers and the commercial side – not just the technical oneStrong communication skills: you talk confidently to customers and engineers alikeAn entrepreneurial mindset: give you an outcome, and you find a way without getting blockedYou stay calm and constructive when your ideas are challengedFluent English Bonus Points Coursework or projects on inference engines or ML systems (e.g. vLLM, SGLang, TensorRT-LLM)Experience with GPU sizing, model serving, benchmarking or performance optimizationA previous internship at an AI infrastructure or inference companyStartup experienceGerman Why Join Us Cutting-edge work: Solve real AI inference problems with real customers from day oneRare mix: Combine technical and commercial work – unusual for an internshipReal ownership: Help build what becomes our product, from GPU matching to customizationsFounder access: Work directly with our product team and the foundersMission-driven team: Build sustainable, 100% renewable compute for the AI era, backed by top-tier investors