LLMOps Engineer - AI Model Evaluation

Thrive It Systems Ltd — Poland · Posted ~2 hours ago

Mid Full-time

Skills

software engineering data analysis AI model evaluation LLMOps LLMs DeepEval AI evaluation

🔓 Log in to save this job, tailor your resume & track your apply process — 7 days free, no card needed.

Log in to add to target list

Summary ✨ AI‑Generated

An AI engineering role focused on creating benchmarks and evaluation solutions that measure advanced model capabilities. The position combines software development with data-driven analysis.

Highlights

Opportunity to build evaluation systems for advanced AI models combining software engineering and analytical skills.

Description

This row requires software engineering skills and data analysis skills. About the Role The engineer would need to develop and own the solution which determines if a model has cyber capabilities. Evaluation of models or AI tooling using frameworks such as DeepEval would be a plus.Set up a benchmark and execute it which would run with an output determining if a model has advanced cyber capabilities.