Senior Staff Software Engineer, AI Agent Platform

Google — Poland · Posted ~4 hours ago

Lead Visa History ✓

Skills

software development ML infrastructure ML system design model deployment model evaluation data processing debugging fine-tuning systems architecture large-scale serving infrastructure technical leadership data structures and algorithms international collaboration machine learning infrastructure

🔓 Log in to save this job, tailor your resume & track your apply process — 7 days free, no card needed.

Log in to add to target list

Summary ✨ AI‑Generated

Take a senior technical leadership role building and shaping large-scale AI agent infrastructure. You will define technical direction, design ML systems, work on model deployment and evaluation, and architect high-scale serving infrastructure while collaborating across global, cross-functional teams.

Highlights

Lead technical strategy and architecture for large-scale AI infrastructure, with substantial influence over ML system design and global initiatives. The role offers senior technical leadership across complex, cross-functional projects.

Description

Minimum qualifications: Bachelor’s degree or equivalent practical experience.8 years of experience in software development.7 years of experience leading technical project strategy, ML design, and working with ML infrastructure (e.g., model deployment, model evaluation, data processing, debugging, fine tuning).5 years of experience in systems architecture and large-scale serving infrastructure.Experience delivering technical solutions in a changing environment.Experience working in multinational or internationalised (i18n) organisations on global initiatives. Preferred qualifications: Master’s degree or PhD in Engineering, Computer Science, or a related technical field.8 years of experience with data structures and algorithms.5 years of experience in a technical leadership role leading project teams and setting technical direction.3 years of experience working in a complex, matrix organisations involving cross-functional, cross-timezone or cross-business projects.Experience in serving Generative AI models using inference frameworks or hardware accelerators (e.g., GPUs, TPUs). About the jobGoogle Cloud’s mission is to make every business successful through AI by combining cutting-edge technology, infrastructure, and talent. AI/ML software engineers in Cloud bridge the gap between pioneering models and a massive product vehicle reaching billions. Our talent density and AI-powered tools drive rapid development, rooted in a culture of empowerment and a bias to action. In this role, you aren’t just building technology; you’re shaping the frontier of enterprise and driving the evolution of advanced models. As a part of the Cloud AI Agent Platform team (previously known as Vertex), you will focus on building highly differentiated, highly scalable, and easy-to-use GenAI products and services that enable customers to transform their business with GenAI. You will be responsible for building the high-performance, hyper-scale backend infrastructure that powers the Vertex AI APIs, serving millions of requests and enabling the next-generation of AI-driven applications. In this role, you will have a unique opportunity to be at the absolute forefront of the Generative AI revolution. You will be building the foundational API layer and infrastructure that serves massive foundation models (like Gemini) to the world. In this role, you will address unprecedented scaling challenges, optimize for ultra-low latency, and design scalable systems that define how developers interact with state-of-the-art AI. You will operate in a highly dynamic, incredibly fast-paced environment where your work will have an immediate and massive impact on Google Cloud's most strategic product area.Individual pay is determined by factors including job-related skills, experience, and relevant education or training. Poland: zł640000 - zł655000 (PLN) + 25% bonus target + equity + benefits Responsibilities Learn more about benefits at Google . Architect, develop, and maintain robust, horizontally scalable APIs and backend infrastructure tailored specifically for Generative AI workloads.Optimize the routing, load balancing, quota management, and data plane mechanisms that connect user API requests to backend ML serving clusters across massive fleets of TPUs and GPUs.Prototype and launch new API products rapidly in a fast-paced, highly collaborative environment, rapidly to expose the latest advancements in Google's foundation models.Collaborate closely with ML researchers, ML serving teams, and product managers to translate complex AI capabilities into intuitive, highly reliable enterprise-grade APIs.Define and implement best practices for API security, high availability, fault tolerance, and comprehensive observability across distributed systems. Google is proud to be an equal opportunity workplace and is an affirmative action employer. We are committed to equal employment opportunity regardless of race, color, ancestry, religion, sex, national origin, sexual orientation, age, citizenship, marital status, disability, gender identity or Veteran status. We also consider qualified applicants regardless of criminal histories, consistent with legal requirements. See also Google's EEO Policy and EEO is the Law. If you have a disability or special need that requires accommodation, please let us know by completing our Accommodations for Applicants form .