Verified from career page · Posted 1w ago

Google

Senior Staff Software Engineer, AI Agent Platform, Vertex

Google · Google

Warsaw

Staff

Machine learningDistributed systemsGCPLLM / GenAI

Last seen 2h ago

Posted
1w ago

Posted on 15 September 2026

Workplace
Not specified

Work model not stated

Salary
Not disclosed

Salary range not shared by the company

Visa sponsorship
Not specified

Visa sponsorship details unknown

Google Cloud’s mission is to make every business successful through AI by combining cutting-edge technology, infrastructure, and talent. AI/ML software engineers in Cloud bridge the gap between pioneering models and a massive product vehicle reaching billions. Our talent density and AI-powered tools drive rapid development, rooted in a culture of empowerment and a bias to action. In this role, you aren’t just building technology; you’re shaping the frontier of enterprise and driving the evolution of advanced models.

As a part of the Cloud AI Agent Platform team (previously known as Vertex), you will focus on building highly differentiated, highly scalable, and easy-to-use GenAI products and services that enable customers to transform their business with GenAI.

You will be responsible for building the high-performance, hyper-scale backend infrastructure that powers the Vertex AI APIs, serving millions of requests and enabling the next-generation of AI-driven applications.

In this role, you will have a unique opportunity to be at the absolute forefront of the Generative AI revolution. You will be building the foundational API layer and infrastructure that serves massive foundation models (like Gemini) to the world. In this role, you will address unprecedented scaling challenges, optimize for ultra-low latency, and design scalable systems that define how developers interact with state-of-the-art AI.

You will operate in a highly dynamic, incredibly fast-paced environment where your work will have an immediate and massive impact on Google Cloud's most strategic product area.

Individual pay is determined by factors including job-related skills, experience, and relevant education or training.

Poland: zł640000 - zł655000 (PLN) + 25% bonus target + equity + benefits

Learn more about benefits at Google.

Architect, develop, and maintain robust, horizontally scalable APIs and backend infrastructure tailored specifically for Generative AI workloads.

Optimize the routing, load balancing, quota management, and data plane mechanisms that connect user API requests to backend ML serving clusters across massive fleets of TPUs and GPUs.

Prototype and launch new API products rapidly in a fast-paced, highly collaborative environment, rapidly to expose the latest advancements in Google's foundation models.

Collaborate closely with ML researchers, ML serving teams, and product managers to translate complex AI capabilities into intuitive, highly reliable enterprise-grade APIs.

Define and implement best practices for API security, high availability, fault tolerance, and comprehensive observability across distributed systems.

Minimum qualifications:

Bachelor’s degree or equivalent practical experience.

8 years of experience in software development.

7 years of experience leading technical project strategy, ML design, and working with ML infrastructure (e.g., model deployment, model evaluation, data processing, debugging, fine tuning).

5 years of experience in systems architecture and large-scale serving infrastructure.

Experience delivering technical solutions in a changing environment.

Experience working in multinational or internationalised (i18n) organisations on global initiatives.

Preferred qualifications:

Master’s degree or PhD in Engineering, Computer Science, or a related technical field.

8 years of experience with data structures and algorithms.

5 years of experience in a technical leadership role leading project teams and setting technical direction.

3 years of experience working in a complex, matrix organisations involving cross-functional, cross-timezone or cross-business projects.

Experience in serving Generative AI models using inference frameworks or hardware accelerators (e.g., GPUs, TPUs).

About Google

Google runs core product engineering out of London, Dublin, Zurich, Warsaw and Munich — not support functions. Zurich is one of its largest engineering sites anywhere, and Warsaw has grown substantially. The bar is high and the process is long, but these are genuine product teams.

Apply at Google

Similar roles

  1. Senior UX Designer, Health Ecosystem New Google Google London Design Top-tier Not disclosed Visa: Not specified yesterday
  2. Semiconductor Material Science Research Scientist, DeepMind New Google DeepMind London Machine learningPython Top-tier Not disclosed Visa: Not specified yesterday
  3. Football Partnership Marketing Manager New Google Google London Marketing Top-tier Not disclosed Visa: Not specified yesterday
  4. Data Analytics Sales Manager, Google Cloud New Google Google London SalesGCP Top-tier Not disclosed Visa: Not specified yesterday
  5. Customer Engineer, Platform, FSI, Google Cloud (Polish) New Google Google Warsaw Solutions & FieldGCP Top-tier Not disclosed Visa: Not specified yesterday
  6. DV360 Sales Manager, Northern Europe LCS New Google Google Dublin Sales Top-tier Not disclosed Visa: Not specified yesterday
1,318 more roles like this. Open the board →