Minimum qualifications:- Bachelor's degree or equivalent practical experience.
- 8 years of software engineering and architecture experience testing, and launching software products.
- Experience integrating generative AI tools or LLM interfaces into workflows.
- Experience with establishing standards for evaluation and quality metrics.
- Experience as a people manager.
Preferred qualifications:- Master's degree or PhD in Engineering, Computer Science, or a related technical field.
- 8 years of experience with data structures and algorithms.
- 3 years of experience working in a complex, matrixed organization involving cross-functional, or cross-business projects.
- Generative AI experience.
- Experience working with conversational AI reliability metrics (e.g., helpfulness, grounding, safety) and agent orchestration harnesses with validation protocols.
- Experience collaborating with Product, Data Science, and user-facing feature teams.
About the jobThe Business Agent team drives Google's Agentic Commerce goals, empowering seamless consumer interactions across Search, YouTube, Ads, and RCS Messaging.We are seeking a Tech Lead Manager to own the technical roadmap for evaluating, observing, and scaling AI agent performance across Google's ecosystem. In this role, you will bridge applied AI and engineering excellence-mentoring a nimble team, building evaluation infrastructure, and scaling LLM reliability for commerce.
People shop on Google more than a billion times a day - and the Commerce team is responsible for building the experiences that serve these users. The mission for Google Commerce is to be an essential part of the shopping journey for consumers - from inspiration to to a simple and secure checkout experience - and the best place for retailers/merchants to connect with consumers. We support and partner with the commerce ecosystem, from large retailers to small local merchants, to give them the tools, technology and scale to thrive in today's digital world.Individual pay is determined by factors including job-related skills, experience, and relevant education or training.
US: $207000 - $300000 (USD) 20% bonus target equity benefits
Learn more about benefits at Google .
Responsibilities - Own the technical goal for evaluating, observing, and iterating on AI agent performance across Google's commerce ecosystem.
- Lead the development of evaluation infrastructure and automated rating systems (AutoRaters) to measure conversation helpfulness and factuality at scale.
- Bridge applied AI and engineering excellence to transition evaluation frameworks into a self-service platform for internal partners.
- Solve complex reliability issues to maintain high quality standards for LLM reasoning and low-latency agent pipelines.
- Provide direction while remaining to mentor and grow a nimble, high-performing engineering team.