The agent-quality flywheel: Using Gemini Enterprise Agent Platform evaluations to optimize agents
Skills:
Agent Foundations90%
Key Takeaways
Optimizes agent quality using Gemini Enterprise Agent Platform evaluations and the Quality Flywheel methodology
Original Description
Treating agent quality as a rigorous engineering discipline is the only way to scale. Stop guessing and start measuring. Join this session for a deep dive into the state-of-the-art techniques Google uses to build our own agents, and learn how to make them a part of your process. We’ll demonstrate how you can adopt the “Quality Flywheel” methodology, which includes bootstrapping effective offline evaluation with synthetic test generation, using LLM-as-a-judge autoraters and trajectory evaluations, performing user and environment simulation, identifying systemic failures in production with multi-turn autoraters and loss-pattern clustering, aligning test coverage with actual usage, and using automated optimization capabilities to scientifically refine performance. Ramp up with confidence.
Watch more: 100+ sessions from Google Cloud Next 26 → https://www.googlecloudevents.com/next-vegas/
Subscribe to Google Cloud Tech → https://goo.gle/GoogleCloudTech
Speakers: Dima Melnyk, Alex Martin, Daniel Lewis
BRK3-023
#GoogleCloudNext
Watch on YouTube ↗
(saves to browser)
Sign in to unlock AI tutor explanation · ⚡30
More on: Agent Foundations
View skill →Related Reads
📰
📰
📰
📰
Building Your First Model Context Protocol (MCP) Server with TypeScript and Zod
Dev.to AI
Beyond Code Generation: How OmniSVG Rethinks Vector Graphics with Vision-Language Models
Dev.to · Shrijith Venkatramana
LangGraph Checkpointing: Three Production Rewrites to Stop Losing State
Dev.to · Elena Revicheva
AI builder essentials: tokens, context windows and RAG 101
Dev.to · Tilde A. Thurium
🎓
Tutor Explanation
DeepCamp AI