Loading…
2026 September 10-11 | Tokyo, Japan
View More Details & Registration

IMPORTANT NOTE: Timing of sessions and room locations are subject to change.
Friday September 11, 2026 10:20 - 10:45 JST
The hardest part of agentic engineering isn't writing the code; It’s proving that a prompt tweak or a new tool schema didn't completely break your agent's routing logic. Traditional unit testing falls short when dealing with non-deterministic agent trajectories.

This session looks at how to build an automated local evaluation and regression testing pipeline for ADK workflows. We will walk through how to programmatically mock tool responses, simulate user edge cases, and run parallel assertions against agent trajectories using open-source evaluation frameworks. Attendees will learn how to catch infinite loops, detect tool-calling degradation, and establish a baseline scoring rubric for agent accuracy before code ever hits a production branch.

___________________________
Presentation Language: English

Captioning will be available for attendees in 50+ languages through Wordly. See instructions in each room to utilize captioning.

Speakers
avatar for Thu Ya Kyaw

Thu Ya Kyaw

Senior Developer Relations Engineer, Google
Thu Ya Kyaw is a Senior Developer Relations Engineer for Google Cloud. At Google, he helps to make learning, developing, deploying, and scaling applications on Google Cloud a delightful experience for everyone. He is passionate about using AI to solve real-world problems, and he is... Read More →
Friday September 11, 2026 10:20 - 10:45 JST
Hall C
  Evals & Testing

Sign up or log in to save this to your schedule, view media, leave feedback and see who's attending!

Share Modal

Share this link via

Or copy link