Google releases AQuA, an open-source agent that diagnoses production AI agent failures
Overview
Google has released AQuA (Ambient Quality Agent), an open-source tool that runs unattended in a customer's own Google Cloud project and samples production agent sessions from Cloud Trace, Cloud Logging, or BigQuery to find recurring failures.
Google's developer blog reports that in a 32-session sweep of a multi-agent travel concierge, AQuA verified six issues and traced two to specific prompt lines; after fixes, full-session passes rose from 5/32 to 13/32. Google notes that verification and diagnosis are model-based and that the tool only proposes edits.
According to Google, transcripts, source snapshots, and BigQuery tables stay inside the customer's project, and AQuA sits outside the request path without writing back to the agent. It never applies an edit or opens a pull request on its own. The release covers only the outer quality loop, offered as composable building blocks that Google says it will refine as it studies how teams run continuous agent quality.
Written by AI from the articles below · updated Oct 8, 7:52 PM ET
Check the sources:
Article timeline
Follow the coverage from different perspectives. Times are ET.
- Google Developers BlogPickGoogle's AQuA agent diagnoses production failures in a multi-agent travel concierge
AIGoogle Developers Blog introduces AQuA, an ambient quality agent that runs in a customer's Google Cloud project and samples production sessions to find recurring agent failures. In a 32-session travel-concierge sweep, it verified six issues and traced two of them to specific prompt lines, and a replay after the fixes raised full-session passes from 5/32 to 13/32. The post notes that verification and diagnosis are model-based, and that the tool proposes edits without applying them.
Heat trend
Not enough continuous observations to show a trend yet.