EveryDev.ai
Subscribe
Home
Tools

4,304+ AI tools

  • New
  • Trending
  • Featured
  • Rate tools
  • Compare
  • Arena
Categories
  • Agents3274
  • Coding2275
  • Infrastructure1000
  • Projects696
  • Marketing636
  • Research587
  • MCP532
  • Design508
  • Analytics506
  • Testing394
  • Security376
  • Data327
  • Integration244
  • Prompts244
  • Communication235
  • Extensions217
  • Voice193
  • Learning190
  • Commerce170
  • DevOps153
  • Web103
  • Finance36
AI Tools by Topic
  • AI Coding Assistants
  • Agent Frameworks
  • MCP Servers
  • AI Prompt Tools
  • Vibe Coding Tools
  • AI Design Tools
  • AI Database Tools
  • AI Website Builders
  • AI Testing Tools
  • LLM Evaluations
Follow Us
  • X / Twitter
  • LinkedIn
  • Reddit
  • Discord
  • Threads
  • Bluesky
  • Mastodon
  • YouTube
  • GitHub
  • Instagram
Get Started
  • Users
  • Rate Tools
  • About
  • Editorial Standards
  • Corrections & Disclosures
  • Community Guidelines
  • Advertise
  • Contact Us
  • Newsletter
  • Submit a Tool
  • Start a Discussion
  • Write A Blog
  • Share A Build
  • Terms of Service
  • Privacy Policy
Explore with AI
  • ChatGPT
  • Gemini
  • Claude
  • Grok
  • Perplexity
Agent Experience
  • llms.txt
Theme
With AI, Everyone is a Dev. EveryDev.ai © 2026
    1. Home
    2. Tools
    3. AQuA (Ambient Quality Agent)
    AQuA (Ambient Quality Agent) icon

    AQuA (Ambient Quality Agent)

    Observability Platforms
    Featured

    An open-source ADK recipe that runs beside a production agent in Google Cloud, clustering and diagnosing failures from production trajectories.

    Visit Website

    At a Glance

    Pricing
    Open Source

    AQuA (Ambient Quality Agent) is released in the open in the google/adk-recipes repository under the Apache License 2.0. Running it in your own Google Cloud project incurs separate Google Cloud and Gemini model costs.

    Engagement

    Available On

    Windows
    Linux
    Web
    API
    CLI

    Resources

    WebsiteDocsGitHubllms.txt

    Topics

    Observability PlatformsLLM EvaluationsAutonomous Systems

    Alternatives

    PluraiFuture AGIPandaProbe
    Developer
    Google Open Source Programs OfficeMountain View, CAEst. 2004

    Listed Oct 2026

    About AQuA (Ambient Quality Agent)

    AQuA (Ambient Quality Agent) is a reference implementation from Google, published in the adk-recipes repository, that runs unattended beside a production agent in your Google Cloud project. It sweeps production trajectories from Cloud Trace, Cloud Logging, or BigQuery, then clusters failures and diagnoses their root causes. Google describes it as released in the open as composable building blocks to run, adapt, and help shape.

    What It Is

    AQuA is an outer-loop agent quality tool for agents built with the Agent Development Kit (ADK). It works on live production traffic, not pre-launch test cases, and it never sits in the request path or writes back to the observed agent. Transcripts, source snapshots, and BigQuery tables stay inside the user's project. The repository README describes the recipes in adk-recipes as demonstration starting points, not production use, and as not an officially supported Google product.

    How a Sweep Works

    Each run follows a five-stage pipeline:

    • Sample: pulls a random sample of up to 1,000 recent sessions.
    • Review: grades each session against a nine-point checklist, with an optional plain-English goal.md steering the review, and writes structured actual/expected findings.
    • Cluster: groups findings that share a failure mechanism.
    • Verify: a separate model checks each cluster against up to three full transcripts and discards unsupported ones.
    • Track: matches surviving clusters against open insights in BigQuery as NEW, RECURRING, or auto-RESOLVED after 14 days unseen.

    By default a single-pass session_review judge is used, and Gemini platform trajectory AutoRaters can be enabled optionally. Deterministic Python custom metrics in eval_config.yaml can run alongside the judge.

    Root-Cause Diagnosis

    From the dashboard chat or agents-cli aqua run, a diagnosis agent reads failing trajectories against the immutable source snapshot captured at deploy time. When the defect is in the repository, it cites file and line ranges and proposes an anchored edit. When the fault is upstream, it attributes the failure to a trajectory step without proposing a diff. It never applies edits or opens pull requests itself. Insights can be fetched with agents-cli aqua get-insight so a coding agent can apply a fix and verify it by replaying sessions.

    Trust and Limits

    The blog states that insights are backed by citations to session IDs and validated line ranges rather than confidence scores, and that skipped or failed work is recorded on the run. Users can permanently dismiss by-design findings. Google notes deliberate trade-offs, including random session sampling, capped verification at 50 clusters per run, and static replay that does not reconstruct external environment state. It also lists directions it is exploring, such as attaching across a fleet of agents.

    AQuA (Ambient Quality Agent) - 1

    Community Discussions

    Be the first to start a conversation about AQuA (Ambient Quality Agent)

    Share your experience with AQuA (Ambient Quality Agent), ask questions, or help others learn from your insights.

    Pricing

    OPEN SOURCE

    Open Source (Apache 2.0)

    AQuA (Ambient Quality Agent) is released in the open in the google/adk-recipes repository under the Apache License 2.0. Running it in your own Google Cloud project incurs separate Google Cloud and Gemini model costs.

    • Apache License 2.0 (declared for the adk-recipes repository)
    • Runs beside your agent in your own Google Cloud project
    • Local dashboard exploration with no cloud project, no credentials, and no model calls
    • Sweeps up to 1,000 sessions per run
    • External costs: Gemini model usage at standard Gemini platform pricing and Google Cloud resources; example sweep cost $0.70 for 96 sessions and $3.76 for 32 multi-agent sessions

    Capabilities

    Key Features

    • Scheduled, post-deployment, or on-demand sweeps of production sessions
    • Random sampling of up to 1,000 sessions per run
    • Nine-point checklist session review with actual/expected findings
    • Failure clustering and transcript-based verification
    • Insight tracking in BigQuery as NEW, RECURRING, or RESOLVED
    • Developer goal (goal.md) and custom Python metrics
    • Optional managed trajectory AutoRaters
    • Root-cause diagnosis anchored to deploy-time source snapshots
    • Dashboard on Cloud Run behind Identity-Aware Proxy
    • agents-cli aqua commands and agents-cli-aqua skill for coding agents
    • Local dashboard demo with synthetic data

    Integrations

    Agent Development Kit (ADK)
    Google Cloud Trace
    Google Cloud Logging
    BigQuery
    Cloud Storage
    Cloud Run
    Identity-Aware Proxy
    Gemini
    agents-cli
    Gemini CLI
    Claude Code
    Cursor
    Antigravity
    API Available
    View Docs

    Ratings & Reviews

    No ratings yet

    Be the first to rate AQuA (Ambient Quality Agent) and help others make informed decisions.

    Rate other tools you’ve used

    Developer

    Google Open Source Programs Office

    Google's Open Source Programs Office (OSPO) has supported open source innovation since 2004, making it one of the first OSPOs in the industry. The office releases and maintains major open source projects including Android, Chromium, Go, Kubernetes, and TensorFlow. OSPO runs programs like Google Summer of Code and Season of Docs to bring new contributors into open source and improve documentation across the ecosystem.

    Founded 2004
    Mountain View, CA

    Used by

    Over 1,000 open source organizations…
    225+ open source projects use OSS-Fuzz
    Hundreds of organizations supported…
    Notable projects supported include:…
    Read more about Google Open Source Programs Office
    WebsiteGitHubX / Twitter
    5 tools in directory

    Similar Tools

    Plurai icon

    Plurai

    Plurai is an AI evaluation and guardrails platform that uses small language models to slash costs and increase accuracy for AI agent deployments at scale.

    Future AGI icon

    Future AGI

    An AI lifecycle platform for building, evaluating, monitoring, and securing generative AI agents with hallucination detection, simulations, and real-time guardrails.

    PandaProbe icon

    PandaProbe

    Open source agent engineering platform providing traces, evals, metrics, and live monitoring to debug and improve AI agents.

    Browse all tools

    Related Topics

    Observability Platforms

    Comprehensive platforms that combine metrics, logs, and traces with AI-powered analytics to provide deep insights into complex distributed systems and application behavior.

    136 tools

    LLM Evaluations

    Platforms and frameworks for evaluating, testing, and benchmarking LLM systems and AI applications. These tools provide evaluators and evaluation models to score AI outputs, measure hallucinations, assess RAG quality, detect failures, and optimize model performance. Features include automated testing with LLM-as-a-judge metrics, component-level evaluation with tracing, regression testing in CI/CD pipelines, custom evaluator creation, dataset curation, and real-time monitoring of production systems. Teams use these solutions to validate prompt effectiveness, compare models side-by-side, ensure answer correctness and relevance, identify bias and toxicity, prevent PII leakage, and continuously improve AI product quality through experiments, benchmarks, and performance analytics.

    141 tools

    Autonomous Systems

    AI agents that can perform complex tasks with minimal human guidance.

    473 tools
    Browse all topics
    Back to all toolsSuggest an edit
    ratings
    discussions