← Back to glossary Glossary

Shadow Testing (Agent)

Reviewed 19 July 2026 Canonical definition Part of: Agent Reliability & Deployment Terms →

Shadow testing runs a new agent version in parallel with the production version on real traffic, capturing its outputs without serving them to end users. It enables direct comparison of new vs. old agent behaviour on production inputs before a live release, dramatically reducing the risk of deploying an agent that performs well on benchmarks but poorly in production.

§01 / QUESTIONSterm: Shadow Testing (Agent)
Questions

Common questions.

What is Shadow Testing (Agent)?

Shadow testing runs a new agent version in parallel with the production version on real traffic, capturing its outputs without serving them to end users.

Why does Shadow Testing (Agent) matter for AI agents?

It enables direct comparison of new vs. old agent behaviour on production inputs before a live release, dramatically reducing the risk of deploying an agent that performs well on benchmarks but poorly in production.

Which terms are related to Shadow Testing (Agent)?

Closely related concepts include Environment Promotion (Agent), Agent CI/CD, Agent-as-a-Service, Canary Deployment (Agent). Each is defined in the Prefactor glossary.

§02 / RELATEDnext: where this fits
Keep reading

Where this fits.

See how every agent performs, and make it better

Prefactor helps teams observe, evaluate, and improve their AI agents in production, across every framework and provider.