← Back to glossary Glossary

Capability Control (AI)

Reviewed 19 July 2026 Canonical definition Part of: Runtime Control & Guardrail Terms →

Capability control is the practice of limiting what an AI agent is technically able to do, through tool restrictions, output filtering, API permissions, and execution environment constraints, rather than relying solely on the agent's own safety training to prevent harmful behaviour. Capability controls are a defence-in-depth strategy: even if an agent's reasoning can be manipulated, the controls ensure it cannot take actions outside its permitted scope.

§01 / QUESTIONSterm: Capability Control (AI)
Questions

Common questions.

What is Capability Control (AI)?

Capability control is the practice of limiting what an AI agent is technically able to do, through tool restrictions, output filtering, API permissions, and execution environment constraints, rather than relying solely on the agent's own safety training to prevent harmful behaviour.

How does Capability Control (AI) work?

Capability controls are a defence-in-depth strategy: even if an agent's reasoning can be manipulated, the controls ensure it cannot take actions outside its permitted scope.

Which terms are related to Capability Control (AI)?

Closely related concepts include Ambient Authority, Jailbreak (AI Agent), Containment Strategy (AI), Agent Context Isolation. Each is defined in the Prefactor glossary.

§02 / RELATEDnext: where this fits
Keep reading

Where this fits.

See how every agent performs, and make it better

Prefactor helps teams observe, evaluate, and improve their AI agents in production, across every framework and provider.