
ActionBench: discovering hidden failure modes of agents
Evaluating how agents decide, use tools, and cross boundaries.
The systems, methods, and experiments that shape how we build and safeguard AI products.

Evaluating how agents decide, use tools, and cross boundaries.

Measuring the real-world performance of action-level guardrails.
Preprint · Security
Coming soon
Where permissioning breaks down when models try to act.
Infrastructure
Coming soon
Keeping control decisions fast when agents run at production scale.
Training
Coming soon
Making reinforcement from review practical for production agents.
Works with
Control AI behavior in real time with Runsphere