- Iterate on scaffolds. Docent returns actionable insights to inform prompt tuning, tool instructions, or orchestration logic.
- Post-train models. Compare behavior across checkpoints or training steps to identify what’s driving shifts in eval results.
- Build better benchmarks. Catch reward hacking, evaluation awareness, broken environments, and ambiguous task specifications.
Get started
Installation
Set up the SDK, coding-agent plugin, and API key.
Get in touch
Join our Slack community to ask questions and chat with our team.

