I built Agent Review Studio: a local-first workbench for agent harness evaluations
I have been building Chaser Agent and other AI systems across the Chase ecosystem. As those systems became more capable, I needed a better way to inspect what an agent actually did—not just look at its final answer. So I built Agent Review Studio . It is an open-source, local-first system for evaluating and refining AI agents and agent harnesses. Why I built it An agent run can produce a final…
We haven't written up this one. Dev.to has the full story — the link below goes straight to it.