Local comparison and human evaluation
Run coding agents locally, inspect blind candidates, and record the outcome you prefer.
Dispatch Cloud · Future direction
Evaluation sync is optional and requires explicit consent. Automatic routing and hosted execution are not available in v0.1.
Run coding agents locally, inspect blind candidates, and record the outcome you prefer.
Preview eligible evaluation data, enable consent, and sync only with a configured token.
Future versions may use opt-in evaluations from real tasks to help choose an agent for a task.
Why compare on your own work
Public benchmark scores leave out your codebase, constraints, and preferences. Dispatch lets you inspect each result and choose the one that fits.
Each run records candidate results and your evaluation.
You choose whether to send eligible evaluation data to Dispatch Cloud.
In v0.1, every run names the agents to use.
Run Codex CLI and Cursor Agent on the same task, compare their changes, and choose the result that fits your codebase.