Agent Run Console
Production / compare
← Runs esc

Compare

A · run_f4a155vsB · run_d40e90code-reviewSwap A and B
Duration
41.3s -33%
A: 1m 1s
Cost
$0.076 -18%
A: $0.093
Tokens
29.9k +76%
A: 17.0k
Steps
7 -36%
A: 11
Model changed
GPT-4.1
A: Claude Sonnet

Steps, aligned

ChangeStepAB
SameInput policy check0.6s0.8s
SamePlan the task10.9s9.1s
Slowersearch_codebase1.8s3.3s
SameRetrieve context · top 82.1s3.1s
SameDecide next step9.8s7.4s
Samerun_tests4.1s4.6s
ChangedWrite final answer4.9s10.5s
RemovedDecide next step7.7s—
Removedget_diff4.3s—
Removedrun_tests3.1s—
RemovedWrite final answer8.3s—

Output diff

− 0 words + 0 words 100% unchanged

Reviewed 14 files. Three issues: the cache key ignores the locale, a test is skipped without a reason, and a server action is missing input validation. Tests pass. Suggested changes posted as comments, one marked blocking.

Keyboard shortcuts

Move down / up
jk
Open run
Enter
Back to runs
Esc
Select for compare
x
Compare selected
c
Re-run
r
Search
/
Next / previous step
↓↑
Replay trace
p
Show shortcuts
?