Agent Run Console
Production / compare
← Runs esc

Compare

A · run_a11225vsB · run_307daacode-reviewSwap A and B
Duration
38.7s +69%
A: 22.9s
Cost
$0.073 +250%
A: $0.021
Tokens
26.5k +162%
A: 10.1k
Steps
7 -36%
A: 11
Model changed
GPT-4.1
A: Gemini 2.5 Pro

Steps, aligned

ChangeStepAB
SlowerInput policy check0.2s0.7s
SlowerPlan the task4.0s7.1s
Slowersearch_codebase0.8s2.2s
SlowerRetrieve context · top 81.0s2.7s
SlowerDecide next step2.7s9.1s
Slowerrun_tests1.1s2.9s
ChangedWrite final answer2.1s11.6s
RemovedDecide next step3.2s—
Removedget_diff0.8s—
Removedrun_tests1.4s—
RemovedWrite final answer4.2s—

Output diff

− 0 words + 0 words 100% unchanged

Reviewed 14 files. Three issues: the cache key ignores the locale, a test is skipped without a reason, and a server action is missing input validation. Tests pass. Suggested changes posted as comments, one marked blocking.

Keyboard shortcuts

Move down / up
jk
Open run
Enter
Back to runs
Esc
Select for compare
x
Compare selected
c
Re-run
r
Search
/
Next / previous step
↓↑
Replay trace
p
Show shortcuts
?