Agent Run Console
Production / compare
← Runs esc

Compare

A · run_a11225vsB · run_4f67e0code-reviewSwap A and B
Duration
38.8s +69%
A: 22.9s
Cost
$0.049 +134%
A: $0.021
Tokens
18.2k +80%
A: 10.1k
Steps
7 -36%
A: 11
Model changed
GPT-4.1
A: Gemini 2.5 Pro

Steps, aligned

ChangeStepAB
SlowerInput policy check0.2s0.7s
SlowerPlan the task4.0s10.8s
Slowersearch_codebase0.8s3.9s
SlowerRetrieve context · top 81.0s3.8s
SlowerDecide next step2.7s7.4s
Slowerrun_tests1.1s3.1s
ChangedWrite final answer2.1s6.8s
RemovedDecide next step3.2s—
Removedget_diff0.8s—
Removedrun_tests1.4s—
RemovedWrite final answer4.2s—

Output diff

− 0 words + 0 words 100% unchanged

Reviewed 14 files. Three issues: the cache key ignores the locale, a test is skipped without a reason, and a server action is missing input validation. Tests pass. Suggested changes posted as comments, one marked blocking.

Keyboard shortcuts

Move down / up
jk
Open run
Enter
Back to runs
Esc
Select for compare
x
Compare selected
c
Re-run
r
Search
/
Next / previous step
↓↑
Replay trace
p
Show shortcuts
?