agent_system

GPT-5.5 toolathlon-verified:GPT-5.5 | Default | reasoning=xhigh

config_14a00cacba7d65e688a5c760163b8a5c43f157cd0ac3c74a8908ada4bf75af4a

Published ranking

This score and claim apply only inside the bound ranking group.

74%

capability score

Rank

7

Evidence families

1

Claim

Outside top set

Uncertainty was not reported

Configuration passport

Canonical name
GPT-5.5
Revision
toolathlon-verified:GPT-5.5 | Default | reasoning=xhigh
Interaction policy
agentic
Passport class
agent-system-v1
Harness
toolathlon-verified
Scaffold
Default

Source evidence

Publication snapshotexplorer_11e78a6760100f2399915d63fb3ad62ea871389a876ddec0a9ba40ed98496461

Methodology2026-10-04.1.macroscopebench-code-review-wave