gpt-oss-120b scores 14.5 on the Respan Index (#124), ahead of Qwen3 235B-A22B at 14.4 (#125). gpt-oss-120b leads on 3 of the 5 benchmarks both report. Qwen3 235B-A22B has the larger context window (262K tokens).
* Estimated: no published scores in that category. How the Respan Index works
| 80.8% |
Higher is better. Each score links to where it was published.