Skip to content
Back to Overall

GPT-5.1 2025-11-13

OpenAI · released 2025-11-13 · updated 10/01 14:05

Overall consensus index—Not on the overall board
Category results1/ 4 categories
Results on record0evaluations
Context window—Token
API input / output · per million tokens¥8.38 / ¥67.05cached ¥0.84list $1.25 / $10Provider's official price

Strengths, seen one at a time.

Each capability is computed on its own. Without enough measurements, it stays empty.

Each category's score reflects ranking support within its own reference group; the scores cannot be added up or used to compare how strong different capabilities are.

Every result has a source.

The model's published aggregate results. Open one to see its run configuration and how it was used.

What is still unknown

Missing evaluations do not count as zero. The rank moves with new evidence; when scores are close, do not read much into small gaps.

How it is calculated