START HERE

Choose your trade-off.

Current pickGPT-6 Astra Low45.8 intelligence4.4k tokens / task

This rail contains only configurations on the measured token frontier.

THE MODEL EXPLORER

More clarity. Less guesswork.

Compare AI capability and usage before you choose a model.

Sources checked
Codex Local · Standard
SELECTED CONFIGURATION
OpenAI · GPT-6 series
Output tokens / benchmark task
4.4ktokens
938 reasoning + 3.5k answer tokens
Intelligence
45.8
Artificial Analysis
Reasoning effort

Measured on Artificial Analysis benchmarks. Output includes reasoning. Tokens are not subscription quota. How we calculate this

Capability vs. token use

Higher is more capable. Further left uses less.

Logarithmic x-axis

Swipe to explore the chart ↔

Output tokens per benchmark task · includes reasoning →

The dotted frontier connects configurations with no higher-scoring alternative using fewer output tokens. This is token efficiency on these benchmarks, not a quota recommendation.

The numbers, side by side.30

Artificial Analysis Intelligence Index v4.3.2

Scroll the table to see every metric ↔

Model & reasoningTokens / taskMeasured outputIntelligenceAA IndexReasoning tokensBenchmarkCompare configuration
2.1k
20.9515
2.5k
21559
3.4k
33.9697
3.9k
27.51k
4k
33.5837
4.4k
45.8938
4.5k
251.5k
5.5k
30.11.9k
6.5k
39.82.4k
7.9k
39.22.7k
9.6k
49.63.4k
10.2k
42.85k
11.2k
29.56.2k
11.3k
34.25.4k
11.8k
50.94.7k
13.3k
42.35.9k
13.9k
32.17.2k
16k
44.19.3k
16.9k
52.48.6k
19.6k
4410.1k
19.8k
32.112.8k
20.4k
3811.5k
23.6k
34.614.7k
27.2k
33.918.7k
27.2k
52.716.7k
29.3k
4717.3k
31.2k
47.521k
38.9k
42.126k
41.2k
37.328.1k
50.5k
37.339.3k
Quota ranges repeat across reasoning levels because effort-specific rates are unpublished.Download data
YOUR ALLOWANCE

Working with a partly used quota?

Scale the published message range to the quota you have left.

UNDERSTAND THE NUMBERS

A comparison you can inspect.

Three different measurements, with their boundaries kept visible.

01

Quota is a range.

OpenAI publishes estimated local messages per five hours. We convert the endpoints with 100 ÷ messages. These are indicative model-level ranges, not per-task measurements or statistical confidence intervals.

02

Capability is benchmarked.

Scores come from the Artificial Analysis Intelligence Index v4.3.2. Higher scores mean stronger performance across its evaluation mix. A score twice as high does not mean twice the intelligence.

03

Your work will vary.

Context, caching, tools, reasoning and subagents change consumption. Benchmark tasks differ from your projects. API prices and purchased-credit promotions do not determine your included quota.

What’s included, and what isn’t +

This release covers six OpenAI models used in Codex, five benchmarked reasoning levels, and Plus, Pro 5× and Pro 20×. Only configurations with comparable, sourced benchmarks appear here.

The quota view covers local Codex messages at Standard speed, including local CLI and IDE use with ChatGPT authentication. It does not estimate cloud tasks, Fast mode, API-key spending or weekly quota. A local message may start a multi-step task; it is not a fixed unit of work.

The same family range is intentionally shown for every reasoning level. Changing effort updates the measured capability and token values, but cannot manufacture an effort-specific quota estimate. Ultra and non-reasoning modes are excluded from this comparison because they are not part of this set of five comparable benchmark levels.

Quota dots use the geometric midpoint purely to position ranges on a logarithmic chart. The selected horizontal line shows the full published range. Overlapping ranges do not establish which configuration uses less quota for the same task. The token frontier is computed only from the visible benchmark configurations.

Scores are rounded to one decimal, token counts to whole tokens. Sources are checked before a dated snapshot is published. After 14 days, the page labels the snapshot as needing review. These data are not live account telemetry.

All controls run in your browser. No login, tracking cookies, API key or account data is required. Share links contain the model, comparison, plan and display filters only.

TRACEABLE BY DESIGN

Go straight to the source.

Snapshot reviewed 26 Sept 2026.