LLM7.io
LLM7.io
Toggle theme

chat comparison

Mistral AI logo

codestral-latestvsgpt-oss:20b

OpenAI logo

Compare the things that matter before you build: current price, capabilities, and latest aggregated LLM7 statistics.

Current pricing

See the cost difference clearly

Directly comparable

codestral-latest

Input $0.01 USD · Output $0.02 USD / 1M tokens

gpt-oss:20b

Input $0.04 USD · Output $0.06 USD / 1M tokens

What's different

Only capabilities that differ between these models are listed. Green highlights the broader supported option or larger context window.

Context window

codestral-latest

32,000 tokens

gpt-oss:20bBetter fit

128,000 tokens

Latest LLM7 statistics

Recent aggregated percentages and response-time metrics from LLM7, not a global model benchmark or traffic disclosure.

Recent stability

codestral-latest

98.63%

gpt-oss:20b

98.89%

Average response time

codestral-latest

3,325.72 ms

gpt-oss:20b

7,324.69 ms

P95 API latency

codestral-latest

30,000 ms

gpt-oss:20b

60,000 ms

Side-by-side trends

Shared charts make the observed differences easier to read. Metrics without enough history stay hidden.

Request success rate

Both models on the same scale

codestral-latest: 96.8%gpt-oss:20b: 99.1%
25 Jul14 Aug

Average API latency

Both models on the same scale

codestral-latest: 2,178 msgpt-oss:20b: 6,991 ms
25 Jul14 Aug

P95 API latency

Both models on the same scale

codestral-latest: 10,000 msgpt-oss:20b: 30,000 ms
25 Jul14 Aug

Estimate your workload

Try the same request volume against both public price lists.

Cost calculator

Adjust the volume to see an instant estimate at the current public price.

1M tokens
10,000 tokens
01M

Input estimate

$0.0001

Output estimate

$0.00005

Estimated total

$0.00015

Cost calculator

Adjust the volume to see an instant estimate at the current public price.

1M tokens
10,000 tokens
01M

Input estimate

$0.0004

Output estimate

$0.00015

Estimated total

$0.00055

Quick take

  • gpt-oss:20b has a larger context window.
  • The recent stability rates are too close to distinguish meaningfully.
  • codestral-latest recorded a lower observed p95 API latency on LLM7 during this period.

Keep exploring