LLM7.io
LLM7.io
Toggle theme

chat comparison

OpenAI logo

gpt-5.4-minivsgpt-oss:20b

OpenAI logo

Compare the things that matter before you build: current price, capabilities, and observed LLM7 usage.

Current pricing

See the cost difference clearly

Directly comparable

gpt-5.4-mini

Input $0.04 USD · Output $0.24 USD / 1M tokens

gpt-oss:20b

Input $0.04 USD · Output $0.06 USD / 1M tokens

What's different

Only capabilities that differ between these models are listed. Green highlights the broader supported option or larger context window.

Context window

gpt-5.4-miniBetter fit

400,000 tokens

gpt-oss:20b

128,000 tokens

Input formats

gpt-5.4-mini

text, image

gpt-oss:20b

text

Vision

gpt-5.4-miniBetter fit

Supported

gpt-oss:20b

Not supported

Reasoning

gpt-5.4-miniBetter fit

Supported

gpt-oss:20b

Not supported

Observed usage comparison

A 30-day LLM7 snapshot, not a global model benchmark.

Requests handled

gpt-5.4-mini

57

gpt-oss:20b

474

Observed success rate

gpt-5.4-mini

43.86%

gpt-oss:20b

99.79%

Average API latency

gpt-5.4-mini

14,631.46 ms

gpt-oss:20b

4,323.41 ms

P95 API latency

gpt-5.4-mini

60,000 ms

gpt-oss:20b

30,000 ms

Estimate your workload

Try the same request volume against both public price lists.

Cost calculator

Adjust the volume to see an instant estimate at the current public price.

1M tokens
10,000 tokens
01M

Input estimate

$0.0004

Output estimate

$0.0006

Estimated total

$0.001

Cost calculator

Adjust the volume to see an instant estimate at the current public price.

1M tokens
10,000 tokens
01M

Input estimate

$0.0004

Output estimate

$0.00015

Estimated total

$0.00055

Quick take

  • gpt-5.4-mini has a larger context window.
  • gpt-5.4-mini supports image input while the other model does not.
  • gpt-oss:20b recorded a higher observed LLM7 request success rate during this period.
  • gpt-oss:20b recorded a lower observed p95 API latency on LLM7 during this period.

Keep exploring