LLM7.io
LLM7.io
Toggle theme

chat comparison

Google logo

gemini-3.8-flash-highvsXiaomiMiMo/MiMo-V2.5

Xiaomi logo

Compare the things that matter before you build: current price, capabilities, and latest aggregated LLM7 statistics.

Current pricing

See the cost difference clearly

Directly comparable

gemini-3.8-flash-high

Input $0.05 USD · Output $0.15 USD / 1M tokens

XiaomiMiMo/MiMo-V2.5

Input $0.40 USD · Output $2.00 USD / 1M tokens

What's different

Only capabilities that differ between these models are listed. Green highlights the broader supported option or larger context window.

Context window

gemini-3.8-flash-highBetter fit

1,000,000 tokens

XiaomiMiMo/MiMo-V2.5

256,000 tokens

Input formats

gemini-3.8-flash-high

text, image

XiaomiMiMo/MiMo-V2.5

text

Vision

gemini-3.8-flash-highBetter fit

Supported

XiaomiMiMo/MiMo-V2.5

Not supported

JSON mode

gemini-3.8-flash-high

Not supported

XiaomiMiMo/MiMo-V2.5Better fit

Supported

Reasoning

gemini-3.8-flash-highBetter fit

Supported

XiaomiMiMo/MiMo-V2.5

Not supported

Latest LLM7 statistics

Recent aggregated percentages and response-time metrics from LLM7, not a global model benchmark or traffic disclosure.

Recent stability

gemini-3.8-flash-high

99.51%

XiaomiMiMo/MiMo-V2.5

61.41%

Average response time

gemini-3.8-flash-high

6,177.44 ms

XiaomiMiMo/MiMo-V2.5

52,158.29 ms

P95 API latency

gemini-3.8-flash-high

30,000 ms

XiaomiMiMo/MiMo-V2.5

60,000 ms

Side-by-side trends

Shared charts make the observed differences easier to read. Metrics without enough history stay hidden.

Request success rate

Both models on the same scale

gemini-3.8-flash-high: 100%XiaomiMiMo/MiMo-V2.5: 100%
5 Aug3 Sept

Average API latency

Both models on the same scale

gemini-3.8-flash-high: 7,122 msXiaomiMiMo/MiMo-V2.5: 14,781 ms
5 Aug3 Sept

P95 API latency

Both models on the same scale

gemini-3.8-flash-high: 30,000 msXiaomiMiMo/MiMo-V2.5: 60,000 ms
5 Aug3 Sept

Estimate your workload

Fixed-price models can be estimated directly; dynamic models use recent references and are billed from actual provider usage.

Cost calculator

Adjust the volume to see an instant estimate at the current public price.

1M tokens
10,000 tokens
01M

Input estimate

$0.0005

Output estimate

$0.000375

Estimated total

$0.000875

Cost calculator

Adjust the volume to see an instant estimate at the current public price.

1M tokens
10,000 tokens
01M

Input estimate

$0.004

Output estimate

$0.005

Estimated total

$0.009

Quick take

  • gemini-3.8-flash-high has a larger context window.
  • gemini-3.8-flash-high supports image input while the other model does not.
  • gemini-3.8-flash-high recorded higher recent LLM7 stability during this period.
  • gemini-3.8-flash-high recorded a lower observed p95 API latency on LLM7 during this period.

Keep exploring