LLM7.io
LLM7.io
Toggle theme

chat comparison

gemma4:31bvsminimax-m2.7

MiniMax logo

Compare the things that matter before you build: current price, capabilities, and latest aggregated LLM7 statistics.

Current pricing

See the cost difference clearly

Directly comparable

gemma4:31b

Input $0.03 USD · Output $0.08 USD / 1M tokens

minimax-m2.7

Input $0.03 USD · Output $0.05 USD / 1M tokens

What's different

Only capabilities that differ between these models are listed. Green highlights the broader supported option or larger context window.

Context window

gemma4:31bBetter fit

262,000 tokens

minimax-m2.7

180,000 tokens

Streaming

gemma4:31bBetter fit

Supported

minimax-m2.7

Not supported

Reasoning

gemma4:31b

Not supported

minimax-m2.7Better fit

Supported

Latest LLM7 statistics

Recent aggregated percentages and response-time metrics from LLM7, not a global model benchmark or traffic disclosure.

Recent stability

gemma4:31b

99.5%

minimax-m2.7

91.22%

Average response time

gemma4:31b

5,771.95 ms

minimax-m2.7

19,274.32 ms

P95 API latency

gemma4:31b

30,000 ms

minimax-m2.7

60,000 ms

Side-by-side trends

Shared charts make the observed differences easier to read. Metrics without enough history stay hidden.

Request success rate

Both models on the same scale

gemma4:31b: 99.8%minimax-m2.7: 92.9%
4 Aug3 Sept

Average API latency

Both models on the same scale

gemma4:31b: 5,495 msminimax-m2.7: 19,196 ms
4 Aug3 Sept

P95 API latency

Both models on the same scale

gemma4:31b: 30,000 msminimax-m2.7: 60,000 ms
4 Aug3 Sept

Estimate your workload

Fixed-price models can be estimated directly; dynamic models use recent references and are billed from actual provider usage.

Cost calculator

Adjust the volume to see an instant estimate at the current public price.

1M tokens
10,000 tokens
01M

Input estimate

$0.0003

Output estimate

$0.0002

Estimated total

$0.0005

Cost calculator

Adjust the volume to see an instant estimate at the current public price.

1M tokens
10,000 tokens
01M

Input estimate

$0.0003

Output estimate

$0.000125

Estimated total

$0.000425

Quick take

  • gemma4:31b has a larger context window.
  • gemma4:31b recorded higher recent LLM7 stability during this period.
  • gemma4:31b recorded a lower observed p95 API latency on LLM7 during this period.

Keep exploring