LLM7.io
LLM7.io
Toggle theme

chat comparison

Google logo

gemini-3.7-flashvsgemma4:31b

Compare the things that matter before you build: current price, capabilities, and latest aggregated LLM7 statistics.

Current pricing

See the cost difference clearly

Directly comparable

gemini-3.7-flash

Input $0.06 USD · Output $0.30 USD / 1M tokens

gemma4:31b

Input $0.07 USD · Output $0.23 USD / 1M tokens

What's different

Only capabilities that differ between these models are listed. Green highlights the broader supported option or larger context window.

Context window

gemini-3.7-flashBetter fit

1,000,000 tokens

gemma4:31b

262,000 tokens

Input formats

gemini-3.7-flash

text, image

gemma4:31b

text

Vision

gemini-3.7-flashBetter fit

Supported

gemma4:31b

Not supported

JSON mode

gemini-3.7-flash

Not supported

gemma4:31bBetter fit

Supported

Latest LLM7 statistics

Recent aggregated percentages and response-time metrics from LLM7, not a global model benchmark or traffic disclosure.

Recent stability

gemini-3.7-flash

90.26%

gemma4:31b

99.56%

Average response time

gemini-3.7-flash

5,073.56 ms

gemma4:31b

5,525.29 ms

P95 API latency

gemini-3.7-flash

30,000 ms

gemma4:31b

30,000 ms

Side-by-side trends

Shared charts make the observed differences easier to read. Metrics without enough history stay hidden.

Request success rate

Both models on the same scale

gemini-3.7-flash: 4%gemma4:31b: 100%
23 Aug22 Sept

Average API latency

Both models on the same scale

gemini-3.7-flash: 2,394 msgemma4:31b: 4,359 ms
23 Aug22 Sept

P95 API latency

Both models on the same scale

gemini-3.7-flash: 500 msgemma4:31b: 30,000 ms
23 Aug22 Sept

Estimate your workload

Fixed-price models can be estimated directly; dynamic models use recent references and are billed from actual provider usage.

Cost calculator

Adjust the volume to see an instant estimate at the current public price.

1M tokens
10,000 tokens
01M

Input estimate

$0.0006

Output estimate

$0.00075

Estimated total

$0.00135

Cost calculator

Adjust the volume to see an instant estimate at the current public price.

1M tokens
10,000 tokens
01M

Input estimate

$0.0007

Output estimate

$0.000575

Estimated total

$0.001275

Quick take

  • gemini-3.7-flash has a larger context window.
  • gemini-3.7-flash supports image input while the other model does not.
  • gemma4:31b recorded higher recent LLM7 stability during this period.

Keep exploring