LLM7.io
LLM7.io
Toggle theme

chat comparison

Google logo

gemini-3.1-flash-litevsgemini-3.5-flash-low

Google logo

Compare the things that matter before you build: current price, capabilities, and latest aggregated LLM7 statistics.

Current pricing

See the cost difference clearly

Directly comparable

gemini-3.1-flash-lite

Input $0.02 USD · Output $0.04 USD / 1M tokens

gemini-3.5-flash-low

Input $0.04 USD · Output $0.20 USD / 1M tokens

What's different

Only capabilities that differ between these models are listed. Green highlights the broader supported option or larger context window.

Context window

gemini-3.1-flash-lite

256,000 tokens

gemini-3.5-flash-lowBetter fit

1,040,000 tokens

Input formats

gemini-3.1-flash-lite

text, image

gemini-3.5-flash-low

text

Vision

gemini-3.1-flash-liteBetter fit

Supported

gemini-3.5-flash-low

Not supported

Reasoning

gemini-3.1-flash-lite

Not supported

gemini-3.5-flash-lowBetter fit

Supported

Latest LLM7 statistics

Recent aggregated percentages and response-time metrics from LLM7, not a global model benchmark or traffic disclosure.

Recent stability

gemini-3.1-flash-lite

98.68%

gemini-3.5-flash-low

99.62%

Average response time

gemini-3.1-flash-lite

4,215.7 ms

gemini-3.5-flash-low

5,589.41 ms

P95 API latency

gemini-3.1-flash-lite

10,000 ms

gemini-3.5-flash-low

30,000 ms

Side-by-side trends

Shared charts make the observed differences easier to read. Metrics without enough history stay hidden.

Request success rate

Both models on the same scale

gemini-3.1-flash-lite: 99.4%gemini-3.5-flash-low: 100%
25 Jul24 Aug

Average API latency

Both models on the same scale

gemini-3.1-flash-lite: 3,161 msgemini-3.5-flash-low: 4,998 ms
25 Jul24 Aug

P95 API latency

Both models on the same scale

gemini-3.1-flash-lite: 10,000 msgemini-3.5-flash-low: 10,000 ms
25 Jul24 Aug

Estimate your workload

Fixed-price models can be estimated directly; dynamic models use recent references and are billed from actual provider usage.

Cost calculator

Adjust the volume to see an instant estimate at the current public price.

1M tokens
10,000 tokens
01M

Input estimate

$0.0002

Output estimate

$0.0001

Estimated total

$0.0003

Cost calculator

Adjust the volume to see an instant estimate at the current public price.

1M tokens
10,000 tokens
01M

Input estimate

$0.0004

Output estimate

$0.0005

Estimated total

$0.0009

Quick take

  • gemini-3.5-flash-low has a larger context window.
  • gemini-3.1-flash-lite supports image input while the other model does not.
  • The recent stability rates are too close to distinguish meaningfully.
  • gemini-3.1-flash-lite recorded a lower observed p95 API latency on LLM7 during this period.

Keep exploring