gemini-3.1-flash-lite
Input $0.02 USD · Output $0.04 USD / 1M tokens
chat comparison
Compare the things that matter before you build: current price, capabilities, and latest aggregated LLM7 statistics.
Current pricing
gemini-3.1-flash-lite
Input $0.02 USD · Output $0.04 USD / 1M tokens
gpt-oss:20b
Input $0.04 USD · Output $0.06 USD / 1M tokens
Only capabilities that differ between these models are listed. Green highlights the broader supported option or larger context window.
Context window
gemini-3.1-flash-liteBetter fit
256,000 tokens
gpt-oss:20b
128,000 tokens
Input formats
gemini-3.1-flash-lite
text, image
gpt-oss:20b
text
Vision
gemini-3.1-flash-liteBetter fit
Supported
gpt-oss:20b
Not supported
JSON mode
gemini-3.1-flash-lite
Not supported
gpt-oss:20bBetter fit
Supported
Recent aggregated percentages and response-time metrics from LLM7, not a global model benchmark or traffic disclosure.
Recent stability
gemini-3.1-flash-lite
98.24%
gpt-oss:20b
98.79%
Average response time
gemini-3.1-flash-lite
4,164.18 ms
gpt-oss:20b
6,761.57 ms
P95 API latency
gemini-3.1-flash-lite
10,000 ms
gpt-oss:20b
60,000 ms
Shared charts make the observed differences easier to read. Metrics without enough history stay hidden.
Both models on the same scale
Both models on the same scale
Both models on the same scale
Try the same request volume against both public price lists.
Adjust the volume to see an instant estimate at the current public price.
Input estimate
$0.0002
Output estimate
$0.0001
Estimated total
$0.0003
Adjust the volume to see an instant estimate at the current public price.
Input estimate
$0.0004
Output estimate
$0.00015
Estimated total
$0.00055