gemini-3.1-flash-lite
Input $0.03 USD · Output $0.08 USD / 1M tokens
chat comparison
Compare the things that matter before you build: current price, capabilities, and observed LLM7 usage.
Current pricing
gemini-3.1-flash-lite
Input $0.03 USD · Output $0.08 USD / 1M tokens
gpt-oss:20b
Input $0.04 USD · Output $0.06 USD / 1M tokens
Only capabilities that differ between these models are listed. Green highlights the broader supported option or larger context window.
Input formats
gemini-3.1-flash-lite
text, image
gpt-oss:20b
text
Vision
gemini-3.1-flash-liteBetter fit
Supported
gpt-oss:20b
Not supported
JSON mode
gemini-3.1-flash-lite
Not supported
gpt-oss:20bBetter fit
Supported
A 30-day LLM7 snapshot, not a global model benchmark.
Requests handled
gemini-3.1-flash-lite
256
gpt-oss:20b
474
Observed success rate
gemini-3.1-flash-lite
100%
gpt-oss:20b
99.79%
Average API latency
gemini-3.1-flash-lite
1,967.98 ms
gpt-oss:20b
4,323.41 ms
P95 API latency
gemini-3.1-flash-lite
5,000 ms
gpt-oss:20b
30,000 ms
Try the same request volume against both public price lists.
Adjust the volume to see an instant estimate at the current public price.
Input estimate
$0.0003
Output estimate
$0.0002
Estimated total
$0.0005
Adjust the volume to see an instant estimate at the current public price.
Input estimate
$0.0004
Output estimate
$0.00015
Estimated total
$0.00055