gemini-3.1-flash-lite
Input $0.02 USD · Output $0.04 USD / 1M tokens
chat comparison
Compare the things that matter before you build: current price, capabilities, and latest aggregated LLM7 statistics.
Current pricing
gemini-3.1-flash-lite
Input $0.02 USD · Output $0.04 USD / 1M tokens
Inkling-Small
Input $0.50 USD · Output $1.20 USD / 1M tokens
Only capabilities that differ between these models are listed. Green highlights the broader supported option or larger context window.
Context window
gemini-3.1-flash-lite
256,000 tokens
Inkling-SmallBetter fit
512,000 tokens
Input formats
gemini-3.1-flash-lite
text, image
Inkling-Small
text
Vision
gemini-3.1-flash-liteBetter fit
Supported
Inkling-Small
Not supported
Reasoning
gemini-3.1-flash-lite
Not supported
Inkling-SmallBetter fit
Supported
Recent aggregated percentages and response-time metrics from LLM7, not a global model benchmark or traffic disclosure.
Recent stability
gemini-3.1-flash-lite
98.62%
Inkling-Small
63.16%
Average response time
gemini-3.1-flash-lite
4,284.02 ms
Inkling-Small
9,909.79 ms
P95 API latency
gemini-3.1-flash-lite
10,000 ms
Inkling-Small
60,000 ms
Shared charts make the observed differences easier to read. Metrics without enough history stay hidden.
Both models on the same scale
Both models on the same scale
Both models on the same scale
Fixed-price models can be estimated directly; dynamic models use recent references and are billed from actual provider usage.
Adjust the volume to see an instant estimate at the current public price.
Input estimate
$0.0002
Output estimate
$0.0001
Estimated total
$0.0003
Adjust the volume to see an instant estimate at the current public price.
Input estimate
$0.005
Output estimate
$0.003
Estimated total
$0.008