gemma4:31b
Input $0.07 USD · Output $0.23 USD / 1M tokens
chat comparison
Compare the things that matter before you build: current price, capabilities, and latest aggregated LLM7 statistics.
Current pricing
gemma4:31b
Input $0.07 USD · Output $0.23 USD / 1M tokens
Inkling
Input $1.00 USD · Output $4.05 USD / 1M tokens
Only capabilities that differ between these models are listed. Green highlights the broader supported option or larger context window.
Context window
gemma4:31b
262,000 tokens
InklingBetter fit
512,000 tokens
Tool calling
gemma4:31bBetter fit
Supported
Inkling
Not supported
Reasoning
gemma4:31b
Not supported
InklingBetter fit
Supported
Recent aggregated percentages and response-time metrics from LLM7, not a global model benchmark or traffic disclosure.
Recent stability
gemma4:31b
99.6%
Inkling
94.07%
Average response time
gemma4:31b
5,132.21 ms
Inkling
9,720.25 ms
P95 API latency
gemma4:31b
30,000 ms
Inkling
60,000 ms
Shared charts make the observed differences easier to read. Metrics without enough history stay hidden.
Both models on the same scale
Both models on the same scale
Both models on the same scale
Fixed-price models can be estimated directly; dynamic models use recent references and are billed from actual provider usage.
Adjust the volume to see an instant estimate at the current public price.
Input estimate
$0.0007
Output estimate
$0.000575
Estimated total
$0.001275
Adjust the volume to see an instant estimate at the current public price.
Input estimate
$0.01
Output estimate
$0.010125
Estimated total
$0.020125