Recent stability
99.04%
latest aggregated LLM7 statistic
chat model
deepseek-v4-flash is a chat model available through the LLM7 API. It supports text input, tool calling, streaming, reasoning. It provides a 1,000,000 tokens context window. It currently costs $0.05 USD input and $0.10 USD output per 1M tokens.
deepseek-v4-flashCurrent LLM7 pricing
$0.05 USD input and $0.10 USD output per 1M tokens
Adjust the volume to see an instant estimate at the current public price.
Input estimate
$0.0005
Output estimate
$0.00025
Estimated total
$0.00075
Recent aggregated percentages and response-time metrics from LLM7, not a global benchmark or traffic disclosure.
Recent stability
99.04%
latest aggregated LLM7 statistic
Average response time
3,655.07 ms
latest observed API timing
P95 latency
30,000 ms
slower observed requests
Only aggregated percentages and timing metrics with enough observed history are shown.
Latest aggregated LLM7 statistic
99.1%latest
Latest aggregated LLM7 statistic
2,164 mslatest
Latest aggregated LLM7 statistic
10,000 mslatest
A verified request for this exact model. Add your API key and run it.
curl https://api.llm7.io/v1/chat/completions \
-H "Authorization: Bearer $LLM7_API_TOKEN" \
-H "Content-Type: application/json" \
-d '{
"model": "deepseek-v4-flash",
"messages": [
{ "role": "user", "content": "Hello!" }
]
}'