Requests
124
observed on LLM7
chat model
deepseek-v4-flash is a chat model available through the LLM7 API. It supports text input, tool calling, streaming, reasoning. It provides a 1,000,000 tokens context window. It currently costs $0.05 USD input and $0.10 USD output per 1M tokens.
deepseek-v4-flashGet an API keyCurrent price
$0.05 USD input and $0.10 USD output per 1M tokens
Adjust the volume to see an instant estimate at the current public price.
Input estimate
$0.0005
Output estimate
$0.00025
Estimated total
$0.00075
A 30-day snapshot of requests handled by LLM7, not a global benchmark.
Requests
124
observed on LLM7
Success rate
96.77%
observed requests
Average latency
19,888.04 ms
API response time
P95 latency
60,000 ms
slower observed requests
A verified request for this exact model. Add your API key and run it.
curl https://api.llm7.io/v1/chat/completions \
-H "Authorization: Bearer $LLM7_API_TOKEN" \
-H "Content-Type: application/json" \
-d '{
"model": "deepseek-v4-flash",
"messages": [
{ "role": "user", "content": "Hello!" }
]
}'