LLM7.io
LLM7.io
Toggle theme

chat model

DeepSeek logo

deepseek-v4-flash

deepseek-v4-flash is a chat model available through the LLM7 API. It supports text input, tool calling, streaming, reasoning. It provides a 1,000,000 tokens context window. It currently costs $0.05 USD input and $0.10 USD output per 1M tokens.

Available
deepseek-v4-flashGet an API key

Current price

Simple, pay-as-you-go pricing

1M tokens

$0.05 USD input and $0.10 USD output per 1M tokens

Cost calculator

Adjust the volume to see an instant estimate at the current public price.

1M tokens
10,000 tokens
01M

Input estimate

$0.0005

Output estimate

$0.00025

Estimated total

$0.00075

Observed usage

A 30-day snapshot of requests handled by LLM7, not a global benchmark.

Requests

124

observed on LLM7

Success rate

96.77%

observed requests

Average latency

19,888.04 ms

API response time

P95 latency

60,000 ms

slower observed requests

Start building

A verified request for this exact model. Add your API key and run it.

Read the docs
curl https://api.llm7.io/v1/chat/completions \
  -H "Authorization: Bearer $LLM7_API_TOKEN" \
  -H "Content-Type: application/json" \
  -d '{
    "model": "deepseek-v4-flash",
    "messages": [
      { "role": "user", "content": "Hello!" }
    ]
  }'

Explore similar models

Compare this model