LLM7.io
LLM7.io
Toggle theme

chat model

DeepSeek logo

deepseek-v4-flash:preview

deepseek-v4-flash:preview is a chat model available through the LLM7 API. It supports text input, tool calling, streaming, reasoning. It provides a 1,000,000 tokens context window. It currently costs $0.05 USD input and $0.10 USD output per 1M tokens.

Retired
deepseek-v4-flash:preview

Current LLM7 pricing

Simple, pay-as-you-go pricing

1M tokens

$0.05 USD input and $0.10 USD output per 1M tokens

Cost calculator

Adjust the volume to see an instant estimate at the current public price.

1M tokens
10,000 tokens
01M

Input estimate

$0.0005

Output estimate

$0.00025

Estimated total

$0.00075

Latest LLM7 statistics

Recent aggregated percentages and response-time metrics from LLM7, not a global benchmark or traffic disclosure.

Recent stability

78.22%

latest aggregated LLM7 statistic

Average response time

7,751.51 ms

latest observed API timing

P95 latency

60,000 ms

slower observed requests

Latest statistic trends

Only aggregated percentages and timing metrics with enough observed history are shown.

Request success rate

Latest aggregated LLM7 statistic

69.9%latest

6 Aug11 Aug

Average API latency

Latest aggregated LLM7 statistic

6,210 mslatest

6 Aug11 Aug

P95 API latency

Latest aggregated LLM7 statistic

30,000 mslatest

6 Aug11 Aug

Start building

A verified request for this exact model. Add your API key and run it.

Read the docs
curl https://api.llm7.io/v1/chat/completions \
  -H "Authorization: Bearer $LLM7_API_TOKEN" \
  -H "Content-Type: application/json" \
  -d '{
    "model": "deepseek-v4-flash:preview",
    "messages": [
      { "role": "user", "content": "Hello!" }
    ]
  }'

Explore similar models