LLM7.io
LLM7.io
Toggle theme

chat model

DeepSeek logo

DeepSeek-V4-Flash-0731

DeepSeek-V4-Flash-0731 is a chat model available through the LLM7 API. It supports text input, tool calling, streaming, JSON mode, reasoning. It provides a 400,000 tokens context window. It currently costs $0.02 USD input and $0.04 USD output per 1M tokens.

Retired
DeepSeek-V4-Flash-0731

Current LLM7 pricing

Simple, pay-as-you-go pricing

1M tokens

$0.02 USD input and $0.04 USD output per 1M tokens

Cost calculator

Adjust the volume to see an instant estimate at the current public price.

1M tokens
10,000 tokens
01M

Input estimate

$0.0002

Output estimate

$0.0001

Estimated total

$0.0003

Latest LLM7 statistics

Recent aggregated percentages and response-time metrics from LLM7, not a global benchmark or traffic disclosure.

Recent stability

76.68%

latest aggregated LLM7 statistic

Average response time

31,611.37 ms

latest observed API timing

P95 latency

60,000 ms

slower observed requests

Latest statistic trends

Only aggregated percentages and timing metrics with enough observed history are shown.

Request success rate

Latest aggregated LLM7 statistic

57.5%latest

14 Aug26 Aug

Average API latency

Latest aggregated LLM7 statistic

57,110 mslatest

14 Aug26 Aug

P95 API latency

Latest aggregated LLM7 statistic

60,000 mslatest

14 Aug26 Aug

Start building

A verified request for this exact model. Add your API key and run it.

Read the docs
curl https://api.llm7.io/v1/chat/completions \
  -H "Authorization: Bearer $LLM7_API_TOKEN" \
  -H "Content-Type: application/json" \
  -d '{
    "model": "DeepSeek-V4-Flash-0731",
    "messages": [
      { "role": "user", "content": "Hello!" }
    ]
  }'

Explore similar models