LLM7.io
LLM7.io
Toggle theme

chat model

Z.ai logo

glm-5.3-flash

glm-5.3-flash is a chat model available through the LLM7 API. It supports text input, image input, tool calling, streaming, reasoning. It provides a 1,048,576 tokens context window. It currently costs $0.04 USD input and $0.15 USD output per 1M tokens.

Available
glm-5.3-flashGet an API key

Current LLM7 pricing

Simple, pay-as-you-go pricing

1M tokens

$0.04 USD input and $0.15 USD output per 1M tokens

Cost calculator

Adjust the volume to see an instant estimate at the current public price.

1M tokens
10,000 tokens
01M

Input estimate

$0.0004

Output estimate

$0.000375

Estimated total

$0.000775

Latest LLM7 statistics

Recent aggregated percentages and response-time metrics from LLM7, not a global benchmark or traffic disclosure.

Recent stability

95.82%

latest aggregated LLM7 statistic

Average response time

26,736.32 ms

latest observed API timing

P95 latency

60,000 ms

slower observed requests

Latest statistic trends

Only aggregated percentages and timing metrics with enough observed history are shown.

Request success rate

Latest aggregated LLM7 statistic

97.8%latest

29 Aug3 Sept

Average API latency

Latest aggregated LLM7 statistic

16,370 mslatest

29 Aug3 Sept

P95 API latency

Latest aggregated LLM7 statistic

60,000 mslatest

29 Aug3 Sept

Start building

A verified request for this exact model. Add your API key and run it.

Read the docs
curl https://api.llm7.io/v1/chat/completions \
  -H "Authorization: Bearer $LLM7_API_TOKEN" \
  -H "Content-Type: application/json" \
  -d '{
    "model": "glm-5.3-flash",
    "messages": [
      { "role": "user", "content": "Hello!" }
    ]
  }'

Explore similar models

Compare this model