LLM7.io
LLM7.io
Toggle theme

LLM7 model catalogue

Find the right model for your build.

Explore available chat, image, and video models with their pricing and capabilities in one place.

chat

37

image

4

video

6

Available models

Use filters to narrow the list to the capabilities you need.

Compare models

image · pro

chroma-v.46-flash

Available
chroma-v.46-flash

$0.015 USD per image

Text inputVisionToolsStreaming
Anthropic logo

chat · pro

claude-fable-5

Available
claude-fable-5

$4.50 USD input and $30.00 USD output per 1M tokens

1,000,000 tokens context

Text inputToolsStreamingJSON mode

Stability98.34%

P95 latency60,000 ms

Anthropic logo

chat · pro

claude-fable-5-1

Available
claude-fable-5-1

$4.00 USD input and $15.00 USD output per 1M tokens

1,000,000 tokens context

Text inputVisionToolsStreaming

Stability98.68%

P95 latency10,000 ms

Anthropic logo

chat · pro

claude-haiku-4-5

Available
claude-haiku-4-5

$0.04 USD input and $0.20 USD output per 1M tokens

200,000 tokens context

Text inputToolsStreamingJSON mode

Stability96.16%

P95 latency10,000 ms

Anthropic logo

chat · pro

claude-opus-4-8

Available
claude-opus-4-8

$2.25 USD input and $12.19 USD output per 1M tokens

1,000,000 tokens context

Text inputVisionToolsStreaming

Stability97.8%

P95 latency60,000 ms

Anthropic logo

chat · pro

claude-opus-5

Available
claude-opus-5

$2.50 USD input and $12.50 USD output per 1M tokens

1,000,000 tokens context

Text inputVisionToolsStreaming

Stability97.17%

P95 latency60,000 ms

Available
claude-sonnet-4-6

$0.12 USD input and $0.45 USD output per 1M tokens

1,000,000 tokens context

Text inputVisionToolsStreaming

Stability98.35%

P95 latency60,000 ms

Anthropic logo

chat · pro

claude-sonnet-5

Available
claude-sonnet-5

$0.45 USD input and $2.25 USD output per 1M tokens

1,000,000 tokens context

Text inputToolsStreamingJSON mode

Stability93.38%

P95 latency60,000 ms

Mistral AI logo

chat · turbo

codestral-latest

Available
codestral-latest

$0.01 USD input and $0.02 USD output per 1M tokens

32,000 tokens context

Text inputToolsStreamingJSON mode

Stability97.96%

P95 latency10,000 ms

image · pro

dark-beast-krea2

Available
dark-beast-krea2

$0.03 USD per image

Text inputVisionToolsStreaming

Stability91.69%

P95 latency60,000 ms

Available
deepseek-v4-flash:0731

$0.06 USD input and $0.18 USD output per 1M tokens

1,024,000 tokens context

Text inputToolsStreamingJSON mode

Stability99.29%

P95 latency60,000 ms

Available
deepseek-v4-flash_0731

$0.18 USD input and $0.60 USD output per 1M tokens

1,000,000 tokens context

Text inputToolsStreamingJSON mode
Available
DeepSeek-V4-Flash-0731

$0.02 USD input and $0.04 USD output per 1M tokens

400,000 tokens context

Text inputToolsStreamingJSON mode

Stability77.16%

P95 latency60,000 ms

DeepSeek logo

chat · pro

deepseek-v4-pro

Available
deepseek-v4-pro

$1.11 USD input and $3.33 USD output per 1M tokens

1,000,000 tokens context

Text inputToolsStreamingJSON mode

Stability98.97%

P95 latency30,000 ms

Available
gemini-3.1-flash-lite

$0.02 USD input and $0.04 USD output per 1M tokens

256,000 tokens context

Text inputVisionToolsStreaming

Stability99.12%

P95 latency10,000 ms

Google logo

chat · pro

gemini-3.7-flash

Available
gemini-3.7-flash

$0.06 USD input and $0.30 USD output per 1M tokens

1,000,000 tokens context

Text inputVisionToolsStreaming

Stability95.63%

P95 latency30,000 ms

Available
gemini-3.8-flash-high

$0.05 USD input and $0.15 USD output per 1M tokens

1,000,000 tokens context

Text inputVisionToolsStreaming

Stability96.63%

P95 latency30,000 ms

Google logo

chat · pro

gemini-3-flash

Available
gemini-3-flash

$0.03 USD input and $0.08 USD output per 1M tokens

1,048,576 tokens context

Text inputToolsStreamingJSON mode

Stability91.09%

P95 latency10,000 ms

Google logo

video · pro

gemini-omni-flash

Available
gemini-omni-flash

Dynamic per-request quote

Billed at actual provider cost with no LLM7 markup.

Text inputVisionVideo generation

chat · pro

gemma4:31b

Available
gemma4:31b

$0.07 USD input and $0.23 USD output per 1M tokens

262,000 tokens context

Text inputToolsStreamingJSON mode

Stability99.59%

P95 latency30,000 ms

Z.ai logo

chat · pro

glm-5.3

Available
glm-5.3

$0.50 USD input and $2.00 USD output per 1M tokens

1,000,000 tokens context

Text inputToolsStreamingJSON mode

Stability91.35%

P95 latency60,000 ms

Z.ai logo

chat · pro

glm-5.3-flash

Available
glm-5.3-flash

$0.15 USD input and $0.55 USD output per 1M tokens

1,048,576 tokens context

Text inputVisionToolsStreaming

Stability97.14%

P95 latency60,000 ms

OpenAI logo

chat · pro

gpt-5.5

Available
gpt-5.5

$0.20 USD input and $1.00 USD output per 1M tokens

1,050,000 tokens context

Text inputVisionToolsStreaming

Stability93.16%

P95 latency60,000 ms

OpenAI logo

chat · pro

gpt-5.6-luna

Available
gpt-5.6-luna

$0.45 USD input and $0.58 USD output per 1M tokens

Text inputVisionToolsStreaming

Stability90.87%

P95 latency60,000 ms

OpenAI logo

chat · pro

gpt-5.6-sol

Available
gpt-5.6-sol

$1.92 USD input and $5.40 USD output per 1M tokens

1,000,000 tokens context

Text inputVisionToolsStreaming

Stability88.05%

P95 latency60,000 ms

OpenAI logo

chat · pro

gpt-5.6-terra

Available
gpt-5.6-terra

$0.24 USD input and $1.00 USD output per 1M tokens

1,000,000 tokens context

Text inputVisionToolsStreaming

Stability93.82%

P95 latency60,000 ms

OpenAI logo

chat · pro

gpt-6-astra

Available
gpt-6-astra

$5.00 USD input and $10.00 USD output per 1M tokens

1,000,000 tokens context

Text inputVisionToolsStreaming

Stability79.54%

P95 latency60,000 ms

OpenAI logo

image · pro

gpt-image-2

Available
gpt-image-2

$0.02 USD per image

Text inputVisionStreamingImage generation

Stability80.57%

P95 latency60,000 ms

OpenAI logo

image · pro

gpt-image-2.5

Available
gpt-image-2.5

$0.025 USD per image

Text inputVisionToolsStreaming

Stability95.24%

P95 latency60,000 ms

xAI logo

chat · pro

grok-4.5

Available
grok-4.5

$0.30 USD input and $1.00 USD output per 1M tokens

500,000 tokens context

Text inputToolsStreamingJSON mode

Stability91.47%

P95 latency60,000 ms

xAI logo

chat · pro

grok-4.6

Available
grok-4.6

$0.40 USD input and $0.50 USD output per 1M tokens

500,000 tokens context

Text inputToolsStreamingJSON mode

Stability83.24%

P95 latency60,000 ms

Inkling logo

chat · pro

Inkling

Available
Inkling

$1.00 USD input and $4.05 USD output per 1M tokens

512,000 tokens context

Text inputStreamingJSON modeReasoning

Stability94.12%

P95 latency60,000 ms

Inkling logo

chat · pro

Inkling-Small

Available
Inkling-Small

$0.50 USD input and $1.20 USD output per 1M tokens

512,000 tokens context

Text inputToolsStreamingJSON mode

Stability53.85%

P95 latency5,000 ms

Moonshot AI logo

chat · pro

kimi-k3

Available
kimi-k3

$2.00 USD input and $10.00 USD output per 1M tokens

1,000,000 tokens context

Text inputVisionToolsStreaming

Stability94.91%

P95 latency60,000 ms

Kling logo

video · pro

kling-v3.0-pro

Available
kling-v3.0-pro

Dynamic per-request quote

Billed at actual provider cost with no LLM7 markup.

Text inputVisionVideo generation
Kling logo

video · pro

kling-v3.0-turbo

Available
kling-v3.0-turbo

Dynamic per-request quote

Billed at actual provider cost with no LLM7 markup.

Text inputVisionVideo generation
Available
L3-8B-Lunaris-v1-Turbo

$0.04 USD input and $0.05 USD output per 1M tokens

8,000 tokens context

Text inputStreaming

Stability100%

P95 latency2,000 ms

chat · pro

llama-4-maverick

Available
llama-4-maverick

$0.19 USD input and $0.75 USD output per 1M tokens

1,048,576 tokens context

Text inputVisionToolsStreaming

Stability100%

P95 latency2,000 ms

MiniMax logo

chat · turbo

minimax-m2.7

Available
minimax-m2.7

$0.03 USD input and $0.05 USD output per 1M tokens

180,000 tokens context

Text inputToolsJSON modeReasoning

Stability91.47%

P95 latency60,000 ms

Available
mistral-Nemo-Instruct-2407

$0.03 USD input and $0.03 USD output per 1M tokens

128,000 tokens context

Text inputStreamingJSON mode

Stability98.91%

P95 latency30,000 ms

mistral-Small-24B-Instruct-2501

$0.06 USD input and $0.08 USD output per 1M tokens

32,000 tokens context

Text inputStreamingJSON mode

Stability100%

P95 latency30,000 ms

ByteDance logo

chat · pro

seed-2.0-mini

Available
seed-2.0-mini

$0.10 USD input and $0.40 USD output per 1M tokens

250,000 tokens context

Text inputToolsStreamingJSON mode

Stability66.67%

P95 latency5,000 ms

Seedance logo

video · pro

seedance-2.0

Available
seedance-2.0

Dynamic per-request quote

Billed at actual provider cost with no LLM7 markup.

Text inputVisionVideo generation
Seedance logo

video · pro

seedance-2.0-fast

Available
seedance-2.0-fast

Typically from $1.55 for 10s (720p)

Billed at actual provider cost with no LLM7 markup.

Text inputVisionVideo generation

Stability100%

Seedance logo

video · pro

seedance-2.0-mini

Available
seedance-2.0-mini

Typically from $0.85 for 10s (720p)

Billed at actual provider cost with no LLM7 markup.

Text inputVisionVideo generation

Stability100%

Available
XiaomiMiMo/MiMo-V2.5

$0.40 USD input and $2.00 USD output per 1M tokens

256,000 tokens context

Text inputToolsStreamingJSON mode

Stability61.84%

P95 latency60,000 ms

Available
XiaomiMiMo/MiMo-V2.5-Pro

$1.00 USD input and $3.00 USD output per 1M tokens

1,024,000 tokens context

Text inputToolsStreamingJSON mode

Stability100%

P95 latency30,000 ms

Retired models

Listed for reference only; they are no longer available through LLM7.

Retired
ByteDance/Seed-2.0-mini

$0.10 USD input and $0.40 USD output per 1M tokens

250,000 tokens context

Text inputStreamingJSON mode
bytedance/seedance-2.0-fast/image-to-video

Typically from $1.55 for 10s (720p)

Billed at actual provider cost with no LLM7 markup.

Text inputVisionToolsStreaming
bytedance/seedance-2.0-fast/text-to-video

Dynamic per-request quote

Billed at actual provider cost with no LLM7 markup.

Text inputVisionToolsStreaming
bytedance/seedance-2.0/image-to-video

Dynamic per-request quote

Billed at actual provider cost with no LLM7 markup.

Text inputVisionToolsStreaming
bytedance/seedance-2.0-mini/image-to-video

Typically from $0.85 for 10s (720p)

Billed at actual provider cost with no LLM7 markup.

Text inputVisionToolsStreaming
bytedance/seedance-2.0-mini/text-to-video

Dynamic per-request quote

Billed at actual provider cost with no LLM7 markup.

Text inputVisionToolsStreaming
bytedance/seedance-2.0/text-to-video

Dynamic per-request quote

Billed at actual provider cost with no LLM7 markup.

Text inputVisionToolsStreaming
deepseek-ai/DeepSeek-V4-Flash-0731

$0.06 USD input and $0.18 USD output per 1M tokens

1,024,000 tokens context

Text inputToolsStreamingJSON mode
DeepSeek logo

chat · turbo

deepseek-v3

Retired
deepseek-v3

$0.01 USD input and $0.02 USD output per 1M tokens

1,048,576 tokens context

Text inputToolsStreamingReasoning

Stability87.98%

P95 latency60,000 ms

DeepSeek logo

chat · turbo

deepseek-v3.2

Retired
deepseek-v3.2

$0.02 USD input and $0.10 USD output per 1M tokens

Text inputJSON modeReasoning

Stability88.75%

P95 latency60,000 ms

Retired
deepseek-v4-flash

$0.11 USD input and $0.25 USD output per 1M tokens

1,048,576 tokens context

Text inputToolsStreamingJSON mode

Stability91.73%

P95 latency30,000 ms

Retired
deepseek-v4-flash:preview

$0.05 USD input and $0.10 USD output per 1M tokens

1,000,000 tokens context

Text inputToolsStreamingReasoning
Retired
firefly-gpt-image-2

$0.03 USD per image

Text inputVisionToolsStreaming

Stability78.21%

P95 latency60,000 ms

Firefly logo

image · pro

firefly-image-5

Retired
firefly-image-5

$0.015 USD per image

Text inputVisionToolsStreaming

Stability81.87%

P95 latency60,000 ms

Firefly logo

video · pro

firefly-video

Retired
firefly-video

$0.02 USD per second

Text inputVisionToolsStreaming
Flux logo

image · pro

flux-klein-2

Retired
flux-klein-2

$0.02 USD per image

Text inputVisionToolsStreaming

Stability30.53%

P95 latency30,000 ms

Flux logo

image · pro

flux-kontext-max

Retired
flux-kontext-max

$0.01 USD per image

Text inputStreamingImage generationImage editing
Retired
gemini-3.5-flash-low

$0.04 USD input and $0.20 USD output per 1M tokens

1,040,000 tokens context

Text inputToolsStreamingJSON mode

Stability99.09%

P95 latency30,000 ms

Google logo

video · pro

gemini-veo31

Retired
gemini-veo31

$0.025 USD per second

Text inputVisionToolsStreaming
Z.ai logo

chat · pro

glm-5.2

Retired
glm-5.2

$0.44 USD input and $2.76 USD output per 1M tokens

976,000 tokens context

Text inputToolsStreamingJSON mode
google/gemini-omni-flash/image-to-video

Dynamic per-request quote

Billed at actual provider cost with no LLM7 markup.

Text inputVisionToolsStreaming
google/gemini-omni-flash/text-to-video

Dynamic per-request quote

Billed at actual provider cost with no LLM7 markup.

Text inputVisionToolsStreaming
Retired
gpt-5.3-codex-spark

$0.11 USD input and $0.84 USD output per 1M tokens

400,000 tokens context

Text inputToolsStreamingJSON mode

Stability91.91%

P95 latency60,000 ms

OpenAI logo

chat · pro

gpt-5.4

Retired
gpt-5.4

$0.15 USD input and $0.80 USD output per 1M tokens

1,050,000 tokens context

Text inputToolsStreamingJSON mode

Stability95.13%

P95 latency60,000 ms

OpenAI logo

chat · pro

gpt-5.4-mini

Retired
gpt-5.4-mini

$0.04 USD input and $0.24 USD output per 1M tokens

400,000 tokens context

Text inputVisionToolsStreaming

Stability81.05%

P95 latency60,000 ms

Retired
gpt-5.5-openai-compact

$0.20 USD input and $1.20 USD output per 1M tokens

Text inputStreamingJSON mode
OpenAI logo

chat · pro

gpt-5-nano

Retired
gpt-5-nano

$0.05 USD input and $0.40 USD output per 1M tokens

400,000 tokens context

Text inputVisionToolsStreaming
OpenAI logo

chat · turbo

gpt-oss

Retired
gpt-oss

$0.05 USD input and $0.18 USD output per 1M tokens

131,072 tokens context

Text inputToolsStreamingJSON mode

Stability96.04%

P95 latency60,000 ms

OpenAI logo

chat · turbo

gpt-oss:20b

Retired
gpt-oss:20b

$0.04 USD input and $0.06 USD output per 1M tokens

128,000 tokens context

Text inputToolsStreamingJSON mode

Stability99.15%

P95 latency30,000 ms

xAI logo

chat · turbo

grok-3-mini

Retired
grok-3-mini

$0.02 USD input and $0.02 USD output per 1M tokens

Text inputToolsStreamingJSON mode

image · pro

imagine-1.5

Retired
imagine-1.5

$0.0182 USD per image

Text inputVisionToolsStreaming
Moonshot AI logo

chat · pro

kimi-k2.6

Retired
kimi-k2.6

$0.05 USD input and $0.07 USD output per 1M tokens

240,000 tokens context

Text inputTools

Stability83.06%

P95 latency60,000 ms

Moonshot AI logo

chat · pro

kimi-k2.7-code

Retired
kimi-k2.7-code

$0.09 USD input and $0.35 USD output per 1M tokens

256,000 tokens context

Text inputToolsStreamingJSON mode

Stability95.35%

P95 latency30,000 ms

kwaivgi/kling-v3.0-pro/image-to-video

Dynamic per-request quote

Billed at actual provider cost with no LLM7 markup.

Text inputVisionToolsStreaming
kwaivgi/kling-v3.0-pro/text-to-video

Dynamic per-request quote

Billed at actual provider cost with no LLM7 markup.

Text inputVisionToolsStreaming
kwaivgi/kling-v3.0-turbo/image-to-video

Dynamic per-request quote

Billed at actual provider cost with no LLM7 markup.

Text inputVisionToolsStreaming
kwaivgi/kling-v3.0-turbo/text-to-video

Dynamic per-request quote

Billed at actual provider cost with no LLM7 markup.

Text inputVisionToolsStreaming
meta-Llama-3.1-8B-Instruct-Turbo

$0.03 USD input and $0.05 USD output per 1M tokens

128,000 tokens context

Text inputToolsStreamingJSON mode

Stability96.34%

P95 latency60,000 ms

meta-llama/Meta-Llama-3.1-8B-Instruct-Turbo

$0.03 USD input and $0.04 USD output per 1M tokens

128,000 tokens context

Text inputStreamingJSON mode
MiniMax logo

chat · pro

minimax-m3

Retired
minimax-m3

$0.08 USD input and $0.30 USD output per 1M tokens

1,048,576 tokens context

Text inputVisionToolsStreaming

Stability85%

P95 latency60,000 ms

mistralai/Mistral-Nemo-Instruct-2407

$0.03 USD input and $0.03 USD output per 1M tokens

128,000 tokens context

Text inputStreamingJSON mode
mistralai/Mistral-Small-24B-Instruct-2501

$0.06 USD input and $0.08 USD output per 1M tokens

32,000 tokens context

Text inputStreamingJSON mode
Sao10K/L3-8B-Lunaris-v1-Turbo

$0.04 USD input and $0.05 USD output per 1M tokens

8,000 tokens context

Text inputStreamingJSON mode
Retired
seedance-2.0-unrestricted

$0.035437 USD per second

Text inputVisionToolsStreaming
Retired
thinkingmachines/Inkling

$1.00 USD input and $4.05 USD output per 1M tokens

512,000 tokens context

Text inputStreamingJSON modeReasoning