LLM7.io
LLM7.io
Toggle theme

LLM7 model catalogue

Find the right model for your build.

Explore available chat, image, and video models with their pricing and capabilities in one place.

chat

29

image

1

video

6

Available models

Use filters to narrow the list to the capabilities you need.

Compare models
Anthropic logo

chat · pro

claude-fable-5

Available
claude-fable-5

$4.50 USD input and $30.00 USD output per 1M tokens

1,000,000 tokens context

Text inputToolsStreamingJSON mode

Stability60.87%

P95 latency60,000 ms

Anthropic logo

chat · pro

claude-opus-4-8

Available
claude-opus-4-8

$2.25 USD input and $12.19 USD output per 1M tokens

1,000,000 tokens context

Text inputVisionToolsStreaming

Stability89.8%

P95 latency60,000 ms

Anthropic logo

chat · pro

claude-opus-5

Available
claude-opus-5

$2.50 USD input and $12.50 USD output per 1M tokens

1,000,000 tokens context

Text inputVisionToolsStreaming

Stability88%

P95 latency60,000 ms

Anthropic logo

chat · pro

claude-sonnet-5

Available
claude-sonnet-5

$0.40 USD input and $2.00 USD output per 1M tokens

1,000,000 tokens context

Text inputToolsStreamingJSON mode

Stability94.7%

P95 latency60,000 ms

Mistral AI logo

chat · turbo

codestral-latest

Available
codestral-latest

$0.01 USD input and $0.02 USD output per 1M tokens

32,000 tokens context

Text inputToolsStreamingJSON mode

Stability98.76%

P95 latency30,000 ms

Available
deepseek-v4-flash:0731

$0.08 USD input and $0.16 USD output per 1M tokens

1,000,000 tokens context

Text inputToolsStreamingJSON mode

Stability60%

P95 latency60,000 ms

Available
deepseek-v4-flash:preview

$0.05 USD input and $0.10 USD output per 1M tokens

1,000,000 tokens context

Text inputToolsStreamingReasoning

Stability86.66%

P95 latency30,000 ms

DeepSeek logo

chat · pro

deepseek-v4-pro

Available
deepseek-v4-pro

$0.37 USD input and $0.74 USD output per 1M tokens

1,000,000 tokens context

Text inputToolsStreamingJSON mode

Stability97.41%

P95 latency30,000 ms

Available
gemini-3.1-flash-lite

$0.02 USD input and $0.04 USD output per 1M tokens

256,000 tokens context

Text inputVisionToolsStreaming

Stability98.24%

P95 latency10,000 ms

Available
gemini-3.5-flash-low

$0.04 USD input and $0.20 USD output per 1M tokens

1,040,000 tokens context

Text inputStreamingJSON modeReasoning

Stability80%

P95 latency60,000 ms

Google logo

video · pro

gemini-omni-flash

Available
gemini-omni-flash

$0.125 USD per second

Text inputVisionVideo generation

Stability100%

chat · turbo

gemma4:31b

Available
gemma4:31b

$0.03 USD input and $0.08 USD output per 1M tokens

262,000 tokens context

Text inputToolsStreamingJSON mode

Stability98.68%

P95 latency60,000 ms

OpenAI logo

chat · pro

gpt-5.4

Available
gpt-5.4

$0.15 USD input and $0.80 USD output per 1M tokens

1,050,000 tokens context

Text inputToolsStreamingJSON mode

Stability95.26%

P95 latency60,000 ms

OpenAI logo

chat · pro

gpt-5.4-mini

Available
gpt-5.4-mini

$0.04 USD input and $0.24 USD output per 1M tokens

400,000 tokens context

Text inputVisionToolsStreaming

Stability91.77%

P95 latency30,000 ms

OpenAI logo

chat · pro

gpt-5.5

Available
gpt-5.5

$0.55 USD input and $3.30 USD output per 1M tokens

1,050,000 tokens context

Text inputVisionToolsStreaming

Stability82%

P95 latency60,000 ms

OpenAI logo

chat · pro

gpt-5.6-sol

Available
gpt-5.6-sol

$2.00 USD input and $6.00 USD output per 1M tokens

1,000,000 tokens context

Text inputVisionToolsStreaming

Stability98.91%

P95 latency30,000 ms

OpenAI logo

chat · pro

gpt-5.6-terra

Available
gpt-5.6-terra

$0.50 USD input and $2.00 USD output per 1M tokens

1,000,000 tokens context

Text inputToolsStreamingJSON mode

Stability93.83%

P95 latency30,000 ms

OpenAI logo

image · pro

gpt-image-2

Available
gpt-image-2

$0.02 USD per image

Text inputVisionStreamingImage generation

Stability28.95%

P95 latency60,000 ms

OpenAI logo

chat · turbo

gpt-oss:20b

Available
gpt-oss:20b

$0.04 USD input and $0.06 USD output per 1M tokens

128,000 tokens context

Text inputToolsStreamingJSON mode

Stability98.79%

P95 latency60,000 ms

Inkling logo

chat · pro

Inkling

Available
Inkling

$1.00 USD input and $4.05 USD output per 1M tokens

512,000 tokens context

Text inputStreamingJSON modeReasoning

Stability93.33%

P95 latency60,000 ms

Inkling logo

chat · pro

Inkling-Small

Available
Inkling-Small

$0.50 USD input and $1.20 USD output per 1M tokens

512,000 tokens context

Text inputToolsStreamingJSON mode

Stability100%

P95 latency1,000 ms

Moonshot AI logo

chat · pro

kimi-k2.6

Available
kimi-k2.6

$0.05 USD input and $0.07 USD output per 1M tokens

240,000 tokens context

Text inputTools

Stability87.5%

P95 latency60,000 ms

Moonshot AI logo

chat · pro

kimi-k2.7-code

Available
kimi-k2.7-code

$0.09 USD input and $0.35 USD output per 1M tokens

256,000 tokens context

Text inputToolsStreamingJSON mode

Stability92.62%

P95 latency30,000 ms

Moonshot AI logo

chat · pro

kimi-k3

Available
kimi-k3

$1.70 USD input and $8.00 USD output per 1M tokens

1,000,000 tokens context

Text inputToolsStreamingJSON mode

Stability86.53%

P95 latency60,000 ms

Kling logo

video · pro

kling-v3.0-pro

Available
kling-v3.0-pro

$0.095 USD per second

Text inputVisionVideo generation

Stability100%

Kling logo

video · pro

kling-v3.0-turbo

Available
kling-v3.0-turbo

$0.095 USD per second

Text inputVisionVideo generation

Stability100%

Available
L3-8B-Lunaris-v1-Turbo

$0.04 USD input and $0.05 USD output per 1M tokens

8,000 tokens context

Text inputStreaming

Stability100%

P95 latency2,000 ms

MiniMax logo

chat · turbo

minimax-m2.7

Available
minimax-m2.7

$0.03 USD input and $0.05 USD output per 1M tokens

180,000 tokens context

Text inputToolsJSON modeReasoning

Stability93.9%

P95 latency60,000 ms

Available
mistral-Nemo-Instruct-2407

$0.03 USD input and $0.03 USD output per 1M tokens

128,000 tokens context

Text inputStreamingJSON mode

Stability95.71%

P95 latency30,000 ms

mistral-Small-24B-Instruct-2501

$0.06 USD input and $0.08 USD output per 1M tokens

32,000 tokens context

Text inputStreamingJSON mode

Stability100%

P95 latency1,000 ms

ByteDance logo

chat · pro

seed-2.0-mini

Available
seed-2.0-mini

$0.10 USD input and $0.40 USD output per 1M tokens

250,000 tokens context

Text inputToolsStreamingJSON mode

Stability100%

P95 latency5,000 ms

Seedance logo

video · pro

seedance-2.0

Available
seedance-2.0

$0.09 USD per second

Text inputVisionVideo generation

Stability100%

Seedance logo

video · pro

seedance-2.0-fast

Available
seedance-2.0-fast

$0.072 USD per second

Text inputVisionVideo generation

Stability100%

Seedance logo

video · pro

seedance-2.0-mini

Available
seedance-2.0-mini

$0.045 USD per second

Text inputVisionVideo generation

Stability100%

Available
XiaomiMiMo/MiMo-V2.5

$0.40 USD input and $2.00 USD output per 1M tokens

256,000 tokens context

Text inputToolsStreamingJSON mode

Stability100%

P95 latency30,000 ms

Available
XiaomiMiMo/MiMo-V2.5-Pro

$1.00 USD input and $3.00 USD output per 1M tokens

1,024,000 tokens context

Text inputToolsStreamingJSON mode

Stability85.71%

P95 latency60,000 ms

Retired models

Listed for reference only; they are no longer available through LLM7.

Retired
ByteDance/Seed-2.0-mini

$0.10 USD input and $0.40 USD output per 1M tokens

250,000 tokens context

Text inputStreamingJSON mode
Retired
deepseek-v4-flash

$0.05 USD input and $0.10 USD output per 1M tokens

1,000,000 tokens context

Text inputToolsStreamingReasoning

Stability99.04%

P95 latency30,000 ms

Firefly logo

video · pro

firefly-video

Retired
firefly-video

$0.02 USD per second

Text inputVisionToolsStreaming
Flux logo

image · pro

flux-kontext-max

Retired
flux-kontext-max

$0.01 USD per image

Text inputStreamingImage generationImage editing

Stability9.09%

P95 latency30,000 ms

Google logo

video · pro

gemini-veo31

Retired
gemini-veo31

$0.025 USD per second

Text inputVisionToolsStreaming

Stability100%

Retired
gpt-5.3-codex-spark

$0.36 USD input and $2.81 USD output per 1M tokens

400,000 tokens context

Text inputToolsStreamingJSON mode

Stability54.51%

P95 latency60,000 ms

OpenAI logo

chat · pro

gpt-5.6-luna

Retired
gpt-5.6-luna

$0.31 USD input and $1.81 USD output per 1M tokens

Text inputVisionToolsStreaming

Stability88.46%

P95 latency60,000 ms

xAI logo

chat · turbo

grok-3-mini

Retired
grok-3-mini

$0.02 USD input and $0.02 USD output per 1M tokens

Text inputToolsStreamingJSON mode

Stability35.11%

P95 latency30,000 ms

xAI logo

chat · pro

grok-4.5

Retired
grok-4.5

$0.31 USD input and $1.00 USD output per 1M tokens

500,000 tokens context

Text inputToolsStreamingJSON mode

Stability29.5%

P95 latency30,000 ms

meta-Llama-3.1-8B-Instruct-Turbo

$0.03 USD input and $0.04 USD output per 1M tokens

128,000 tokens context

Text inputToolsStreamingJSON mode

Stability96.64%

P95 latency30,000 ms

meta-llama/Meta-Llama-3.1-8B-Instruct-Turbo

$0.03 USD input and $0.04 USD output per 1M tokens

128,000 tokens context

Text inputStreamingJSON mode

Stability100%

P95 latency2,000 ms

mistralai/Mistral-Nemo-Instruct-2407

$0.03 USD input and $0.03 USD output per 1M tokens

128,000 tokens context

Text inputStreamingJSON mode

Stability100%

P95 latency5,000 ms

mistralai/Mistral-Small-24B-Instruct-2501

$0.06 USD input and $0.08 USD output per 1M tokens

32,000 tokens context

Text inputStreamingJSON mode
Sao10K/L3-8B-Lunaris-v1-Turbo

$0.04 USD input and $0.05 USD output per 1M tokens

8,000 tokens context

Text inputStreamingJSON mode
Retired
thinkingmachines/Inkling

$1.00 USD input and $4.05 USD output per 1M tokens

512,000 tokens context

Text inputStreamingJSON modeReasoning