Recent stability
96.64%
latest aggregated LLM7 statistic
chat model
meta-Llama-3.1-8B-Instruct-Turbo is a chat model available through the LLM7 API. It supports text input, tool calling, streaming, JSON mode. It provides a 128,000 tokens context window. It currently costs $0.03 USD input and $0.04 USD output per 1M tokens.
meta-Llama-3.1-8B-Instruct-TurboCurrent LLM7 pricing
$0.03 USD input and $0.04 USD output per 1M tokens
Adjust the volume to see an instant estimate at the current public price.
Input estimate
$0.0003
Output estimate
$0.0001
Estimated total
$0.0004
Recent aggregated percentages and response-time metrics from LLM7, not a global benchmark or traffic disclosure.
Recent stability
96.64%
latest aggregated LLM7 statistic
Average response time
5,691.88 ms
latest observed API timing
P95 latency
30,000 ms
slower observed requests
Only aggregated percentages and timing metrics with enough observed history are shown.
Latest aggregated LLM7 statistic
90.5%latest
Latest aggregated LLM7 statistic
4,726 mslatest
Latest aggregated LLM7 statistic
30,000 mslatest
A verified request for this exact model. Add your API key and run it.
curl https://api.llm7.io/v1/chat/completions \
-H "Authorization: Bearer $LLM7_API_TOKEN" \
-H "Content-Type: application/json" \
-d '{
"model": "meta-Llama-3.1-8B-Instruct-Turbo",
"messages": [
{ "role": "user", "content": "Hello!" }
]
}'