Llama 4 Scout 17B 16E Instruct API
Run meta-llama/Llama-4-Scout-17B-16E-Instruct through FastInfra's OpenAI-compatible API.
Pay per token, no subscription, routed to the cheapest available provider.
Llama 4 Scout 17B 16E Instruct pricing
Billed per token used. Prices sync automatically from wholesale providers.
| Direction | Price per 1M tokens |
|---|---|
| Input | $0.11 |
| Output | $0.32 |
Served by 2 providers
Requests route to DeepInfra by default (cheapest).
Pin a specific provider with the :provider suffix.
| Provider | Upstream model ID |
|---|---|
| together | meta-llama/Llama-4-Scout-17B-16E-Instruct |
| deepinfra | meta-llama/Llama-4-Scout-17B-16E-Instruct |
Call Llama 4 Scout 17B 16E Instruct in 30 seconds
Works with any OpenAI SDK — change the base URL and API key only.
Python
from openai import OpenAI
client = OpenAI(api_key="YOUR_API_KEY", base_url="https://api.fastinfra.ai/v1")
response = client.chat.completions.create(
model="meta-llama/Llama-4-Scout-17B-16E-Instruct",
messages=[{"role": "user", "content": "Hello!"}]
)
print(response.choices[0].message.content)
curl
curl https://api.fastinfra.ai/v1/chat/completions \
-H "Authorization: Bearer YOUR_API_KEY" \
-H "Content-Type: application/json" \
-d '{
"model": "meta-llama/Llama-4-Scout-17B-16E-Instruct",
"messages": [{"role": "user", "content": "Hello!"}]
}'
Llama 4 Scout 17B 16E Instruct — common questions
How much does the meta-llama/Llama-4-Scout-17B-16E-Instruct API cost?
On FastInfra, meta-llama/Llama-4-Scout-17B-16E-Instruct costs $0.105/1M input, $0.315/1M output tokens. Billing is per token used, with no subscription.
Is meta-llama/Llama-4-Scout-17B-16E-Instruct compatible with the OpenAI SDK?
Yes. FastInfra exposes meta-llama/Llama-4-Scout-17B-16E-Instruct through an OpenAI-compatible endpoint at https://api.fastinfra.ai/v1 — point any OpenAI SDK at that base URL with a FastInfra API key and keep your existing code.
Which providers serve meta-llama/Llama-4-Scout-17B-16E-Instruct?
meta-llama/Llama-4-Scout-17B-16E-Instruct is available from 2 provider(s): together, deepinfra. Requests route to DeepInfra by default (cheapest); append ":provider" to the model ID to pin one.
Related models
Compare side by side: Llama 4 Scout 17B 16E Instruct vs Llama3.1:8b · Llama 4 Scout 17B 16E Instruct vs Llama3.2:1b · Llama 4 Scout 17B 16E Instruct vs Llama3.2:3b