Qwen2.5:14b API
Run qwen2.5:14b through FastInfra's OpenAI-compatible API.
Free tier — no per-token charge.
Qwen2.5:14b pricing
Billed per token used. Prices sync automatically from wholesale providers.
| Direction | Price per 1M tokens |
|---|---|
| Input | Free |
| Output | Free |
Served by 1 provider
Requests route to Ollama (free tier) by default (cheapest).
Pin a specific provider with the :provider suffix.
| Provider | Upstream model ID |
|---|---|
| ollama | qwen2.5:14b |
Call Qwen2.5:14b in 30 seconds
Works with any OpenAI SDK — change the base URL and API key only.
Python
from openai import OpenAI
client = OpenAI(api_key="YOUR_API_KEY", base_url="https://api.fastinfra.ai/v1")
response = client.chat.completions.create(
model="qwen2.5:14b",
messages=[{"role": "user", "content": "Hello!"}]
)
print(response.choices[0].message.content)
curl
curl https://api.fastinfra.ai/v1/chat/completions \
-H "Authorization: Bearer YOUR_API_KEY" \
-H "Content-Type: application/json" \
-d '{
"model": "qwen2.5:14b",
"messages": [{"role": "user", "content": "Hello!"}]
}'
Qwen2.5:14b — common questions
How much does the qwen2.5:14b API cost?
qwen2.5:14b is on FastInfra's free tier — $0 per token.
Is qwen2.5:14b compatible with the OpenAI SDK?
Yes. FastInfra exposes qwen2.5:14b through an OpenAI-compatible endpoint at https://api.fastinfra.ai/v1 — point any OpenAI SDK at that base URL with a FastInfra API key and keep your existing code.
Which providers serve qwen2.5:14b?
qwen2.5:14b is available from 1 provider(s): ollama. Requests route to Ollama (free tier) by default (cheapest); append ":provider" to the model ID to pin one.
Related models
Compare side by side: Qwen2.5:14b vs Qwen2.5 Coder:1.5b · Qwen2.5:14b vs Qwen2.5:0.5b · Qwen2.5:14b vs Qwen2.5:1.5b