Qwen3 Coder 480B A35B Instruct FP8 API

Run Qwen/Qwen3-Coder-480B-A35B-Instruct-FP8 through FastInfra's OpenAI-compatible API. Pay per token, no subscription, routed to the cheapest available provider.

Read the docs
Pricing

Qwen3 Coder 480B A35B Instruct FP8 pricing

Billed per token used. Prices sync automatically from wholesale providers.

Direction Price per 1M tokens
Input$2.10
Output$2.10
Availability

Served by 1 provider

Requests route to Together AI by default (cheapest). Pin a specific provider with the :provider suffix.

Provider Upstream model ID
together Qwen/Qwen3-Coder-480B-A35B-Instruct-FP8
Quickstart

Call Qwen3 Coder 480B A35B Instruct FP8 in 30 seconds

Works with any OpenAI SDK — change the base URL and API key only.

Python

from openai import OpenAI

client = OpenAI(api_key="YOUR_API_KEY", base_url="https://api.fastinfra.ai/v1")

response = client.chat.completions.create(
    model="Qwen/Qwen3-Coder-480B-A35B-Instruct-FP8",
    messages=[{"role": "user", "content": "Hello!"}]
)
print(response.choices[0].message.content)

curl

curl https://api.fastinfra.ai/v1/chat/completions \
  -H "Authorization: Bearer YOUR_API_KEY" \
  -H "Content-Type: application/json" \
  -d '{
    "model": "Qwen/Qwen3-Coder-480B-A35B-Instruct-FP8",
    "messages": [{"role": "user", "content": "Hello!"}]
  }'
FAQ

Qwen3 Coder 480B A35B Instruct FP8 — common questions

How much does the Qwen/Qwen3-Coder-480B-A35B-Instruct-FP8 API cost?

On FastInfra, Qwen/Qwen3-Coder-480B-A35B-Instruct-FP8 costs $2.1/1M input, $2.1/1M output tokens. Billing is per token used, with no subscription.

Is Qwen/Qwen3-Coder-480B-A35B-Instruct-FP8 compatible with the OpenAI SDK?

Yes. FastInfra exposes Qwen/Qwen3-Coder-480B-A35B-Instruct-FP8 through an OpenAI-compatible endpoint at https://api.fastinfra.ai/v1 — point any OpenAI SDK at that base URL with a FastInfra API key and keep your existing code.

Which providers serve Qwen/Qwen3-Coder-480B-A35B-Instruct-FP8?

Qwen/Qwen3-Coder-480B-A35B-Instruct-FP8 is available from 1 provider(s): together. Requests route to Together AI by default (cheapest); append ":provider" to the model ID to pin one.