This model is a mixed int4 model with group_size 64 and symmetric quantization of Qwen/Qwen3-Coder-480B-A35B-Instruct generated by intel/auto-round via **RTN** (no algorithm tuning). Non expert layers fallback to 8 bits and group_size 128. mlp.gate layers fallback to 16 bits to ensure runing successfully on vLLM.
# Paste into Claude Code, Cursor, or Codex
Help me set up the Respan gateway: fetch and follow https:/
4 providers serve this model through the Respan gateway, cheapest first. Copy a model ID to call that route directly.
| Provider | Input | Output |
|---|---|---|
TeraBYOK only tera/Qwen/Qwen3-Coder-480B-A35B-Instruct | $0.22/M | $1.80/M |
NovitaSave 10% novita/qwen/qwen3-coder-480b-a35b-instruct | $0.34/M | $1.40/M |
IO.net ionet/Intel/Qwen3-Coder-480B-A35B-Instruct-int4-mixed-ar | $0.53/M | $2.60/M |
TogetherAIBYOK only together_ai/Qwen/Qwen3-Coder-480B-A35B-Instruct-FP8 | $2.00/M | $2.00/M |
Built to meet the security and privacy standards that enterprise and healthcare teams require.
Is a provider down? Together AI status
Averaged across 4 providers, in USD per 1M tokens after gateway discounts.
Total cost from 0 to 1T tokens on the cheapest and most expensive provider.
30-day averages per provider from requests through the Respan gateway, best first. 2 of 4 providers had traffic in that window.
10tokens/sbest, on IO.net
0.9sbest, on IO.net
Use Qwen3 Coder 480B A35B Instruct in the tools you already run. Each guide points the app at the Respan gateway, then you pick this model.
Setup guides for more apps are in the integration docs.
ISO 27001
The internationally recognized standard for information security management.
SOC 2
Secure, compliant management of your data across all of our systems.
GDPR
Operated under GDPR, the world's strictest standard for data privacy.
HIPAA
HIPAA compliant, with a BAA available for healthcare teams.