Model Library/Qwen3 Coder 480B A35B Instruct
Qwen

Qwen3 Coder 480B A35B Instruct

qwen/qwen3-coder-480b-a35b-instruct
Qwen3-Coder-480B-A35B-Instruct is a cutting-edge open coding model from Qwen, matching Claude Sonnet’s performance in agentic programming, browser automation, and core development tasks. With native 256K context (extendable to 1M tokens via YaRN), it excels at repository-scale analysis and features specialized function-call support for platforms like Qwen Code and CLINE—making it ideal for complex, real-world development workflows.

Características

API serverless

Documentación

qwen/qwen3-coder-480b-a35b-instruct is available via Novita's serverless API, where you pay per token. There are several ways to call the API, including OpenAI-compatible endpoints with exceptional reasoning performance.

Implementaciones bajo demanda

Documentación

On-demand deployments allow you to use qwen/qwen3-coder-480b-a35b-instruct on dedicated GPUs with high-performance serving stack with high reliability and no rate limits.

Serverless disponible

Ejecuta consultas de inmediato, paga solo por el uso

Entrada$0.38 / M Tokens
Salida$1.55 / M Tokens

Usa los siguientes ejemplos de código para integrarte con nuestra API:

1from openai import OpenAI
2
3client = OpenAI(
4    api_key="<Your API Key>",
5    base_url="https://api.novita.ai/openai"
6)
7
8response = client.chat.completions.create(
9    model="qwen/qwen3-coder-480b-a35b-instruct",
10    messages=[
11        {"role": "system", "content": "You are a helpful assistant."},
12        {"role": "user", "content": "Hello, how are you?"}
13    ],
14    max_tokens=65536,
15    temperature=0.7
16)
17
18print(response.choices[0].message.content)

Información

Proveedor
Qwen
Cuantización
fp8

Funcionalidad compatible

Longitud del contexto
262144
Salida máxima
65536
Serverless
Compatible
Function Calling
Compatible
Structured Output
Compatible
API de Anthropic
Compatible
Capacidades de entrada
text
Capacidades de salida
text

Todo lo que necesitas para crear IA de producción.

Más de 200 modelos, GPUs bajo demanda y entornos de ejecución seguros para agentes, unificados bajo una API. Gratis para empezar, escala a medida que creces.