- Contexto
- 1M
- Entrada
- $1.32 /Mt· · Leitura de cache $0.132 /Mt
- Saída
- $3.96 /Mt
Explore os preços das nossas APIs de modelos e recursos de GPU. Encontre o plano certo para atender às suas necessidades, com tarifas transparentes e opções flexíveis.
A inferência em lote está disponível com um desconto introdutório de 50% nos tokens de entrada e saída para modelos compatíveis. Saiba mais
Advanced AI models from DeepSeek, offering cutting-edge reasoning capabilities and competitive pricing for enterprise and research applications.
Advanced AI models from DeepSeek, offering cutting-edge reasoning capabilities and competitive pricing for enterprise and research applications.
| Nome do modelo | Contexto | Entrada | Saída | Ações |
|---|---|---|---|---|
| DeepSeek V4 Pro 0813 | 1M | $1.32 /Mt· · Leitura de cache $0.132 /Mt | $3.96 /Mt | Mais |
| Deepseek V4 Flash 0731 | 1M | $0.44 /Mt· · Leitura de cache $0.028 /Mt | $1.32 /Mt | Mais |
| Deepseek V4 Flash | 1M | $0.14 /Mt· · Leitura de cache $0.028 /Mt | $0.28 /Mt | Mais |
| Deepseek V4 Pro | 1M | $1.6 /Mt· · Leitura de cache $0.135 /Mt | $3.2 /Mt | Mais |
| Deepseek V3.2 | 160K | $0.269 /Mt· · Leitura de cache $0.1345 /Mt | $0.4 /Mt | Mais |
| DeepSeek-OCR 2 | 8K | $0.03 /Mt | $0.03 /Mt | Mais |
| Deepseek V3.2 Exp | 160K | $0.27 /Mt | $0.41 /Mt | Mais |
| Deepseek V3.1 Terminus | 128K | $0.27 /Mt· · Leitura de cache $0.135 /Mt | $1 /Mt | Mais |
| DeepSeek V3.1 | 128K | $0.27 /Mt· · Leitura de cache $0.135 /Mt | $1 /Mt | Mais |
| DeepSeek V3 0324 | 160K | $0.27 /Mt· · Leitura de cache $0.135 /Mt | $1.12 /Mt | Mais |
| DeepSeek R1 0528 | 160K | $0.7 /Mt· · Leitura de cache $0.35 /Mt | $2.5 /Mt | Mais |
| DeepSeek R1 Distill LLama 70B | 8K | $0.8 /Mt | $0.8 /Mt | Mais |
| DeepSeek V3 (Turbo) | 63K | $0.4 /Mt | $1.3 /Mt | Mais |
| DeepSeek R1 (Turbo) | 63K | $0.7 /Mt | $2.5 /Mt | Mais |
Qwen series models offering efficient language processing with various parameter sizes, from lightweight to enterprise-grade solutions.
Qwen series models offering efficient language processing with various parameter sizes, from lightweight to enterprise-grade solutions.
Baidu's ERNIE models providing advanced Chinese language understanding and multimodal capabilities, optimized for Chinese applications with competitive pricing.
Baidu's ERNIE models providing advanced Chinese language understanding and multimodal capabilities, optimized for Chinese applications with competitive pricing.
| Nome do modelo | Contexto | Entrada | Saída | Ações |
|---|---|---|---|---|
| CoBuddy | 128K | $0.28 /Mt· · Leitura de cache $0.07 /Mt | $1.13 /Mt | Mais |
| ERNIE 4.5 VL 424B A47B | 120K | $0.42 /Mt | $1.25 /Mt | Mais |
| ERNIE 4.5 21B A3B | 117K | $0.07 /Mt | $0.28 /Mt | Mais |
GLM series models from Tsinghua University, featuring advanced Chinese language understanding and generation capabilities.
GLM series models from Tsinghua University, featuring advanced Chinese language understanding and generation capabilities.
| Nome do modelo | Contexto | Entrada | Saída | Ações |
|---|---|---|---|---|
| GLM 5.2 | 1M | $1.4 /Mt· · Leitura de cache $0.26 /Mt | $4.4 /Mt | Mais |
| GLM-5.1 | 200K | $1.38 /Mt· · Leitura de cache $0.26 /Mt | $4.4 /Mt | Mais |
| GLM-5 | 198K | $1 /Mt· · Leitura de cache $0.2 /Mt | $3.2 /Mt | Mais |
| GLM-4.7-Flash | 195K | $0.07 /Mt· · Leitura de cache $0.01 /Mt | $0.4 /Mt | Mais |
| GLM-4.7 | 200K | $0.6 /Mt· · Leitura de cache $0.11 /Mt | $2.2 /Mt | Mais |
| AutoGLM-Phone-9B-Multilingual | 64K | $0.035 /Mt | $0.138 /Mt | Mais |
| GLM 4.6V | 128K | $0.3 /Mt· · Leitura de cache $0.055 /Mt | $0.9 /Mt | Mais |
| GLM 4.6 | 200K | $0.55 /Mt· · Leitura de cache $0.11 /Mt | $2.2 /Mt | Mais |
| GLM 4.5V | 64K | $0.6 /Mt· · Leitura de cache $0.11 /Mt | $1.8 /Mt | Mais |
| zai-org/glm-4.5-air | 128K | $0.13 /Mt· · Leitura de cache $0.025 /Mt | $0.85 /Mt | Mais |
Specialized fine-tuned models optimized for creative and roleplay applications with enhanced storytelling capabilities.
Specialized fine-tuned models optimized for creative and roleplay applications with enhanced storytelling capabilities.
| Nome do modelo | Contexto | Entrada | Saída | Ações |
|---|---|---|---|---|
| Sao10k L3 8B Lunaris | 8K | $0.05 /Mt | $0.05 /Mt | Mais |
| L3 8B Stheno V3.2 | 8K | $0.05 /Mt | $0.05 /Mt | Mais |
| L31 70B Euryale V2.2 | 8K | $1.48 /Mt | $1.48 /Mt | Mais |
--
--
| Nome do modelo | Contexto | Entrada | Saída | Ações |
|---|---|---|---|---|
| Kimi K3 | 1M | $3 /Mt· · Leitura de cache $0.3 /Mt | $15 /Mt | Mais |
| Kimi K2.7 Code | 256K | $0.95 /Mt· · Leitura de cache $0.19 /Mt | $4 /Mt | Mais |
| Kimi K2.6 | 256K | $0.8 /Mt· · Leitura de cache $0.16 /Mt | $3.4 /Mt | Mais |
| Kimi K2.5 | 256K | $0.6 /Mt· · Leitura de cache $0.1 /Mt | $3 /Mt | Mais |
| Kimi K2 Thinking | 256K | $0.6 /Mt· · Leitura de cache $0.15 /Mt | $2.5 /Mt | Mais |
| Kimi K2 0905 | 256K | $0.6 /Mt | $2.5 /Mt | Mais |
| Kimi K2 Instruct | 128K | $0.57 /Mt | $2.3 /Mt | Mais |
--
--
| Nome do modelo | Contexto | Entrada | Saída | Ações |
|---|---|---|---|---|
| Macaron V1 Venti | 1M | $1.5 /Mt· · Leitura de cache $0.3 /Mt | $4.5 /Mt | Mais |
| Macaron V1 Tall | 256K | $0.45 /Mt· · Leitura de cache $0.08 /Mt | $2.6 /Mt | Mais |
Minimax AI's advanced language models delivering robust conversational AI capabilities with optimized performance for customer service, content generation, and creative applications, featuring strong multilingual support and enterprise-ready scalability.
Minimax AI's advanced language models delivering robust conversational AI capabilities with optimized performance for customer service, content generation, and creative applications, featuring strong multilingual support and enterprise-ready scalability.
| Nome do modelo | Contexto | Entrada | Saída | Ações |
|---|---|---|---|---|
| MiniMax M3 | 977K | - | Mais | |
| MiniMax M2.7 | 200K | $0.3 /Mt· · Leitura de cache $0.06 /Mt | $1.2 /Mt | Mais |
| MiniMax M2.5-highspeed | 200K | $0.6 /Mt· · Leitura de cache $0.03 /Mt | $2.4 /Mt | Mais |
| MiniMax M2.5 | 200K | $0.3 /Mt· · Leitura de cache $0.03 /Mt | $1.2 /Mt | Mais |
| Minimax M2.1 | 200K | $0.3 /Mt· · Leitura de cache $0.03 /Mt | $1.2 /Mt | Mais |
| MiniMax-M2 | 200K | $0.3 /Mt· · Leitura de cache $0.03 /Mt | $1.2 /Mt | Mais |
| MiniMax M1 | 977K | $0.55 /Mt | $2.2 /Mt | Mais |
--
--
| Nome do modelo | Contexto | Entrada | Saída | Ações |
|---|---|---|---|---|
| Ling 3.0 Flash Fast | 256K | $0.06 /Mt· · Leitura de cache $0.012 /Mt | $0.18 /Mt | Mais |
| Ling 3.0 Flash | 256K | $0.06 /Mt· · Leitura de cache $0.012 /Mt | $0.18 /Mt | Mais |
| Ling-2.6-flash | 256K | $0.1 /Mt· · Leitura de cache $0.02 /Mt | $0.3 /Mt | Mais |
| Ling-2.6-1T | 256K | $0.3 /Mt· · Leitura de cache $0.06 /Mt | $2.5 /Mt | Mais |
--
--
| Nome do modelo | Contexto | Entrada | Saída | Ações |
|---|---|---|---|---|
| Step 3.7 Flash | 256K | $0.2 /Mt· · Leitura de cache $0.04 /Mt | $1.15 /Mt | Mais |
--
--
| Nome do modelo | Contexto | Entrada | Saída | Ações |
|---|---|---|---|---|
| Nemotron 3 Nano 30B A3B | 256K | $0.05 /Mt | $0.2 /Mt | Mais |
Google's Gemma models offering high-quality language processing with excellent performance for various NLP tasks.
Google's Gemma models offering high-quality language processing with excellent performance for various NLP tasks.
| Nome do modelo | Contexto | Entrada | Saída | Ações |
|---|---|---|---|---|
| Gemma 4 26B A4B | 256K | $0.13 /Mt | $0.4 /Mt | Mais |
| Gemma 4 31B | 256K | $0.14 /Mt | $0.4 /Mt | Mais |
| Gemma 3 27B | 96K | $0.119 /Mt | $0.2 /Mt | Mais |
--
--
| Nome do modelo | Contexto | Entrada | Saída | Ações |
|---|---|---|---|---|
| Kat Coder Pro | 250K | $0.3 /Mt· · Leitura de cache $0.06 /Mt | $1.2 /Mt | Mais |
--
--
| Nome do modelo | Contexto | Entrada | Saída | Ações |
|---|---|---|---|---|
| OpenAI GPT OSS 120B | 128K | $0.05 /Mt | $0.25 /Mt | Mais |
| OpenAI: GPT OSS 20B | 128K | $0.04 /Mt | $0.15 /Mt | Mais |
Meta's Llama models providing state-of-the-art language understanding with open architecture designed for diverse applications.
Meta's Llama models providing state-of-the-art language understanding with open architecture designed for diverse applications.
| Nome do modelo | Contexto | Entrada | Saída | Ações |
|---|---|---|---|---|
| Llama 3.1 8B Instruct | 16K | $0.02 /Mt | $0.05 /Mt | Mais |
| Llama 3.3 70B Instruct | 12K | $0.135 /Mt | $0.4 /Mt | Mais |
| Llama 4 Maverick Instruct | 1M | $0.27 /Mt | $0.85 /Mt | Mais |
| Llama 4 Scout Instruct | 128K | $0.18 /Mt | $0.59 /Mt | Mais |
--
--
| Nome do modelo | Contexto | Entrada | Saída | Ações |
|---|---|---|---|---|
| Mistral Nemo | 59K | $0.04 /Mt | $0.17 /Mt | Mais |
--
--
| Nome do modelo | Contexto | Entrada | Saída | Ações |
|---|---|---|---|---|
| XiaomiMiMo/MiMo-V2.5 | 1M | $0.168 /Mt· · Leitura de cache $0.0034 /Mt | $0.336 /Mt | Mais |
| XiaomiMiMo/MiMo-V2.5-Pro | 1M | $0.522 /Mt· · Leitura de cache $0.0043 /Mt | $1.044 /Mt | Mais |
| Wizardlm 2 8x22B | 64K | $0.62 /Mt | $0.62 /Mt | Mais |
| Ring-2.6-1T | 256K | $0.3 /Mt· · Leitura de cache $0.06 /Mt | $2.5 /Mt | Mais |
| Nome do modelo | Contexto | Entrada |
|---|---|---|
qwen/qwen3-embedding-0.6b | 32K | $0.07 /Mt |
Qwen3 Embedding 8B | 32K | $0.07 /Mt |
BAAI:BGE-M3 | 8K | $0.01 /Mt |
Pricing may vary based on image dimensions, inference steps, and upscaling factors. Use theCalculadora de preçosfor an estimate.
| Nome da API | Modo | Largura&Altura | Preços |
|---|---|---|---|
Flux.1 Kontext Dev | - | - | $0.0225 /image |
| fast_mode | - | $0.018 /image | |
Flux.1 Kontext Max | - | - | $0.072 /image |
Flux.1 Kontext Pro | - | - | $0.36 /image |
Qwen-Image Edit | - | - | $0.02 /image |
Qwen-Image Text to Image | - | - | $0.02 /image |
Pricing may vary based on the number of frames, chosen model, and inference steps. Use theCalculadora de preçosfor an estimate.
| Nome da API | Modo | Duração | Resolução | Preços |
|---|---|---|---|---|
Kling v3.0 Pro Image-to-Video | No Audio | - | $0.112 /s | |
| Audio | - | $0.168 /s | ||
Kling v3.0 Pro Text-to-Video | No Audio | - | $0.112 /s | |
| Audio | - | $0.168 /s | ||
Kling v3.0 Standard Image-to-Video | No Audio | - | $0.084 /s | |
| Audio | - | $0.126 /s | ||
Kling v3.0 Standard Text-to-Video | No Audio | - | $0.084 /s | |
| Audio | - | $0.126 /s | ||
Minimax Hailuo 2.3 Fast Image to Video | - | 6s | 768P | $0.19 /video |
| - | 10s | 768P | $0.32 /video | |
| - | 6s | 1080P | $0.33 /video | |
Minimax Hailuo 2.3 Image to Video | - | 6s | 768P | $0.28 /video |
| - | 10s | 768P | $0.56 /video | |
| - | 6s | 1080P | $0.49 /video | |
Minimax Hailuo 2.3 Text to Video | - | 6s | 768P | $0.28 /video |
| - | 10s | 768P | $0.56 /video | |
| - | 6s | 1080P | $0.49 /video | |
Wan 2.5 Image to Video | - | 5s | 480P | $0.25 /video |
| - | 10s | 480P | $0.50 /video | |
| - | 5s | 720P | $0.50 /video | |
| - | 10s | 720P | $1.00 /video | |
| - | 5s | 1080P | $0.75 /video | |
| - | 10s | 1080P | $1.50 /video | |
Wan 2.5 Text to Video | - | 5s | 480P | $0.25 /video |
| - | 10s | 480P | $0.50 /video | |
| - | 5s | 720P | $0.50 /video | |
| - | 10s | 720P | $1.00 /video | |
| - | 5s | 1080P | $0.75 /video | |
| - | 10s | 1080P | $1.50 /video | |
Wan 2.6 Image to Video | - | 5s | 720P | $0.50 /video |
| - | 10s | 720P | $1.00 /video | |
| - | 15s | 720P | $1.50 /video | |
| - | 5s | 1080P | $0.75 /video | |
| - | 10s | 1080P | $1.50 /video | |
| - | 15s | 1080P | $2.25 /video | |
Wan 2.6 Reference to Video | - | 5s | 720P | $0.50 /video |
| - | 10s | 720P | $1.00 /video | |
| - | 5s | 1080P | $0.75 /video | |
| - | 10s | 1080P | $1.50 /video | |
Wan 2.6 Text to Video | - | 5s | 720P | $0.50 /video |
| - | 10s | 720P | $1.00 /video | |
| - | 15s | 720P | $1.50 /video | |
| - | 5s | 1080P | $0.75 /video | |
| - | 10s | 1080P | $1.50 /video | |
| - | 15s | 1080P | $2.25 /video |
| Nome da API | Modelo | Etapas | Preços |
|---|---|---|---|
Image to Video | SVD-XT | 20 | $0.024 /video |
| SVD | 20 | $0.0134 /video |
| Nome da API | Modo | Preços |
|---|---|---|
Fish Audio Text to Speech | - | $15 /1M characters |
Fish Audio Voice Cloning | - | $0.1 /voice |
MiniMax speech-2.6-hd | T2A / T2A Async | $100 /1M characters |
MiniMax speech-2.6-turbo | T2A / T2A Async | $60 /1M characters |
MiniMax Voice-Cloning | - | $1.5 /voice |
Text to Speech | - | $15 /1M characters |
| API Name | Mode | Pricing |
|---|---|---|
| EXA | neuralSearch | $0.007/request |
| deepSearch | $0.012/request | |
| deepReasoningSearch | $0.015/request | |
| additional_result | $0.001/result (beyond 10 results) | |
| answer | $0.005/request | |
| contentText | $0.001/item | |
| contentHighlight | $0.001/item | |
| contentSummary | $0.001/item | |
| Tavily | basicSearch | $0.008/request |
| advancedSearch | $0.016/request | |
| basicExtract | $0.0016/url | |
| advancedExtract | $0.0032/url | |
| regularMapping | $0.0008/page | |
| instructedMapping | $0.0016/page | |
| Crawl | Extract + Mapping Cost |
Mais de 200 modelos, GPUs sob demanda e ambientes de execução de agentes seguros — unificados em uma única API. Grátis para começar, escala conforme você cresce.