Model Library/MiMo V2.6 Flash
MiMo V2.6 Flash

MiMo V2.6 Flash

xiaomimimo/mimo-v2.6-flash
MiMo V2.6 Flash is Xiaomi's cost-efficient open-source reasoning model for high-frequency calls and large-scale professional workflows. It combines native text, image, video, and audio understanding with a 1M-token context window and up to 128K output tokens, while supporting controllable deep thinking, tool calling, JSON mode, streaming, and prompt caching.

機能

サーバーレス API

ドキュメント

xiaomimimo/mimo-v2.6-flash is available via Novita's serverless API, where you pay per token. There are several ways to call the API, including OpenAI-compatible endpoints with exceptional reasoning performance.

利用可能なサーバーレス

クエリをすぐに実行し、使用した分だけお支払い

入力$0.14 / M Tokens
キャッシュ読み取り$0.0028 / M Tokens
出力$0.28 / M Tokens

以下のコード例を使用して、当社の API と統合してください:

1from openai import OpenAI
2
3client = OpenAI(
4    api_key="<Your API Key>",
5    base_url="https://api.novita.ai/openai"
6)
7
8response = client.chat.completions.create(
9    model="xiaomimimo/mimo-v2.6-flash",
10    messages=[
11        {"role": "system", "content": "You are a helpful assistant."},
12        {"role": "user", "content": "Hello, how are you?"}
13    ],
14    max_tokens=131072,
15    temperature=0.7
16)
17
18print(response.choices[0].message.content)

情報

プロバイダー
Xiaomi
量子化
fp8

サポートされている機能

コンテキスト長
1M
最大出力
128K
Serverless
サポートされています
Function Calling
サポートされています
Structured Output
サポートされています
Reasoning
サポートされています
Anthropic API
サポートされています
入力機能
text, image, video, audio
出力機能
text

本番環境向けAIを構築するために必要なすべて。

200以上のモデル、オンデマンド GPUs、安全なエージェントランタイムを、1つの API に統合。無料で始められ、成長に合わせてスケールできます。