Model Library/Macaron V1 Venti
Mind Lab

Macaron V1 Venti

mindai/macaron-v1-venti
Macaron-V1-Venti is a 748B-parameter flagship model in the Macaron-V1 family, built for personal intelligence, tool use, coding workflows, and code-native Generative UI. The model uses a Mixture of LoRA (MoL) architecture on top of GLM-5.2, consisting of a 744B-parameter base model and four 1B-parameter LoRA specialists. The specialists cover chat, personal-agent tasks, coding, and GenUI, with an L0 router selecting the most suitable specialist for each new user request.

Funktionen

Serverless API

Dokumentation

mindai/macaron-v1-venti is available via Novita's serverless API, where you pay per token. There are several ways to call the API, including OpenAI-compatible endpoints with exceptional reasoning performance.

Verfügbare Serverless

Abfragen sofort ausführen, nur für die Nutzung bezahlen

Eingabe$0 / M Tokens
Ausgabe$0 / M Tokens

Verwenden Sie die folgenden Codebeispiele, um unsere API zu integrieren:

1from openai import OpenAI
2
3client = OpenAI(
4    api_key="<Your API Key>",
5    base_url="https://api.novita.ai/openai"
6)
7
8response = client.chat.completions.create(
9    model="mindai/macaron-v1-venti",
10    messages=[
11        {"role": "system", "content": "You are a helpful assistant."},
12        {"role": "user", "content": "Hello, how are you?"}
13    ],
14    max_tokens=131072,
15    temperature=0.7
16)
17
18print(response.choices[0].message.content)

Info

Anbieter
Mind Lab
Quantisierung
-

Unterstützte Funktionalität

Kontextlänge
1048576
Maximale Ausgabe
131072
Serverless
Unterstützt
Function Calling
Unterstützt
Reasoning
Unterstützt
Anthropic API
Unterstützt
Eingabefähigkeiten
text
Ausgabefähigkeiten
text

Alles, was Sie brauchen, um produktionsreife AI zu entwickeln.

Über 200 Modelle, GPUs auf Abruf und sichere Agent-Runtimes — vereint unter einer API. Kostenlos zum Einstieg, skaliert mit Ihrem Wachstum.