Model Library/GLM 4.6V
zai-org/glm-4.6v

GLM 4.6V

zai-org/glm-4.6v
GLM-4.6V represents a significant multimodal advancement in the GLM series, featuring a 128k-token training context window and achieving state-of-the-art visual understanding accuracy for models of its parameter scale. Notably, it's the first visual model to natively integrate Function Call capabilities directly into its architecture, creating a seamless pathway from visual perception to executable actions. This breakthrough establishes a unified technical foundation for deploying multimodal agents in real-world business applications.

Fonctionnalités

API sans serveur

Documentation

zai-org/glm-4.6v is available via Novita's serverless API, where you pay per token. There are several ways to call the API, including OpenAI-compatible endpoints with exceptional reasoning performance.

Sans serveur disponible

Exécutez des requêtes immédiatement, ne payez que pour l’utilisation

Entrée$0.3 / M Tokens
Lecture du cache$0.055 / M Tokens
Sortie$0.9 / M Tokens

Utilisez les exemples de code suivants pour intégrer notre API :

1from openai import OpenAI
2
3client = OpenAI(
4    api_key="<Your API Key>",
5    base_url="https://api.novita.ai/openai"
6)
7
8response = client.chat.completions.create(
9    model="zai-org/glm-4.6v",
10    messages=[
11        {"role": "system", "content": "You are a helpful assistant."},
12        {"role": "user", "content": "Hello, how are you?"}
13    ],
14    max_tokens=32768,
15    temperature=0.7
16)
17
18print(response.choices[0].message.content)

Infos

Fournisseur
Zai-org
Quantification
bf16

Fonctionnalités prises en charge

Longueur du contexte
131072
Sortie maximale
32768
Serverless
Pris en charge
Function Calling
Pris en charge
Structured Output
Pris en charge
Reasoning
Pris en charge
API Anthropic
Pris en charge
Capacités d’entrée
text, video, image
Capacités de sortie
text