Model Library/GLM 4.5V
GLM 4.5V

GLM 4.5V

zai-org/glm-4.5v
Z.ai's GLM-4.5V sets a new standard in visual reasoning, achieving SOTA performance across 42 benchmarks among open-source models. Beyond benchmarks, it excels in real-world applications through hybrid training, enabling comprehensive visual understanding—from image/video analysis and GUI interaction to complex document processing and precise visual grounding. In China's GeoGuessr challenge, GLM-4.5V surpassed 99% of 21,000 human players within 16 hours, reaching 66th place in a week. Built on the GLM-4.5-Air foundation and inheriting GLM-4.1V-Thinking's approach, it leverages a 106B-parameter MoE architecture for scalable, efficient performance. This model bridges advanced AI research with practical deployment, delivering unmatched visual intelligence

Características

API serverless

Documentación

zai-org/glm-4.5v is available via Novita's serverless API, where you pay per token. There are several ways to call the API, including OpenAI-compatible endpoints with exceptional reasoning performance.

Implementaciones bajo demanda

Documentación

On-demand deployments allow you to use zai-org/glm-4.5v on dedicated GPUs with high-performance serving stack with high reliability and no rate limits.

Serverless disponible

Ejecuta consultas de inmediato, paga solo por el uso

Entrada$0.6 / M Tokens
Lectura de caché$0.11 / M Tokens
Salida$1.8 / M Tokens

Usa los siguientes ejemplos de código para integrarte con nuestra API:

1from openai import OpenAI
2
3client = OpenAI(
4    api_key="<Your API Key>",
5    base_url="https://api.novita.ai/openai"
6)
7
8response = client.chat.completions.create(
9    model="zai-org/glm-4.5v",
10    messages=[
11        {"role": "system", "content": "You are a helpful assistant."},
12        {"role": "user", "content": "Hello, how are you?"}
13    ],
14    max_tokens=16384,
15    temperature=0.7
16)
17
18print(response.choices[0].message.content)

Información

Proveedor
Z.ai
Cuantización
fp8

Funcionalidad compatible

Longitud del contexto
64K
Salida máxima
16K
Serverless
Compatible
Function Calling
Compatible
Structured Output
Compatible
Reasoning
Compatible
Capacidades de entrada
text, video, image
Capacidades de salida
text

Todo lo que necesitas para crear IA de producción.

Más de 200 modelos, GPUs bajo demanda y entornos de ejecución seguros para agentes, unificados bajo una API. Gratis para empezar, escala a medida que creces.