atrás

GLM 5.2 FP8

glm-5.2-fp8 · Z.ai

754B FP8 en 8× B200. Generalista y coding con tool calling.

Servido en GPU NVIDIA B200 · España (UE)

Fuente del modelo

Hugging Facezai-org/GLM-5.2-FP8

Especificaciones

Parámetros
743B · 39B active MoE
Hardware
8× B200
Contexto
202.752 tokens
Cuantización
FP8
Modalidades
texto
Capacidades
chat · coding · tools

Precio

entrada

1,00

salida

3,00

lectura de caché

0,19

€ / millón de tokens · IVA aparte

Ejemplo (curl)

Compatible con el SDK de OpenAI. Cambia la URL base y la API key.

curl https://api.gpuflow.ai/v1/chat/completions \
  -H "Authorization: Bearer $GPUFLOW_API_KEY" \
  -d '{ "model": "glm-5.2-fp8",
        "messages": [{ "role": "user", "content": "Hola" }] }'

Pruébalo con 5 € gratis

Crea tu cuenta, apunta tu SDK a GPU Flow y llama a este modelo en GPU NVIDIA B200, en España.