The high-speed variant of Z.ai's GLM-5.3-Flash, a natively multimodal model delivering inference speeds of up to 200 tokens per second. Built on the same hybrid sparse and linear attention architecture, with image and video input, reasoning, tool use, and a 1M-token context window.
Сортировка по общей стоимости входных и выходных токенов за 1 млн. Выберите строку, чтобы открыть страницу провайдера.
| Провайдер | Цена за 1 млн | Лимиты | Регионы | Состояние | Задержка |
|---|---|---|---|---|---|
Вход: $0.37Выход: $1.25 | 60 RPM / 200K TPM | us-east-1 | Работает | 0ms |
Используйте эту модель через OpenRouter с OpenAI-совместимым SDK.
import OpenAI from "openai";
const client = new OpenAI({
baseURL: "https://openrouter.ai/api/v1",
apiKey: process.env.OPENROUTER_API_KEY,
});
const response = await client.chat.completions.create({
model: "z-ai/glm-5.3-flashx",
messages: [
{ role: "user", content: "Hello!" }
],
});
console.log(response.choices[0].message.content);API OpenRouter · OpenAI-совместимый SDK
Все зафиксированные цены на модель по провайдерам. Цены указаны за 1 млн токенов в USD.
С 1 окт. 2026 г. изменений цены не зафиксировано.
| Дата | Провайдер | Вход | Выход |
|---|---|---|---|
| 1 окт. 2026 г.Добавлена · Текущая | OpenRouter | $0.37 | $1.25 |