Модели

4 канонические LLM-модели от всех провайдеров

Показаны модели 1–4 из 4

GLM-5.2

Китай

Z.ai's (formerly Zhipu AI) flagship open-weight coding model with a 1M-token context window. Mixture-of-Experts architecture with 753B total parameters and ~40B active per request, featuring two cost-balancing reasoning modes. Tops several coding benchmarks while remaining a fraction of the cost of comparable proprietary frontier models. MIT-licensed weights.

Контекст
1.0M
Добавлена
июнь 2026 г.

DeepSeek V4 Flash

Китай

DeepSeek's efficient V4 model with 284B total parameters (13B activated). Optimized for speed and cost-efficiency while maintaining strong performance. Supports 1M token context window.

Контекст
1.0M
Добавлена
апр. 2026 г.

DeepSeek V4 Pro

Китай

DeepSeek's flagship V4 model with 1.6T total parameters (49B activated). MoE architecture supporting 1M token context. Closes the gap with frontier proprietary models on reasoning and coding benchmarks.

Контекст
1.0M
Добавлена
апр. 2026 г.

MiMo-V2.5-Pro

Китай

Xiaomi's flagship 1.02T-parameter Mixture-of-Experts model with 42B active parameters, built on a hybrid-attention architecture with 3-layer Multi-Token Prediction. Designed for complex agentic tasks, software engineering, and long-horizon instruction following with a 1M-token context window.

Контекст
1.0M
Добавлена
апр. 2026 г.