Модели

11 канонических LLM-моделей от всех провайдеров

Некоторые описания переведены автоматически в рамках пилота и пока не проверены редактором.

Показаны модели 1–11 из 11

GPT-5.6 Sol

Соединенные Штаты

Флагманская модель OpenAI серии GPT-5.6, развивающая программирование, научное рассуждение, долгосрочное планирование и агентные рабочие процессы и одновременно повышающая надёжность и эффективность на сложных практических задачах. Добавляет максимальный уровень глубины рассуждения и режим ultra, который запускает субагентов для сложной многоэтапной работы.

Контекст
1.0M
Добавлена
июль 2026 г.

Nemotron 3 Ultra

Соединенные Штаты

NVIDIA's flagship open 550B-parameter Mixture-of-Experts model with 55B active parameters, built for frontier reasoning and orchestration in long-running agentic systems. Features hybrid Mamba-Transformer architecture, LatentMoE routing, multi-token prediction, and NVFP4 precision for 5x higher throughput. Achieves 30% lower cost-to-task-completion on agentic benchmarks. Supports 1M+ token context window with 95% accuracy on Ruler@1M.

Контекст
1.0M
Добавлена
июнь 2026 г.

Gemini 3.5 Flash

Соединенные Штаты

Google DeepMind's balanced Gemini 3.5 model that pairs Pro-line reasoning quality with Flash-line latency and cost. Natively multimodal across text, image, audio, and video with a 1M-token context window, configurable thinking levels, and streaming function calling, tuned for high-throughput production workloads.

Контекст
1.0M
Добавлена
май 2026 г.

Gemini 3 Flash

Соединенные Штаты

Google's balanced model combining Gemini 3 Pro's reasoning capabilities with the Flash line's latency, efficiency, and cost. Features configurable thinking levels, multimodal function responses, and streaming function calling for complex agentic workflows.

Контекст
1.0M
Добавлена
май 2026 г.

Gemini 3.1 Flash-Lite

Соединенные Штаты

Google's most cost-efficient Gemini model optimized for high-volume, low-latency use cases. Delivers 2.5x faster time to first token versus Gemini 2.5 Flash with full multimodal support. Ideal for agentic tasks, data extraction, translation, and classification.

Контекст
1.0M
Добавлена
май 2026 г.

GPT-5.4

Соединенные Штаты

OpenAI's frontier reasoning model combining advances in coding, reasoning, and agentic workflows. Features 1.1M token context window and strong performance on complex multi-step problems.

Контекст
1.1M
Добавлена
март 2026 г.

Gemini 3.1 Pro

Соединенные Штаты

Новейшая флагманская мультимодальная модель Google с передовыми результатами в рассуждении, программировании и понимании разных типов данных. Поддерживает встроенное использование инструментов, привязку к источникам и контекстное окно на миллион токенов.

Контекст
2.0M
Добавлена
март 2026 г.

Gemini 2.5 Flash

Соединенные Штаты

Google's cost-effective model optimized for high throughput tasks. Balances speed and intelligence with strong multimodal capabilities and 1M token context window.

Контекст
1.0M
Добавлена
июнь 2025 г.

Gemini 2.5 Pro

Соединенные Штаты

Google's high-capability reasoning model with adaptive thinking for complex agentic and multimodal challenges. Features 1M token context window and strong performance on coding and scientific tasks.

Контекст
1.0M
Добавлена
июнь 2025 г.

Llama 4 Maverick

Соединенные Штаты

Meta's quality-focused MoE model with 17B active parameters (400B total, 128 experts). Targets quality-critical tasks with benchmark scores competitive with GPT-4o and Gemini 2.5 Pro.

Контекст
1.0M
Добавлена
апр. 2025 г.

Llama 4 Scout

Соединенные Штаты

Meta's efficient MoE model with 17B active parameters (109B total, 16 experts). Supports up to 10M token context — the longest of any production model. Strong performance on reasoning and multilingual tasks.

Контекст
10.0M
Добавлена
апр. 2025 г.