Модели
19 канонических LLM-моделей от всех провайдеров
Некоторые описания переведены автоматически в рамках пилота и пока не проверены редактором.
Флагманская модель Gemini от Google DeepMind, заново построенная на новой основе с контекстным окном 2 млн токенов и режимом рассуждения Deep Think для самых сложных задач по математике, программированию и мультимодальной обработке. Нативно работает с текстом, изображениями, аудио и видео, поддерживает потоковый вызов функций и надёжную работу с длинным контекстом.
Сбалансированная модель серии GPT-5.6 от OpenAI: немного уступает по максимальному качеству, но обеспечивает заметно меньшую задержку и стоимость. Сохраняет сильные возможности рассуждения, программирования и агентного использования инструментов с настраиваемой глубиной рассуждения, поэтому подходит как основной вариант для масштабных производственных нагрузок, которым нужны передовые возможности.
Флагманская модель OpenAI серии GPT-5.6, развивающая программирование, научное рассуждение, долгосрочное планирование и агентные рабочие процессы и одновременно повышающая надёжность и эффективность на сложных практических задачах. Добавляет максимальный уровень глубины рассуждения и режим ultra, который запускает субагентов для сложной многоэтапной работы.
Самая мощная модель Anthropic класса Sonnet, переносящая передовые возможности программирования, агентной и профессиональной работы в средний сегмент и сокращающая разрыв с Opus 4.8 при меньшей цене. Поддерживает адаптивное мышление с выбором глубины рассуждения, контекстное окно 1 млн токенов и входные данные в виде текста, изображений и файлов. Кодовое имя — Fennec.
MiniMax's frontier open-weight model with 1M-token context window, native multimodality (text, image, video), and strong coding capabilities. Built on MiniMax Sparse Attention (MSA) architecture, achieving 59% on SWE-Bench Pro with significantly improved efficiency at long context.
Writer's most advanced adaptive reasoning model with a 1 million token context window. Processes full million-token prompts in approximately 22 seconds with multi-turn function calls in 300ms. Optimized for enterprise agentic AI workflows at 3-4x lower cost than GPT-4.1.
Google DeepMind's balanced Gemini 3.5 model that pairs Pro-line reasoning quality with Flash-line latency and cost. Natively multimodal across text, image, audio, and video with a 1M-token context window, configurable thinking levels, and streaming function calling, tuned for high-throughput production workloads.
Google's balanced model combining Gemini 3 Pro's reasoning capabilities with the Flash line's latency, efficiency, and cost. Features configurable thinking levels, multimodal function responses, and streaming function calling for complex agentic workflows.
Google's most cost-efficient Gemini model optimized for high-volume, low-latency use cases. Delivers 2.5x faster time to first token versus Gemini 2.5 Flash with full multimodal support. Ideal for agentic tasks, data extraction, translation, and classification.
OpenAI's most capable model designed for complex real-world work including coding, online research, information analysis, and document creation. Features advanced agentic capabilities with tool search and multi-step task execution.
OpenAI's compact reasoning model optimized for coding, computer use, and subagent tasks. Approaches GPT-5.4 performance on several benchmarks while running more than 2x faster.
xAI's latest and most intelligent model with strong agentic tool calling, minimal hallucinations, and configurable reasoning. Supports 1M token context window with competitive pricing.
OpenAI's frontier reasoning model combining advances in coding, reasoning, and agentic workflows. Features 1.1M token context window and strong performance on complex multi-step problems.
Moonshot AI's latest model with ultra-long context window support, strong reasoning capabilities, and excellent performance on complex multi-step tasks. Known for reliable long-document understanding.
Новейшая флагманская мультимодальная модель Google с передовыми результатами в рассуждении, программировании и понимании разных типов данных. Поддерживает встроенное использование инструментов, привязку к источникам и контекстное окно на миллион токенов.
xAI's multi-agent capable model with 2M token context window. Available in reasoning, non-reasoning, and multi-agent variants for diverse enterprise workloads.
xAI's fast and cost-effective model with 2M token context window. Offers both reasoning and non-reasoning modes at significantly lower pricing than flagship models.
Google's cost-effective model optimized for high throughput tasks. Balances speed and intelligence with strong multimodal capabilities and 1M token context window.
Google's high-capability reasoning model with adaptive thinking for complex agentic and multimodal challenges. Features 1M token context window and strong performance on coding and scientific tasks.