Модели

46 канонических LLM-моделей от всех провайдеров

Некоторые описания переведены автоматически в рамках пилота и пока не проверены редактором.

Показаны модели 1–24 из 46

Muse Spark 1.1512K ctx

Meta Superintelligence Labs' updated flagship, building on Muse Spark with stronger agentic reasoning, more reliable multi-agent orchestration, and improved multimodal understanding across voice, text, and image. Extends the context window and reduces latency and reasoning token usage while raising coding and tool-use accuracy. Powers Meta AI across its product ecosystem.

Gemini 3.5 Pro2.0M ctx

Флагманская модель Gemini от Google DeepMind, заново построенная на новой основе с контекстным окном 2 млн токенов и режимом рассуждения Deep Think для самых сложных задач по математике, программированию и мультимодальной обработке. Нативно работает с текстом, изображениями, аудио и видео, поддерживает потоковый вызов функций и надёжную работу с длинным контекстом.

GPT-5.6 Luna400K ctx

The fast, low-cost tier of OpenAI's GPT-5.6 series, optimized for high-volume, latency-sensitive tasks such as classification, extraction, routing, and lightweight agentic steps. Approaches the larger GPT-5.6 tiers on many benchmarks while running several times faster at a fraction of the price.

GPT-5.6 Sol1.0M ctx

Флагманская модель OpenAI серии GPT-5.6, развивающая программирование, научное рассуждение, долгосрочное планирование и агентные рабочие процессы и одновременно повышающая надёжность и эффективность на сложных практических задачах. Добавляет максимальный уровень глубины рассуждения и режим ultra, который запускает субагентов для сложной многоэтапной работы.

GPT-5.6 Terra1.0M ctx

Сбалансированная модель серии GPT-5.6 от OpenAI: немного уступает по максимальному качеству, но обеспечивает заметно меньшую задержку и стоимость. Сохраняет сильные возможности рассуждения, программирования и агентного использования инструментов с настраиваемой глубиной рассуждения, поэтому подходит как основной вариант для масштабных производственных нагрузок, которым нужны передовые возможности.

Grok 4.5500K ctx

Самая мощная на текущий момент модель xAI, созданная для программирования, агентных задач и интеллектуальной работы и разрабатывавшаяся совместно с инструментами для реальной разработки ПО. Поддерживает доступ к актуальной информации, расширенное рассуждение и работу с инструментами в большом контексте через API, совместимый с OpenAI.

Claude Sonnet 51.0M ctx

Самая мощная модель Anthropic класса Sonnet, переносящая передовые возможности программирования, агентной и профессиональной работы в средний сегмент и сокращающая разрыв с Opus 4.8 при меньшей цене. Поддерживает адаптивное мышление с выбором глубины рассуждения, контекстное окно 1 млн токенов и входные данные в виде текста, изображений и файлов. Кодовое имя — Fennec.

Sakana Fugu256K ctx

Sakana AI's multi-agent orchestration model from Tokyo, delivered as a single OpenAI-compatible API. Fugu is itself a language model trained to call a pool of specialist LLMs (and recursive instances of itself), handling model selection, delegation, verification, and synthesis behind one endpoint. Built on Sakana AI's TRINITY and Conductor research, its routing intelligence is learned in model weights rather than hand-configured.

Sakana Fugu Ultra256K ctx

The higher-quality tier of Sakana AI's Fugu multi-agent orchestration system, tuned for the hardest coding, reasoning, science, and agentic tasks. Coordinates a swappable pool of frontier LLMs through one OpenAI-compatible endpoint, delegating sub-tasks, verifying intermediate work, and synthesizing a single answer. Sakana reports strong vendor benchmarks including 93.2 on LiveCodeBench, 73.7 on SWE-Bench Pro, and 82.1 on TerminalBench.

Claude Mythos 5300K ctx

Anthropic's frontier Mythos-class model — the same underlying model as Claude Fable 5 but with safeguards lifted in some areas. It has the strongest cybersecurity capabilities of any model in the world, alongside state-of-the-art performance in software engineering, knowledge work, vision, and scientific research. Access is restricted to a small group of trusted cyberdefenders and infrastructure providers through Project Glasswing.

Claude Fable 5300K ctx

Первая общедоступная модель Anthropic класса Mythos, превосходящая по возможностям все модели, которые компания ранее выпускала для широкого доступа. Показывает передовые результаты почти во всех протестированных бенчмарках и особенно сильна в разработке ПО, интеллектуальной работе, обработке изображений и научных исследованиях. Её преимущество растёт на более длинных и сложных задачах. Встроенные меры защиты направляют чувствительные запросы по кибербезопасности, биологии, химии и дистилляции в Claude Opus 4.8.

MiniMax M31.0M ctx

MiniMax's frontier open-weight model with 1M-token context window, native multimodality (text, image, video), and strong coding capabilities. Built on MiniMax Sparse Attention (MSA) architecture, achieving 59% on SWE-Bench Pro with significantly improved efficiency at long context.

Claude Opus 4.8300K ctx

Самая продвинутая модель Anthropic, развивающая Opus 4.7 и улучшающая результаты в программировании, агентных навыках, рассуждении и интеллектуальной работе. Отличается повышенной честностью ответов, более эффективным использованием инструментов, поддержкой динамических рабочих процессов и улучшенным выравниванием.

Palmyra X51.0M ctx

Writer's most advanced adaptive reasoning model with a 1 million token context window. Processes full million-token prompts in approximately 22 seconds with multi-turn function calls in 300ms. Optimized for enterprise agentic AI workflows at 3-4x lower cost than GPT-4.1.

Solar Pro 3128K ctx

Upstage's powerful Mixture-of-Experts language model with 102B total parameters and 12B active parameters per forward pass. Optimized for Korean with strong English and Japanese support. Excels at complex reasoning, structured output generation, and agentic workflows.

Qwen 3.7 Max131K ctx

Alibaba's flagship proprietary model engineered for advanced agentic coding, complex reasoning, and long-horizon task execution. Ranked

Qwen 3.7 Plus131K ctx

Alibaba's multimodal variant in the Qwen 3.7 family, optimized for vision understanding and multimodal tasks. Ranked

Gemini 3.5 Flash1.0M ctx

Google DeepMind's balanced Gemini 3.5 model that pairs Pro-line reasoning quality with Flash-line latency and cost. Natively multimodal across text, image, audio, and video with a 1M-token context window, configurable thinking levels, and streaming function calling, tuned for high-throughput production workloads.

Gemini 3.1 Flash-Lite1.0M ctx

Google's most cost-efficient Gemini model optimized for high-volume, low-latency use cases. Delivers 2.5x faster time to first token versus Gemini 2.5 Flash with full multimodal support. Ideal for agentic tasks, data extraction, translation, and classification.

Gemini 3 Flash1.0M ctx

Google's balanced model combining Gemini 3 Pro's reasoning capabilities with the Flash line's latency, efficiency, and cost. Features configurable thinking levels, multimodal function responses, and streaming function calling for complex agentic workflows.

Laguna M.1128K ctx

Poolside AI's flagship agentic coding model with 225B total parameters and 23B active (MoE). Trained from scratch in-house on 30T tokens across 6,144 NVIDIA Hopper GPUs. Optimized for complex multi-step software engineering tasks including codebase exploration, file editing, test running, and iterative debugging.

GPT-5.51.0M ctx

OpenAI's most capable model designed for complex real-world work including coding, online research, information analysis, and document creation. Features advanced agentic capabilities with tool search and multi-step task execution.

GPT-5.4 Mini1.1M ctx

OpenAI's compact reasoning model optimized for coding, computer use, and subagent tasks. Approaches GPT-5.4 performance on several benchmarks while running more than 2x faster.

Muse Spark256K ctx

Meta Superintelligence Labs' first model, featuring advanced reasoning, multimodal understanding, and agentic capabilities. Processes voice, text, and image inputs with tool use and multi-agent orchestration. Powers Meta AI across its product ecosystem.