Модели
45 канонических LLM-моделей от всех провайдеров
Некоторые описания переведены автоматически в рамках пилота и пока не проверены редактором.
Inkling
Thinking Machines Lab's open-weights general-purpose multimodal Mixture-of-Experts model with 975B total parameters and 41B active parameters. Inkling accepts text, image, and audio inputs, produces text, and is designed for agentic and tool-use systems, coding assistants, chatbots, and retrieval-augmented generation.
Muse Spark 1.1
Meta Superintelligence Labs' updated flagship, building on Muse Spark with stronger agentic reasoning, more reliable multi-agent orchestration, and improved multimodal understanding across voice, text, and image. Extends the context window and reduces latency and reasoning token usage while raising coding and tool-use accuracy. Powers Meta AI across its product ecosystem.
Gemini 3.5 Pro
Флагманская модель Gemini от Google DeepMind, заново построенная на новой основе с контекстным окном 2 млн токенов и режимом рассуждения Deep Think для самых сложных задач по математике, программированию и мультимодальной обработке. Нативно работает с текстом, изображениями, аудио и видео, поддерживает потоковый вызов функций и надёжную работу с длинным контекстом.
GPT-5.6 Luna
The fast, low-cost tier of OpenAI's GPT-5.6 series, optimized for high-volume, latency-sensitive tasks such as classification, extraction, routing, and lightweight agentic steps. Approaches the larger GPT-5.6 tiers on many benchmarks while running several times faster at a fraction of the price.
GPT-5.6 Sol
Флагманская модель OpenAI серии GPT-5.6, развивающая программирование, научное рассуждение, долгосрочное планирование и агентные рабочие процессы и одновременно повышающая надёжность и эффективность на сложных практических задачах. Добавляет максимальный уровень глубины рассуждения и режим ultra, который запускает субагентов для сложной многоэтапной работы.
GPT-5.6 Terra
Сбалансированная модель серии GPT-5.6 от OpenAI: немного уступает по максимальному качеству, но обеспечивает заметно меньшую задержку и стоимость. Сохраняет сильные возможности рассуждения, программирования и агентного использования инструментов с настраиваемой глубиной рассуждения, поэтому подходит как основной вариант для масштабных производственных нагрузок, которым нужны передовые возможности.
Grok 4.5
Самая мощная на текущий момент модель xAI, созданная для программирования, агентных задач и интеллектуальной работы и разрабатывавшаяся совместно с инструментами для реальной разработки ПО. Поддерживает доступ к актуальной информации, расширенное рассуждение и работу с инструментами в большом контексте через API, совместимый с OpenAI.
Claude Sonnet 5
Самая мощная модель Anthropic класса Sonnet, переносящая передовые возможности программирования, агентной и профессиональной работы в средний сегмент и сокращающая разрыв с Opus 4.8 при меньшей цене. Поддерживает адаптивное мышление с выбором глубины рассуждения, контекстное окно 1 млн токенов и входные данные в виде текста, изображений и файлов. Кодовое имя — Fennec.
Claude Fable 5
Первая общедоступная модель Anthropic класса Mythos, превосходящая по возможностям все модели, которые компания ранее выпускала для широкого доступа. Показывает передовые результаты почти во всех протестированных бенчмарках и особенно сильна в разработке ПО, интеллектуальной работе, обработке изображений и научных исследованиях. Её преимущество растёт на более длинных и сложных задачах. Встроенные меры защиты направляют чувствительные запросы по кибербезопасности, биологии, химии и дистилляции в Claude Opus 4.8.
Claude Mythos 5
Anthropic's frontier Mythos-class model — the same underlying model as Claude Fable 5 but with safeguards lifted in some areas. It has the strongest cybersecurity capabilities of any model in the world, alongside state-of-the-art performance in software engineering, knowledge work, vision, and scientific research. Access is restricted to a small group of trusted cyberdefenders and infrastructure providers through Project Glasswing.
Gemma 4 12B
Google's medium-size open-weight model with 12 billion parameters from the Gemma 4 family. Encoder-free unified multimodal architecture that natively processes text, image, audio, and video inputs without dedicated encoders. Features a 256K context window and supports 140+ languages. First medium-sized model capable of natively ingesting audio. Suitable for local deployment on GPUs with 16GB VRAM.
Claude Opus 4.8
Самая продвинутая модель Anthropic, развивающая Opus 4.7 и улучшающая результаты в программировании, агентных навыках, рассуждении и интеллектуальной работе. Отличается повышенной честностью ответов, более эффективным использованием инструментов, поддержкой динамических рабочих процессов и улучшенным выравниванием.
Gemini 3.5 Flash
Google DeepMind's balanced Gemini 3.5 model that pairs Pro-line reasoning quality with Flash-line latency and cost. Natively multimodal across text, image, audio, and video with a 1M-token context window, configurable thinking levels, and streaming function calling, tuned for high-throughput production workloads.
Gemini 3 Flash
Google's balanced model combining Gemini 3 Pro's reasoning capabilities with the Flash line's latency, efficiency, and cost. Features configurable thinking levels, multimodal function responses, and streaming function calling for complex agentic workflows.
Gemini 3.1 Flash-Lite
Google's most cost-efficient Gemini model optimized for high-volume, low-latency use cases. Delivers 2.5x faster time to first token versus Gemini 2.5 Flash with full multimodal support. Ideal for agentic tasks, data extraction, translation, and classification.
GPT-5.5
OpenAI's most capable model designed for complex real-world work including coding, online research, information analysis, and document creation. Features advanced agentic capabilities with tool search and multi-step task execution.
GPT-5.4 Mini
OpenAI's compact reasoning model optimized for coding, computer use, and subagent tasks. Approaches GPT-5.4 performance on several benchmarks while running more than 2x faster.
Muse Spark
Meta Superintelligence Labs' first model, featuring advanced reasoning, multimodal understanding, and agentic capabilities. Processes voice, text, and image inputs with tool use and multi-agent orchestration. Powers Meta AI across its product ecosystem.
Gemma 4 31B
Google's flagship open-weight dense model with 31 billion parameters from the Gemma 4 family. All parameters active per forward pass with top-tier performance on reasoning benchmarks including AIME 2026 and MMLU Pro. Supports vision and extended 256K context window.
Gemma 4 31B
Google's flagship open-weight dense model with 31B parameters. All parameters active per forward pass. Ranks among top open models with strong performance on AIME 2026 (89.2%) and MMLU Pro (85.2%). Supports vision and extended context.
Gemma 4 26B
Google's high-performance open-weight dense model with 26 billion parameters from the Gemma 4 family. Supports multimodal inputs including text and images with a 256K extended context window. Strong reasoning and code generation capabilities with all parameters active per forward pass.
Grok 4.3
xAI's latest and most intelligent model with strong agentic tool calling, minimal hallucinations, and configurable reasoning. Supports 1M token context window with competitive pricing.
Claude Opus 4.7
Anthropic's latest and most advanced model with state-of-the-art reasoning, coding, and analysis capabilities. Features improved tool use, extended thinking, and enhanced safety alignment.