Модели

25 канонических LLM-моделей от всех провайдеров

Некоторые описания переведены автоматически в рамках пилота и пока не проверены редактором.

Показаны модели 1–24 из 25

Gemini 3.6 Flash

Соединенные Штаты

Google DeepMind's workhorse Flash model that builds on Gemini 3.5 Flash with better coding, knowledge work, and multimodal performance while reducing output token usage by roughly 17% per the Artificial Analysis Index. Natively multimodal across text, image, audio, and video with a 1M-token context window, configurable thinking levels, and built-in computer use, tuned for scaling agentic workflows at a lower cost per output token.

Контекст
1.0M
Добавлена
июль 2026 г.

Gemini 3.5 Flash-Lite

Соединенные Штаты

Google's fastest and most cost-effective Gemini 3.5-class model, delivering around 350 output tokens per second per the Artificial Analysis Index. Designed for low-latency and high-throughput agentic workflows such as agentic search and document processing, with configurable thinking levels, built-in computer use, and full multimodal support across a 1M-token context window.

Контекст
1.0M
Добавлена
июль 2026 г.

Gemini 3.5 Flash Cyber

Соединенные Штаты

A specialized, cyber-focused Gemini model built on top of Gemini 3.5 Flash and fine-tuned for finding and fixing cybersecurity vulnerabilities at a lower price per token than larger models. Deployed within Google's CodeMender code security agent, where multiple 3.5 Flash Cyber agents collaborate to reach competitive frontier performance on benchmarks like CyberGym. Given its dual-use nature, it is available exclusively to governments and trusted partners via CodeMender as a limited-access pilot.

Контекст
1.0M
Добавлена
июль 2026 г.

Inkling

Соединенные Штаты

Thinking Machines Lab's open-weights general-purpose multimodal Mixture-of-Experts model with 975B total parameters and 41B active parameters. Inkling accepts text, image, and audio inputs, produces text, and is designed for agentic and tool-use systems, coding assistants, chatbots, and retrieval-augmented generation.

Контекст
1.0M
Добавлена
июль 2026 г.

GPT-5.6 Terra

Соединенные Штаты

Сбалансированная модель серии GPT-5.6 от OpenAI: немного уступает по максимальному качеству, но обеспечивает заметно меньшую задержку и стоимость. Сохраняет сильные возможности рассуждения, программирования и агентного использования инструментов с настраиваемой глубиной рассуждения, поэтому подходит как основной вариант для масштабных производственных нагрузок, которым нужны передовые возможности.

Контекст
1.0M
Добавлена
июль 2026 г.

Gemini 3.5 Pro

Соединенные Штаты

Флагманская модель Gemini от Google DeepMind, заново построенная на новой основе с контекстным окном 2 млн токенов и режимом рассуждения Deep Think для самых сложных задач по математике, программированию и мультимодальной обработке. Нативно работает с текстом, изображениями, аудио и видео, поддерживает потоковый вызов функций и надёжную работу с длинным контекстом.

Контекст
2.0M
Добавлена
июль 2026 г.

GPT-5.6 Sol

Соединенные Штаты

Флагманская модель OpenAI серии GPT-5.6, развивающая программирование, научное рассуждение, долгосрочное планирование и агентные рабочие процессы и одновременно повышающая надёжность и эффективность на сложных практических задачах. Добавляет максимальный уровень глубины рассуждения и режим ultra, который запускает субагентов для сложной многоэтапной работы.

Контекст
1.0M
Добавлена
июль 2026 г.

Claude Sonnet 5

Соединенные Штаты

Самая мощная модель Anthropic класса Sonnet, переносящая передовые возможности программирования, агентной и профессиональной работы в средний сегмент и сокращающая разрыв с Opus 4.8 при меньшей цене. Поддерживает адаптивное мышление с выбором глубины рассуждения, контекстное окно 1 млн токенов и входные данные в виде текста, изображений и файлов. Кодовое имя — Fennec.

Контекст
1.0M
Добавлена
июнь 2026 г.

Nemotron 3 Ultra

Соединенные Штаты

NVIDIA's flagship open 550B-parameter Mixture-of-Experts model with 55B active parameters, built for frontier reasoning and orchestration in long-running agentic systems. Features hybrid Mamba-Transformer architecture, LatentMoE routing, multi-token prediction, and NVFP4 precision for 5x higher throughput. Achieves 30% lower cost-to-task-completion on agentic benchmarks. Supports 1M+ token context window with 95% accuracy on Ruler@1M.

Контекст
1.0M
Добавлена
июнь 2026 г.

Palmyra X5

Соединенные Штаты

Writer's most advanced adaptive reasoning model with a 1 million token context window. Processes full million-token prompts in approximately 22 seconds with multi-turn function calls in 300ms. Optimized for enterprise agentic AI workflows at 3-4x lower cost than GPT-4.1.

Контекст
1.0M
Добавлена
май 2026 г.

Gemini 3.1 Flash-Lite

Соединенные Штаты

Google's most cost-efficient Gemini model optimized for high-volume, low-latency use cases. Delivers 2.5x faster time to first token versus Gemini 2.5 Flash with full multimodal support. Ideal for agentic tasks, data extraction, translation, and classification.

Контекст
1.0M
Добавлена
май 2026 г.

Gemini 3 Flash

Соединенные Штаты

Google's balanced model combining Gemini 3 Pro's reasoning capabilities with the Flash line's latency, efficiency, and cost. Features configurable thinking levels, multimodal function responses, and streaming function calling for complex agentic workflows.

Контекст
1.0M
Добавлена
май 2026 г.

Gemini 3.5 Flash

Соединенные Штаты

Google DeepMind's balanced Gemini 3.5 model that pairs Pro-line reasoning quality with Flash-line latency and cost. Natively multimodal across text, image, audio, and video with a 1M-token context window, configurable thinking levels, and streaming function calling, tuned for high-throughput production workloads.

Контекст
1.0M
Добавлена
май 2026 г.

GPT-5.5

Соединенные Штаты

OpenAI's most capable model designed for complex real-world work including coding, online research, information analysis, and document creation. Features advanced agentic capabilities with tool search and multi-step task execution.

Контекст
1.0M
Добавлена
апр. 2026 г.

GPT-5.4 Mini

Соединенные Штаты

OpenAI's compact reasoning model optimized for coding, computer use, and subagent tasks. Approaches GPT-5.4 performance on several benchmarks while running more than 2x faster.

Контекст
1.1M
Добавлена
апр. 2026 г.

Nemotron 3 Super 120B

Соединенные Штаты

NVIDIA's open hybrid Mamba-Transformer MoE model with 120B total parameters (12B active). Features 1M token context window and excels at agentic reasoning, coding, planning, and tool calling.

Контекст
1.0M
Добавлена
апр. 2026 г.

Grok 4.3

Соединенные Штаты

xAI's latest and most intelligent model with strong agentic tool calling, minimal hallucinations, and configurable reasoning. Supports 1M token context window with competitive pricing.

Контекст
1.0M
Добавлена
апр. 2026 г.

GPT-5.4

Соединенные Штаты

OpenAI's frontier reasoning model combining advances in coding, reasoning, and agentic workflows. Features 1.1M token context window and strong performance on complex multi-step problems.

Контекст
1.1M
Добавлена
март 2026 г.

Gemini 3.1 Pro

Соединенные Штаты

Новейшая флагманская мультимодальная модель Google с передовыми результатами в рассуждении, программировании и понимании разных типов данных. Поддерживает встроенное использование инструментов, привязку к источникам и контекстное окно на миллион токенов.

Контекст
2.0M
Добавлена
март 2026 г.

Grok 4.20

Соединенные Штаты

xAI's multi-agent capable model with 2M token context window. Available in reasoning, non-reasoning, and multi-agent variants for diverse enterprise workloads.

Контекст
2.0M
Добавлена
февр. 2026 г.

Grok 4.1 Fast

Соединенные Штаты

xAI's fast and cost-effective model with 2M token context window. Offers both reasoning and non-reasoning modes at significantly lower pricing than flagship models.

Контекст
2.0M
Добавлена
нояб. 2025 г.

Gemini 2.5 Pro

Соединенные Штаты

Google's high-capability reasoning model with adaptive thinking for complex agentic and multimodal challenges. Features 1M token context window and strong performance on coding and scientific tasks.

Контекст
1.0M
Добавлена
июнь 2025 г.

Gemini 2.5 Flash

Соединенные Штаты

Google's cost-effective model optimized for high throughput tasks. Balances speed and intelligence with strong multimodal capabilities and 1M token context window.

Контекст
1.0M
Добавлена
июнь 2025 г.

Llama 4 Scout

Соединенные Штаты

Meta's efficient MoE model with 17B active parameters (109B total, 16 experts). Supports up to 10M token context — the longest of any production model. Strong performance on reasoning and multilingual tasks.

Контекст
10.0M
Добавлена
апр. 2025 г.