Модели

126 канонических LLM-моделей от всех провайдеров

Некоторые описания переведены автоматически в рамках пилота и пока не проверены редактором.

Показаны модели 25–48 из 126

Claude Opus 4.8

Соединенные Штаты

Самая продвинутая модель Anthropic, развивающая Opus 4.7 и улучшающая результаты в программировании, агентных навыках, рассуждении и интеллектуальной работе. Отличается повышенной честностью ответов, более эффективным использованием инструментов, поддержкой динамических рабочих процессов и улучшенным выравниванием.

Контекст
300K
Добавлена
май 2026 г.

Falcon-H1

ОАЭ

TII's hybrid Mamba-Transformer model that outperforms comparable offerings from Meta's Llama and Alibaba's Qwen in the 30-70B parameter range. Designed for real-world AI on everyday devices and resource-limited settings with state-of-the-art efficiency.

Контекст
131K
Добавлена
май 2026 г.

DBRX

Соединенные Штаты

Databricks' open-source 132B parameter Mixture-of-Experts transformer model with 36B active parameters per input. Released under Databricks Open Model License, optimized for enterprise workloads including SQL generation and coding tasks.

Контекст
33K
Добавлена
май 2026 г.

Ring-2.6-1T

Китай

InclusionAI's (Ant Group) trillion-parameter open-weights reasoning model with 63B active parameters per token. Built for real-world agent workflows with adaptive reasoning-effort modes. Features hybrid linear and MLA attention architecture with MIT license.

Контекст
131K
Добавлена
май 2026 г.

Snowflake Arctic

Соединенные Штаты

Snowflake's enterprise-focused open LLM with 480B total parameters using a fine-grained MoE architecture with only 17B active parameters per input. Apache 2.0 licensed, excels at SQL generation, coding, and enterprise intelligence tasks with breakthrough training efficiency.

Контекст
4K
Добавлена
май 2026 г.

Jamba Large 1.7

Израиль

AI21's latest hybrid SSM-Transformer model with Mixture-of-Experts architecture. Features a 256K context window, improved grounding and instruction-following. 94B total parameters with 398B active, optimized for enterprise long-context tasks.

Контекст
262K
Добавлена
май 2026 г.

Falcon 3 10B

ОАЭ

TII's open-source 10B parameter model from the Falcon 3 family. Achieved number one position on Hugging Face's LLM leaderboard in its size category, outperforming Meta's Llama variants and other models under 13B parameters.

Контекст
33K
Добавлена
май 2026 г.

Yi-Lightning

Китай

01.AI's flagship large language model with enhanced Mixture-of-Experts architecture. Ranked 6th on Chatbot Arena with particularly strong results in Chinese, Math, Coding, and Hard Prompts categories. Features advanced expert segmentation and optimized KV-caching.

Контекст
131K
Добавлена
май 2026 г.

StableLM 2 12B

Великобритания

Stability AI's 12.1 billion parameter decoder-only language model pre-trained on 2 trillion tokens of diverse multilingual and code datasets. Supports multiple languages and offers strong performance for its compact size with instruction-tuned chat variant available.

Контекст
4K
Добавлена
май 2026 г.

Palmyra X5

Соединенные Штаты

Writer's most advanced adaptive reasoning model with a 1 million token context window. Processes full million-token prompts in approximately 22 seconds with multi-turn function calls in 300ms. Optimized for enterprise agentic AI workflows at 3-4x lower cost than GPT-4.1.

Контекст
1.0M
Добавлена
май 2026 г.

Alloma 8B Instruct

Узбекистан

Uzbek LLM Lab's 8B parameter instruction-tuned model optimized for the Uzbek language. Built on Llama architecture with a custom tokenizer averaging 1.7 tokens per Uzbek word versus 3.5 in original Llama, enabling 2x faster inference. Trained on 3.6B tokens with 4096 context length.

Контекст
4K
Добавлена
май 2026 г.

Solar Pro 3

Республика Корея

Upstage's powerful Mixture-of-Experts language model with 102B total parameters and 12B active parameters per forward pass. Optimized for Korean with strong English and Japanese support. Excels at complex reasoning, structured output generation, and agentic workflows.

Контекст
128K
Добавлена
май 2026 г.

K2 Think

ОАЭ

A 32 billion parameter open-weights reasoning model by LLM360/MBZUAI, built on Qwen2.5-32B. Trained with reinforcement learning and verifiable rewards for long chain-of-thought reasoning, agentic planning, and complex problem solving in math, science, and code.

Контекст
131K
Добавлена
май 2026 г.

Qwen 3.7 Max

Китай

Alibaba's flagship proprietary model engineered for advanced agentic coding, complex reasoning, and long-horizon task execution. Ranked

Контекст
131K
Добавлена
май 2026 г.

Qwen 3.7 Plus

Китай

Alibaba's multimodal variant in the Qwen 3.7 family, optimized for vision understanding and multimodal tasks. Ranked

Контекст
131K
Добавлена
май 2026 г.

Gemini 3.5 Flash

Соединенные Штаты

Google DeepMind's balanced Gemini 3.5 model that pairs Pro-line reasoning quality with Flash-line latency and cost. Natively multimodal across text, image, audio, and video with a 1M-token context window, configurable thinking levels, and streaming function calling, tuned for high-throughput production workloads.

Контекст
1.0M
Добавлена
май 2026 г.

Gemini 3.1 Flash-Lite

Соединенные Штаты

Google's most cost-efficient Gemini model optimized for high-volume, low-latency use cases. Delivers 2.5x faster time to first token versus Gemini 2.5 Flash with full multimodal support. Ideal for agentic tasks, data extraction, translation, and classification.

Контекст
1.0M
Добавлена
май 2026 г.

Gemini 3 Flash

Соединенные Штаты

Google's balanced model combining Gemini 3 Pro's reasoning capabilities with the Flash line's latency, efficiency, and cost. Features configurable thinking levels, multimodal function responses, and streaming function calling for complex agentic workflows.

Контекст
1.0M
Добавлена
май 2026 г.

MiniCPM-V 4.6

Китай

Ultra-efficient multimodal language model from OpenBMB built on SigLIP2-400M and Qwen3.5-0.8B (~1B parameters). Supports single-image, multi-image, and video understanding with mixed 4x/16x visual token compression. Designed for edge deployment on iOS, Android, and HarmonyOS.

Контекст
256K
Добавлена
май 2026 г.

Granite 4.1 30B

Соединенные Штаты

IBM's largest dense decoder-only 30B parameter language model from the Granite 4.1 family. Trained on approximately 15T tokens with long-context extension up to 512K tokens. Supports tool calling, RAG, code generation, multilingual tasks across 12 languages. Released under Apache 2.0.

Контекст
524K
Добавлена
май 2026 г.

Granite 4.1 8B

Соединенные Штаты

IBM's dense decoder-only 8B parameter language model from the Granite 4.1 family. Supports 131K-token context, tool calling, RAG, code generation with fill-in-the-middle, text summarization, classification, and extraction across 12 languages. Released under Apache 2.0.

Контекст
131K
Добавлена
май 2026 г.

Laguna M.1

Соединенные Штаты

Poolside AI's flagship agentic coding model with 225B total parameters and 23B active (MoE). Trained from scratch in-house on 30T tokens across 6,144 NVIDIA Hopper GPUs. Optimized for complex multi-step software engineering tasks including codebase exploration, file editing, test running, and iterative debugging.

Контекст
128K
Добавлена
апр. 2026 г.

GPT-5.5

Соединенные Штаты

OpenAI's most capable model designed for complex real-world work including coding, online research, information analysis, and document creation. Features advanced agentic capabilities with tool search and multi-step task execution.

Контекст
1.0M
Добавлена
апр. 2026 г.

DeepSeek V4 Pro

Китай

DeepSeek's flagship V4 model with 1.6T total parameters (49B activated). MoE architecture supporting 1M token context. Closes the gap with frontier proprietary models on reasoning and coding benchmarks.

Контекст
1.0M
Добавлена
апр. 2026 г.