Модели

58 канонических LLM-моделей от всех провайдеров

Показаны модели 25–48 из 58

Hy3 Preview

Китай

Tencent's flagship open-weight Mixture-of-Experts model from the Hunyuan family with 295B total parameters and 21B active. Integrates fast and slow thinking modes with configurable reasoning effort. Designed for agentic workflows, cross-file code refactoring, long-document analysis, and multi-step tool use.

Контекст
256K
Добавлена
апр. 2026 г.

MiMo-V2.5-Pro

Китай

Xiaomi's flagship 1.02T-parameter Mixture-of-Experts model with 42B active parameters, built on a hybrid-attention architecture with 3-layer Multi-Token Prediction. Designed for complex agentic tasks, software engineering, and long-horizon instruction following with a 1M-token context window.

Контекст
1.0M
Добавлена
апр. 2026 г.

Qwen 3.6 27B

Китай

Alibaba's dense 27B parameter model that outperforms its own 397B MoE predecessor on agentic coding benchmarks. Strong multilingual and reasoning capabilities released under Apache 2.0.

Контекст
131K
Добавлена
апр. 2026 г.

Qwen 3.6 35B-A3B

Китай

Alibaba's efficient Mixture-of-Experts model with 35B total parameters and 3B active per token. Frontier-level agentic coding performance with 73.4% on SWE-bench Verified and 92.7 on AIME 2026. Released under Apache 2.0.

Контекст
131K
Добавлена
апр. 2026 г.

GPT-5.4 Mini

Соединенные Штаты

OpenAI's compact reasoning model optimized for coding, computer use, and subagent tasks. Approaches GPT-5.4 performance on several benchmarks while running more than 2x faster.

Контекст
1.1M
Добавлена
апр. 2026 г.

Qwen 3.6 Plus

Китай

Alibaba's proprietary flagship model in the Qwen 3.6 family, targeting enterprise AI workflows with stronger agentic coding capability, visual coding support, and end-to-end enterprise engineering features.

Контекст
131K
Добавлена
апр. 2026 г.

Gemma 4 31B

Соединенные Штаты

Google's flagship open-weight dense model with 31B parameters. All parameters active per forward pass. Ranks among top open models with strong performance on AIME 2026 (89.2%) and MMLU Pro (85.2%). Supports vision and extended context.

Контекст
262K
Добавлена
апр. 2026 г.

Nemotron 3 Super 120B

Соединенные Штаты

NVIDIA's open hybrid Mamba-Transformer MoE model with 120B total parameters (12B active). Features 1M token context window and excels at agentic reasoning, coding, planning, and tool calling.

Контекст
1.0M
Добавлена
апр. 2026 г.

GPT-OSS 20B

Соединенные Штаты

OpenAI's compact open-weight model with 20 billion parameters. Released under Apache 2.0 license, designed for efficient deployment on consumer hardware while maintaining strong coding and reasoning capabilities.

Контекст
131K
Добавлена
апр. 2026 г.

MiniMax M2.7

Китай

MiniMax's latest large language model with strong multilingual and multimodal capabilities. Competitive pricing with high-quality text generation and improved reasoning performance.

Контекст
200K
Добавлена
март 2026 г.

GLM-5.1

Китай

Zhipu AI's latest bilingual model with strong Chinese and English capabilities. Features improved reasoning, coding, and tool use with competitive performance on academic benchmarks.

Контекст
131K
Добавлена
март 2026 г.

Qwen 3.6

Китай

Alibaba's latest Qwen model with enhanced reasoning, multilingual capabilities, and improved instruction following. Features strong performance on coding, math, and general knowledge benchmarks.

Контекст
131K
Добавлена
март 2026 г.

Kimi K2.6

Китай

Moonshot AI's latest model with ultra-long context window support, strong reasoning capabilities, and excellent performance on complex multi-step tasks. Known for reliable long-document understanding.

Контекст
1.0M
Добавлена
март 2026 г.

Mistral Small 4

Франция

Mistral AI's efficient hybrid model unifying instruct, reasoning, and coding in a single model. Open-weight under Apache 2.0 with strong performance for its size class.

Контекст
128K
Добавлена
март 2026 г.

Mistral Medium 3.5

Франция

Mistral AI's balanced model offering strong multilingual performance with excellent price-performance ratio. Optimized for production workloads requiring reliable quality across European and global languages.

Контекст
128K
Добавлена
февр. 2026 г.

DeepSeek V4

Китай

DeepSeek's fourth-generation model with improved mixture-of-experts architecture, enhanced reasoning and coding capabilities, and stronger multilingual performance. Competitive with frontier proprietary models.

Контекст
256K
Добавлена
февр. 2026 г.

Devstral 2

Франция

Mistral AI's frontier code agents model designed for solving software engineering tasks. Open-weight model optimized for agentic coding workflows and complex development tasks.

Контекст
128K
Добавлена
дек. 2025 г.

Grok 4.1 Fast

Соединенные Штаты

xAI's fast and cost-effective model with 2M token context window. Offers both reasoning and non-reasoning modes at significantly lower pricing than flagship models.

Контекст
2.0M
Добавлена
нояб. 2025 г.

Claude Haiku 4.5

Соединенные Штаты

Anthropic's fastest model with near-frontier intelligence. Optimized for high-throughput, low-latency applications requiring quick responses at minimal cost. Supports extended thinking.

Контекст
200K
Добавлена
окт. 2025 г.

GLM-4.7

Китай

Zhipu AI's multilingual agentic coding model with strong reasoning, tool use, and UI generation capabilities. Predecessor to GLM-5.1 with competitive performance on coding benchmarks.

Контекст
131K
Добавлена
окт. 2025 г.

AlemLLM

Казахстан

Kazakhstan's flagship Mixture-of-Experts language model developed by Astana Hub with technical support from 01.AI. Features 247B total parameters with 22B active per token, achieving state-of-the-art results on Kazakh, Russian, and English benchmarks. Outperforms GPT-4o on Kazakh language tasks.

Контекст
131K
Добавлена
авг. 2025 г.

Gemini 2.5 Flash

Соединенные Штаты

Google's cost-effective model optimized for high throughput tasks. Balances speed and intelligence with strong multimodal capabilities and 1M token context window.

Контекст
1.0M
Добавлена
июнь 2025 г.

Nemotron Nano 9B v2

Соединенные Штаты

NVIDIA's compact 9B parameter model trained from scratch for both reasoning and non-reasoning tasks. Generates reasoning traces before final responses. Efficient for edge and on-device deployment.

Контекст
131K
Добавлена
июнь 2025 г.

Llama 4 Scout

Соединенные Штаты

Meta's efficient MoE model with 17B active parameters (109B total, 16 experts). Supports up to 10M token context — the longest of any production model. Strong performance on reasoning and multilingual tasks.

Контекст
10.0M
Добавлена
апр. 2025 г.