Модели

19 канонических LLM-моделей от всех провайдеров

Показаны модели 1–19 из 19

Sarvam-M

Индия

Sarvam AI's 24B-parameter instruction-tuned model derived from Mistral-Small-3.1-24B, post-trained on English plus eleven major Indic languages (bn, hi, kn, gu, mr, ml, or, pa, ta, te). Delivers large relative gains on Indian-language, math, and programming benchmarks over its base model, with a hybrid reasoning mode for complex tasks.

Контекст
131K
Добавлена
июнь 2026 г.

Sarvam-30B

Индия

Sarvam AI's 30B-parameter Mixture-of-Experts reasoning model trained from scratch with only 2.4B active parameters per token. Optimized for real-time deployment and Indian languages, delivering strong reasoning, coding, and conversational performance while remaining efficient to serve. Open-weights.

Контекст
128K
Добавлена
июнь 2026 г.

Sarvam-105B

Индия

Sarvam AI's sovereign 105B-parameter Mixture-of-Experts model activating ~9B parameters per token, with a 128K-token context window. Trained on 12 trillion tokens across 22 Indian languages using 128 sparse experts with Multi-head Latent Attention and a custom low-fertility Indic tokenizer. Wins the majority of pairwise comparisons on Indian-language and STEM benchmarks.

Контекст
128K
Добавлена
июнь 2026 г.

DiffusionGemma

Соединенные Штаты

Google DeepMind's experimental diffusion-based member of the Gemma 4 open model family. Unlike autoregressive models that generate text one token at a time, DiffusionGemma denoises a canvas of placeholder tokens to produce up to 256 tokens in parallel, finalizing output in one block. A Mixture-of-Experts model with 26B total parameters and 3.8B active per inference, delivering roughly 4x the throughput of similarly sized autoregressive Gemma models on local hardware. Excels at non-linear tasks like in-line editing, molecular sequencing, mathematical graphing, and self-correcting puzzles.

Контекст
262K
Добавлена
июнь 2026 г.

Gemma 4 12B

Соединенные Штаты

Google's medium-size open-weight model with 12 billion parameters from the Gemma 4 family. Encoder-free unified multimodal architecture that natively processes text, image, audio, and video inputs without dedicated encoders. Features a 256K context window and supports 140+ languages. First medium-sized model capable of natively ingesting audio. Suitable for local deployment on GPUs with 16GB VRAM.

Контекст
262K
Добавлена
июнь 2026 г.

Falcon-H1

ОАЭ

TII's hybrid Mamba-Transformer model that outperforms comparable offerings from Meta's Llama and Alibaba's Qwen in the 30-70B parameter range. Designed for real-world AI on everyday devices and resource-limited settings with state-of-the-art efficiency.

Контекст
131K
Добавлена
май 2026 г.

K2 Think

ОАЭ

A 32 billion parameter open-weights reasoning model by LLM360/MBZUAI, built on Qwen2.5-32B. Trained with reinforcement learning and verifiable rewards for long chain-of-thought reasoning, agentic planning, and complex problem solving in math, science, and code.

Контекст
131K
Добавлена
май 2026 г.

Granite 4.1 30B

Соединенные Штаты

IBM's largest dense decoder-only 30B parameter language model from the Granite 4.1 family. Trained on approximately 15T tokens with long-context extension up to 512K tokens. Supports tool calling, RAG, code generation, multilingual tasks across 12 languages. Released under Apache 2.0.

Контекст
524K
Добавлена
май 2026 г.

Granite 4.1 8B

Соединенные Штаты

IBM's dense decoder-only 8B parameter language model from the Granite 4.1 family. Supports 131K-token context, tool calling, RAG, code generation with fill-in-the-middle, text summarization, classification, and extraction across 12 languages. Released under Apache 2.0.

Контекст
131K
Добавлена
май 2026 г.

Qwen 3.6 27B

Китай

Alibaba's dense 27B parameter model that outperforms its own 397B MoE predecessor on agentic coding benchmarks. Strong multilingual and reasoning capabilities released under Apache 2.0.

Контекст
131K
Добавлена
апр. 2026 г.

Qwen 3.6 35B-A3B

Китай

Alibaba's efficient Mixture-of-Experts model with 35B total parameters and 3B active per token. Frontier-level agentic coding performance with 73.4% on SWE-bench Verified and 92.7 on AIME 2026. Released under Apache 2.0.

Контекст
131K
Добавлена
апр. 2026 г.

Gemma 4 31B

Соединенные Штаты

Google's flagship open-weight dense model with 31B parameters. All parameters active per forward pass. Ranks among top open models with strong performance on AIME 2026 (89.2%) and MMLU Pro (85.2%). Supports vision and extended context.

Контекст
262K
Добавлена
апр. 2026 г.

GPT-OSS 20B

Соединенные Штаты

OpenAI's compact open-weight model with 20 billion parameters. Released under Apache 2.0 license, designed for efficient deployment on consumer hardware while maintaining strong coding and reasoning capabilities.

Контекст
131K
Добавлена
апр. 2026 г.

Mistral Small 4

Франция

Mistral AI's efficient hybrid model unifying instruct, reasoning, and coding in a single model. Open-weight under Apache 2.0 with strong performance for its size class.

Контекст
128K
Добавлена
март 2026 г.

Qwen 3.6

Китай

Alibaba's latest Qwen model with enhanced reasoning, multilingual capabilities, and improved instruction following. Features strong performance on coding, math, and general knowledge benchmarks.

Контекст
131K
Добавлена
март 2026 г.

Devstral 2

Франция

Mistral AI's frontier code agents model designed for solving software engineering tasks. Open-weight model optimized for agentic coding workflows and complex development tasks.

Контекст
128K
Добавлена
дек. 2025 г.

AlemLLM

Казахстан

Kazakhstan's flagship Mixture-of-Experts language model developed by Astana Hub with technical support from 01.AI. Features 247B total parameters with 22B active per token, achieving state-of-the-art results on Kazakh, Russian, and English benchmarks. Outperforms GPT-4o on Kazakh language tasks.

Контекст
131K
Добавлена
авг. 2025 г.

Qwen3 32B

Китай

Alibaba's Qwen3 32B dense language model with strong reasoning and multilingual capabilities, supporting function calling and code generation across diverse tasks.

Контекст
131K
Добавлена
апр. 2025 г.

Qwen3 Coder

Китай

Alibaba's Qwen3 Coder model optimized for software development tasks including code generation, debugging, code review, and technical documentation with strong multilingual programming support.

Контекст
131K
Добавлена
апр. 2025 г.