Модели

105 канонических LLM-моделей от всех провайдеров

Некоторые описания переведены автоматически в рамках пилота и пока не проверены редактором.

Показаны модели 97–105 из 105

DeepSeek R1

Китай

Модель DeepSeek, ориентированная на рассуждение и обученная с подкреплением для сложных многоэтапных задач. Особенно хорошо справляется с математикой, естественными науками и программированием, где требуется последовательное рассуждение.

Контекст
131K
Добавлена
янв. 2025 г.

Codestral

Франция

Mistral AI's cutting-edge code generation model specializing in low-latency, high-frequency tasks such as fill-in-the-middle (FIM), code completion, correction, and test generation. Features efficient architecture with 2x faster generation than its predecessor.

Контекст
256K
Добавлена
янв. 2025 г.

ISSAI KazLLM 1.0 70B

Казахстан

Large language model developed by ISSAI (Nazarbayev University) customized from Llama 3.1 70B to improve helpfulness of responses in the Kazakh language. Part of Kazakhstan's initiative to ensure the country benefits from generative AI advancements.

Контекст
128K
Добавлена
дек. 2024 г.

Llama 3.3 70B Instruct

Соединенные Штаты

Meta's flagship open-weight model with 70 billion parameters. Strong multilingual capabilities with competitive performance on reasoning and coding benchmarks. Available for self-hosting and through various inference providers.

Контекст
131K
Добавлена
дек. 2024 г.

Command R7B

Соединенные Штаты

Cohere's compact 7B parameter model optimized for RAG, tool use, and code tasks. Delivers top-tier speed and efficiency on commodity GPUs and edge devices with 128K context window.

Контекст
128K
Добавлена
дек. 2024 г.

DeepSeek V3

Китай

DeepSeek's third-generation large language model featuring mixture-of-experts architecture, strong multilingual capabilities, and competitive performance on reasoning and coding benchmarks.

Контекст
128K
Добавлена
дек. 2024 г.

Llama 3.1 8B Instruct

Соединенные Штаты

Meta's efficient open-weight model with 8 billion parameters from the Llama 3.1 family. Optimized for instruction following with strong performance on general tasks, coding, and multilingual benchmarks. Ideal for cost-effective deployment and edge inference scenarios.

Контекст
131K
Добавлена
июль 2024 г.

Claude 3 Opus

Соединенные Штаты

Anthropic's most powerful model in the Claude 3 family, excelling at complex analysis, nuanced content generation, scientific reasoning, and code generation with extended context support.

Контекст
200K
Добавлена
март 2024 г.

GPT-4

Соединенные Штаты

OpenAI's flagship large language model with advanced reasoning, instruction following, and code generation capabilities. Supports multimodal inputs including text and images.

Контекст
128K
Добавлена
янв. 2024 г.