Modelos

Explora 37 modelos LLM canónicos de todos los proveedores

Mostrando 25–37 de 37 modelos

AlemLLM

Kazajistán

Kazakhstan's flagship Mixture-of-Experts language model developed by Astana Hub with technical support from 01.AI. Features 247B total parameters with 22B active per token, achieving state-of-the-art results on Kazakh, Russian, and English benchmarks. Outperforms GPT-4o on Kazakh language tasks.

Contexto
131K
Publicado
ago 2025

Trendyol LLM 8B T1

Turquía

Turkish-optimized 8B chat model developed by Trendyol, Turkey's largest e-commerce platform. Built on Qwen3-8B and fine-tuned on large-scale Turkish e-commerce datasets. Features advanced chain-of-thought reasoning in Turkish with dual operation modes (/think and /no_think), strong instruction following, summarization, coding, and attribute extraction for catalogue enrichment. English reasoning capabilities are preserved alongside Turkish.

Contexto
33K
Publicado
jul 2025

YandexGPT 5 Lite

Rusia

Yandex's compact 8B parameter language model trained on 15T tokens of primarily Russian and English text. Features 32K context window with strong performance on web, code, and mathematics tasks. Open-weight release.

Contexto
32K
Publicado
jun 2025

Qwen3 32B

China

Alibaba's Qwen3 32B dense language model with strong reasoning and multilingual capabilities, supporting function calling and code generation across diverse tasks.

Contexto
131K
Publicado
abr 2025

Qwen3 Coder

China

Alibaba's Qwen3 Coder model optimized for software development tasks including code generation, debugging, code review, and technical documentation with strong multilingual programming support.

Contexto
131K
Publicado
abr 2025

Qwen3 235B

China

Alibaba's Qwen3 235B mixture-of-experts model delivering frontier-level performance with advanced reasoning, function calling, and code generation capabilities at massive scale.

Contexto
131K
Publicado
abr 2025

Gemma 3 12B

Estados Unidos

Google's mid-size open-weight model with 12 billion parameters from the Gemma 3 family. Supports multimodal inputs including text and images with a 128K context window. Strong performance on reasoning and code generation tasks at moderate compute cost.

Contexto
131K
Publicado
mar 2025

Gemma 3 1B

Estados Unidos

Google's lightweight open-weight model with 1 billion parameters from the Gemma 3 family. Designed for on-device and resource-constrained deployments. Supports text-only tasks with a 32K context window. Efficient for chat and basic completion workloads.

Contexto
33K
Publicado
mar 2025

Gemma 3 4B

Estados Unidos

Google's compact open-weight model with 4 billion parameters from the Gemma 3 family. Supports multimodal inputs including text and images with a 128K context window. Balances efficiency and capability for vision and language tasks.

Contexto
131K
Publicado
mar 2025

Gemma 3 27B

Estados Unidos

Google's largest open-weight model in the Gemma 3 family with 27 billion parameters. Supports multimodal inputs including text and images with a 128K context window. Delivers strong performance across reasoning, code generation, and vision tasks, competitive with larger proprietary models.

Contexto
131K
Publicado
mar 2025

Mistral Small 3.1

Francia

Mistral AI's Small 3.1 model with 24B parameters offering efficient multimodal capabilities including vision, function calling, and code generation with a large 128K context window.

Contexto
128K
Publicado
mar 2025

QwQ 32B

China

Alibaba's QwQ 32B reasoning-focused model designed for complex problem solving, mathematical reasoning, and step-by-step logical analysis with strong chain-of-thought capabilities.

Contexto
131K
Publicado
mar 2025

Cotype Nano

Rusia

MTS AI's lightweight 1.5B parameter language model optimized for resource-constrained environments. Excels at Russian and English language tasks including content creation, translation, and text analysis. Runs efficiently on both CPU and GPU, including laptops and smartphones.

Contexto
32K
Publicado
nov 2024