Modelos
Explora 21 modelos LLM canónicos de todos los proveedores
Algunas descripciones forman parte del piloto de traducción automática y aún no han sido revisadas.
Google DeepMind's workhorse Flash model that builds on Gemini 3.5 Flash with better coding, knowledge work, and multimodal performance while reducing output token usage by roughly 17% per the Artificial Analysis Index. Natively multimodal across text, image, audio, and video with a 1M-token context window, configurable thinking levels, and built-in computer use, tuned for scaling agentic workflows at a lower cost per output token.
Google's fastest and most cost-effective Gemini 3.5-class model, delivering around 350 output tokens per second per the Artificial Analysis Index. Designed for low-latency and high-throughput agentic workflows such as agentic search and document processing, with configurable thinking levels, built-in computer use, and full multimodal support across a 1M-token context window.
El modelo insignia de la serie GPT-5.6 de OpenAI, que mejora la programación, el razonamiento científico, la planificación a largo plazo y los flujos de trabajo con agentes, al tiempo que aumenta la fiabilidad y la eficiencia en tareas reales exigentes. Añade un nivel máximo de esfuerzo de razonamiento y un modo ultra que inicia subagentes para trabajos complejos de varios pasos.
El modelo más potente de xAI hasta la fecha, diseñado para destacar en programación, tareas con agentes y trabajo del conocimiento, y desarrollado junto con herramientas de programación para ingeniería de software real. Ofrece acceso a información en tiempo real, razonamiento ampliado y uso de herramientas con contextos grandes mediante una API compatible con OpenAI.
MiniMax's frontier open-weight model with 1M-token context window, native multimodality (text, image, video), and strong coding capabilities. Built on MiniMax Sparse Attention (MSA) architecture, achieving 59% on SWE-Bench Pro with significantly improved efficiency at long context.
Upstage's powerful Mixture-of-Experts language model with 102B total parameters and 12B active parameters per forward pass. Optimized for Korean with strong English and Japanese support. Excels at complex reasoning, structured output generation, and agentic workflows.
Google's most cost-efficient Gemini model optimized for high-volume, low-latency use cases. Delivers 2.5x faster time to first token versus Gemini 2.5 Flash with full multimodal support. Ideal for agentic tasks, data extraction, translation, and classification.
Google DeepMind's balanced Gemini 3.5 model that pairs Pro-line reasoning quality with Flash-line latency and cost. Natively multimodal across text, image, audio, and video with a 1M-token context window, configurable thinking levels, and streaming function calling, tuned for high-throughput production workloads.
Google's balanced model combining Gemini 3 Pro's reasoning capabilities with the Flash line's latency, efficiency, and cost. Features configurable thinking levels, multimodal function responses, and streaming function calling for complex agentic workflows.
OpenAI's frontier reasoning model combining advances in coding, reasoning, and agentic workflows. Features 1.1M token context window and strong performance on complex multi-step problems.
MiniMax's latest large language model with strong multilingual and multimodal capabilities. Competitive pricing with high-quality text generation and improved reasoning performance.
Moonshot AI's latest model with ultra-long context window support, strong reasoning capabilities, and excellent performance on complex multi-step tasks. Known for reliable long-document understanding.
El último modelo multimodal insignia de Google, con rendimiento de vanguardia en razonamiento, programación y comprensión multimodal. Incluye uso nativo de herramientas, fundamentación y una ventana de contexto de un millón de tokens.
Zhipu AI's latest bilingual model with strong Chinese and English capabilities. Features improved reasoning, coding, and tool use with competitive performance on academic benchmarks.
Mistral AI's balanced model offering strong multilingual performance with excellent price-performance ratio. Optimized for production workloads requiring reliable quality across European and global languages.
Anthropic's most capable model in the Claude 4 family, excelling at complex analysis, extended reasoning, scientific research, and advanced code generation. Features significantly improved accuracy and reduced hallucinations.
Anthropic's balanced model offering strong performance at lower cost and latency than Opus. Excellent for everyday coding, analysis, and content generation tasks with good reasoning capabilities.
Anthropic's fastest model with near-frontier intelligence. Optimized for high-throughput, low-latency applications requiring quick responses at minimal cost. Supports extended thinking.
Google's high-capability reasoning model with adaptive thinking for complex agentic and multimodal challenges. Features 1M token context window and strong performance on coding and scientific tasks.
Google's cost-effective model optimized for high throughput tasks. Balances speed and intelligence with strong multimodal capabilities and 1M token context window.
OpenAI's fifth-generation flagship model with significant improvements in reasoning, multimodal understanding, and code generation. Features enhanced instruction following and expanded context window.