Modelos

Explora 6 modelos LLM canónicos de todos los proveedores

Algunas descripciones forman parte del piloto de traducción automática y aún no han sido revisadas.

Mostrando 1–6 de 6 modelos

GLM-5.21.0M ctx

Z.ai's (formerly Zhipu AI) flagship open-weight coding model with a 1M-token context window. Mixture-of-Experts architecture with 753B total parameters and ~40B active per request, featuring two cost-balancing reasoning modes. Tops several coding benchmarks while remaining a fraction of the cost of comparable proprietary frontier models. MIT-licensed weights.

DeepSeek V4 Flash1.0M ctx

DeepSeek's efficient V4 model with 284B total parameters (13B activated). Optimized for speed and cost-efficiency while maintaining strong performance. Supports 1M token context window.

DeepSeek V4 Pro1.0M ctx

DeepSeek's flagship V4 model with 1.6T total parameters (49B activated). MoE architecture supporting 1M token context. Closes the gap with frontier proprietary models on reasoning and coding benchmarks.

MiMo-V2.5-Pro1.0M ctx

Xiaomi's flagship 1.02T-parameter Mixture-of-Experts model with 42B active parameters, built on a hybrid-attention architecture with 3-layer Multi-Token Prediction. Designed for complex agentic tasks, software engineering, and long-horizon instruction following with a 1M-token context window.

DeepSeek V4256K ctx

DeepSeek's fourth-generation model with improved mixture-of-experts architecture, enhanced reasoning and coding capabilities, and stronger multilingual performance. Competitive with frontier proprietary models.

DeepSeek R1131K ctx

Modelo de DeepSeek centrado en el razonamiento y entrenado mediante aprendizaje por refuerzo para tareas complejas de varios pasos. Destaca en problemas de matemáticas, ciencia y programación que requieren razonamiento en cadena.