模型

浏览来自所有提供商的 17 个标准化 LLM 模型

部分描述为试点机器翻译内容,尚未经过人工审核。

显示第 1–17 项,共 17 个模型

MiniMax M3

中国

MiniMax's frontier open-weight model with 1M-token context window, native multimodality (text, image, video), and strong coding capabilities. Built on MiniMax Sparse Attention (MSA) architecture, achieving 59% on SWE-Bench Pro with significantly improved efficiency at long context.

上下文
1.0M
发布日期
2026年6月

Ring-2.6-1T

中国

InclusionAI's (Ant Group) trillion-parameter open-weights reasoning model with 63B active parameters per token. Built for real-world agent workflows with adaptive reasoning-effort modes. Features hybrid linear and MLA attention architecture with MIT license.

上下文
131K
发布日期
2026年5月

DeepSeek V4 Flash

中国

DeepSeek's efficient V4 model with 284B total parameters (13B activated). Optimized for speed and cost-efficiency while maintaining strong performance. Supports 1M token context window.

上下文
1.0M
发布日期
2026年4月

DeepSeek V4 Pro

中国

DeepSeek's flagship V4 model with 1.6T total parameters (49B activated). MoE architecture supporting 1M token context. Closes the gap with frontier proprietary models on reasoning and coding benchmarks.

上下文
1.0M
发布日期
2026年4月

Hy3 Preview

中国

Tencent's flagship open-weight Mixture-of-Experts model from the Hunyuan family with 295B total parameters and 21B active. Integrates fast and slow thinking modes with configurable reasoning effort. Designed for agentic workflows, cross-file code refactoring, long-document analysis, and multi-step tool use.

上下文
256K
发布日期
2026年4月

Qwen 3.6 27B

中国

Alibaba's dense 27B parameter model that outperforms its own 397B MoE predecessor on agentic coding benchmarks. Strong multilingual and reasoning capabilities released under Apache 2.0.

上下文
131K
发布日期
2026年4月

Qwen 3.6 35B-A3B

中国

Alibaba's efficient Mixture-of-Experts model with 35B total parameters and 3B active per token. Frontier-level agentic coding performance with 73.4% on SWE-bench Verified and 92.7 on AIME 2026. Released under Apache 2.0.

上下文
131K
发布日期
2026年4月

Qwen 3.6

中国

Alibaba's latest Qwen model with enhanced reasoning, multilingual capabilities, and improved instruction following. Features strong performance on coding, math, and general knowledge benchmarks.

上下文
131K
发布日期
2026年3月

MiniMax M2.7

中国

MiniMax's latest large language model with strong multilingual and multimodal capabilities. Competitive pricing with high-quality text generation and improved reasoning performance.

上下文
200K
发布日期
2026年3月

GLM-5.1

中国

Zhipu AI's latest bilingual model with strong Chinese and English capabilities. Features improved reasoning, coding, and tool use with competitive performance on academic benchmarks.

上下文
131K
发布日期
2026年3月

Kimi K2.6

中国

Moonshot AI's latest model with ultra-long context window support, strong reasoning capabilities, and excellent performance on complex multi-step tasks. Known for reliable long-document understanding.

上下文
1.0M
发布日期
2026年3月

DeepSeek V4

中国

DeepSeek's fourth-generation model with improved mixture-of-experts architecture, enhanced reasoning and coding capabilities, and stronger multilingual performance. Competitive with frontier proprietary models.

上下文
256K
发布日期
2026年2月

GLM-4.7

中国

Zhipu AI's multilingual agentic coding model with strong reasoning, tool use, and UI generation capabilities. Predecessor to GLM-5.1 with competitive performance on coding benchmarks.

上下文
131K
发布日期
2025年10月

Qwen3 Coder

中国

Alibaba's Qwen3 Coder model optimized for software development tasks including code generation, debugging, code review, and technical documentation with strong multilingual programming support.

上下文
131K
发布日期
2025年4月

Qwen3 32B

中国

Alibaba's Qwen3 32B dense language model with strong reasoning and multilingual capabilities, supporting function calling and code generation across diverse tasks.

上下文
131K
发布日期
2025年4月

DeepSeek R1

中国

DeepSeek 面向推理的模型,通过强化学习训练以处理复杂的多步骤推理任务。尤其擅长需要思维链推理的数学、科学和编程问题。

上下文
131K
发布日期
2025年1月

DeepSeek V3

中国

DeepSeek's third-generation large language model featuring mixture-of-experts architecture, strong multilingual capabilities, and competitive performance on reasoning and coding benchmarks.

上下文
128K
发布日期
2024年12月