模型

浏览来自所有提供商的 77 个标准化 LLM 模型

部分描述为试点机器翻译内容,尚未经过人工审核。

显示第 73–77 项,共 77 个模型

Llama 4 Maverick

美国

Meta's quality-focused MoE model with 17B active parameters (400B total, 128 experts). Targets quality-critical tasks with benchmark scores competitive with GPT-4o and Gemini 2.5 Pro.

上下文
1.0M
发布日期
2025年4月

Qwen3 32B

中国

Alibaba's Qwen3 32B dense language model with strong reasoning and multilingual capabilities, supporting function calling and code generation across diverse tasks.

上下文
131K
发布日期
2025年4月

Command A

美国

Cohere's flagship 111B parameter model optimized for demanding enterprises requiring fast, secure, and high-quality AI. Excels at RAG, tool use, and multilingual tasks with strong reasoning capabilities.

上下文
256K
发布日期
2025年3月

DeepSeek R1

中国

DeepSeek 面向推理的模型,通过强化学习训练以处理复杂的多步骤推理任务。尤其擅长需要思维链推理的数学、科学和编程问题。

上下文
131K
发布日期
2025年1月

DeepSeek V3

中国

DeepSeek's third-generation large language model featuring mixture-of-experts architecture, strong multilingual capabilities, and competitive performance on reasoning and coding benchmarks.

上下文
128K
发布日期
2024年12月