模型

浏览来自所有提供商的 39 个标准化 LLM 模型

显示第 25–39 项,共 39 个模型

Devstral 2

法国

Mistral AI's frontier code agents model designed for solving software engineering tasks. Open-weight model optimized for agentic coding workflows and complex development tasks.

上下文
128K
发布日期
2025年12月

Mistral Large 3

法国

Mistral AI's largest open-weight model with 41B active parameters (675B total MoE). State-of-the-art general-purpose multimodal model with 256K context window and powerful agentic capabilities. Released under Apache 2.0.

上下文
256K
发布日期
2025年12月

AlemLLM

哈萨克斯坦

Kazakhstan's flagship Mixture-of-Experts language model developed by Astana Hub with technical support from 01.AI. Features 247B total parameters with 22B active per token, achieving state-of-the-art results on Kazakh, Russian, and English benchmarks. Outperforms GPT-4o on Kazakh language tasks.

上下文
131K
发布日期
2025年8月

Trendyol LLM 8B T1

土耳其

Turkish-optimized 8B chat model developed by Trendyol, Turkey's largest e-commerce platform. Built on Qwen3-8B and fine-tuned on large-scale Turkish e-commerce datasets. Features advanced chain-of-thought reasoning in Turkish with dual operation modes (/think and /no_think), strong instruction following, summarization, coding, and attribute extraction for catalogue enrichment. English reasoning capabilities are preserved alongside Turkish.

上下文
33K
发布日期
2025年7月

YandexGPT 5 Lite

俄罗斯

Yandex's compact 8B parameter language model trained on 15T tokens of primarily Russian and English text. Features 32K context window with strong performance on web, code, and mathematics tasks. Open-weight release.

上下文
32K
发布日期
2025年6月

Qwen3 Coder

中国

Alibaba's Qwen3 Coder model optimized for software development tasks including code generation, debugging, code review, and technical documentation with strong multilingual programming support.

上下文
131K
发布日期
2025年4月

Qwen3 235B

中国

Alibaba's Qwen3 235B mixture-of-experts model delivering frontier-level performance with advanced reasoning, function calling, and code generation capabilities at massive scale.

上下文
131K
发布日期
2025年4月

Qwen3 32B

中国

Alibaba's Qwen3 32B dense language model with strong reasoning and multilingual capabilities, supporting function calling and code generation across diverse tasks.

上下文
131K
发布日期
2025年4月

Gemma 3 1B

美国

Google's lightweight open-weight model with 1 billion parameters from the Gemma 3 family. Designed for on-device and resource-constrained deployments. Supports text-only tasks with a 32K context window. Efficient for chat and basic completion workloads.

上下文
33K
发布日期
2025年3月

Gemma 3 12B

美国

Google's mid-size open-weight model with 12 billion parameters from the Gemma 3 family. Supports multimodal inputs including text and images with a 128K context window. Strong performance on reasoning and code generation tasks at moderate compute cost.

上下文
131K
发布日期
2025年3月

Gemma 3 27B

美国

Google's largest open-weight model in the Gemma 3 family with 27 billion parameters. Supports multimodal inputs including text and images with a 128K context window. Delivers strong performance across reasoning, code generation, and vision tasks, competitive with larger proprietary models.

上下文
131K
发布日期
2025年3月

Gemma 3 4B

美国

Google's compact open-weight model with 4 billion parameters from the Gemma 3 family. Supports multimodal inputs including text and images with a 128K context window. Balances efficiency and capability for vision and language tasks.

上下文
131K
发布日期
2025年3月

Mistral Small 3.1

法国

Mistral AI's Small 3.1 model with 24B parameters offering efficient multimodal capabilities including vision, function calling, and code generation with a large 128K context window.

上下文
128K
发布日期
2025年3月

QwQ 32B

中国

Alibaba's QwQ 32B reasoning-focused model designed for complex problem solving, mathematical reasoning, and step-by-step logical analysis with strong chain-of-thought capabilities.

上下文
131K
发布日期
2025年3月

Cotype Nano

俄罗斯

MTS AI's lightweight 1.5B parameter language model optimized for resource-constrained environments. Excels at Russian and English language tasks including content creation, translation, and text analysis. Runs efficiently on both CPU and GPU, including laptops and smartphones.

上下文
32K
发布日期
2024年11月