覆盖整个生态系统的提供商
127
模型
52
提供商
199
提供商映射
$8.04
每百万 Token 平均价格
热门模型
按相关性和提供商可用性排序的热门模型
Claude Fable 5
Anthropic's first publicly available Mythos-class model, exceeding the capabilities of any model the company has previously made generally available. State-of-the-art on nearly all tested benchmarks, with exceptional performance in software engineering, knowledge work, vision, and scientific research. Its lead grows on longer and more complex tasks. Ships with built-in safeguards that route sensitive cybersecurity, biology, chemistry, and distillation queries to Claude Opus 4.8.
GPT-5.6 Terra
The balanced tier of OpenAI's GPT-5.6 series, trading a small amount of peak quality for markedly lower latency and cost. Retains strong reasoning, coding, and agentic tool use with configurable reasoning effort, making it a default choice for production workloads that need frontier capability at scale.
Grok 4.5
xAI's strongest model to date, built to excel at coding, agentic tasks, and knowledge work and co-developed alongside coding tools for real-world software engineering. Features real-time information access, extended reasoning, and large-context tool use with an OpenAI-compatible API.
Claude Opus 4.8
Anthropic's most advanced model, building on Opus 4.7 with improvements across benchmarks in coding, agentic skills, reasoning, and knowledge work. Features enhanced honesty, better tool use efficiency, dynamic workflows support, and improved alignment.
GPT-5.5 Pro
OpenAI's premium tier model with extended reasoning capabilities, higher accuracy on complex tasks, and priority access. Optimized for professional and enterprise workloads requiring maximum quality.
Gemini 3.1 Pro
Google's latest flagship multimodal model with state-of-the-art performance on reasoning, coding, and multimodal understanding. Features native tool use, grounding, and million-token context window.
DeepSeek R1
DeepSeek's reasoning-focused model trained with reinforcement learning for complex multi-step reasoning. Excels at math, science, and coding problems requiring chain-of-thought reasoning.
DeepSeek V4
DeepSeek's fourth-generation model with improved mixture-of-experts architecture, enhanced reasoning and coding capabilities, and stronger multilingual performance. Competitive with frontier proprietary models.
Nemotron 3 Ultra
NVIDIA's flagship open 550B-parameter Mixture-of-Experts model with 55B active parameters, built for frontier reasoning and orchestration in long-running agentic systems. Features hybrid Mamba-Transformer architecture, LatentMoE routing, multi-token prediction, and NVFP4 precision for 5x higher throughput. Achieves 30% lower cost-to-task-completion on agentic benchmarks. Supports 1M+ token context window with 95% accuracy on Ruler@1M.
Gemma 4 31B
Google's flagship open-weight dense model with 31B parameters. All parameters active per forward pass. Ranks among top open models with strong performance on AIME 2026 (89.2%) and MMLU Pro (85.2%). Supports vision and extended context.
最新洞察
LLM 生态系统的分析、基准与比较
此内容目前仅提供英文版本。

GPT-5.6 and ChatGPT Work: From AI Assistant to AI Worker
OpenAI is no longer positioning ChatGPT as a conversational assistant. With GPT-5.6 and ChatGPT Work, the company is moving toward a full work execution layer across apps, files, code, and business workflows.

The AI Race Is Shifting From IQ to Agentic Economics
The AI race is shifting from benchmark scores to agentic economics. Why inference costs, latency, and open-weight models are reshaping the industry in 2026.

Stanford AI Index 2026: AI Is Scaling Faster Than Society Can Adapt
The release of the 2026 AI Index Report by Stanford HAI paints a very clear picture: artificial intelligence is no longer an emerging technology — it has become global infrastructure.