模型
浏览来自所有提供商的 9 个标准化 LLM 模型
部分描述为试点机器翻译内容,尚未经过人工审核。
GPT-5.6 Sol
OpenAI GPT-5.6 系列旗舰模型,在提升可靠性和效率的同时,推进了编程、科学推理、长周期规划和智能体工作流能力。新增最高推理强度设置,以及可为复杂多步骤任务启动子智能体的 ultra 模式。
Gemma 4 12B
Google's medium-size open-weight model with 12 billion parameters from the Gemma 4 family. Encoder-free unified multimodal architecture that natively processes text, image, audio, and video inputs without dedicated encoders. Features a 256K context window and supports 140+ languages. First medium-sized model capable of natively ingesting audio. Suitable for local deployment on GPUs with 16GB VRAM.
Gemini 3.1 Flash-Lite
Google's most cost-efficient Gemini model optimized for high-volume, low-latency use cases. Delivers 2.5x faster time to first token versus Gemini 2.5 Flash with full multimodal support. Ideal for agentic tasks, data extraction, translation, and classification.
Gemini 3 Flash
Google's balanced model combining Gemini 3 Pro's reasoning capabilities with the Flash line's latency, efficiency, and cost. Features configurable thinking levels, multimodal function responses, and streaming function calling for complex agentic workflows.
Gemini 3.5 Flash
Google DeepMind's balanced Gemini 3.5 model that pairs Pro-line reasoning quality with Flash-line latency and cost. Natively multimodal across text, image, audio, and video with a 1M-token context window, configurable thinking levels, and streaming function calling, tuned for high-throughput production workloads.
Gemini 3.1 Pro
Google 最新的旗舰多模态模型,在推理、编程和多模态理解方面达到领先水平。支持原生工具使用、信息溯源以及百万 token 上下文窗口。
Gemini 2.5 Flash
Google's cost-effective model optimized for high throughput tasks. Balances speed and intelligence with strong multimodal capabilities and 1M token context window.