Модели

4 канонические LLM-модели от всех провайдеров

Показаны модели 1–4 из 4

Gemini 3.1 Flash-Lite

Соединенные Штаты

Google's most cost-efficient Gemini model optimized for high-volume, low-latency use cases. Delivers 2.5x faster time to first token versus Gemini 2.5 Flash with full multimodal support. Ideal for agentic tasks, data extraction, translation, and classification.

Контекст
1.0M
Добавлена
май 2026 г.

Gemini 3.5 Flash

Соединенные Штаты

Google DeepMind's balanced Gemini 3.5 model that pairs Pro-line reasoning quality with Flash-line latency and cost. Natively multimodal across text, image, audio, and video with a 1M-token context window, configurable thinking levels, and streaming function calling, tuned for high-throughput production workloads.

Контекст
1.0M
Добавлена
май 2026 г.

Gemini 3 Flash

Соединенные Штаты

Google's balanced model combining Gemini 3 Pro's reasoning capabilities with the Flash line's latency, efficiency, and cost. Features configurable thinking levels, multimodal function responses, and streaming function calling for complex agentic workflows.

Контекст
1.0M
Добавлена
май 2026 г.

Gemini 2.5 Flash

Соединенные Штаты

Google's cost-effective model optimized for high throughput tasks. Balances speed and intelligence with strong multimodal capabilities and 1M token context window.

Контекст
1.0M
Добавлена
июнь 2025 г.