Modeller

Tüm sağlayıcılardaki 96 kanonik LLM modelini inceleyin

Bazı açıklamalar otomatik çevrilmiş pilot içeriklerdir ve henüz editör tarafından incelenmemiştir.

96 modelden 25–48 arası gösteriliyor

Gemma 4 12B

Amerika Birleşik Devletleri

Google's medium-size open-weight model with 12 billion parameters from the Gemma 4 family. Encoder-free unified multimodal architecture that natively processes text, image, audio, and video inputs without dedicated encoders. Features a 256K context window and supports 140+ languages. First medium-sized model capable of natively ingesting audio. Suitable for local deployment on GPUs with 16GB VRAM.

Bağlam
262K
Yayımlanma
Haz 2026

Nemotron 3 Ultra

Amerika Birleşik Devletleri

NVIDIA's flagship open 550B-parameter Mixture-of-Experts model with 55B active parameters, built for frontier reasoning and orchestration in long-running agentic systems. Features hybrid Mamba-Transformer architecture, LatentMoE routing, multi-token prediction, and NVFP4 precision for 5x higher throughput. Achieves 30% lower cost-to-task-completion on agentic benchmarks. Supports 1M+ token context window with 95% accuracy on Ruler@1M.

Bağlam
1.0M
Yayımlanma
Haz 2026

MiniMax M3

Çin

MiniMax's frontier open-weight model with 1M-token context window, native multimodality (text, image, video), and strong coding capabilities. Built on MiniMax Sparse Attention (MSA) architecture, achieving 59% on SWE-Bench Pro with significantly improved efficiency at long context.

Bağlam
1.0M
Yayımlanma
Haz 2026

Claude Opus 4.8

Amerika Birleşik Devletleri

Anthropic'in en gelişmiş modeli; Opus 4.7'nin üzerine kodlama, agent yetenekleri, akıl yürütme ve bilgi çalışmasında daha güçlü performans ekler. Daha dürüst yanıtlar, daha verimli araç kullanımı, dinamik iş akışı desteği ve iyileştirilmiş uyum sunar.

Bağlam
300K
Yayımlanma
May 2026

Falcon-H1

Birleşik Arap Emirlikleri

TII's hybrid Mamba-Transformer model that outperforms comparable offerings from Meta's Llama and Alibaba's Qwen in the 30-70B parameter range. Designed for real-world AI on everyday devices and resource-limited settings with state-of-the-art efficiency.

Bağlam
131K
Yayımlanma
May 2026

Yi-Lightning

Çin

01.AI's flagship large language model with enhanced Mixture-of-Experts architecture. Ranked 6th on Chatbot Arena with particularly strong results in Chinese, Math, Coding, and Hard Prompts categories. Features advanced expert segmentation and optimized KV-caching.

Bağlam
131K
Yayımlanma
May 2026

Ring-2.6-1T

Çin

InclusionAI's (Ant Group) trillion-parameter open-weights reasoning model with 63B active parameters per token. Built for real-world agent workflows with adaptive reasoning-effort modes. Features hybrid linear and MLA attention architecture with MIT license.

Bağlam
131K
Yayımlanma
May 2026

Palmyra X5

Amerika Birleşik Devletleri

Writer's most advanced adaptive reasoning model with a 1 million token context window. Processes full million-token prompts in approximately 22 seconds with multi-turn function calls in 300ms. Optimized for enterprise agentic AI workflows at 3-4x lower cost than GPT-4.1.

Bağlam
1.0M
Yayımlanma
May 2026

Jamba Large 1.7

İsrail

AI21's latest hybrid SSM-Transformer model with Mixture-of-Experts architecture. Features a 256K context window, improved grounding and instruction-following. 94B total parameters with 398B active, optimized for enterprise long-context tasks.

Bağlam
262K
Yayımlanma
May 2026

Solar Pro 3

Güney Kore

Upstage's powerful Mixture-of-Experts language model with 102B total parameters and 12B active parameters per forward pass. Optimized for Korean with strong English and Japanese support. Excels at complex reasoning, structured output generation, and agentic workflows.

Bağlam
128K
Yayımlanma
May 2026

K2 Think

Birleşik Arap Emirlikleri

A 32 billion parameter open-weights reasoning model by LLM360/MBZUAI, built on Qwen2.5-32B. Trained with reinforcement learning and verifiable rewards for long chain-of-thought reasoning, agentic planning, and complex problem solving in math, science, and code.

Bağlam
131K
Yayımlanma
May 2026

Qwen 3.7 Max

Çin

Alibaba's flagship proprietary model engineered for advanced agentic coding, complex reasoning, and long-horizon task execution. Ranked

Bağlam
131K
Yayımlanma
May 2026

Qwen 3.7 Plus

Çin

Alibaba's multimodal variant in the Qwen 3.7 family, optimized for vision understanding and multimodal tasks. Ranked

Bağlam
131K
Yayımlanma
May 2026

Gemini 3.5 Flash

Amerika Birleşik Devletleri

Google DeepMind's balanced Gemini 3.5 model that pairs Pro-line reasoning quality with Flash-line latency and cost. Natively multimodal across text, image, audio, and video with a 1M-token context window, configurable thinking levels, and streaming function calling, tuned for high-throughput production workloads.

Bağlam
1.0M
Yayımlanma
May 2026

Gemini 3 Flash

Amerika Birleşik Devletleri

Google's balanced model combining Gemini 3 Pro's reasoning capabilities with the Flash line's latency, efficiency, and cost. Features configurable thinking levels, multimodal function responses, and streaming function calling for complex agentic workflows.

Bağlam
1.0M
Yayımlanma
May 2026

Granite 4.1 30B

Amerika Birleşik Devletleri

IBM's largest dense decoder-only 30B parameter language model from the Granite 4.1 family. Trained on approximately 15T tokens with long-context extension up to 512K tokens. Supports tool calling, RAG, code generation, multilingual tasks across 12 languages. Released under Apache 2.0.

Bağlam
524K
Yayımlanma
May 2026

Laguna M.1

Amerika Birleşik Devletleri

Poolside AI's flagship agentic coding model with 225B total parameters and 23B active (MoE). Trained from scratch in-house on 30T tokens across 6,144 NVIDIA Hopper GPUs. Optimized for complex multi-step software engineering tasks including codebase exploration, file editing, test running, and iterative debugging.

Bağlam
128K
Yayımlanma
Nis 2026

GPT-5.5

Amerika Birleşik Devletleri

OpenAI's most capable model designed for complex real-world work including coding, online research, information analysis, and document creation. Features advanced agentic capabilities with tool search and multi-step task execution.

Bağlam
1.0M
Yayımlanma
Nis 2026

DeepSeek V4 Pro

Çin

DeepSeek's flagship V4 model with 1.6T total parameters (49B activated). MoE architecture supporting 1M token context. Closes the gap with frontier proprietary models on reasoning and coding benchmarks.

Bağlam
1.0M
Yayımlanma
Nis 2026

Qwen 3.6 27B

Çin

Alibaba's dense 27B parameter model that outperforms its own 397B MoE predecessor on agentic coding benchmarks. Strong multilingual and reasoning capabilities released under Apache 2.0.

Bağlam
131K
Yayımlanma
Nis 2026

Hy3 Preview

Çin

Tencent's flagship open-weight Mixture-of-Experts model from the Hunyuan family with 295B total parameters and 21B active. Integrates fast and slow thinking modes with configurable reasoning effort. Designed for agentic workflows, cross-file code refactoring, long-document analysis, and multi-step tool use.

Bağlam
256K
Yayımlanma
Nis 2026

MiMo-V2.5-Pro

Çin

Xiaomi's flagship 1.02T-parameter Mixture-of-Experts model with 42B active parameters, built on a hybrid-attention architecture with 3-layer Multi-Token Prediction. Designed for complex agentic tasks, software engineering, and long-horizon instruction following with a 1M-token context window.

Bağlam
1.0M
Yayımlanma
Nis 2026

Qwen 3.6 35B-A3B

Çin

Alibaba's efficient Mixture-of-Experts model with 35B total parameters and 3B active per token. Frontier-level agentic coding performance with 73.4% on SWE-bench Verified and 92.7 on AIME 2026. Released under Apache 2.0.

Bağlam
131K
Yayımlanma
Nis 2026

GPT-5.4 Mini

Amerika Birleşik Devletleri

OpenAI's compact reasoning model optimized for coding, computer use, and subagent tasks. Approaches GPT-5.4 performance on several benchmarks while running more than 2x faster.

Bağlam
1.1M
Yayımlanma
Nis 2026