Proveedores compatibles con OpenAI

Proveedores de inferencia con una API compatible con OpenAI: cambia la URL base y conserva tu SDK.

41 proveedores

01.AI

Chinese AI company founded by Kai-Fu Lee, developing the Yi family of large language models. Offers Yi-Lightning and other models through their API platform with strong performance in Chinese, math, and coding tasks.

Compatible con OpenAIbearer

Alibaba Model Studio

Alibaba Cloud AI platform providing access to Qwen family models and other large language models through DashScope API, offering OpenAI-compatible endpoints with multi-region availability.

Compatible con OpenAIapi-key

Amazon Bedrock

Fully managed AWS service offering foundation models from leading AI companies including Anthropic, Meta, and Mistral through a unified API. Supports OpenAI-compatible endpoints via Mantle inference engine with cross-region inference capabilities.

Compatible con OpenAIbearer

Anyscale

Serverless inference platform built on Ray, offering high-throughput access to popular open-weight models. OpenAI-compatible API with competitive pricing and enterprise-grade scalability.

Compatible con OpenAIbearer

Azure AI

Microsoft's cloud AI platform providing access to OpenAI models, open-source models, and enterprise AI services through Azure AI Studio. Offers global deployment with enterprise-grade security and compliance.

Compatible con OpenAIapi-key

Baseten

ML infrastructure platform for deploying and serving machine learning models at scale, offering managed GPU inference with auto-scaling and OpenAI-compatible API endpoints for popular LLMs.

Compatible con OpenAIapi-key

Cerebras

Ultra-fast inference provider powered by the Cerebras Wafer Scale Engine, known for extremely high tokens/sec throughput. Offers OpenAI-compatible API with free tier access.

Compatible con OpenAIapi-key

Cloudflare Workers AI

Serverless AI inference platform running on Cloudflare's global edge network. Offers 10,000 neurons/day free allocation with OpenAI-compatible API and wide model selection including vision models.

Compatible con OpenAIapi-key

Deep Infra

Serverless inference platform offering fast and cost-effective access to popular open-weight models. OpenAI-compatible API with pay-per-token pricing and no minimum commitments.

Compatible con OpenAIbearer

DeepSeek

Chinese AI research company providing direct API access to their DeepSeek family of models with competitive pricing and strong performance on coding and reasoning tasks.

Compatible con OpenAIbearer

Featherless

Serverless LLM inference platform hosting 20,000+ open-source models from Hugging Face with flat-rate subscription pricing and unlimited token usage. OpenAI-compatible API with no per-token billing — access any model up to a given size based on subscription tier. Largest Hugging Face inference provider by model count.

Compatible con OpenAIbearer

Fireworks

High-speed AI inference platform optimized for low-latency serving of open-source models, offering OpenAI-compatible API endpoints with custom model deployment and fine-tuning capabilities.

Compatible con OpenAIapi-key

Google AI Studio

Google's developer platform for accessing Gemini and Gemma models via OpenAI-compatible API. Free tier available with generous rate limits. Data may be used for training outside EU/EEA/UK/CH regions.

Compatible con OpenAIapi-key

Groq

Ultra-fast LPU (Language Processing Unit) inference provider offering extremely low latency. Supports streaming, function calling, and audio transcription via Whisper models. Per-model rate limits apply.

Compatible con OpenAIbearer

Hugging Face Inference

Inference API and dedicated endpoints for open-source models hosted on the Hugging Face Hub. Offers serverless inference for popular models and dedicated GPU endpoints for production workloads.

Compatible con OpenAIbearer

Hyperbolic

Decentralized AI compute platform offering affordable GPU inference for open-source models, providing OpenAI-compatible API endpoints with competitive pricing and global availability.

Compatible con OpenAIapi-key

InclusionAI

Ant Group's AI research lab focused on open-source large language models. Offers inference via ZenMux platform with OpenAI-compatible API. Develops the Ling (non-thinking) and Ring (reasoning) model families at trillion-parameter scale.

Compatible con OpenAIbearer

Inference.net

Distributed AI inference network providing affordable access to open-source language models through a decentralized GPU marketplace, offering OpenAI-compatible API endpoints with competitive per-token pricing.

Compatible con OpenAIapi-key

Lambda

GPU cloud and inference provider offering on-demand access to NVIDIA GPUs for AI training and inference. Provides both cloud instances and managed inference API for open-source LLMs with competitive pricing.

Compatible con OpenAIbearer

MiniMax

Chinese AI company providing large language models with strong multilingual and multimodal capabilities. Known for competitive pricing and high-quality text generation.

Compatible con OpenAIapi-key

Mistral AI

European AI company providing high-performance language models with strong multilingual capabilities. Offers both open-weight and proprietary models through an OpenAI-compatible API.

Compatible con OpenAIbearer

Modal

Serverless cloud platform for running AI workloads with on-demand GPU access, offering custom model deployment and OpenAI-compatible inference endpoints with automatic scaling and pay-per-second pricing.

Compatible con OpenAIbearer

Moonshot AI

Chinese AI company behind the Kimi series of models, known for ultra-long context windows and strong reasoning capabilities. Offers OpenAI-compatible API access.

Compatible con OpenAIapi-key

Nebius

Cloud AI platform providing scalable GPU infrastructure and managed inference services for large language models, with data centers in Europe and competitive pricing for open-source model hosting.

Compatible con OpenAIapi-key

Novita

AI model inference platform providing affordable access to open-source language models with OpenAI-compatible API endpoints, offering pay-per-token pricing and global availability.

Compatible con OpenAIapi-key

NVIDIA NIM

NVIDIA's inference microservice platform providing optimized deployment of LLMs on GPU infrastructure. Offers free endpoints for select models and partner endpoints through Deep Infra, Together AI, Bitdeer, GMI Cloud, and CoreWeave.

Compatible con OpenAIapi-key

OpenAI

Leading AI research company providing API access to GPT-4, GPT-3.5, DALL-E, and other foundation models through a developer-friendly REST API with global availability.

Compatible con OpenAIbearer

OpenRouter

Unified API gateway providing access to hundreds of models from multiple providers through a single OpenAI-compatible endpoint. Free models available with shared quota, up to 1000 requests/day.

Compatible con OpenAIapi-key

Perplexity

AI-powered answer engine offering API access to proprietary and open-source models with built-in web search grounding. Specializes in providing accurate, cited responses with real-time information access.

Compatible con OpenAIbearer

Sakana AI

Tokyo-based AI research lab building nature-inspired and evolutionary approaches to foundation models. Provides the Fugu family of multi-agent orchestration models through a single OpenAI-compatible API that coordinates a pool of specialist LLMs behind one endpoint, available via pay-as-you-go and subscription plans.

Compatible con OpenAIapi-key

SambaNova

AI hardware and software platform offering high-performance inference services powered by custom DataScale systems, providing OpenAI-compatible API endpoints for open-source models with industry-leading throughput.

Compatible con OpenAIapi-key

Sarvam AI

Indian AI company building sovereign language models and a full-stack GenAI platform optimized for Indian languages. Provides an OpenAI-compatible API for the Sarvam model family (Sarvam-M, Sarvam-1, Sarvam-30B, Sarvam-105B) along with speech and translation services, hosted on India-based infrastructure.

Compatible con OpenAIapi-key

Scaleway

European cloud provider offering managed AI inference endpoints with GPU instances across European data centers, providing OpenAI-compatible API access to popular open-source models.

Compatible con OpenAIapi-key

SiliconFlow

High-performance AI inference platform offering ultra-low latency and cost-effective access to open-source models. Supports models with up to 1M token context windows and OpenAI-compatible API endpoints.

Compatible con OpenAIbearer

Tinker

Thinking Machines Lab's cloud service for fine-tuning and inference. Its OpenAI-compatible inference API is currently in beta and intended for testing, evaluation, and internal workflows; production-grade inference is not yet generally available.

Compatible con OpenAIapi-key

Together AI

Cloud platform for running and fine-tuning open-source AI models, offering competitive pricing and OpenAI-compatible API endpoints for popular open-weight models.

Compatible con OpenAIbearer

Upstage

South Korean AI company providing enterprise-grade language models optimized for Korean, English, and Japanese. Offers the Solar model family through a direct API with competitive pricing and high throughput.

Compatible con OpenAIapi-key

xAI

AI company founded by Elon Musk providing the Grok family of models. Known for real-time information access and strong reasoning capabilities with OpenAI-compatible API.

Compatible con OpenAIbearer

Xiaomi MiMo

Xiaomi's AI inference platform providing access to the MiMo family of models via an OpenAI-compatible API endpoint. Offers flagship agentic and multimodal models with competitive pricing.

Compatible con OpenAIapi-key

Yandex Cloud

Russian cloud platform providing access to YandexGPT foundation models through Yandex Cloud AI Studio. Offers OpenAI-compatible API endpoints with strong Russian and English language capabilities.

Compatible con OpenAIapi-key

Zhipu AI

Chinese AI company providing the GLM family of models with strong bilingual (Chinese/English) capabilities. Known for competitive performance on reasoning and coding benchmarks.

Compatible con OpenAIapi-key