Providers¶
Companies and platforms that offer LLM inference, APIs, managed AI services, foundation models, and neural retrieval.
For a cross-provider pricing and free-tier tracker, see API Pricing & Free Tier Matrix.
Contents¶
| Provider | What it offers |
|---|---|
| Anthropic | Claude model family, API, and safety-focused AI research |
| Aurora80K | Aurora80K Release |
| AWS Bedrock | Managed foundation model platform from multiple AI providers on AWS |
| Azure OpenAI | Enterprise-grade OpenAI models with Azure enterprise security and compliance |
| Baseten | Serverless inference platform for deploying and scaling custom open LLMs and diffusion models |
| BigSwitch | Directory of European tech alternatives and localized sovereign providers |
| Codestral | Mistral AI's flagship open-weights generative model family optimized for code tasks |
| Cohere | Enterprise-focused LLMs (Command R series) and specialized enterprise RAG models |
| DeepSeek | High-performance open-weights reasoning models (V3, R1) from DeepSeek |
| Exa AI | Neural search engine and API built specifically for AI agents and LLM retrieval |
| EXAONE | LG AI Research's open-weights foundation model ecosystem |
| Fireworks AI | Low-latency API inference platform for popular open-source foundation models |
| Gemini | Google's state-of-the-art multimodal model family and Vertex AI |
| GLM / Zhipu AI | High-performance GLM model family featuring long context and multi-lingual capabilities |
| Groq | Ultra-high-speed LLM inference API powered by LPU architecture |
| Hugging Face | The primary central hub for open models, datasets, space demos, and inference endpoints |
| InternLM | Open-source multilingual foundation model family developed by Shanghai AI Laboratory |
| Kat Coder Air | Specialized low-latency coding model optimized for real-time completion |
| LFM Encoders | Liquid AI's non-transformer dynamical architecture encoders for sequence modeling |
| Ling3.0 | Ling3.0 Release |
| Microsoft Graph API | Unified enterprise API gateway for Microsoft 365 data and services |
| MiniMax | Coding-optimized foundation models (M2.5) and managed multimodal AI platform |
| Mistral AI | Pioneer of high-efficiency open and commercial foundation models |
| Monolith | Large-scale recommendation system framework and inference platform |
| Moonshot AI | Kimi models featuring pioneering long-context processing capabilities |
| NVIDIA | High-performance microservices (NIM), GPU infrastructure, and model catalog |
| OpenAI | Industry-leading GPT model family and developer ecosystem |
| OpenPangu | Huawei's industrial-grade foundation model series for multimodal and scientific domain tasks |
| OpenRouter | Unified API gateway for access to 100+ open and proprietary models |
| Perplexity | AI-powered search engine and retrieval API with real-time web access |
| Poolside | AI research company building frontier foundation models for software engineering |
| Portkey AI Gateway | Control plane for production AI apps providing routing, tracing, and caching |
| Replicate | Cloud platform for running open-source machine learning models via API |
| SenseNova U15Lite | SenseNova U15Lite Release |
| Solar Pro | Upstage's high-efficiency frontier LLM model family for specialized enterprise inference |
| SOOFI | Domain-specific AI inference provider targeting sovereign data operations |
| Tavily | Search and information extraction API optimized specifically for autonomous AI agents |
| Together AI | High-performance cloud inference platform for large-scale open foundation models |
| Vercel AI Gateway | Streamlined interface and caching gateway across multiple AI providers |
| xAI Grok | Grok model family focusing on real-time search grounding and reasoning |
Related tools / concepts¶
- API Pricing & Free Tier Matrix
- Model Routing Guide
- OpenRouter Logs Backlog
- Infrastructure
- Frameworks
Contribution Metadata¶
- Last reviewed: 2027-01-07
- Confidence: high