AI Model Providers
Every AI model provider tracked by CloudPrice - frontier labs, inference platforms and self-hosted open-weights shortcuts. Click into any provider to see their full catalog with live pricing, benchmarks and capabilities.
OpenRouter
The unified interface for every model. Find the best models & prices for your prompts
Vercel AI Gateway
Everything you need to build, run, and govern frontier agents. From the framework to model routing, secure execution, and identity.
Fireworks AI
Fireworks’ state of the art training and inference platform take you beyond the frontier, transforming open models into your specialized intelligence.
Azure AI Foundry
Microsoft Foundry
Google Gemini
Build with Gemini 2.0 Flash, 2.5 Pro, and Gemma using the Gemini API and Google AI Studio.
DeepInfra
DeepInfra offers cost-effective, scalable, easy-to-deploy, and production-ready machine-learning models and infrastructures for deep-learning models.
Google Vertex AI
Gemini Enterprise Agent Platform (formerly Vertex AI) is a comprehensive platform for developers to build, scale, govern and optimize agents.
Alibaba Qwen
Supercharge Your AI Journey Effortlessly With Industry-Leading GenAI Models
Novita
Novita AI provides 200+ Model APIs, custom deployment, GPU Instances, and Serverless GPUs. Scale AI, optimize performance, and innovate with ease and efficiency.
Amazon Bedrock
Amazon Bedrock: The platform for building generative AI applications and agents at production scale
OpenAI
Creator of GPT-4o, o3, and the GPT model family. Offers text, vision, audio, image generation, speech, and embedding models via a REST API. Pioneered the modern LLM API interface now widely adopted as the de-facto standard.
Together AI
Build what's next on the AI Native Cloud. Full-stack AI platform for inference, fine-tuning, and GPU clusters — powered by cutting-edge research.
Hugging Face
We’re on a journey to advance and democratize artificial intelligence through open source and open science.
Snowflake
Snowflake's managed AI layer running inside the Snowflake Data Cloud. Cortex LLM Functions give SQL-native access to Llama, Mistral, Snowflake Arctic, and other models without data leaving the Snowflake environment. Arctic is Snowflake's own enterprise-focused open-weight model. REST API follows the OpenAI specification.
Cloudflare Workers AI
Run machine learning models, powered by serverless GPUs, on Cloudflare's global network.
Perplexity
AI company best known for its search assistant. Also offers the Sonar model family via API: Sonar (fast, grounded), Sonar Pro (more capable), and Sonar Reasoning (chain-of-thought). All models include real-time web search grounding by default.
Aihubmix
Databricks
Databricks offers a unified platform for data, analytics and AI. Build better AI with a data-centric approach. Simplify ETL, data warehousing, governance and AI on the Data + AI Platform.
Nebius
Build and scale faster on the purpose-built AI cloud, engineered from silicon to API.
Mistral AI
The most powerful AI platform for enterprises. Customize, fine-tune, and deploy AI assistants, autonomous agents, and multimodal AI with open models.
xAI
Elon Musk's AI company and creators of the Grok model family. Grok models offer large context windows, real-time knowledge via X (Twitter) integration, vision, and multimodal output. Grok-4 is their frontier reasoning model.
Deepgram
Power enterprise voice solutions with Deepgram’s Speech-to-Text, Text-to-Speech, and Voice Agent APIs. Real-time, accurate, and built for scale.
Oracle Cloud (OCI)
Transform your business with generative AI, and unlock a new era of productivity with task automation and end-to-end AI solutions for enterprise customers.
Replicate
Run open-source machine learning models with a cloud API
Cohere
Cohere builds powerful models and AI solutions enabling enterprises to automate processes, empower employees, and turn fragmented data into actionable insights.
Qwen AI Platform
Qwencloud
IBM watsonx
IBM watsonx is a portfolio of AI products that accelerates the impact of generative AI in core workflows to drive productivity.
Voyage
Voyage AI provides cutting-edge embedding models and rerankers for search and retrieval
Weights & Biases
GitHub Models
GitHub is where people build software. More than 150 million people use GitHub to discover, fork, and contribute to over 420 million projects.
Baseten
fal.ai
Easiest & most cost-effective way to use Gen AI. fal.ai is how devs integrate dozens of generative media models. FLUX, Kling, Hailuo +1000 more
Stability AI
Stability AI is the enterprise-ready creative partner for teams and creators, delivering professional-grade generative AI tools and solutions for content production at scale.
Groq
Groq is the premier neocloud for fast inference. One fully integrated platform for infrastructure, inference, and control. Millions of developers run trillions of tokens on Groq every week.
Lambda
Train and scale AI on NVIDIA VR200 NVL 72, GB300 NVL 72, B300, B200, H200, H100, and and more GPUs. Launch on-demand instances or reserve a cluster. Get started.
Anthropic
Anthropic is an AI safety and research company that's working to build reliable, interpretable, and steerable AI systems.
Moonshot AI (Kimi)
Chinese AI company creator of the Kimi model family. Kimi k1.5 and k2 are strong reasoning models with long context. Kimi VL handles vision tasks. Known for efficient long-document processing. Direct API available globally via Moonshot's platform.
SambaNova
Discover SambaNova - the complete AI platform delivering the fastest AI inference, fine-tuning, and scalable solutions for agentic AI easily integrated into existing data center infrastructures.
Scaleway
European cloud and GPU-inference provider offering serverless LLM inference via an OpenAI-compatible API with GDPR-compliant EU data residency. Hosts open-source models from Meta, Qwen, Mistral, and others.
GMI Cloud
Singapore-based GPU cloud offering serverless inference for a broad catalog of open-weight and frontier models. Hosts Qwen, MiniMax, DeepSeek, Llama, and others with OpenAI-compatible endpoints. Focuses on Asia-Pacific availability and competitive pricing.
Hyperbolic
Hyperbolic is the Open-Access AI Cloud. Get on-demand H100, H200 & B200 GPUs and reserve dedicated multi-node clusters for production-scale AI training and fine-tuning.
Nscale
Nscale full-stack AI cloud platform and services are designed for scale, resilience, and speed.
Runway
Runway is building foundational Real-World Intelligence that can understand, simulate and act in the world. We offer products and services built on-top of this intelligence to empower individuals and organizations to do more in the world.
Z AI (Zhipu)
Meet Z.ai, the AI assistant powered by GLM-5.3-Flash. Build websites, write code, handle long-horizon tasks, and get instant answers. Fast, smart, and reliable.
LlamaGate
OVHcloud
Discover the generative AI APIs offered by OVHcloud. High-performing, easy to integrate, and secure, for application power.
AI/ML API
Access 1000+ AI models through a single API for text, image, video, audio, code, embeddings, and more. Compare pricing, latency, and capabilities—all in one place with AIMLAPI.
Gradient AI
Hyperagent is the system of agents that does real work, learns how your organization operates, and deploys across your entire team.
Volcengine (ByteDance)
ByteDance's cloud platform offering the Doubao model family and third-party models via the Ark inference service. Doubao-Pro and Doubao-Lite cover a range of cost and capability trade-offs. Dominant in China; international access available via the VolcEngine API.
Anyscale
Libertai
AI21 Labs
Israeli AI company offering the Jamba model family. Jamba combines SSM (Mamba) and Transformer architecture for very long context (256K tokens) at high throughput. Also offers Wordtune AI writing tools. Available on Amazon Bedrock and Azure. API is OpenAI-compatible.
FriendliAI
FriendliAI is the fastest inference cloud for agents, built to run frontier open-weight models in production at scale. It delivers up to 7x faster output token speed, up to 90% lower inference costs, and 99.99% uptime across the most demanding agent workloads — long-context inference, real-time streaming, and accurate tool calling.
MiniMax
Building AGI with our mission Intelligence with Everyone. Global leader in multi-modal models and AI-native products with over 200 million users.
Tensormesh
Cerebras
Cerebras powers the world's fastest AI inference on the biggest wafer chip. Cerebras CS-4 delivers up to 30x faster inference than GPUs.
Black Forest Labs
Black Forest Labs is building visual intelligence: models that understand, reason, and act in the world. Use FLUX models via our API.
DeepSeek
DeepSeek is an AI research company focused on building world-leading general artificial intelligence. We develop and open-source frontier LLMs including DeepSeek-V4, DeepSeek-R1, and DeepSeek-Coder. Chat with DeepSeek AI or integrate via API.
Crusoe
Meta
Pinstripes
ElevenLabs
Create lifelike speech with our AI voice generator and voice agents platform. Access 5,000+ voices in 70+ languages with secure APIs and SDKs.
Amazon Nova
Amazon Nova is a family of foundation models and services that delivers frontier intelligence and industry-leading price performance.
AWS Polly
Parallel AI
Cognition
Tencent
Typesafe
Apiserpent
AssemblyAI
Darkbloom
Inception
Linkup
Morph
Recraft
Recraft is a top-ranked text-to-image model and design platform for photorealism, vector generation, custom styles, mockups, and more
Scx AI
Soniox
Tavily
Xiaomi Mimo
Bing Grounding
DataForSEO
Google Pse
Jina AI
Nimble
NLP Cloud
API platform for deploying NLP and LLM models including text generation, summarisation, sentiment analysis, and named entity recognition. Hosts open-source models and provides fine-tuning capabilities. Simple REST API with pay-as-you-go pricing.
Serper
Transcribe