AI Model Providers
Every AI model provider tracked by CloudPrice - frontier labs, inference platforms and self-hosted open-weights shortcuts. Click into any provider to see their full catalog with live pricing, benchmarks and capabilities.
Vercel AI Gateway
Deploy AI apps in seconds with Vercel's AI SDK and Frontend Cloud. Built-in adapters, streaming UI helpers, and zero-config deployments.
OpenRouter
The unified interface for LLMs. Find the best models & prices for your prompts
Azure AI Foundry
Microsoft Foundry
Fireworks AI
Fireworks’ state of the art training and inference platform take you beyond the frontier, transforming open models into your specialized intelligence.
Google Vertex AI
Gemini Enterprise Agent Platform (formerly Vertex AI) is a comprehensive platform for developers to build, scale, govern and optimize agents.
Google Gemini
Build with Gemini 2.0 Flash, 2.5 Pro, and Gemma using the Gemini API and Google AI Studio.
Amazon Bedrock
Amazon Bedrock: The platform for building generative AI applications and agents at production scale
Alibaba Qwen
Supercharge Your AI Journey Effortlessly With Industry-Leading GenAI Models
OpenAI
Creator of GPT-4o, o3, and the GPT model family. Offers text, vision, audio, image generation, speech, and embedding models via a REST API. Pioneered the modern LLM API interface now widely adopted as the de-facto standard.
Hugging Face
We’re on a journey to advance and democratize artificial intelligence through open source and open science.
Novita
Novita AI provides 200+ Model APIs, custom deployment, GPU Instances, and Serverless GPUs. Scale AI, optimize performance, and innovate with ease and efficiency.
Snowflake
Discover Snowflake Arctic, a breakthrough LLM built for enterprise AI. Enterprise intelligence. Breakthrough efficiency. Truly open.
DeepInfra
DeepInfra offers cost-effective, scalable, easy-to-deploy, and production-ready machine-learning models and infrastructures for deep-learning models.
Cloudflare Workers AI
Run machine learning models, powered by serverless GPUs, on Cloudflare's global network.
Oracle Cloud (OCI)
Transform your business with generative AI, and unlock a new era of productivity with task automation and end-to-end AI solutions for enterprise customers.
Mistral AI
The most powerful AI platform for enterprises. Customize, fine-tune, and deploy AI assistants, autonomous agents, and multimodal AI with open models.
Replicate
Run open-source machine learning models with a cloud API
Deepgram
Power enterprise voice solutions with Deepgram’s Speech-to-Text, Text-to-Speech, and Voice Agent APIs. Real-time, accurate, and built for scale.
xAI
Elon Musk's AI company and creators of the Grok model family. Grok models offer large context windows, real-time knowledge via X (Twitter) integration, vision, and multimodal output. Grok-4 is their frontier reasoning model.
IBM watsonx
IBM watsonx is a portfolio of AI products that accelerates the impact of generative AI in core workflows to drive productivity.
Nebius
Build and scale faster on the purpose-built AI cloud, engineered from silicon to API.
Databricks
Databricks offers a unified platform for data, analytics and AI. Build better AI with a data-centric approach. Simplify ETL, data warehousing, governance and AI on the Data + AI Platform.
Together AI
Build what's next on the AI Native Cloud. Full-stack AI platform for inference, fine-tuning, and GPU clusters — powered by cutting-edge research.
Perplexity
AI company best known for its search assistant. Also offers the Sonar model family via API: Sonar (fast, grounded), Sonar Pro (more capable), and Sonar Reasoning (chain-of-thought). All models include real-time web search grounding by default.
Cohere
Cohere builds powerful models and AI solutions enabling enterprises to automate processes, empower employees, and turn fragmented data into actionable insights.
Stability AI
Stability AI is the enterprise-ready creative partner for teams and creators, delivering professional-grade generative AI tools and solutions for content production at scale.
Anthropic
Anthropic is an AI safety and research company that's working to build reliable, interpretable, and steerable AI systems.
Lambda
Train and scale AI on NVIDIA VR200 NVL 72, GB300 NVL 72, B300, B200, H200, H100, and and more GPUs. Launch on-demand instances or reserve a cluster. Get started.
Voyage
Voyage AI provides cutting-edge embedding models and rerankers for search and retrieval
Moonshot AI (Kimi)
Chinese AI company creator of the Kimi model family. Kimi k1.5 and k2 are strong reasoning models with long context. Kimi VL handles vision tasks. Known for efficient long-document processing. Direct API available globally via Moonshot's platform.
GMI Cloud
Singapore-based GPU cloud offering serverless inference for a broad catalog of open-weight and frontier models. Hosts Qwen, MiniMax, DeepSeek, Llama, and others with OpenAI-compatible endpoints. Focuses on Asia-Pacific availability and competitive pricing.
SambaNova
Discover SambaNova - the complete AI platform delivering the fastest AI inference, fine-tuning, and scalable solutions for agentic AI easily integrated into existing data center infrastructures.
Hyperbolic
Hyperbolic is the Open-Access AI Cloud. Get on-demand H100, H200 & B200 GPUs and reserve dedicated multi-node clusters for production-scale AI training and fine-tuning.
Scaleway
European cloud and GPU-inference provider offering serverless LLM inference via an OpenAI-compatible API with GDPR-compliant EU data residency. Hosts open-source models from Meta, Qwen, Mistral, and others.
Groq
Groq is the premier neocloud for fast inference. One fully integrated platform for infrastructure, inference, and control. Millions of developers run trillions of tokens on Groq every week.
Nscale
Nscale full-stack AI cloud platform and services are designed for scale, resilience, and speed.
OVHcloud
Discover the generative AI APIs offered by OVHcloud. High-performing, easy to integrate, and secure, for application power.
AI/ML API
Access 1000+ AI models through a single API for text, image, video, audio, code, embeddings, and more. Compare pricing, latency, and capabilities—all in one place with AIMLAPI.
Z AI (Zhipu)
Meet Z.ai, the AI assistant powered by GLM-5.2. Build websites, write code, handle long-horizon tasks, and get instant answers. Fast, smart, and reliable.
AI21 Labs
AI21 builds Foundation Models and AI Systems for the enterprise. Power your most critical enterprise workflows with accurate, reliable, and scalable AI.
Gradient AI
Hyperagent is the system of agents that does real work, learns how your organization operates, and deploys across your entire team.
MiniMax
Building AGI with our mission Intelligence with Everyone. Global leader in multi-modal models and AI-native products with over 200 million users.
fal.ai
Serverless inference platform specialising in generative media — image, video, audio, and 3D. Hosts FLUX, Stable Diffusion, Kling, and many community diffusion models. Async job queues with webhooks, real-time streaming, and LoRA support.
Black Forest Labs
Black Forest Labs is building visual intelligence: models that understand, reason, and act in the world. Use FLUX models via our API.
Cerebras
Cerebras is the go-to platform for fast and effortless AI training. Learn more at cerebras.ai.
DeepSeek
DeepSeek, unravel the mystery of AGI with curiosity. Answer the essential question with long-termism.
Runway
Runway is building foundational Real-World Intelligence that can understand, simulate and act in the world. We offer products and services built on-top of this intelligence to empower individuals and organizations to do more in the world.
Amazon Nova
Amazon Nova is a family of foundation models and services that delivers frontier intelligence and industry-leading price performance.
ElevenLabs
Create lifelike speech with our AI voice generator and voice agents platform. Access 5,000+ voices in 70+ languages with secure APIs and SDKs.
FriendliAI
FriendliAI is the fastest inference cloud for agents, built to run frontier open-weight models in production at scale. It delivers up to 7x faster output token speed, up to 90% lower inference costs, and 99.99% uptime across the most demanding agent workloads — long-context inference, real-time streaming, and accurate tool calling.
GitHub Models
GitHub is where people build software. More than 150 million people use GitHub to discover, fork, and contribute to over 420 million projects.
Recraft
Recraft is a top-ranked text-to-image model and design platform for photorealism, vector generation, custom styles, mockups, and more
NLP Cloud
API platform for deploying NLP and LLM models including text generation, summarisation, sentiment analysis, and named entity recognition. Hosts open-source models and provides fine-tuning capabilities. Simple REST API with pay-as-you-go pricing.