Skip to index
☁️

AI Platforms

Platforms for hosting, deploying, and running AI models

Hugging Face

Hugging Face

Verified

The largest open-source community for AI models and datasets, offering model hosting, inference APIs, and collaboration tools.

4.8(15K)
Freemium
Open sourceModelsCommunity

Replicate

Replicate

Verified

A cloud platform for running AI models, letting you run open-source models without managing infrastructure.

4.4(3.2K)
Pay-as-you-go
CloudAPIModel hosting

Ollama

Ollama

Verified

The standard tool for running open-source LLMs locally: one command downloads and serves models like Llama, Qwen and DeepSeek on your own machine, with a built-in API.

4.7(5.8K)
Free
Local AIOpen sourceSelf-hosted

Together AI

Together AI

Verified

A cloud platform for training, fine-tuning and serving open-source models at scale, with fast inference APIs across hundreds of checkpoints.

4.3(1.2K)
Paid
InferenceOpen sourceAPI

Groq

Groq

Verified

An inference platform built on custom LPU hardware, famous for generating open-model responses at exceptional speeds through a simple API.

4.6(2K)
Freemium
InferenceHardwareAPI

OpenRouter

OpenRouter

Verified

A unified API that routes requests to hundreds of models from every major lab, with price and latency comparisons and automatic fallbacks.

4.5(2.5K)
Pay-as-you-go
APIRoutingOpen source

DeepInfra

DeepInfra

Verified

A pay-per-token inference cloud serving hundreds of open models — text, image, speech and embedding — through a simple standardized API.

4.4(1.2K)
Pay-as-you-go
InferenceAPIOpen source

Fireworks AI

Fireworks

Verified

A fast inference cloud for open models with production-grade reliability, fine-tuning and compositional multi-model deployments.

4.4(1.1K)
Paid
InferenceOpen sourceAPI

Baseten

Baseten

Verified

An inference platform for deploying custom and open models on dedicated GPUs, with autoscaling, low-latency serving and model privacy.

4.4(900)
Paid
InferenceGPUDeployment

Modal

Modal Labs

Verified

Serverless cloud for AI and machine learning: run inference jobs, fine-tuning and batch workloads on autoscaling GPUs with code-defined infrastructure.

4.6(1.5K)
Pay-as-you-go
ServerlessGPUInference

RunPod

RunPod

Verified

A GPU cloud for AI workloads: serverless inference, on-demand and spot GPUs, and pod environments for training and hosting open models.

4.4(2K)
Pay-as-you-go
GPUInferenceCloud

Vast.ai

Vast.ai

Verified

A marketplace for renting community-run GPUs at the lowest prices, supporting training, inference and rendering workloads with flexible bidding.

4(1.5K)
Pay-as-you-go
GPUMarketplaceTraining

Novita AI

Novita

Verified

An affordable inference cloud for hundreds of open models — LLMs, image and audio generation — with per-token and per-image pricing.

4.2(1K)
Pay-as-you-go
InferenceOpen sourceAPI

SiliconFlow

SiliconFlow

Verified

A China-rooted inference cloud serving leading open models — GLM, Qwen, DeepSeek and more — through standardized, low-cost APIs.

4.2(900)
Pay-as-you-go
InferenceOpen sourceAPI

Perplexity Sonar API

Perplexity

Verified

Perplexity's search-grounded inference API: LLM responses with real-time web citations, built for applications that need factual, sourced answers.

4.4(1.8K)
Paid
APISearchGrounded

Azure OpenAI

Microsoft

Verified

Microsoft Azure's enterprise service for OpenAI models: GPT-4o, o-series and embedding models with enterprise security, compliance and regional availability.

4.3(3K)
Paid
EnterpriseAPICompliance

Cohere

Cohere

Verified

An enterprise LLM platform built for security-conscious organizations: the Command model family, retrieval-grade Embed and Rerank models, and private deployment on your own cloud or on-premises.

4.4(850)
Pay-as-you-go
EnterpriseLLM APIRAG

Writer

Writer

Verified

A full-stack enterprise generative AI platform: Palmyra LLMs, a no-code agent builder, Knowledge Graph grounding and governance controls for rolling out AI across regulated organizations.

4.6(1.3K)
Paid
EnterpriseAI platformGovernance

vLLM

vLLM Project

Verified

The open-source inference engine behind much of modern LLM serving: PagedAttention delivers state-of-the-art throughput with an OpenAI-compatible API and broad model support on your own GPUs.

4.8(3.1K)
Free
InferenceOpen sourceSelf-hosted

ModelScope

Alibaba (Dammo Academy)

Verified

Alibaba's open-source model community and MaaS platform: hundreds of thousands of models and datasets, free online inference notebooks, and the home hosting of the Qwen family.

4.4(2.3K)
Free
Model hubOpen sourceFree

Dify

Dify (LangGenius)

Verified

The open-source LLM app platform: build RAG assistants and agent workflows visually, back them with your choice of 100+ models, and run them on your own infrastructure or Dify Cloud.

4.6(3.8K)
Freemium
Open sourceLLM platformSelf-hosted

Exa

Exa AI

Verified

A search engine built for AI: an API that returns clean, semantically relevant web content — page contents, not ad-laced result lists — designed to be read by models rather than humans.

4.5(900)
Pay-as-you-go
Search APIAgentsDeveloper

Tavily

Tavily

Verified

A search API purpose-built for LLMs and RAG: one call returns ranked, content-ready web results optimized for retrieval — the plug-and-play web tool of countless agent frameworks.

4.5(1.1K)
Pay-as-you-go
Search APIRAGDeveloper

Serper

Serper Dev

Verified

A fast, low-cost Google Search results API built for AI: structured SERP data (organic, news, images, maps) at a fraction of legacy SERP-API pricing — the workhorse behind many research agents.

4.4(800)
Pay-as-you-go
SERP APIGoogle dataDeveloper

fal.ai

fal

Verified

Generative-media cloud infrastructure: host and serve image, video and audio models through one API, famous for sub-second FLUX generation and day-one hosting of new open models.

4.6(1.2K)
Pay-as-you-go
Inference cloudAPIMedia generation

Runware

Runware

Verified

Ultra-low-cost image generation infrastructure: a sub-second API claiming the industry's lowest per-image price by owning its GPU pipeline end to end — built for apps generating at massive volume.

4.5(500)
Pay-as-you-go
Inference APILow costImage generation

Vapi

Vapi

Verified

The developer platform for voice agents: assemble sub-second phone-call AI from best-in-class STT, LLMs and TTS — with phone numbers, transfers, and integrations — without building the telephony plumbing yourself.

4.5(600)
Pay-as-you-go
Voice agentsPhone AIDeveloper

Brave Search API

Brave Software

Verified

An independent search index exposed as an API: privacy-first Google-alternative results with web, news, image and local endpoints — the independent-data option in the AI search stack.

4.4(650)
Pay-as-you-go
Search APIIndependent indexPrivacy

Open WebUI

Open WebUI

Verified

The self-hosted ChatGPT-style interface for local and private models: a polished chat UI over Ollama, OpenAI-compatible APIs or any backend — with RAG, multi-user management and tool calling built in.

4.7(3.9K)
Free
Self-hostedOpen sourceChat UI

LibreChat

LibreChat

Verified

An open-source, multi-provider chat interface: one self-hosted UI for OpenAI, Anthropic, Google, local models and more — with agents, RAG, artifacts and multi-user management, no per-vendor lock-in.

4.6(2.4K)
Free
Open sourceMulti-providerSelf-hosted

Apple Intelligence

Apple

Verified

Apple's on-device and Private-Cloud AI layer: writing tools, notification summaries, image generation and a redesigned Siri — engineered so personal data never has to leave your device.

4.2(5.2K)
Free
On-devicePrivacyOS-integrated

Pinecone

Pinecone

Verified

The managed vector database that defined the category: millisecond similarity search over billions of embeddings, serverless scaling, and the retrieval backbone of countless production RAG systems.

4.5(1.2K)
Pay-as-you-go
Vector databaseManagedRAG