Directory

Providers

Compare inference API providers by features, pricing model, and supported models.

DE

DeepInfra

Serverless Featured

Serverless inference for open-source LLMs and generative models. Pay-per-token with fast cold starts.

13 models
OpenAI Compat Streaming Finetuning Embeddings Vision Free Tier
GO

Google

Proprietary Featured

Official Gemini API via Google AI Studio and Vertex AI. Direct access to Gemini, Imagen, and Gemma models.

19 models
Streaming Embeddings Functions Vision Free Tier
GR

Groq

Serverless Featured

Fastest LLM inference powered by custom LPU chips. OpenAI-compatible API with sub-second latency.

4 models
OpenAI Compat Streaming Functions Vision Free Tier
KI

KIE AI

Aggregator Featured

Affordable AI API aggregator offering 259+ models across chat, image, video, and music at discounted prices.

55 models
OpenAI Compat Streaming Vision
MU

Muapi

Aggregator Featured

AI API aggregator with 315+ model endpoints across text, image, video, and audio at competitive prices.

75 models
OpenAI Compat Streaming Vision
NO

Novita AI

Aggregator Featured

Budget AI inference platform with broad model catalog across LLM, image, video, and audio. Very competitive per-token pricing.

41 models
OpenAI Compat Streaming Vision Free Tier
OP

OpenAI

Proprietary Featured

Official OpenAI API. Direct access to GPT, DALL-E, Whisper, and embedding models.

18 models
OpenAI Compat Streaming Embeddings Functions Vision Free Tier
SI

SiliconFlow

Serverless Featured

Fast and affordable AI inference platform. 2.3x faster speeds and 32% lower latency than major cloud platforms. Supports LLM, image, video, and audio models.

20 models
OpenAI Compat Streaming Vision Free Tier
FA

fal.ai

Serverless Featured

Fast inference platform for generative media — image, video, audio, and 3D models with serverless GPU infrastructure.

70 models
Streaming Vision Free Tier
AI

AIMLAPI

Aggregator

Unified API for 400+ AI models across text, image, video, and audio. OpenAI-compatible with serverless inference.

25 models
OpenAI Compat Streaming Embeddings Vision Free Tier
CL

Cloudflare Workers AI

Serverless

Edge AI inference across 200+ cities worldwide. Serverless, pay-per-use with OpenAI-compatible API.

12 models
OpenAI Compat Streaming Embeddings Vision Free Tier