Replicate
Run open-source models with an API.
Run and fine-tune thousands of community models — especially image, video and audio — via a simple API.
- Vendor
- Replicate
- Category
- Inference & model hosting
- Pricing
- Paid
- Open source
- No
- Platforms
- API
- Launched
- 2020
- Website
- replicate.com
Features
- Thousands of models
- Cog packaging
- Pay per second
Best for
- Image generation
- Multimodal
More inference & model hosting
- Ollama — Ollama. Run open models locally with one command.
- OpenRouter — OpenRouter. One API for hundreds of models.
- vLLM — vLLM project. High-throughput LLM serving engine.
- llama.cpp — ggml. LLM inference in C/C++ on any hardware.
- Amazon Bedrock — AWS. Foundation models on AWS.
- Vertex AI — Google Cloud. Google Cloud’s AI platform.
- LM Studio — Element Labs. Desktop app to discover and run local LLMs.
- Azure AI Foundry — Microsoft. Build and run AI apps on Azure.