Models & platforms
Replicate
A cloud API for running, fine-tuning, and deploying machine-learning models without managing servers.

Key features
- API access to a catalog of hosted models
- Custom model deployment and usage-based inference
Pros and tradeoffs
Strengths
- Removes much of the infrastructure work required to serve models
Consider before choosing
- Latency, cost, and availability depend on the selected model deployment
What people use Replicate for
- Add image, audio, video, or language models to an application through an API
Frequently asked questions
Does Replicate require managing GPU servers?
No. Replicate manages the serving infrastructure and exposes models through hosted web and API interfaces.