What is Replicate?
Platform for running, fine-tuning, and deploying AI models (LLMs, image generation, video, speech, etc.) via API. Thousands of open-source and proprietary models available, plus the ability to deploy custom models.
Replicate pricing
Plans, per-tier features and add-ons, dated and linked to live pricing. Pricing changes often; always verify at source before you rely on it.
Usage-based: public models from $0.01–$0.25 per output token/image/second; private models from $0.09–$40.32/hr. Free tier available.
Public Models
Thousands of open-source and proprietary models billed by execution time or output tokens/images
- Run public models
- Text-to-image generation
- Image editing and restoration
- Video generation
- Speech generation
- Large language models
- Music generation
Private Models
Deploy custom models on dedicated hardware. Billed for all uptime (setup, idle, and processing)
- Dedicated hardware (no shared queue)
- Auto-scaling
- Custom model deployment
- Hourly billing available
Fast Booting Fine-Tunes
Private fine-tuned models billed only for active processing time (no idle time charges)
- Fine-tuned model deployment
- No idle time billing
Add-ons
- Multi-GPU A100 Capacity$0.01/4x Nvidia A100 (80GB)/second + $0.01/8x Nvidia A100 (80GB)/second
- Multi-GPU H100 Capacity$0.00/2x Nvidia H100/second + $0.01/4x Nvidia H100/second + $0.01/8x Nvidia H100/second
What Replicate does
The capabilities that matter for mlops & llmops tools, normalised so it lines up with every alternative. “-” means we haven't confirmed it, not that it's missing.
- Tool role
- Model serving
- Self-hostable / OSS core
- ✓
- Managed cloud available
- ✓
- On-prem / VPC deployment
- ✗
- LLM tracing / observability
- ✗
- Evaluation (offline / LLM-judge / human)
- ✗
- Prompt management + versioning
- ✗
- Experiment tracking / model registry
- ✗
- Model serving / inference endpoint
- ✓
- OpenTelemetry / OpenLLMetry compatible
- ✗
- Framework-agnostic
- ✓
- Multi-provider model support
- ✓
- No-train-on-customer-data guarantee
- -
Platform & deployment
Independently observed- CLI
- Web
- Cloud / SaaS
- Self-hosted
Integrations (9)
Independently observed- Node.js
- Python
- Next.js
- SwiftUI
- Discord
- ComfyUI
- Cloudflare
- Val Town
- OpenAI
Replicate alternatives
Other mlops & llmops tools we track, ranked by the same independent score.
- OllamaThe easiest way to build with open modelslow · 22%
- OllamaThe easiest way to build with open modelslow · 45%
- BasetenInference is everythinglow · 47%
- PortkeyProduction Stack for Gen AI Builderslow · 46%
- MLflowOpen Source AI Platform for Agents, LLMs & Modelslow · 47%
- Weights & BiasesThe AI developer platformlow · 25%
The Vioscale score: one lens on the evidence
Not user reviews and not a paid placement: a confidence-weighted blend of the independent signals below (adoption, activity, security posture, and more), which you can sort and re-weight yourself. Vendors can correct their listing but can never move their rank, and stars are weighted low as a vanity metric. It is one way to read the evidence for Replicate, not the verdict.
| Signal | Score | Weight | Contribution | Evidence |
|---|---|---|---|---|
| Capabilities | 75 | 32.00 | 2391.1 | ✓ |
| Pricing Transparency | 100 | 16.00 | 1600.0 | ✓ |
| Price Level | 80 | 10.00 | 800.0 | ✓ |
| Reliability | 50 | 12.00 | 600.0 | ✓ |
| Integrations | 29 | 18.00 | 517.6 | ✓ |
| Security Posture | 0 | 12.00 | 0.0 | - |
Computed . Re-weight it by intent, or see the full method.
All data & sourcesshow ↓
Every value we hold, with its source, retrieval date, and confidence. This is the evidence behind the score: don't trust it, verify it.
Features
| Attribute | Value | Evidence |
|---|---|---|
| Capabilities | {"role":"serving","evaluation":false,"managed_cloud":true,"model_serving":true,"self_hostable":true,"multi_provider":true,"vpc_deployment":false,"otel_compatible":false,"llm_observability":false,"prompt_management":false,"framework_agnostic":true,"experiment_tracking":false} | mediumsource · 2026-08-03 · 60% |
Integrations
| Attribute | Value | Evidence |
|---|---|---|
| Count | 9 | mediumsource · 2026-08-03 · 60% |
Pricing
Reliability
| Attribute | Value | Evidence |
|---|---|---|
| Status page | Yes | mediumsource · 2026-08-03 · 60% |