Replicate
A cloud platform for running machine learning models through an API and deploying custom models to production
- Also known as
- replicate
Available worldwide · Popular in: US
What is Replicate?
Execute pre-built or custom machine learning models through a unified API. Replicate handles infrastructure scaling and provides tools for fine-tuning models with custom data and monitoring model performance.
Replicate pricing
Plans, per-tier features and add-ons, dated and linked to live pricing. Pricing changes often; always verify at source before you rely on it.
Usage-based: public models from $0.01–$0.25 per output token/image/second; private models from $0.09–$40.32/hr. Free tier available.
Public Models
Thousands of open-source and proprietary models billed by execution time or output tokens/images
- Run public models
- Text-to-image generation
- Image editing and restoration
- Video generation
- Speech generation
- Large language models
- Music generation
Private Models
Deploy custom models on dedicated hardware. Billed for all uptime (setup, idle, and processing)
- Dedicated hardware (no shared queue)
- Auto-scaling
- Custom model deployment
- Hourly billing available
Fast Booting Fine-Tunes
Private fine-tuned models billed only for active processing time (no idle time charges)
- Fine-tuned model deployment
- No idle time billing
Add-ons
- Multi-GPU A100 Capacity$0.01/4x Nvidia A100 (80GB)/second + $0.01/8x Nvidia A100 (80GB)/second
- Multi-GPU H100 Capacity$0.00/2x Nvidia H100/second + $0.01/4x Nvidia H100/second + $0.01/8x Nvidia H100/second
What Replicate does
The capabilities that matter for mlops & llmops tools, normalised so it lines up with every alternative. “-” means we haven't confirmed it, not that it's missing.
- Tool role
- Model serving
- Self-hostable / OSS core
- ✓
- Managed cloud available
- ✓
- On-prem / VPC deployment
- ✓
- LLM tracing / observability
- ✓
- Evaluation (offline / LLM-judge / human)
- -
- Prompt management + versioning
- -
- Experiment tracking / model registry
- ✓
- Model serving / inference endpoint
- ✓
- OpenTelemetry / OpenLLMetry compatible
- -
- Framework-agnostic
- ✓
- Multi-provider model support
- ✓
- No-train-on-customer-data guarantee
- Unknown
Platform & deployment
Independently observed- Web
- Cloud / SaaS
- Self-hosted
Integrations (10)
Independently observed- OpenAI
- ByteDance
- Black Forest Labs
- HuggingFace
- Anthropic
- Alibaba
- Krea
- GitHub
- Docker
Replicate FAQ
Common questions about Replicate, answered from independent, dated evidence.
What is Replicate?
Execute pre-built or custom machine learning models through a unified API. Replicate handles infrastructure scaling and provides tools for fine-tuning models with custom data and monitoring model performance. It is indexed under MLOps & LLMOps Tools.
Source: https://replicate.com
Is Replicate free?
Replicate offers a free tier, so you can start without paying. Paid plans start at $0.00 per CPU (Small) per second. Pricing changes often, so verify at source before relying on it.
Source: https://replicate.com
What platforms does Replicate support?
We have confirmed browser-based access to Replicate. That is the extent of what we could verify from public sources, so it may well offer desktop or mobile clients we have not indexed.
Source: https://replicate.com
Can Replicate be self-hosted?
Yes. Replicate can be deployed cloud / SaaS and self-hosted, so it does not have to run on the vendor's infrastructure.
Source: https://replicate.com
What does Replicate integrate with?
We have confirmed 10 integrations for Replicate, including Google, OpenAI, ByteDance, Black Forest Labs, HuggingFace, Anthropic, Alibaba and Krea, plus 2 more. This is what we could verify from public sources, so the vendor may support others we have not indexed.
Source: https://replicate.com
Where is Replicate available?
Replicate is available worldwide. Its primary market is the United States.
Source: https://replicate.com
Replicate alternatives
Other mlops & llmops tools we track, ranked by the same independent score.
- PortkeyAn API gateway for routing requests across thousands of language models with built-in safety guardrailshigh · 76%
- LangChainA platform for building, testing, and operating AI agents at scalemedium · 70%
- BasetenAn inference platform for deploying and running AI models at scalemedium · 73%
- Together AICloud platform for deploying and running open-source AI models with optimized inferencemedium · 61%
- OllamaA platform for running open-source language models locally or in the cloud with cost-effective access.medium · 57%
- RayDistributed computing infrastructure for scaling machine learning applicationsmedium · 65%
The vioscaleAI score: one lens on the evidence
Not user reviews and not a paid placement: a confidence-weighted blend of the independent signals below (adoption, activity, security posture, and more), which you can sort and re-weight yourself. Vendors can correct their listing but can never move their rank, and stars are weighted low as a vanity metric. It is one way to read the evidence for Replicate, not the verdict.
| Signal | Score | Weight | Contribution | Evidence |
|---|---|---|---|---|
| Capabilities | 100 | 0.04 | 4.0 | ✓ |
| Pricing transparency | 45 | 0.08 | 3.8 | ✓ |
| Reliability | 50 | 0.06 | 3.2 | ✓ |
| Integrations | 27 | 0.05 | 1.3 | ✓ |
| Price level | 0 | 0.05 | 0.0 | - |
| Security posture | 0 | 0.06 | 0.0 | - |
Computed . Re-weight it by intent, or see the full method.
All data & sourcesshow ↓
Every value we hold, with its source, retrieval date, and confidence. This is the evidence behind the score: don't trust it, verify it.
Content
| Attribute | Value | Evidence |
|---|---|---|
| Faq | 6 items | mediumsource · 2026-09-10 · 58% |
Features
| Attribute | Value | Evidence |
|---|---|---|
| Capabilities | Role: serving · Compute model: serverless_per_token · Managed cloud: Yes · Model serving: Yes · Pricing model: per_compute_hour · Self hostable: Yes | mediumsource · 2026-09-06 · 60% |
Integrations
| Attribute | Value | Evidence |
|---|---|---|
| Count | 8 | mediumsource · 2026-08-21 · 60% |
Market
| Attribute | Value | Evidence |
|---|---|---|
| Availability | PrimaryMarkets: … · AvailabilityScope: global · AvailableCountries: … · NotAvailableCountries: … | mediumsource · 2026-08-21 · 50% |
Pricing
Reliability
| Attribute | Value | Evidence |
|---|---|---|
| Status page | Yes | mediumsource · 2026-08-03 · 60% |
Is Replicate the right choice for you?
Tell us the job, the constraints and what you weigh most, and we will rank Replicate against the rest of the mlops & llmops tools we index, using the same dated evidence weighted your way.
Free to run, no account needed to start. How the evaluation works
Is this your product?
This profile was built from public sources without asking you. You can take the badge below and use it anywhere, and you can claim the profile to correct anything we got wrong. Both are free, and neither moves Replicate up or down: nobody can buy rank here, including you.
Take the badge
Live, always current, and free to use on your own site. It shows Replicate's independent score and links back to this profile.
<a href="https://www.vioscale.ai/software/replicate" target="_blank" rel="noopener">
<img src="https://www.vioscale.ai/badge/software/replicate.svg" alt="Replicate, verified on vioscaleAI" width="330" height="76" loading="lazy" />
</a>Markdown, for a README →
[](https://www.vioscale.ai/software/replicate)Claim the profile
Verify you control the domain and you can correct the facts, add the sources we should be reading, and see how AI assistants are describing Replicate. Free, and it does not change the score.
- Correct anything wrong, with evidence
- Point our crawler at the pages that matter
- See which AI systems are reading this profile
Not the owner? How vendor profiles work