# Arize AI vs Replicate

| Attribute | Arize AI | Replicate |
|---|---|---|
| **Vioscale score** | 70.5 (32% (low)) | 52 (41% (low)) |
| deployment.options | `{"cloud":true,"on_prem":true,"self_hosted":true}` | `{"cloud":true,"on_prem":true,"self_hosted":true}` |
| description.long | An AI engineering platform that enables teams to observe agent behavior end-to-end, run evaluations at scale, and systematically improve agents through testing and experimentation, available as both a managed cloud service and self-hosted deployment. | Replicate is an infrastructure platform that lets you run state-of-the-art AI models through a simple API or deploy your own custom models. It automatically handles containerization, scaling, and resource allocation, charging only for compute used. |
| features.capabilities | `{"role":"platform","evaluation":true,"managed_cloud":true,"self_hostable":true,"multi_provider":true,"vpc_deployment":true,"otel_compatible":true,"llm_observability":true,"prompt_management":true,"framework_agnostic":true,"experiment_tracking":true}` | `{"role":"serving","compute_model":"serverless_per_token","managed_cloud":true,"model_serving":true,"pricing_model":"per_compute_hour","self_hostable":true,"multi_provider":true,"vpc_deployment":true,"no_train_on_data":"unknown","llm_observability":true,"framework_agnostic":true,"experiment_tracking":true,"bring_your_own_weights_byow_hosting":true,"multimodal_vision_and_audio_in_out_support":true,"lpu_or_custom_silicon_for_extreme_low_latency":false}` |
| integrations.count | 25 | 8 |
| integrations.list | `[{"name":"OpenAI"},{"name":"Anthropic"},{"name":"Azure OpenAI"},{"name":"AWS Bedrock"},{"name":"Vertex AI"},{"name":"Google GenAI"},{"name":"NVIDIA NIM"},{"name":"Gemini"},{"name":"OpenRouter"},{"name":"LiteLLM"},{"name":"Claude Code"},{"name":"Cursor"},{"name":"OpenCode"},{"name":"LangGraph"},{"name":"Vercel AI SDK"},{"name":"Mastra"},{"name":"CrewAI"},{"name":"LlamaIndex"},{"name":"DSPy"},{"name":"OpenAI Agents SDK"},{"name":"Claude Agent SDK"},{"name":"BigQuery"},{"name":"Databricks"},{"name":"Snowflake"},{"name":"OpenTelemetry"}]` | `[{"name":"Google"},{"name":"OpenAI"},{"name":"ByteDance"},{"name":"Black Forest Labs"},{"name":"HuggingFace"},{"name":"Anthropic"},{"name":"Alibaba"},{"name":"Krea"},{"name":"GitHub"},{"name":"Docker"}]` |
| market.availability | - | `{"primaryMarkets":["US"],"availabilityScope":"global","availableCountries":[],"notAvailableCountries":[]}` |
| platform.support | `{"cli":true,"web":true}` | `{"cli":true,"web":true}` |
| pricing | - | `{"type":"usage","plans":[{"name":"Public Models","summary":"Per-token or per-output pricing varies by model","features":["Run public models","Text-to-image generation","Image editing and restoration","Video generation","Speech generation","Large language models","Music generation"],"components":[{"per":{"qty":1000,"unit":"output tokens"},"kind":"metered","amount":0.015,"currency":"USD"},{"per":{"qty":1000000,"unit":"input tokens"},"kind":"metered","amount":3,"currency":"USD"},{"kind":"per_unit","unit":"image","amount":0.04,"period":"per_output","currency":"USD"},{"kind":"per_unit","unit":"video","amount":0.09,"period":"per_second","currency":"USD"}],"description":"Thousands of open-source and proprietary models billed by execution time or output tokens/images","contactSales":false},{"free":false,"name":"Private Models","summary":"$0.09–$20.16/hr depending on hardware","features":["Dedicated hardware (no shared queue)","Auto-scaling","Custom model deployment","Hourly billing available"],"components":[{"kind":"per_unit","unit":"CPU (Small)","amount":0.000025,"period":"second","currency":"USD"},{"kind":"per_unit","unit":"CPU","amount":0.0001,"period":"second","currency":"USD"},{"kind":"per_unit","unit":"Nvidia A100 (80GB)","amount":0.0014,"period":"second","currency":"USD"},{"kind":"per_unit","unit":"Nvidia H100","amount":0.001525,"period":"second","currency":"USD"},{"kind":"per_unit","unit":"Nvidia L40S","amount":0.000975,"period":"second","currency":"USD"},{"kind":"per_unit","unit":"Nvidia T4","amount":0.000225,"period":"second","currency":"USD"}],"description":"Deploy custom models on dedicated hardware. Billed for all uptime (setup, idle, and processing)","contactSales":false},{"free":false,"name":"Fast Booting Fine-Tunes","summary":"Active processing time only","features":["Fine-tuned model deployment","No idle time billing"],"description":"Private fine-tuned models billed only for active processing time (no idle time charges)","contactSales":false}],"addOns":[{"name":"Multi-GPU A100 Capacity","components":[{"kind":"per_unit","unit":"4x Nvidia A100 (80GB)","amount":0.0056,"period":"second","currency":"USD"},{"kind":"per_unit","unit":"8x Nvidia A100 (80GB)","amount":0.0112,"period":"second","currency":"USD"}]},{"name":"Multi-GPU H100 Capacity","components":[{"kind":"per_unit","unit":"2x Nvidia H100","amount":0.00305,"period":"second","currency":"USD"},{"kind":"per_unit","unit":"4x Nvidia H100","amount":0.0061,"period":"second","currency":"USD"},{"kind":"per_unit","unit":"8x Nvidia H100","amount":0.0122,"period":"second","currency":"USD"}]}],"summary":"Usage-based: public models from $0.01–$0.25 per output token/image/second; private models from $0.09–$40.32/hr. Free tier available.","currency":"USD","freeTier":true,"sourceUrl":"https://replicate.com","retrievedAt":"2026-08-03T22:46:21.215Z","startingPrice":{"amount":0.000025,"period":"month","currency":"USD"}}` |
| pricing.free_tier | - | yes |
| pricing.model | commercial | commercial |
| pricing.price_level | - | unknown |
| pricing.transparent | - | no |
| reliability.status_page | yes | yes |
| security.gdpr | yes | - |
| security.hipaa | yes | - |
| security.iso27001 | yes | - |
| security.pci | yes | - |
| security.soc2 | yes | - |

## Capabilities (MLOps & LLMOps Tools)

| Capability | Arize AI | Replicate |
|---|:--:|:--:|
| **Core** |  |  |
| Tool role | End-to-end ML platform | Model serving |
| **Deployment** |  |  |
| Self-hostable / OSS core | ✓ | ✓ |
| Managed cloud available | ✓ | ✓ |
| On-prem / VPC deployment | ✓ | ✓ |
| **Observability** |  |  |
| LLM tracing / observability | ✓ | ✓ |
| Evaluation (offline / LLM-judge / human) | ✓ | - |
| **Dev** |  |  |
| Prompt management + versioning | ✓ | - |
| **Tracking** |  |  |
| Experiment tracking / model registry | ✓ | ✓ |
| **Serving** |  |  |
| Model serving / inference endpoint | - | ✓ |
| **Interop** |  |  |
| OpenTelemetry / OpenLLMetry compatible | ✓ | - |
| Framework-agnostic | ✓ | ✓ |
| **Gateway** |  |  |
| Multi-provider model support | ✓ | ✓ |
| **Data** |  |  |
| No-train-on-customer-data guarantee | - | Unknown |

*Source: Vioscale. Generated 2026-09-01T15:24:58.151Z. "-" = undocumented, not absent.*
