# Azure AI Speech vs Kokoro

**Leader by Vioscale score:** Azure AI Speech

| Attribute | Azure AI Speech | Kokoro |
|---|---|---|
| **Vioscale score** | 53.3 (51% (medium)) | 14 (15% (low)) |
| activity.commits_last_30d | - | 0 |
| adoption.dependent_repos | - | 0 |
| adoption.github_stars | - | 8,636 |
| deployment.options | `{"cloud":true}` | `{"cloud":true,"on_prem":true,"self_hosted":true}` |
| language.primary | - | JavaScript |
| license.spdx | - | Apache-2.0 |
| market.availability | `{"primaryMarkets":[],"availabilityScope":"global","availableCountries":[],"notAvailableCountries":[]}` | - |
| platform.support | `{"cli":true,"web":true}` | `{"cli":true,"linux":true,"windows":true}` |
| pricing | `{"type":"subscription","summary":"Free tier available; 20+ services free for 12 months, 65+ services free always, $200 credit in first 30 days","currency":"USD","freeTier":true,"sourceUrl":"https://azure.microsoft.com/en-us/pricing/calculator/","retrievedAt":"2026-08-31T15:11:07.113Z"}` | `{"type":"open_source","freeTier":true,"sourceUrl":"https://github.com/pricing","retrievedAt":"2026-08-31T15:11:47.802Z"}` |
| pricing.free_tier | yes | yes |
| pricing.model | freemium | open_source |
| pricing.price_level | low | free |
| reliability.status_page | yes | - |
| security.certifications | `[{"name":"100+ compliance offerings","status":"active"}]` | - |
| security.gdpr | yes | - |
| security.iso27001 | yes | - |
| security.scorecard | - | 2.9 |
| security.soc2 | yes | - |
| security.trust_center | https://azure.microsoft.com/en-us/explore/trusted-cloud/ | https://github.com/trust-center |
| security.vulnerabilities | - | `{"count":0,"source":"https://advisories.ecosyste.ms/api/v1/advisories?ecosystem=npm&package_name=%40somniumism%2Fkokoro-js-ja&per_page=100","last_12m":0,"max_severity":null}` |

## Capabilities (AI Voiceover)

| Capability | Azure AI Speech | Kokoro |
|---|:--:|:--:|
| **Capabilities** |  |  |
| Architecture model | - | - |
| Zero shot instant voice cloning from short samples | - | - |
| Emotional prosody directing anger whisper joy | - | - |
| Speech to speech sts acting and accent transfer | - | - |
| Automatic translated dubbing retaining original voice | - | - |
| Sub 500ms ultra low latency streaming API available | - | - |
| Phoneme level pronunciation dictionary overrides | - | - |
| Multi speaker timeline editor for conversations | - | - |
| Voice isolation and background noise removal | - | - |
| Watermarking and AI generation detection safety | - | - |
| SOC2 type ii | - | - |
| ISO 27001 | - | - |
| Pricing model | - | - |

*Source: Vioscale. Generated 2026-09-01T15:15:25.960Z. "-" = undocumented, not absent.*
