# AssemblyAI vs SpeechBrain

**Leader by Vioscale score:** AssemblyAI

| Attribute | AssemblyAI | SpeechBrain |
|---|---|---|
| **Vioscale score** | 80.2 (60% (medium)) | 26.8 (23% (low)) |
| activity.commits_last_30d | - | 11 |
| adoption.dependent_repos | - | 0 |
| adoption.github_stars | - | 11,785 |
| deployment.options | `{"cloud":true}` | `{"self_hosted":true}` |
| description.long | A platform providing speech-to-text APIs for both pre-recorded and live audio, voice agent capabilities, and audio analysis features like speaker identification and sentiment analysis. Serves millions of developers with focus on accuracy and scalability. | A framework for developing conversational AI with support for speech recognition, synthesis, enhancement, and language understanding. Designed with flexibility and transparency for research and production use. |
| integrations.count | 5 | 2 |
| integrations.list | `[{"name":"Semantic Kernel .NET"},{"name":"Vercel AI SDK"},{"name":"OpenAI GPT-5.6 Luna"},{"name":"OpenAI GPT-5.6 Terra"},{"name":"Google Gemini"}]` | `[{"name":"HuggingFace"},{"name":"Orion"}]` |
| language.primary | - | Python |
| license.spdx | - | Apache-2.0 |
| market.availability | `{"primaryMarkets":["US","GB"],"availabilityScope":"global","availableCountries":[],"notAvailableCountries":[]}` | - |
| platform.support | `{"web":true}` | - |
| pricing | `{"type":"usage","plans":[{"free":false,"name":"Async Speech-to-Text","summary":"$0.37/hour","components":[{"per":{"qty":1,"unit":"hour"},"kind":"metered","amount":0.37,"currency":"USD"}],"description":"Pre-recorded audio transcription","contactSales":false},{"free":false,"name":"Real-time Speech-to-Text","summary":"$0.47/hour","components":[{"per":{"qty":1,"unit":"hour"},"kind":"metered","amount":0.47,"currency":"USD"}],"description":"Live audio stream transcription","contactSales":false},{"free":false,"name":"Sync Speech-to-Text","summary":"$0.45/hour","components":[{"per":{"qty":1,"unit":"hour"},"kind":"metered","amount":0.45,"currency":"USD"}],"description":"Same-request transcription for short clips","contactSales":false}],"addOns":[{"name":"Speaker Diarization","components":[{"per":{"qty":1,"unit":"hour"},"kind":"metered","amount":0.12,"currency":"USD"}]},{"name":"LLM Prompting","components":[{"per":{"qty":1,"unit":"hour"},"kind":"metered","amount":0.05,"currency":"USD"}]},{"name":"Voice Isolation","components":[{"per":{"qty":1,"unit":"hour"},"kind":"metered","amount":0.1,"currency":"USD"}]},{"name":"SSO","components":[{"kind":"fixed","amount":199,"period":"month","currency":"USD"}]}],"summary":"From $0.37/hour for async transcription. Free tier available.","currency":"USD","freeTier":true,"sourceUrl":"https://www.assemblyai.com/blog/lower-latency-new-pricing","retrievedAt":"2026-08-19T21:32:28.755Z","startingPrice":{"unit":"hour","amount":0.37,"currency":"USD"},"billingPeriods":["month"]}` | `{"type":"open_source","summary":"Free and open source (Apache License)","freeTier":true,"sourceUrl":"https://speechbrain.github.io/","retrievedAt":"2026-08-20T14:05:43.501Z"}` |
| pricing.free_tier | yes | yes |
| pricing.model | freemium | open_source |
| pricing.price_level | low | free |
| pricing.starting_price | `{"amount":0.37,"currency":"USD"}` | - |
| pricing.transparent | yes | yes |
| release.cadence_days | - | 110 |
| release.history | - | `[{"url":"https://github.com/speechbrain/speechbrain/releases/tag/v1.1.0","date":"2026-03-30T14:41:48Z","type":"stable","version":"v1.1.0"},{"url":"https://github.com/speechbrain/speechbrain/releases/tag/v1.0.3","date":"2025-04-07T17:05:07Z","type":"stable","version":"v1.0.3"},{"url":"https://github.com/speechbrain/speechbrain/releases/tag/v1.0.2","date":"2024-10-31T10:06:07Z","type":"stable","version":"v1.0.2"},{"url":"https://github.com/speechbrain/speechbrain/releases/tag/v1.0.1","date":"2024-10-01T10:55:15Z","type":"stable","version":"v1.0.1"},{"url":"https://github.com/speechbrain/speechbrain/releases/tag/v1.0.0","date":"2024-10-01T10:54:56Z","type":"stable","version":"v1.0.0"},{"url":"https://github.com/speechbrain/speechbrain/releases/tag/v0.5.16","date":"2023-11-22T02:28:28Z","type":"stable","version":"v0.5.16"},{"url":"https://github.com/speechbrain/speechbrain/releases/tag/v0.5.15","date":"2023-07-22T18:07:47Z","type":"stable","version":"v0.5.15"},{"url":"https://github.com/speechbrain/speechbrain/releases/tag/v0.5.14","date":"2023-03-24T17:40:58Z","type":"stable","version":"v0.5.14"},{"url":"https://github.com/speechbrain/speechbrain/releases/tag/v0.5.13","date":"2022-08-29T16:25:16Z","type":"stable","version":"v0.5.13"},{"url":"https://github.com/speechbrain/speechbrain/releases/tag/v0.5.12","date":"2022-06-26T20:19:45Z","type":"stable","version":"v0.5.12"},{"url":"https://github.com/speechbrain/speechbrain/releases/tag/v0.5.11","date":"2021-12-20T04:22:27Z","type":"stable","version":"v0.5.11"},{"url":"https://github.com/speechbrain/speechbrain/releases/tag/0.5.10","date":"2021-09-11T22:34:16Z","type":"stable","version":"0.5.10"},{"url":"https://github.com/speechbrain/speechbrain/releases/tag/v0.5.9","date":"2021-06-17T01:25:19Z","type":"stable","version":"v0.5.9"},{"url":"https://github.com/speechbrain/speechbrain/releases/tag/v0.5.8","date":"2021-06-06T01:42:20Z","type":"stable","version":"v0.5.8"},{"url":"https://github.com/speechbrain/speechbrain/releases/tag/0.5.7","date":"2021-04-29T17:12:43Z","type":"stable","version":"0.5.7"}]` |
| reliability.sla_pct | 99.9 | - |
| reliability.status_page | yes | - |
| security.disclosure_policy | yes | - |
| security.gdpr | yes | - |
| security.hipaa | yes | - |
| security.pci | yes | - |
| security.soc2 | yes | - |
| security.vulnerabilities | - | `{"count":0,"source":"https://advisories.ecosyste.ms/api/v1/advisories?ecosystem=go&package_name=github.com%2Fspeechbrain%2Fspeechbrain&per_page=100","last_12m":0,"max_severity":null}` |

## Capabilities (Speech To Text)

| Capability | AssemblyAI | SpeechBrain |
|---|:--:|:--:|
| **Capabilities** |  |  |
| Real time streaming | - | - |
| Speaker diarization | - | - |
| Custom vocabulary | - | - |
| Fine tuning | - | - |
| Word timestamps | - | - |
| Punctuation formatting | - | - |
| Pii redaction | - | - |
| Vertical model | - | - |
| Offline on device | - | - |
| Open source | - | - |

*Source: Vioscale. Generated 2026-09-01T17:20:28.046Z. "-" = undocumented, not absent.*
