A-05 · Category market reportarchive-recorded · snapshot 2026-09-17
Audio & Voice AI Tools: Category Market Report
Archive-derived market report for the Audio & Voice category: 55 tool archives, median overall score 79.7/100 across five recorded dimensions, 44 free-tier listings and 165 recorded use-case entries. All figures are archive-recorded, computed from data/tools with the method and N published below.
Key numbers
Archives
55
data/tools · N=55
Overall median
79.7
score.overall · N=55
Free-tier listings
44
pricing.model · N=55
Median mo starting price
$10/mo
pricing.startingPrice.amount (unit=mo) · N=24
Recorded use-case entries
165
finder.useCases · N=55
Dated in snapshot window
53
updatedOn · N=55
Top 12 of 55 archives by overall score
score.overall as recorded; ties broken by slug
| # | Tool | Overall | Quality | Features | Usability | Performance | Value | Starting price | Updated |
|---|---|---|---|---|---|---|---|---|---|
| 1 | PlayHT | 89.2 | 91 | 89 | 87 | 89 | 89 | $9.99/mo | 2026-09-14 |
| 2 | ElevenLabs | 87.8 | 95 | 90 | 88 | 84 | 80 | $6/mo | 2026-09-06 |
| 3 | Cartesia | 85.8 | 88 | 80 | 84 | 90 | 86 | $5/mo | 2026-09-06 |
| 4 | Adobe Podcast | 85.6 | 88 | 76 | 88 | 84 | 92 | $0/mo | 2026-09-13 |
| 5 | Deepgram | 85.1 | 88 | 86 | 82 | 86 | 82 | PAYG STT from $0.0043/min · TTS from $0.015/1k chars · … | 2026-09-15 |
| 6 | AssemblyAI | 84.5 | 88 | 86 | 78 | 86 | 82 | $0.15/hr | 2026-09-13 |
| 7 | Auphonic | 84.2 | 86 | 82 | 82 | 84 | 86 | Free 2h processed audio/mo · recurring credit plans (9h… | 2026-09-17 |
| 8 | Wispr Flow | 82.8 | 84 | 78 | 92 | 86 | 76 | $15/user | 2026-09-13 |
| 9 | Chatterbox | 82.6 | 83 | 76 | 72 | 84 | 95 | Free open source (MIT) · community edition on GitHub | 2026-09-17 |
| 10 | Hume AI | 82.6 | 84 | 82 | 80 | 82 | 84 | $3/mo | 2026-09-13 |
| 11 | Kokoro TTS | 82.5 | 84 | 70 | 74 | 86 | 96 | Open-weight (Apache 2.0) · hosted API serving under $1 … | 2026-09-17 |
| 12 | Rime | 82.5 | 84 | 76 | 82 | 86 | 84 | Starter usage-based (Mist v3 $0.03/1k chars, Coda $0.05… | 2026-09-15 |
Dimension medians
Median of each recorded dimension across the 55 Audio & Voice archives (score.dims, scale 0–100).
Method
- Population: all archives in data/tools with category = "audio" (N=55). No sampling.
- Medians use the standard median (middle element, or mean of the two middle elements rounded to 1 decimal) over score.overall and score.dims as recorded; scores are never recomputed or re-weighted.
- Starting price statistics use only numeric pricing.startingPrice.amount values in unit "mo" (N=24); archives with null amounts are excluded and displayed via pricing.label fallback instead.
- Benchmark coverage in this category: 0 of 55 archives carry a benchmark entry; T1 benchmarks recorded: 3 / 448 across the whole library.
- Freshness: 53 of 55 archives carry updatedOn inside the ten-day snapshot window 2026-09-06 → 2026-09-17.