Synthszr Charts — die großen AI-Marken im Wettkampf ums Podium
synthszr charts
cartesia

Cartesia · v3.6 · since 17. August 2026 · 2× · last seen Aug 20, 2026

36
Momentum

Sonic-3.6 is Cartesia's current real-time text-to-speech model, released on August 17, 2026 as the successor to Sonic-3.5. The model runs on state space models rather than transformers and achieves a claimed sub-90ms time-to-first-audio latency. It supports 44 languages, offers instant and professional voice cloning, and currently leads both Artificial Analysis Speech Arena leaderboards (Provider Voice and Controlled Voice). It is available as a hosted API in beta, not as self-hostable weights.

Momentum trend
22.05.20.08.

Features

Real-Time StreamingYes, streaming TTS model based on state space models rather than transformers
LatencySub-90ms time-to-first-audio (vendor claim)
LicenseCommercial use included from Pro plan ($5/month); Free tier has no commercial license
PlatformHosted API (beta), not self-hosted weights
PriceFrom $5/month (Pro plan, commercial license); Free tier $0 (20,000 credits/month, non-commercial); Startup $49/month; Scale $299/month
Release DateAugust 17, 2026
Languages44 languages
Voice CloningInstant voice cloning (from Pro plan) and professional voice cloning (from Startup plan) available

More products in this category: Text-to-Speech (TTS)

Sources (2)

Subscribe free. Unsubscribe the second it sucks.

High-signal news across AI, business, UX, and tech. Every morning.