Synthszr Charts — die großen AI-Marken im Wettkampf ums Podium
synthszr charts
cartesia

Cartesia · v3.5 · siet 2026-06-16 · 9× · tolest 30. Juni 2026

14
Momentum

Sonic-3.5 is Cartesia's Text-to-Speech model for real-time speech synthesis in Voice-Agents. It was released on June 16, 2026 together with the transcription model Ink-2 as a joint Voice-Stack. According to the manufacturer, the model achieves speech output latency of under 90ms, natively supports 42 languages, and is listed as a leading TTS model on Benchmark platforms such as Artificial Analysis. It offers Instant Voice-Cloning, Voice-Changer, and localization features and is available through the Cartesia API as well as partners like LiveKit.

Momentum-Verloop
16.05.14.08.

Features

Real-Time StreamingYes, designed for real-time voice agents with bidirectional streaming
LatencySub-90ms speech latency (time-to-first-audio); around 82ms per Artificial Analysis
LicenseCommercial use license included from the Pro plan (paid); Free tier not intended for commercial use
PlatformCartesia's cloud API (model ID sonic-3.5), also available via partners like LiveKit; deployable in cloud, on-premise, and on-device
PriceFrom $5/month (Pro plan, ~133 min TTS/month); Free tier $0/month (20,000 credits); Startup $49/month; Scale $299/month; API usage e.g. $0.03/min or $50 per 1M characters (via LiveKit)
Release DateJune 16, 2026
Languages42 languages natively, including English, Hindi, Spanish, French, German, Japanese, Hebrew
Voice CloningInstant voice cloning with just 10 seconds of audio; Professional Voice Cloning also available for higher quality

Mehr Produkten in disse Kategorie: Spraaksynthese (TTS)

Belege (9)

Subscribe free. Unsubscribe the second it sucks.

High-signal news across AI, business, UX, and tech. Every morning.