TLDRocket
Sign in

Cartesia releases Sonic-3.6 text-to-speech model, achieving top rankings on Artificial Analysis benchmarks

Model release Provisional 95% confidence first seen

Cartesia released Sonic-3.6, a streaming text-to-speech model built on state space models that achieved first-place rankings on Artificial Analysis speech leaderboards with sub-90ms latency to first audio. The model supports 44 languages and is available as a hosted API starting at $5 per month, priced at $49 per million characters, undercutting competitors like ElevenLabs.

Source coverage

The daily briefing

Every AI story that matters, in your inbox by 8am.

TLDRocket reads all relevant sources, removes duplicate coverage, and summarises the day in two minutes. Follow companies and topics for alerts, or get the briefing in Slack. Free, no spam, unsubscribe anytime.