Cartesia releases Sonic-3.6, leading Artificial Analysis speech leaderboards
Cartesia has released Sonic-3.6, the latest version of its real-time text-to-speech model, which now holds the #1 position on both Artificial Analysis speech leaderboards. The update focuses on improved naturalness and runs on state space models rather than transformers to achieve sub-90ms time-to-first-audio latency.