---
title: "Cartesia's new TTS model tops both speech leaderboards"
url: https://www.parallelquant.com/posts/cartesia-s-new-tts-model-tops-both-speech-leaderboards-6cb27f
source_name: "MarkTechPost"
source_url: https://www.marktechpost.com/2026/08/18/cartesia-ships-sonic-3-6-a-streaming-tts-model-that-now-leads-both-artificial-analysis-speech-arenas/
published: 2026-08-18T10:37:49.000Z
topics: ["research", "products"]
publisher: "Parallel Quant"
---

# Cartesia's new TTS model tops both speech leaderboards

*2026-08-18 · Source: [MarkTechPost](https://www.marktechpost.com/2026/08/18/cartesia-ships-sonic-3-6-a-streaming-tts-model-that-now-leads-both-artificial-analysis-speech-arenas/)*

Cartesia released Sonic-3.6, a streaming text-to-speech model built on state space models instead of transformers. It now ranks #1 on both Artificial Analysis speech arenas, with sub-90 millisecond time-to-first-audio, and is available in beta on Cartesia's API.

**Why it matters:** The result adds to evidence that state-space architectures can outperform transformers for latency-sensitive tasks like real-time voice, a niche where response speed matters more than raw scale. Sub-90ms first-audio latency pushes streaming TTS closer to feeling truly conversational, relevant for voice agents and real-time assistants.

**Topics:** research, products

---
Read the original: https://www.marktechpost.com/2026/08/18/cartesia-ships-sonic-3-6-a-streaming-tts-model-that-now-leads-both-artificial-analysis-speech-arenas/
Canonical: https://www.parallelquant.com/posts/cartesia-s-new-tts-model-tops-both-speech-leaderboards-6cb27f
Published by Parallel Quant — https://www.parallelquant.com
