unsloth
Spark-TTS is an advanced text-to-speech system that uses the power of large language models (LLM) for highly accurate and natural-sounding voice synthesis. It is designed to be efficient, flexible, and powerful for both research and production use.
Running this yourself: can likely run on your own machine.
---
Quality Score
---
Arena ELO
500M
Parameters
---
Context
This measures the amount of verifiable public evidence we have, not how capable the model is. A missing field means it has not been verified yet, not that its value is zero.
10 of 22 public signals
Sign in to join the discussion
586
Downloads
17
Likes
May 2025
Released
2/5 signals
0/4 signals
3/5 signals
3/4 signals
2/4 signals
Parameters
—
Training compute
Not reported
Dataset scale
Not reported
Base model
Not reported
Gaps we are still tracking
Metadata sources
This section gives the plain-English basics first: what the model is, how large it is, how much context it can handle, and whether it is still actively supported.
This tells you how you can use the model in practice: whether weights are open, whether an API exists, and what kinds of input or output it supports.
Capabilities