Model
Trends
.ai
Model
Trends
.ai
the best open source ai models
AD
Install / Run on Graydient
Requires membership
/wf /run:audio-miso
About This Model
MISOTTS · Workflows
Miso Text-to-Speech S 8B is a text-to-speech model based on the Sesame CSM architecture. It generates Mimi audio codes from text and optional audio context, using a large Llama 3.2-style backbone and a smaller autoregressive audio decoder. The model is designed for high-quality conversational speech generation and voice continuation from prompt audio.
ModelTrends.ai Model ID
#28002
Flag content
Similar Models
No similar models found
AI Models for Image, Video & Text Generation
audio miso is a workflow for the MisoTTS family listed on ModelTrends.ai, a read-only catalog of open source AI image, video and text models. Compare it with other MisoTTS models, check its heat score to see how it is trending, and open it on PirateDiffusion or BitVector to try it.
Browse by Model Family
Browse all models