Models & LLMs

Suno's New Speech Feature Blends Audio with Music

Suno introduces a new feature that combines spoken text with background music in a single audio track, aimed at creating content like poems and bedtime stories.

The Decoder · Oct 02, 2026

What happened

  • Suno is adding a new feature called 'Speech' to its AI music generator.
  • The feature allows users to input text and describe voice and music style preferences.
  • Suno has not disclosed how the model was trained.

Why it matters

Suno's new Speech feature represents a significant advancement in AI-generated audio, blending spoken text with music. However, the lack of transparency about training methods and ongoing issues like accent misclassification raise questions about reliability and ethical use of AI in content creation.

The Elephant take

🐘 prehistoric 🐘 Suno’s new 'Speech' feature is a step forward in AI audio, but the company’s silence on training methods and the persistent bugs suggest a fragile experiment. The legal issues with copyright also loom large.

Who should care

  • AI developers
  • Content creators
  • Legal professionals

What to do next

  1. Test the feature for accuracy
  2. Monitor legal developments
  3. Evaluate ethical implications of AI-generated content

Keep in mind

The feature is still in beta and has known issues like accent misclassification.

Read the original reporting at The Decoder ↗