All open source
Open SourceVoice & Media38k

Bark

Transformer text-to-audio model that speaks, sings, laughs, and makes sound effects.

Bark, from Suno, is a transformer-based text-to-audio model that generates highly realistic multilingual speech as well as nonverbal sounds like laughter, sighing, music, and noise. Prompted with bracketed cues and speaker presets, it can produce expressive narration and ambient audio rather than purely clean read-aloud TTS.

Repository

suno-ai/bark

Language

Python

What you'd build with it

  • Generate expressive, emotional narration with laughter, pauses, and tonal cues
  • Produce multilingual voiceovers or character dialogue for games and video
  • Create simple sound effects, music snippets, or ambient audio from text prompts

Tags

ttstext-to-audiomultilingualexpressive