Skip to content

Entry

Bark

Appears in 6 awesome lists

Bark is a transformer-based text-to-audio model created by Suno. Bark can generate highly realistic, multilingual speech as well as other audio - including music, background noise and simple sound effects.

Open github.comsuno-ai/bark

Found in these lists

AI Game DevTools (AI-GDT)

Section: Speech · Text-Prompted Generative Audio Model.

ActiveScore 77

Awesome AI Tools

Section: Speech · A transformer-based text-to-audio model. #opensource

SlowScore 67

Awesome AI Tools for Game Developers

Section: Voice Generation · (Open Source): Bark can generate highly realistic, multilingual speech as well as other audio - including music, background noise and simple sound effects.

StaleScore 52

awesome-ChatGPT-repositories

Section: Prompts · 🔊 Text-Prompted Generative Audio Model

FreshScore 87

Awesome LLMOps

Section: Audio Foundation Model · Bark is a transformer-based text-to-audio model created by Suno. Bark can generate highly realistic, multilingual speech as well as other audio - including music, background noise and simple sound effects.

ActiveScore 75

Awesome Generative AI

Section: Text-to-speech · A transformer-based text-to-audio model. #opensource

FreshScore 90

Whisper

Whisper is a general-purpose speech recognition model that can be run locally offline. It can transcribe audio from and to multiple languages.

In 11 listsDetails

VibeVoice

VibeVoice is a novel framework designed for generating expressive, long-form, multi-speaker conversational audio, such as podcasts, from text. It addresses significant challenges in traditional Text-to-Speech (TTS) systems, particularly in scalability, speaker consistency, and natural turn-taking.

In 5 listsDetails

TorToiSe

A multi-voice TTS system trained with an emphasis on quality github | research paper | demo

In 5 listsDetails

kittentts

Kitten TTS is an open-source realistic text-to-speech model with just 15 million parameters, designed for lightweight deployment and high-quality voice synthesis.

In 4 listsDetails

TTS

🐸💬 - a deep learning toolkit for Text-to-Speech, battle-tested in research and production

In 4 listsDetails

VoxCPM

Open-sourced tokenizer-free multilingual speech synthesis model with high-quality TTS and style transfer workflows.

In 4 listsDetails

ChatTTS

Generative TTS model optimized for natural, expressive daily dialogue with fine-grained prosody control.

In 4 listsDetails

Fliki

Create text to video and text to speech content with ai powered voices in minutes.

In 3 lists