Audio / Text-to-Speech / Speech Recognition

VibeVoice

Open-source voice AI from Microsoft with both long-form text-to-speech and speech recognition models. The repo highlights 90-minute multi-speaker TTS, 60-minute single-pass ASR, multilingual support, hotwording, and links to docs, Hugging Face, playground, finetuning, and papers.

Clear28/30
Useful28/30
Specific18/20
Complete16/20
VibeVoice screenshot

Why it was accepted

The page clearly presents an AI voice product/research repo with concrete capabilities, model names, and usage links. It shows both TTS and ASR functionality, long-form audio support, multilingual features, and integration paths that make it useful for AI builders and users.

Weakness

The crawl does not show installation steps, API details, licensing terms beyond the repo metadata, or enough hands-on examples to tell how easy it is to run locally versus through the linked model pages.

Review status

118 days ago #78 ↓ -58

Last evaluated 118 days ago. Current rank #78. Down 58 spots in the rankings.

Score history

9291899188919190

Related listings

Curia screenshot
#1467 Curia
79

Audio / AI Podcasts

Curia turns a user’s saved links, articles, and threads into produced audio shows and AI podcasts built around their interests.

Demades screenshot
#1471 Demades
79

Audio / Podcasting

Demades lets users turn any topic into a roughly 60-second AI-generated podcast clip with three hosts, without sign-up.