Directory

Search results

Published directory entries matching your search.

Search results

5,110 listings
github.com

Open-source project that uses AI for speech, voice, transcription, music, or other audio workflows.

AI Audio & Speech 167
github.com

Open-source project that uses AI for speech, voice, transcription, music, or other audio workflows.

AI Audio & Speech 160
github.com

Open-source project that uses AI for speech, voice, transcription, music, or other audio workflows.

AI Audio & Speech 266
github.com

Open-source project that uses AI for speech, voice, transcription, music, or other audio workflows.

AI Audio & Speech 275
github.com

Open-source project that uses AI for speech, voice, transcription, music, or other audio workflows.

AI Audio & Speech 162
github.com

Open-source project that uses AI for speech, voice, transcription, music, or other audio workflows.

AI Audio & Speech 221
github.com

Speech recognition toolkit with streaming ASR, VAD, punctuation, speaker diarization, and OpenAI-compatible serving for voice AI applications.

AI Audio & Speech 249

Port of OpenAI's Whisper model in C/C++. It can be executed locally.

AI Audio & Speech 345
github.com

On-device voice dictation for macOS — transcribes a 5-minute clip in 2.8 s; noise-robust, paste at cursor. 99 languages, ~8 MB, no telemetry. MIT.

AI Audio & Speech 161

Accelerates transcription with the combination of OpenAI's Whisper Large v2, HF Transformers, Optimum, and flash attention.

AI Audio & Speech 184
github.com

A Whisper CLI client compatible with the original OpenAI client, using CTranslate2 for faster inference. opensource.

AI Audio & Speech 206
github.com

Port of OpenAI's Whisper model in C/C++. opensource.

AI Audio & Speech 291
github.com

You can also find more comprehensive list on and There's an AI AI Voice Cloning list.

AI Audio & Speech 237
github.com

awesome-rss-feeds helps download, manage, or archive music, audio, podcasts, or karaoke media.

RSS & Feed Tools 232
github.com

Fast inference engine for whisper in C++ using CTranslate2.

AI Audio & Speech 317
github.com

headphones helps download, manage, or archive music, audio, podcasts, or karaoke media.

Music Torrents 204
github.com

Deploy and scale machine learning models on Kubernetes. Built for LLMs, embeddings, and speech-to-text.

AI Audio & Speech 157

Implementation of MusicLM, Google's new SOTA model for music generation using attention networks, in Pytorch.

AI Audio & Speech 276

musescore-downloader helps download, manage, or archive music, audio, podcasts, or karaoke media.

Media & Direct Downloaders 230

EPUB to audiobook converter, optimized for Audiobookshelf.

AI Video Generation 152
github.com

podgrab helps download, manage, or archive music, audio, podcasts, or karaoke media.

Podcasts 170
github.com

qobuz-dl helps download, manage, or archive music, audio, podcasts, or karaoke media.

Music Downloads 321