Implementation of MusicLM, Google's new SOTA model for music generation using attention networks, in Pytorch.
Directory
Search results
Published directory entries matching your search.
Mixxx is a GitHub repository or organization with source code, releases, documentation, or project resources.
A simple notebook demonstrating prompt-based music generation via Mubert API.
musescore-downloader helps download, manage, or archive music, audio, podcasts, or karaoke media.
Open-source project that uses AI for speech, voice, transcription, music, or other audio workflows.
An open-source framework by NVIDIA for building speech AI systems, including automatic speech recognition and text-to-speech. opensource.
On-device voice dictation for macOS — transcribes a 5-minute clip in 2.8 s; noise-robust, paste at cursor. 99 languages, ~8 MB, no telemetry. MIT.
Open-source project that uses AI for speech, voice, transcription, music, or other audio workflows.
EPUB to audiobook converter, optimized for Audiobookshelf.
PodcastBulkDownloader is a GitHub repository or organization with source code, releases, documentation, or project resources.
podgrab helps download, manage, or archive music, audio, podcasts, or karaoke media.
qobuz-dl helps download, manage, or archive music, audio, podcasts, or karaoke media.
QobuzDownloaderX is a GitHub repository or organization with source code, releases, documentation, or project resources.
QobuzDownloaderX-MOD is a GitHub repository or organization with source code, releases, documentation, or project resources.
Open-source project that uses AI for speech, voice, transcription, music, or other audio workflows.
scdl helps download, manage, or archive music, audio, podcasts, or karaoke media.
An Optimized Speech-to-Text Pipeline for the Whisper Model.
streamrip helps download, manage, or archive music, audio, podcasts, or karaoke media.
TagEditor is a GitHub repository or organization with source code, releases, documentation, or project resources.
Tidal-Media-Downloader helps download, manage, or archive music, audio, podcasts, or karaoke media.
Open-source project that uses AI for speech, voice, transcription, music, or other audio workflows.
Open-source project that uses AI for speech, voice, transcription, music, or other audio workflows.
Open-source project that uses AI for speech, voice, transcription, music, or other audio workflows.
Accelerates transcription with the combination of OpenAI's Whisper Large v2, HF Transformers, Optimum, and flash attention.