Open-source project that uses AI for speech, voice, transcription, music, or other audio workflows.
Directory
Search results
Published directory entries matching your search.
Audio.Z.AI uses AI for speech, voice, transcription, music, or other audio workflows.
AudioArena uses AI for speech, voice, transcription, music, or other audio workflows.
AudioCraft is a single-stop code base for all your generative audio needs: music, sound effects, and compression after training on raw audio signals.
Text-to-Audio Generation with Latent Diffusion Models - Speech Research.
You can also find more comprehensive list on and There's an AI AI Voice Cloning list.
Balabolka is a text-to-speech application (freeware).
Open-source project that uses AI for speech, voice, transcription, music, or other audio workflows.
Bespoke Sounds uses AI for speech, voice, transcription, music, or other audio workflows.
Boomy uses AI for speech, voice, transcription, music, or other audio workflows.
Theano based library for deep and recurrent neural networks.
Works with speech, voice, music or other audio using machine-learning models.
Cartesia uses AI for speech, voice, transcription, music, or other audio workflows.
Text-to-speech solutions with character.
Chatterbox uses AI for speech, voice, transcription, music, or other audio workflows.
Open-source project that uses AI for speech, voice, transcription, music, or other audio workflows.
CMU Sphinx uses AI for speech, voice, transcription, music, or other audio workflows.
A comparative framework for multimodal recommender systems with a focus on models leveraging auxiliary data.
Free speech-to-text tool for content creators that accurately transcribes audio & video files up to 2GB.
Course materials and notes for Stanford class CS231n: Deep Learning for Computer Vision.
Works with speech, voice, music or other audio using machine-learning models.
Discover amazing ML apps made by the community.
A python library for accurate and scalable fuzzy matching, record deduplication and entity-resolution.
Demo uses AI for speech, voice, transcription, music, or other audio workflows.