CMU Sphinx uses AI for speech, voice, transcription, music, or other audio workflows.
Directory
Search results
Published directory entries matching your search.
Architecture of voice AI, from speech recognition to emotional intelligence, and learn how to build, scale, and evaluate them.
Works with speech, voice, music or other audio using machine-learning models.
Open-source project that uses AI for speech, voice, transcription, music, or other audio workflows.
Speech recognition toolkit with streaming ASR, VAD, punctuation, speaker diarization, and OpenAI-compatible serving for voice AI applications.
Good Tape uses AI for speech, voice, transcription, music, or other audio workflows.
Img To Music uses AI for speech, voice, transcription, music, or other audio workflows.
Open-source project that uses AI for speech, voice, transcription, music, or other audio workflows.
Works with speech, voice, music or other audio using machine-learning models.
Letterfork uses AI for speech, voice, transcription, music, or other audio workflows.
Murf AI uses AI for speech, voice, transcription, music, or other audio workflows.
MuseGen uses AI for speech, voice, transcription, music, or other audio workflows.
MusicGen uses AI for speech, voice, transcription, music, or other audio workflows.
NeMo Parakeet ASR Models attain strong speech recognition accuracy while being efficient for inference. Available in CTC and RNN-Transducer variants.
Works with speech, voice, music or other audio using machine-learning models.
Read AI uses AI for speech, voice, transcription, music, or other audio workflows.
ReadSpeaker uses AI for speech, voice, transcription, music, or other audio workflows.
Replica Studios uses AI for speech, voice, transcription, music, or other audio workflows.
Resemble AI uses AI for speech, voice, transcription, music, or other audio workflows.
Works with speech, voice, music or other audio using machine-learning models.
Works with speech, voice, music or other audio using machine-learning models.
Audio jammer generator that masks speech with noise for privacy or sound experiments.
SpeechChat is a social media utility for platform customization, archives, embeds, or community tools.
The Prompter vicc Substack uses AI for speech, voice, transcription, music, or other audio workflows.