Open-source project that uses AI for speech, voice, transcription, music, or other audio workflows.
Directory
Search results
Published directory entries matching your search.
Speech recognition toolkit with streaming ASR, VAD, punctuation, speaker diarization, and OpenAI-compatible serving for voice AI applications.
Good Tape uses AI for speech, voice, transcription, music, or other audio workflows.
Img To Music uses AI for speech, voice, transcription, music, or other audio workflows.
High-quality text-to-speech and voice recognition.
Open-source project that uses AI for speech, voice, transcription, music, or other audio workflows.
Deploy and scale machine learning models on Kubernetes. Built for LLMs, embeddings, and speech-to-text.
Works with speech, voice, music or other audio using machine-learning models.
Letterfork uses AI for speech, voice, transcription, music, or other audio workflows.
Tts for Lojban using VITS TTS models.
Murf AI uses AI for speech, voice, transcription, music, or other audio workflows.
MuseGen uses AI for speech, voice, transcription, music, or other audio workflows.
MusicGen uses AI for speech, voice, transcription, music, or other audio workflows.
NeMo Parakeet ASR Models attain strong speech recognition accuracy while being efficient for inference. Available in CTC and RNN-Transducer variants.
Works with speech, voice, music or other audio using machine-learning models.
Read AI uses AI for speech, voice, transcription, music, or other audio workflows.
ReadSpeaker uses AI for speech, voice, transcription, music, or other audio workflows.
Replica Studios uses AI for speech, voice, transcription, music, or other audio workflows.
Resemble AI uses AI for speech, voice, transcription, music, or other audio workflows.
Works with speech, voice, music or other audio using machine-learning models.
An Optimized Speech-to-Text Pipeline for the Whisper Model.
Works with speech, voice, music or other audio using machine-learning models.
Works with speech, voice, music or other audio using machine-learning models.
The Prompter vicc Substack uses AI for speech, voice, transcription, music, or other audio workflows.