Architecture of voice AI, from speech recognition to emotional intelligence, and learn how to build, scale, and evaluate them.
Directory
Search results
Published directory entries matching your search.
Improve productivity by getting a personalized daily audio briefing on updates from your favorite sites/apps.
Discover amazing ML apps made by the community.
Demo uses AI for speech, voice, transcription, music, or other audio workflows.
Eapy uses AI for speech, voice, transcription, music, or other audio workflows.
Ebook2audiobook uses AI for speech, voice, transcription, music, or other audio workflows.
Open-source project that uses AI for speech, voice, transcription, music, or other audio workflows.
ElevenLabs uses AI for speech, voice, transcription, music, or other audio workflows.
Open-source project that uses AI for speech, voice, transcription, music, or other audio workflows.
ElevenReader uses AI for speech, voice, transcription, music, or other audio workflows.
Open-source project that uses AI for speech, voice, transcription, music, or other audio workflows.
En uses AI for speech, voice, transcription, music, or other audio workflows.
Personalized soundscapes to help you focus, relax, and sleep. Backed by neuroscience.
Ezstems uses AI for speech, voice, transcription, music, or other audio workflows.
FakeYou uses AI for speech, voice, transcription, music, or other audio workflows.
Fast inference engine for whisper in C++ using CTranslate2.
FreeTTS uses AI for speech, voice, transcription, music, or other audio workflows.
Speech recognition toolkit with streaming ASR, VAD, punctuation, speaker diarization, and OpenAI-compatible serving for voice AI applications.
Port of OpenAI's Whisper model in C/C++. It can be executed locally.
Google Flow Music uses AI for speech, voice, transcription, music, or other audio workflows.
Open-source project that uses AI for speech, voice, transcription, music, or other audio workflows.
Hume uses AI for speech, voice, transcription, music, or other audio workflows.
Distribution, publishing, funding, marketing, and a hands-on team for independent artists. We.
High-quality text-to-speech and voice recognition.