Works with speech, voice, music or other audio using machine-learning models.
Recently added
Recently added to the directory
Browse dated community launches and imported or editorial directory additions.
AI-driven sound and music generation.
This premium domain name is available for purchase!
Works with speech, voice, music or other audio using machine-learning models.
Works with speech, voice, music or other audio using machine-learning models.
Turn any idea into scroll-stopping Shorts, Reels, and TikToks with AI visuals, studio voiceovers and synced captions.
An Optimized Speech-to-Text Pipeline for the Whisper Model.
Robust Speech Recognition Leaderboard 2022 uses AI for speech, voice, transcription, music, or other audio workflows.
Works with speech, voice, music or other audio using machine-learning models.
Rev AI, part of the Rev family, is a developer-first API that delivers industry- accuracy and fast performance at global scale. Click to learn more.
Professional AI voice generator for real production workflows, with free testing and flexible integration for studios and media teams.
Resemble AI uses AI for speech, voice, transcription, music, or other audio workflows.
Replica Studios uses AI for speech, voice, transcription, music, or other audio workflows.
Works with speech, voice, music or other audio using machine-learning models.
ReadSpeaker uses AI for speech, voice, transcription, music, or other audio workflows.
Read AI uses AI for speech, voice, transcription, music, or other audio workflows.
Works with speech, voice, music or other audio using machine-learning models.
NeMo Parakeet ASR Models attain strong speech recognition accuracy while being efficient for inference. Available in CTC and RNN-Transducer variants.
Otter AI Meeting Agent supports real-time transcription, live chat, automated summaries, insights, and action items.
On-device voice dictation for macOS — transcribes a 5-minute clip in 2.8 s; noise-robust, paste at cursor. 99 languages, ~8 MB, no telemetry. MIT.
APIs for messaging, voice, and phone verification.
An open-source framework by NVIDIA for building speech AI systems, including automatic speech recognition and text-to-speech. opensource.
Discover amazing ML apps made by the community.
MusicGen uses AI for speech, voice, transcription, music, or other audio workflows.