Space emoji (emoji-only character allowed).
Directory
AI Audio & Speech
Explore AI Audio & Speech resources in AI & Machine Learning.
Vocova uses AI for speech, voice, transcription, music, or other audio workflows.
Works with speech, voice, music or other audio using machine-learning models.
Voice Dream uses AI for speech, voice, transcription, music, or other audio workflows.
Voicery uses AI for speech, voice, transcription, music, or other audio workflows.
VoiceSphere AI: Streamline document handling for PDFs, DOCs, PPTs, videos, texts. Fast, precise answers. #VoiceSphereAI #AIChat #DocumentManagement.
Turn your voice into any instrument with AI. Upload vocals and convert to piano, guitar, violin, sax and 100+ instruments. Free online tool.
An offline speech recognition toolkit with C++ support, designed for low-resource devices and multiple languages.
A simple and efficient end-to-end Automatic Speech Recognition (ASR) system from Facebook AI Research.
Works with speech, voice, music or other audio using machine-learning models.
USE BAZEL VERSION=5.0.0./bazelisk-linux-amd64 build wavegru mod -c opt --copt=-march=native.
WellSaid uses AI for speech, voice, transcription, music, or other audio workflows.
Whisper API uses AI for speech, voice, transcription, music, or other audio workflows.
Works with speech, voice, music or other audio using machine-learning models.
A Whisper CLI client compatible with the original OpenAI client, using CTranslate2 for faster inference. opensource.
Port of OpenAI's Whisper model in C/C++. opensource.
Works with speech, voice, music or other audio using machine-learning models.
Works with speech, voice, music or other audio using machine-learning models.
YobiYoba uses AI for speech, voice, transcription, music, or other audio workflows.
Zenmic.com uses AI for speech, voice, transcription, music, or other audio workflows.
Audio.Z.AI uses AI for speech, voice, transcription, music, or other audio workflows.
AudioArena uses AI for speech, voice, transcription, music, or other audio workflows.
Fish Audio uses AI for speech, voice, transcription, music, or other audio workflows.
Paper2Audio uses AI for speech, voice, transcription, music, or other audio workflows.