Voicery uses AI for speech, voice, transcription, music, or other audio workflows.
Directory
Search results
Published directory entries matching your search.
Turn your voice into any instrument with AI. Upload vocals and convert to piano, guitar, violin, sax and 100+ instruments. Free online tool.
Works with speech, voice, music or other audio using machine-learning models.
WellSaid uses AI for speech, voice, transcription, music, or other audio workflows.
Whisper API uses AI for speech, voice, transcription, music, or other audio workflows.
Works with speech, voice, music or other audio using machine-learning models.
Works with speech, voice, music or other audio using machine-learning models.
WolframTones uses AI for speech, voice, transcription, music, or other audio workflows.
Works with speech, voice, music or other audio using machine-learning models.
YobiYoba uses AI for speech, voice, transcription, music, or other audio workflows.
Zenmic.com uses AI for speech, voice, transcription, music, or other audio workflows.
Zyphra uses AI for speech, voice, transcription, music, or other audio workflows.
Port of OpenAI's Whisper model in C/C++. It can be executed locally.
We are a community-driven organization releasing open-source generative audio tools to make music production more accessible and fun for everyone.
An open-source framework by NVIDIA for building speech AI systems, including automatic speech recognition and text-to-speech. opensource.
Open Source framework for voice and multimodal conversational AI.
Accelerates transcription with the combination of OpenAI's Whisper Large v2, HF Transformers, Optimum, and flash attention.
A Whisper CLI client compatible with the original OpenAI client, using CTranslate2 for faster inference. opensource.
Port of OpenAI's Whisper model in C/C++. opensource.
↗ external - Open-source dictation that types where you talk.
AI Mastering is an automated online audio mastering service using AI. Free mastering is available.
Audio generation using diffusion models, in PyTorch.
(USA) A API company for advanced Speech-to-Text, offering highly accurate transcription, summarization, and audio intelligence.
AudioCraft is a single-stop code base for all your generative audio needs: music, sound effects, and compression after training on raw audio signals.