Open-source project that uses AI for speech, voice, transcription, music, or other audio workflows.
Directory
Search results
Published directory entries matching your search.
Voicery uses AI for speech, voice, transcription, music, or other audio workflows.
Turn your voice into any instrument with AI. Upload vocals and convert to piano, guitar, violin, sax and 100+ instruments. Free online tool.
Works with speech, voice, music or other audio using machine-learning models.
Open-source project that uses AI for speech, voice, transcription, music, or other audio workflows.
WellSaid uses AI for speech, voice, transcription, music, or other audio workflows.
Whisper API uses AI for speech, voice, transcription, music, or other audio workflows.
Works with speech, voice, music or other audio using machine-learning models.
Works with speech, voice, music or other audio using machine-learning models.
WolframTones uses AI for speech, voice, transcription, music, or other audio workflows.
Works with speech, voice, music or other audio using machine-learning models.
YobiYoba uses AI for speech, voice, transcription, music, or other audio workflows.
Zenmic.com uses AI for speech, voice, transcription, music, or other audio workflows.
Open-source project that uses AI for speech, voice, transcription, music, or other audio workflows.
Zyphra uses AI for speech, voice, transcription, music, or other audio workflows.
Encrypted messaging app for private texts, voice calls, video calls, and group chats.
Discord invite for joining a related community, support server, updates, or project discussion.
An alternative Discord client with voice support made with C++ and GTK 3
AI Mastering is an automated online audio mastering service using AI. Free mastering is available.
YumCut - free AI video generator to turn a prompt into ready vertical videos for TikTok, Reels and YouTube Shorts. Auto script, scenes, voiceover, subtitles and watermark. Built with Next.js. Local-first pipeline + templates, batch rendering and API hooks for creators and indie makers. Self-hosted, FFmpeg-ready, multi-language output. Low cost fast
Audio generation using diffusion models, in PyTorch.
(USA) A API company for advanced Speech-to-Text, offering highly accurate transcription, summarization, and audio intelligence.
AudioCraft is a single-stop code base for all your generative audio needs: music, sound effects, and compression after training on raw audio signals.
Text-to-Audio Generation with Latent Diffusion Models - Speech Research.