Voicery uses AI for speech, voice, transcription, music, or other audio workflows.
Directory
Search results
Published directory entries matching your search.
An offline speech recognition toolkit with C++ support, designed for low-resource devices and multiple languages.
A simple and efficient end-to-end Automatic Speech Recognition (ASR) system from Facebook AI Research.
Works with speech, voice, music or other audio using machine-learning models.
USE BAZEL VERSION=5.0.0./bazelisk-linux-amd64 build wavegru mod -c opt --copt=-march=native.
WellSaid uses AI for speech, voice, transcription, music, or other audio workflows.
Whisper API uses AI for speech, voice, transcription, music, or other audio workflows.
Works with speech, voice, music or other audio using machine-learning models.
Works with speech, voice, music or other audio using machine-learning models.
YobiYoba uses AI for speech, voice, transcription, music, or other audio workflows.
Zenmic.com uses AI for speech, voice, transcription, music, or other audio workflows.
"transforming the future of music creation".
A free AI voice generator that generates natural sounding text-to-speech voice overs.
Examples for ICASSP2024 paper “StemGen: A music generation model that listens”.
Google AI Mode is an AI tool or resource for chat, research, automation, media generation, or model development.
Google Whitepaper is an AI tool or resource for chat, research, automation, media generation, or model development.
Audio generation using diffusion models, in PyTorch.
A simple notebook demonstrating prompt-based music generation via Mubert API.
AI-driven sound and music generation.
Google tool for building, testing, and experimenting with Gemini prompts, models, and API workflows.
Google tool for building, testing, and experimenting with Gemini prompts, models, and API workflows.
AI Mastering is an automated online audio mastering service using AI. Free mastering is available.
AudioCraft is a single-stop code base for all your generative audio needs: music, sound effects, and compression after training on raw audio signals.
You can also find more comprehensive list on and There's an AI AI Voice Cloning list.