Examples for ICASSP2024 paper “StemGen: A music generation model that listens”.
Directory
Search results
Published directory entries matching your search.
Suno uses AI for speech, voice, transcription, music, or other audio workflows.
NaturalReader online reader converts text into spoken audio for listening and accessibility.
Open-source project that uses AI for speech, voice, transcription, music, or other audio workflows.
Treblo uses AI for speech, voice, transcription, music, or other audio workflows.
Open-source project that uses AI for speech, voice, transcription, music, or other audio workflows.
TTS Online uses AI for speech, voice, transcription, music, or other audio workflows.
Open-source project that uses AI for speech, voice, transcription, music, or other audio workflows.
TTS-WebUI uses AI for speech, voice, transcription, music, or other audio workflows.
Vocal Remover uses AI for speech, voice, transcription, music, or other audio workflows.
Accelerates transcription with the combination of OpenAI's Whisper Large v2, HF Transformers, Optimum, and flash attention.
Space emoji (emoji-only character allowed).
VanillaVoice uses AI for speech, voice, transcription, music, or other audio workflows.
Vocali.se uses AI for speech, voice, transcription, music, or other audio workflows.
VocalRemover uses AI for speech, voice, transcription, music, or other audio workflows.
Open-source project that uses AI for speech, voice, transcription, music, or other audio workflows.
Open-source project that uses AI for speech, voice, transcription, music, or other audio workflows.
Voice Models uses AI for speech, voice, transcription, music, or other audio workflows.
Open-source project that uses AI for speech, voice, transcription, music, or other audio workflows.
VoiceSphere AI: Streamline document handling for PDFs, DOCs, PPTs, videos, texts. Fast, precise answers. #VoiceSphereAI #AIChat #DocumentManagement.
Turn your voice into any instrument with AI. Upload vocals and convert to piano, guitar, violin, sax and 100+ instruments. Free online tool.
An offline speech recognition toolkit with C++ support, designed for low-resource devices and multiple languages.
A simple and efficient end-to-end Automatic Speech Recognition (ASR) system from Facebook AI Research.
Open-source project that uses AI for speech, voice, transcription, music, or other audio workflows.