Implementation of MusicLM, Google's new SOTA model for music generation using attention networks, in Pytorch.
Directory
Search results
Published directory entries matching your search.
"transforming the future of music creation".
A simple notebook demonstrating prompt-based music generation via Mubert API.
Works with speech, voice, music or other audio using machine-learning models.
An AI-powered voiceover tool that provides realistic voices for videos, podcasts, and presentations.
Discover amazing ML apps made by the community.
An open-source framework by NVIDIA for building speech AI systems, including automatic speech recognition and text-to-speech. opensource.
APIs for messaging, voice, and phone verification.
On-device voice dictation for macOS — transcribes a 5-minute clip in 2.8 s; noise-robust, paste at cursor. 99 languages, ~8 MB, no telemetry. MIT.
NeMo Parakeet ASR Models attain strong speech recognition accuracy while being efficient for inference. Available in CTC and RNN-Transducer variants.
Works with speech, voice, music or other audio using machine-learning models.
Works with speech, voice, music or other audio using machine-learning models.
Professional AI voice generator for real production workflows, with free testing and flexible integration for studios and media teams.
Rev AI, part of the Rev family, is a developer-first API that delivers industry- accuracy and fast performance at global scale. Click to learn more.
Works with speech, voice, music or other audio using machine-learning models.
An Optimized Speech-to-Text Pipeline for the Whisper Model.
Turn any idea into scroll-stopping Shorts, Reels, and TikToks with AI visuals, studio voiceovers and synced captions.
Works with speech, voice, music or other audio using machine-learning models.
Works with speech, voice, music or other audio using machine-learning models.
This premium domain name is available for purchase!
AI-driven sound and music generation.
Works with speech, voice, music or other audio using machine-learning models.
Works with speech, voice, music or other audio using machine-learning models.
Speechify reads anything aloud to you. Listen to books, PDFs, or web pages anytime with natural voices. Try Speechify free.