MuseGen uses AI for speech, voice, transcription, music, or other audio workflows.
Recently added
Recently added to the directory
Browse dated community launches and imported or editorial directory additions.
An AI-powered voiceover tool that provides realistic voices for videos, podcasts, and presentations.
Murf AI uses AI for speech, voice, transcription, music, or other audio workflows.
Works with speech, voice, music or other audio using machine-learning models.
A simple notebook demonstrating prompt-based music generation via Mubert API.
"transforming the future of music creation".
Implementation of MusicLM, Google's new SOTA model for music generation using attention networks, in Pytorch.
Tts for Lojban using VITS TTS models.
Letterfork uses AI for speech, voice, transcription, music, or other audio workflows.
Works with speech, voice, music or other audio using machine-learning models.
Deploy and scale machine learning models on Kubernetes. Built for LLMs, embeddings, and speech-to-text.
Open-source project that uses AI for speech, voice, transcription, music, or other audio workflows.
High-quality text-to-speech and voice recognition.
Distribution, publishing, funding, marketing, and a hands-on team for independent artists. We.
Img To Music uses AI for speech, voice, transcription, music, or other audio workflows.
We are a community-driven organization releasing open-source generative audio tools to make music production more accessible and fun for everyone.
Good Tape uses AI for speech, voice, transcription, music, or other audio workflows.
Port of OpenAI's Whisper model in C/C++. It can be executed locally.
Speech recognition toolkit with streaming ASR, VAD, punctuation, speaker diarization, and OpenAI-compatible serving for voice AI applications.
結婚式スピーチ・弔辞・年賀状・退職挨拶・お礼状・お詫び文など、冠婚葬祭と日常の手紙・挨拶文をAIが作成します。シチュエーションを選んで項目を埋めるだけで、そのまま使える文面が完成。82種類のツールを登録不要・完全無料で今すぐ使えます。.
Fast inference engine for whisper in C++ using CTranslate2.
Open-source project that uses AI for speech, voice, transcription, music, or other audio workflows.
Personalized soundscapes to help you focus, relax, and sleep. Backed by neuroscience.
Works with speech, voice, music or other audio using machine-learning models.