FreeTTS uses AI for speech, voice, transcription, music, or other audio workflows.
Directory
Search results
Published directory entries matching your search.
Speech recognition toolkit with streaming ASR, VAD, punctuation, speaker diarization, and OpenAI-compatible serving for voice AI applications.
Port of OpenAI's Whisper model in C/C++. It can be executed locally.
Google Flow Music uses AI for speech, voice, transcription, music, or other audio workflows.
Open-source project that uses AI for speech, voice, transcription, music, or other audio workflows.
headphones helps download, manage, or archive music, audio, podcasts, or karaoke media.
Hume uses AI for speech, voice, transcription, music, or other audio workflows.
Distribution, publishing, funding, marketing, and a hands-on team for independent artists. We.
High-quality text-to-speech and voice recognition.
A curated list of resources of audio-driven talking face generation.
KikiVoice uses AI for speech, voice, transcription, music, or other audio workflows.
Open-source project that uses AI for speech, voice, transcription, music, or other audio workflows.
Open-source project that uses AI for speech, voice, transcription, music, or other audio workflows.
Kokoro TTS uses AI for speech, voice, transcription, music, or other audio workflows.
Deploy and scale machine learning models on Kubernetes. Built for LLMs, embeddings, and speech-to-text.
Kyutai TTS uses AI for speech, voice, transcription, music, or other audio workflows.
LazyPy uses AI for speech, voice, transcription, music, or other audio workflows.
Tts for Lojban using VITS TTS models.
LOVO uses AI for speech, voice, transcription, music, or other audio workflows.
Implementation of MusicLM, Google's new SOTA model for music generation using attention networks, in Pytorch.
"transforming the future of music creation".
Mazmazika uses AI for speech, voice, transcription, music, or other audio workflows.
MMAudio uses AI for speech, voice, transcription, music, or other audio workflows.
Moe TTS uses AI for speech, voice, transcription, music, or other audio workflows.