Deploy and scale machine learning models on Kubernetes. Built for LLMs, embeddings, and speech-to-text.
Directory
Search results
Published directory entries matching your search.
Works with speech, voice, music or other audio using machine-learning models.
Letterfork uses AI for speech, voice, transcription, music, or other audio workflows.
The state of the art AI image generation engine.
Tts for Lojban using VITS TTS models.
"transforming the future of music creation".
AI-powered design tools including image generation, background removal, and creative templates.
Creates or edits images with generative models and visual controls.
Works with speech, voice, music or other audio using machine-learning models.
Murf AI uses AI for speech, voice, transcription, music, or other audio workflows.
An AI-powered voiceover tool that provides realistic voices for videos, podcasts, and presentations.
MuseGen uses AI for speech, voice, transcription, music, or other audio workflows.
MusicGen uses AI for speech, voice, transcription, music, or other audio workflows.
Discover amazing ML apps made by the community.
An open-source framework by NVIDIA for building speech AI systems, including automatic speech recognition and text-to-speech. opensource.
APIs for messaging, voice, and phone verification.
NightCafe Creator is an AI Art Generator app with multiple methods of AI art generation.
On-device voice dictation for macOS — transcribes a 5-minute clip in 2.8 s; noise-robust, paste at cursor. 99 languages, ~8 MB, no telemetry. MIT.
Otter AI Meeting Agent supports real-time transcription, live chat, automated summaries, insights, and action items.
NeMo Parakeet ASR Models attain strong speech recognition accuracy while being efficient for inference. Available in CTC and RNN-Transducer variants.
Pixray is an image generation system.
Creates or edits images with generative models and visual controls.
Works with speech, voice, music or other audio using machine-learning models.
Read AI uses AI for speech, voice, transcription, music, or other audio workflows.