YumCut - free AI video generator to turn a prompt into ready vertical videos for TikTok, Reels and YouTube Shorts. Auto script, scenes, voiceover, subtitles and watermark. Built with Next.js. Local-first pipeline + templates, batch rendering and API hooks for creators and indie makers. Self-hosted, FFmpeg-ready, multi-language output. Low cost fast
Directory
Search results
Published directory entries matching your search.
Audio generation using diffusion models, in PyTorch.
(USA) A API company for advanced Speech-to-Text, offering highly accurate transcription, summarization, and audio intelligence.
AudioCraft is a single-stop code base for all your generative audio needs: music, sound effects, and compression after training on raw audio signals.
Text-to-Audio Generation with Latent Diffusion Models - Speech Research.
Balabolka is a text-to-speech application (freeware).
Generate full songs with AI for free. Describe your idea — Boppy writes lyrics and creates a complete track in minutes. No signup, no credit card.
Text-to-speech solutions with character.
Free speech-to-text tool for content creators that accurately transcribes audio & video files up to 2GB.
Discover amazing ML apps made by the community.
Discord invite for joining a related community, support server, updates, or project discussion.
Personalized soundscapes to help you focus, relax, and sleep. Backed by neuroscience.
Fast inference engine for whisper in C++ using CTranslate2.
Turn text, scripts, and blog posts into videos with 2,000+ AI voices in 80+ languages. Free AI video generator - no camera, no editing skills needed.
Port of OpenAI's Whisper model in C/C++. It can be executed locally.
Discord invite for joining a related community, support server, updates, or project discussion.
Distribution, publishing, funding, marketing, and a hands-on team for independent artists. We.
Discord invite for joining a related community, support server, updates, or project discussion.
Voice-first AI Assistant for online meetings that can actively participate and solve tasks live during the meeting.
Deploy and scale machine learning models on Kubernetes. Built for LLMs, embeddings, and speech-to-text.
Tts for Lojban using VITS TTS models.
Implementation of MusicLM, Google's new SOTA model for music generation using attention networks, in Pytorch.
Fine-tuned on AMD MI300X for the AMD Developer Hackathon 2026 (Fine-Tuning Track).
"transforming the future of music creation".