Creates, edits or enhances video and animation with generative models.
Directory
Search results
Published directory entries matching your search.
Gen-2 by Runway uses AI to create, edit, transform, or animate video content.
Turn text, scripts, and blog posts into videos with 2,000+ AI voices in 80+ languages. Free AI video generator - no camera, no editing skills needed.
Creates, edits or enhances video and animation with generative models.
Depth-Aware Video Frame Interpolation (CVPR 2019).
Open-source project that uses AI for speech, voice, transcription, music, or other audio workflows.
Open-source project that uses AI for speech, voice, transcription, music, or other audio workflows.
Kokoro-82M supports machine learning models, deployment, inspection, datasets, or AI development workflows.
Open-source project that uses AI for speech, voice, transcription, music, or other audio workflows.
Discord invite for joining a related community, support server, updates, or project discussion.
Ebook2audiobook uses AI for speech, voice, transcription, music, or other audio workflows.
Discord invite for joining a related community, support server, updates, or project discussion.
Paper2Audio uses AI for speech, voice, transcription, music, or other audio workflows.
A Whisper CLI client compatible with the original OpenAI client, using CTranslate2 for faster inference. opensource.
A simple and efficient end-to-end Automatic Speech Recognition (ASR) system from Facebook AI Research.
Space emoji (emoji-only character allowed).
Accelerates transcription with the combination of OpenAI's Whisper Large v2, HF Transformers, Optimum, and flash attention.
Examples for ICASSP2024 paper “StemGen: A music generation model that listens”.
An Optimized Speech-to-Text Pipeline for the Whisper Model.
Robust Speech Recognition Leaderboard 2022 uses AI for speech, voice, transcription, music, or other audio workflows.
On-device voice dictation for macOS — transcribes a 5-minute clip in 2.8 s; noise-robust, paste at cursor. 99 languages, ~8 MB, no telemetry. MIT.
結婚式スピーチ・弔辞・年賀状・退職挨拶・お礼状・お詫び文など、冠婚葬祭と日常の手紙・挨拶文をAIが作成します。シチュエーションを選んで項目を埋めるだけで、そのまま使える文面が完成。82種類のツールを登録不要・完全無料で今すぐ使えます。.
Fast inference engine for whisper in C++ using CTranslate2.
Free speech-to-text tool for content creators that accurately transcribes audio & video files up to 2GB.