Accelerates transcription with the combination of OpenAI's Whisper Large v2, HF Transformers, Optimum, and flash attention.
Directory
Search results
Published directory entries matching your search.
Tunisian Speech Recognition uses AI for speech, voice, transcription, music, or other audio workflows.
AI transcription & translation for audio/video. Two services: transcribe or translate in 50+ languages. 98% accuracy, pay-as-you-go.
The Prompter vicc Substack uses AI for speech, voice, transcription, music, or other audio workflows.
Examples for ICASSP2024 paper “StemGen: A music generation model that listens”.
Speechify reads anything aloud to you. Listen to books, PDFs, or web pages anytime with natural voices. Try Speechify free.
Robust Speech Recognition Leaderboard 2022 uses AI for speech, voice, transcription, music, or other audio workflows.
Professional AI voice generator for real production workflows, with free testing and flexible integration for studios and media teams.
Resemble AI uses AI for speech, voice, transcription, music, or other audio workflows.
Replica Studios uses AI for speech, voice, transcription, music, or other audio workflows.
ReadSpeaker uses AI for speech, voice, transcription, music, or other audio workflows.
Read AI uses AI for speech, voice, transcription, music, or other audio workflows.
Otter AI Meeting Agent supports real-time transcription, live chat, automated summaries, insights, and action items.
On-device voice dictation for macOS — transcribes a 5-minute clip in 2.8 s; noise-robust, paste at cursor. 99 languages, ~8 MB, no telemetry. MIT.
APIs for messaging, voice, and phone verification.
MusicGen uses AI for speech, voice, transcription, music, or other audio workflows.
MuseGen uses AI for speech, voice, transcription, music, or other audio workflows.
An AI-powered voiceover tool that provides realistic voices for videos, podcasts, and presentations.
Murf AI uses AI for speech, voice, transcription, music, or other audio workflows.
Letterfork uses AI for speech, voice, transcription, music, or other audio workflows.
Deploy and scale machine learning models on Kubernetes. Built for LLMs, embeddings, and speech-to-text.
Open-source project that uses AI for speech, voice, transcription, music, or other audio workflows.
Img To Music uses AI for speech, voice, transcription, music, or other audio workflows.
We are a community-driven organization releasing open-source generative audio tools to make music production more accessible and fun for everyone.