Fast inference engine for whisper in C++ using CTranslate2.
Directory
Search results
Published directory entries matching your search.
Turn text, scripts, and blog posts into videos with 2,000+ AI voices in 80+ languages. Free AI video generator - no camera, no editing skills needed.
結婚式スピーチ・弔辞・年賀状・退職挨拶・お礼状・お詫び文など、冠婚葬祭と日常の手紙・挨拶文をAIが作成します。シチュエーションを選んで項目を埋めるだけで、そのまま使える文面が完成。82種類のツールを登録不要・完全無料で今すぐ使えます。.
Port of OpenAI's Whisper model in C/C++. It can be executed locally.
We are a community-driven organization releasing open-source generative audio tools to make music production more accessible and fun for everyone.
Distribution, publishing, funding, marketing, and a hands-on team for independent artists. We.
Deploy and scale machine learning models on Kubernetes. Built for LLMs, embeddings, and speech-to-text.
Tts for Lojban using VITS TTS models.
Implementation of MusicLM, Google's new SOTA model for music generation using attention networks, in Pytorch.
"transforming the future of music creation".
An AI-powered voiceover tool that provides realistic voices for videos, podcasts, and presentations.
Discover amazing ML apps made by the community.
An open-source framework by NVIDIA for building speech AI systems, including automatic speech recognition and text-to-speech. opensource.
Voice AI that turns calls into outcomes - and gets sharper with each one. Our own model, built for the phone. Sub-400ms response. Live in days.
NeMo Parakeet ASR Models attain strong speech recognition accuracy while being efficient for inference. Available in CTC and RNN-Transducer variants.
A "machine learning framework to automate text-and voice-based conversations.".
Rev AI, part of the Rev family, is a developer-first API that delivers industry- accuracy and fast performance at global scale. Click to learn more.
Never miss another call with our AI-powered call answering service. 10x better than voicemail. 10x cheaper than a traditional phone answering service.
An Optimized Speech-to-Text Pipeline for the Whisper Model.
Turn any idea into scroll-stopping Shorts, Reels, and TikToks with AI visuals, studio voiceovers and synced captions.
Enterprise-grade Voice AI simulation SDK for scenario-driven stress testing of multimodal and agentic systems.
This premium domain name is available for purchase!
AI-driven sound and music generation.
Speechify reads anything aloud to you. Listen to books, PDFs, or web pages anytime with natural voices. Try Speechify free.