Deploy and scale machine learning models on Kubernetes. Built for LLMs, embeddings, and speech-to-text.
Directory
Search results
Published directory entries matching your search.
Tts for Lojban using VITS TTS models.
"transforming the future of music creation".
An AI-powered voiceover tool that provides realistic voices for videos, podcasts, and presentations.
Discover amazing ML apps made by the community.
An open-source framework by NVIDIA for building speech AI systems, including automatic speech recognition and text-to-speech. opensource.
Otter AI Meeting Agent supports real-time transcription, live chat, automated summaries, insights, and action items.
NeMo Parakeet ASR Models attain strong speech recognition accuracy while being efficient for inference. Available in CTC and RNN-Transducer variants.
Rev AI, part of the Rev family, is a developer-first API that delivers industry- accuracy and fast performance at global scale. Click to learn more.
An Optimized Speech-to-Text Pipeline for the Whisper Model.
Turn any idea into scroll-stopping Shorts, Reels, and TikToks with AI visuals, studio voiceovers and synced captions.
This premium domain name is available for purchase!
Speechify reads anything aloud to you. Listen to books, PDFs, or web pages anytime with natural voices. Try Speechify free.
Create AI generated videos from text with the most advanced AI avatars and voiceovers in 160+ languages. Try our free AI video generator now!
AI transcription & translation for audio/video. Two services: transcribe or translate in 50+ languages. 98% accuracy, pay-as-you-go.
Accelerates transcription with the combination of OpenAI's Whisper Large v2, HF Transformers, Optimum, and flash attention.
Space emoji (emoji-only character allowed).
VoiceSphere AI: Streamline document handling for PDFs, DOCs, PPTs, videos, texts. Fast, precise answers. #VoiceSphereAI #AIChat #DocumentManagement.
An offline speech recognition toolkit with C++ support, designed for low-resource devices and multiple languages.
A simple and efficient end-to-end Automatic Speech Recognition (ASR) system from Facebook AI Research.
USE BAZEL VERSION=5.0.0./bazelisk-linux-amd64 build wavegru mod -c opt --copt=-march=native.
A Whisper CLI client compatible with the original OpenAI client, using CTranslate2 for faster inference. opensource.
Port of OpenAI's Whisper model in C/C++. opensource.
Al Watan Voice publishes newspaper reporting, headlines and current-affairs coverage in Palestine.