NeMo Parakeet ASR Models attain strong speech recognition accuracy while being efficient for inference. Available in CTC and RNN-Transducer variants.
Directory
Search results
Published directory entries matching your search.
Open Source framework for voice and multimodal conversational AI.
A "machine learning framework to automate text-and voice-based conversations.".
Rev AI, part of the Rev family, is a developer-first API that delivers industry- accuracy and fast performance at global scale. Click to learn more.
Never miss another call with our AI-powered call answering service. 10x better than voicemail. 10x cheaper than a traditional phone answering service.
An Optimized Speech-to-Text Pipeline for the Whisper Model.
Turn any idea into scroll-stopping Shorts, Reels, and TikToks with AI visuals, studio voiceovers and synced captions.
Enterprise-grade Voice AI simulation SDK for scenario-driven stress testing of multimodal and agentic systems.
This premium domain name is available for purchase!
Speechify reads anything aloud to you. Listen to books, PDFs, or web pages anytime with natural voices. Try Speechify free.
Create AI generated videos from text with the most advanced AI avatars and voiceovers in 160+ languages. Try our free AI video generator now!
NaturalReader online reader converts text into spoken audio for listening and accessibility.
Simple Python script to interact with the TikTok TTS API
AI transcription & translation for audio/video. Two services: transcribe or translate in 50+ languages. 98% accuracy, pay-as-you-go.
Accelerates transcription with the combination of OpenAI's Whisper Large v2, HF Transformers, Optimum, and flash attention.
Space emoji (emoji-only character allowed).
Nonprofit newsroom covering San Diego government, education, housing, infrastructure and community issues.
VoiceSphere AI: Streamline document handling for PDFs, DOCs, PPTs, videos, texts. Fast, precise answers. #VoiceSphereAI #AIChat #DocumentManagement.
An offline speech recognition toolkit with C++ support, designed for low-resource devices and multiple languages.
A simple and efficient end-to-end Automatic Speech Recognition (ASR) system from Facebook AI Research.
USE BAZEL VERSION=5.0.0./bazelisk-linux-amd64 build wavegru mod -c opt --copt=-march=native.
Port of OpenAI's Whisper model in C/C++. opensource.
↗ external - Open-source dictation that types where you talk.
A sexy achievement file parser with real-time notification, automatic screenshot and playtime tracking. View every achievements earned on your PC whether it's coming from Steam, a Steam emulator, and more.