Enterprise-grade Voice AI simulation SDK for scenario-driven stress testing of multimodal and agentic systems.
Directory
Search results
Published directory entries matching your search.
This premium domain name is available for purchase!
AI-driven sound and music generation.
Speechify reads anything aloud to you. Listen to books, PDFs, or web pages anytime with natural voices. Try Speechify free.
Examples for ICASSP2024 paper “StemGen: A music generation model that listens”.
Create AI generated videos from text with the most advanced AI avatars and voiceovers in 160+ languages. Try our free AI video generator now!
NaturalReader online reader converts text into spoken audio for listening and accessibility.
Generate TikTok Text-to-Speech voices in your browser
Simple Python script to interact with the TikTok TTS API
AI transcription & translation for audio/video. Two services: transcribe or translate in 50+ languages. 98% accuracy, pay-as-you-go.
Accelerates transcription with the combination of OpenAI's Whisper Large v2, HF Transformers, Optimum, and flash attention.
Space emoji (emoji-only character allowed).
Nonprofit newsroom covering San Diego government, education, housing, infrastructure and community issues.
VoiceSphere AI: Streamline document handling for PDFs, DOCs, PPTs, videos, texts. Fast, precise answers. #VoiceSphereAI #AIChat #DocumentManagement.
An offline speech recognition toolkit with C++ support, designed for low-resource devices and multiple languages.
A simple and efficient end-to-end Automatic Speech Recognition (ASR) system from Facebook AI Research.
USE BAZEL VERSION=5.0.0./bazelisk-linux-amd64 build wavegru mod -c opt --copt=-march=native.
A Whisper CLI client compatible with the original OpenAI client, using CTranslate2 for faster inference. opensource.
Port of OpenAI's Whisper model in C/C++. opensource.
↗ external - Open-source dictation that types where you talk.