Open source voice chat software for low-latency group communication and self-hosted servers.
Directory
Search results
Published directory entries matching your search.
An open-source framework by NVIDIA for building speech AI systems, including automatic speech recognition and text-to-speech. opensource.
Open Source framework for voice and multimodal conversational AI.
A "machine learning framework to automate text-and voice-based conversations.".
An Optimized Speech-to-Text Pipeline for the Whisper Model.
Enterprise-grade Voice AI simulation SDK for scenario-driven stress testing of multimodal and agentic systems.
Generate TikTok Text-to-Speech voices in your browser
Simple Python script to interact with the TikTok TTS API
Accelerates transcription with the combination of OpenAI's Whisper Large v2, HF Transformers, Optimum, and flash attention.
An offline speech recognition toolkit with C++ support, designed for low-resource devices and multiple languages.
A simple and efficient end-to-end Automatic Speech Recognition (ASR) system from Facebook AI Research.
A Whisper CLI client compatible with the original OpenAI client, using CTranslate2 for faster inference. opensource.
Port of OpenAI's Whisper model in C/C++. opensource.
↗ external - Open-source dictation that types where you talk.