Robust Speech Recognition Leaderboard 2022 uses AI for speech, voice, transcription, music, or other audio workflows.
Directory
Search results
Published directory entries matching your search.
Tunisian Speech Recognition uses AI for speech, voice, transcription, music, or other audio workflows.
Works with speech, voice, music or other audio using machine-learning models.
Works with speech, voice, music or other audio using machine-learning models.
Architecture of voice AI, from speech recognition to emotional intelligence, and learn how to build, scale, and evaluate them.
Face recognition library that recognizes and manipulates faces from Python or from the command line.
Speech recognition toolkit with streaming ASR, VAD, punctuation, speaker diarization, and OpenAI-compatible serving for voice AI applications.
The Gesture Recognition Toolkit (GRT) is a cross-platform, open-source, C++ machine learning library designed for real-time gesture recognition.
High-quality text-to-speech and voice recognition.
An open-source framework by NVIDIA for building speech AI systems, including automatic speech recognition and text-to-speech. opensource.
NeMo Parakeet ASR Models attain strong speech recognition accuracy while being efficient for inference. Available in CTC and RNN-Transducer variants.
This package contains the matlab implementation of the algorithms described in the book Pattern Recognition and Machine Learning by C. Bishop.
Space emoji (emoji-only character allowed).
An offline speech recognition toolkit with C++ support, designed for low-resource devices and multiple languages.
A simple and efficient end-to-end Automatic Speech Recognition (ASR) system from Facebook AI Research.
Open-source project that uses AI for speech, voice, transcription, music, or other audio workflows.
Google Illuminate uses AI for speech, voice, transcription, music, or other audio workflows.
Acapela Group uses AI for speech, voice, transcription, music, or other audio workflows.
Acapella Extractor uses AI for speech, voice, transcription, music, or other audio workflows.
AceTagGen uses AI for speech, voice, transcription, music, or other audio workflows.
AI Dictation uses AI for speech, voice, transcription, music, or other audio workflows.
AI Mastering is an automated online audio mastering service using AI. Free mastering is available.
Works with speech, voice, music or other audio using machine-learning models.
Helps teams build, deploy, observe or operate machine-learning systems.