Directory

Search results

Published directory entries matching your search.

Search results

42 listings
github.com

You can also find more comprehensive list on and There's an AI AI Voice Cloning list.

AI Audio & Speech 0
github.com

Fast inference engine for whisper in C++ using CTranslate2.

AI Audio & Speech 0
github.com

Speech recognition toolkit with streaming ASR, VAD, punctuation, speaker diarization, and OpenAI-compatible serving for voice AI applications.

AI Audio & Speech 0

Port of OpenAI's Whisper model in C/C++. It can be executed locally.

AI Audio & Speech 0
github.com

Deploy and scale machine learning models on Kubernetes. Built for LLMs, embeddings, and speech-to-text.

AI Audio & Speech 0

Implementation of MusicLM, Google's new SOTA model for music generation using attention networks, in Pytorch.

AI Audio & Speech 0
github.com

An open-source framework by NVIDIA for building speech AI systems, including automatic speech recognition and text-to-speech. opensource.

AI Audio & Speech 0
github.com

On-device voice dictation for macOS — transcribes a 5-minute clip in 2.8 s; noise-robust, paste at cursor. 99 languages, ~8 MB, no telemetry. MIT.

AI Audio & Speech 0
github.com

Works with speech, voice, music or other audio using machine-learning models.

AI Audio & Speech 0
github.com

An Optimized Speech-to-Text Pipeline for the Whisper Model.

AI Audio & Speech 0
github.com

Works with speech, voice, music or other audio using machine-learning models.

AI Audio & Speech 0
github.com

Free tool to create viral videos from YouTube, generating clips optimized for TikTok and Instagram with automatic transcription and 9:16 editing.

TikTok 0
github.com

An offline speech recognition toolkit with C++ support, designed for low-resource devices and multiple languages.

AI Audio & Speech 0
github.com

A simple and efficient end-to-end Automatic Speech Recognition (ASR) system from Facebook AI Research.

AI Audio & Speech 0
github.com

A Whisper CLI client compatible with the original OpenAI client, using CTranslate2 for faster inference. opensource.

AI Audio & Speech 0
github.com

Port of OpenAI's Whisper model in C/C++. opensource.

AI Audio & Speech 0