Directory

Search results

Published directory entries matching your search.

Search results

6,853 listings
github.com

Open-source project that uses AI for speech, voice, transcription, music, or other audio workflows.

AI Audio & Speech 0
github.com

Open-source project that uses AI for speech, voice, transcription, music, or other audio workflows.

AI Audio & Speech 0
github.com

Open-source project that uses AI for speech, voice, transcription, music, or other audio workflows.

AI Audio & Speech 0
github.com

Open-source project that uses AI for speech, voice, transcription, music, or other audio workflows.

AI Audio & Speech 0
github.com

Open-source project that uses AI for speech, voice, transcription, music, or other audio workflows.

AI Audio & Speech 0
github.com

Open-source project that uses AI for speech, voice, transcription, music, or other audio workflows.

AI Audio & Speech 0
github.com

Applio is a GitHub repository or organization with source code, releases, documentation, or project resources.

AI Audio & Speech 0
github.com

Ace Step 1.5 is a GitHub repository or organization with source code, releases, documentation, or project resources.

AI Audio & Speech 0
github.com

Mmaudio is a GitHub repository or organization with source code, releases, documentation, or project resources.

AI Audio & Speech 0
github.com

Open-source project that uses AI for speech, voice, transcription, music, or other audio workflows.

AI Audio & Speech 0
github.com

Open-source project that uses AI for speech, voice, transcription, music, or other audio workflows.

AI Audio & Speech 0
github.com

Port of OpenAI's Whisper model in C/C++. opensource.

AI Audio & Speech 0
github.com

A Whisper CLI client compatible with the original OpenAI client, using CTranslate2 for faster inference. opensource.

AI Audio & Speech 0
github.com

A simple and efficient end-to-end Automatic Speech Recognition (ASR) system from Facebook AI Research.

AI Audio & Speech 0
github.com

An offline speech recognition toolkit with C++ support, designed for low-resource devices and multiple languages.

AI Audio & Speech 0

Accelerates transcription with the combination of OpenAI's Whisper Large v2, HF Transformers, Optimum, and flash attention.

AI Audio & Speech 0
github.com

Works with speech, voice, music or other audio using machine-learning models.

AI Audio & Speech 0
github.com

An Optimized Speech-to-Text Pipeline for the Whisper Model.

AI Audio & Speech 0
github.com

Works with speech, voice, music or other audio using machine-learning models.

AI Audio & Speech 0
github.com

On-device voice dictation for macOS — transcribes a 5-minute clip in 2.8 s; noise-robust, paste at cursor. 99 languages, ~8 MB, no telemetry. MIT.

AI Audio & Speech 0
github.com

An open-source framework by NVIDIA for building speech AI systems, including automatic speech recognition and text-to-speech. opensource.

AI Audio & Speech 0

Implementation of MusicLM, Google's new SOTA model for music generation using attention networks, in Pytorch.

AI Audio & Speech 0
github.com

Deploy and scale machine learning models on Kubernetes. Built for LLMs, embeddings, and speech-to-text.

AI Audio & Speech 0