Open-source project that uses AI for speech, voice, transcription, music, or other audio workflows.
Directory
Search results
Published directory entries matching your search.
A curated list of resources of audio-driven talking face generation.
Deploy and scale machine learning models on Kubernetes. Built for LLMs, embeddings, and speech-to-text.
Open-source project that supports software development with code generation, analysis, debugging, or documentation.
Open-source project that supports software development with code generation, analysis, debugging, or documentation.
Open-source project that supports software development with code generation, analysis, debugging, or documentation.
Open-source project that supports software development with code generation, analysis, debugging, or documentation.
Open-source project that supports software development with code generation, analysis, debugging, or documentation.
Open-source project that supports software development with code generation, analysis, debugging, or documentation.
An open-source framework by NVIDIA for building speech AI systems, including automatic speech recognition and text-to-speech. opensource.
Open-source project that supports software development with code generation, analysis, debugging, or documentation.
Open-source project that supports software development with code generation, analysis, debugging, or documentation.
Open-source project that supports software development with code generation, analysis, debugging, or documentation.
Open-source project that supports software development with code generation, analysis, debugging, or documentation.
On-device voice dictation for macOS — transcribes a 5-minute clip in 2.8 s; noise-robust, paste at cursor. 99 languages, ~8 MB, no telemetry. MIT.
Vanna.ai - An open-source Python RAG framework for SQL generation and related functionality.
Open-source project that supports software development with code generation, analysis, debugging, or documentation.
Open-source project that supports software development with code generation, analysis, debugging, or documentation.
Works with speech, voice, music or other audio using machine-learning models.
An Optimized Speech-to-Text Pipeline for the Whisper Model.
Open-source project that supports software development with code generation, analysis, debugging, or documentation.
Large Language Model Text Generation Inference.
Works with speech, voice, music or other audio using machine-learning models.
Open-source project that supports software development with code generation, analysis, debugging, or documentation.