A curated list of resources of audio-driven talking face generation.
Directory
Search results
Published directory entries matching your search.
Deploy and scale machine learning models on Kubernetes. Built for LLMs, embeddings, and speech-to-text.
Open-source project that supports software development with code generation, analysis, debugging, or documentation.
Open-source project that supports software development with code generation, analysis, debugging, or documentation.
Open-source project that supports software development with code generation, analysis, debugging, or documentation.
Open-source project that supports software development with code generation, analysis, debugging, or documentation.
Open-source project that supports software development with code generation, analysis, debugging, or documentation.
Open-source project that supports software development with code generation, analysis, debugging, or documentation.
An open-source framework by NVIDIA for building speech AI systems, including automatic speech recognition and text-to-speech. opensource.
Open-source project that supports software development with code generation, analysis, debugging, or documentation.
Open-source project that supports software development with code generation, analysis, debugging, or documentation.
Open-source project that supports software development with code generation, analysis, debugging, or documentation.
Open-source project that supports software development with code generation, analysis, debugging, or documentation.
On-device voice dictation for macOS — transcribes a 5-minute clip in 2.8 s; noise-robust, paste at cursor. 99 languages, ~8 MB, no telemetry. MIT.
Vanna.ai - An open-source Python RAG framework for SQL generation and related functionality.
Open-source project that supports software development with code generation, analysis, debugging, or documentation.
Pixray is an image generation system.
Open-source project that supports software development with code generation, analysis, debugging, or documentation.
Works with speech, voice, music or other audio using machine-learning models.
An Optimized Speech-to-Text Pipeline for the Whisper Model.
Open-source project that supports software development with code generation, analysis, debugging, or documentation.
Large Language Model Text Generation Inference.
A Collection of Awesome Generative AI Applications.
A system of bots that collects clips automatically via custom made filters, lets you easily browse these clips, and puts them together into a compilation video ready to be uploaded straight to any social media platform. Full VPS support is provided, along with an accounts system so multiple users can use the bot at once. This bot is split up into three separate programs. The server. The client. The video generator. These programs perform different functions that when combined creates a very powe