Directory

Search results

Published directory entries matching your search.

Search results

4,408 listings
github.com

Works with speech, voice, music or other audio using machine-learning models.

AI Audio & Speech 0
github.com

Open-source project that uses AI for speech, voice, transcription, music, or other audio workflows.

AI Audio & Speech 0
github.com

Works with speech, voice, music or other audio using machine-learning models.

AI Audio & Speech 0
github.com

Open-source project that uses AI for speech, voice, transcription, music, or other audio workflows.

AI Audio & Speech 0
github.com

Open-source project that uses AI for speech, voice, transcription, music, or other audio workflows.

AI Audio & Speech 0
github.com

Open-source project that uses AI for speech, voice, transcription, music, or other audio workflows.

AI Audio & Speech 0
github.com

A comprehensive testing and evaluation framework for voice agents across language models, prompts, and agent personas.

Models & Machine Learning 0
github.com

Open-source project that uses AI for speech, voice, transcription, music, or other audio workflows.

AI Audio & Speech 0
github.com

Open-source project that uses AI for speech, voice, transcription, music, or other audio workflows.

AI Audio & Speech 0
github.com

Open-source project that uses AI for speech, voice, transcription, music, or other audio workflows.

AI Audio & Speech 0
github.com

Open source voice chat software for low-latency group communication and self-hosted servers.

Messaging & Chat 0
github.com

An alternative Discord client with voice support made with C++ and GTK 3

Discord 0
github.com

YumCut - free AI video generator to turn a prompt into ready vertical videos for TikTok, Reels and YouTube Shorts. Auto script, scenes, voiceover, subtitles and watermark. Built with Next.js. Local-first pipeline + templates, batch rendering and API hooks for creators and indie makers. Self-hosted, FFmpeg-ready, multi-language output. Low cost fast

TikTok 0
github.com

Fast inference engine for whisper in C++ using CTranslate2.

AI Audio & Speech 0

Port of OpenAI's Whisper model in C/C++. It can be executed locally.

AI Audio & Speech 0
github.com

Voice-first AI Assistant for online meetings that can actively participate and solve tasks live during the meeting.

AI Assistants & Chatbots 0
github.com

Deploy and scale machine learning models on Kubernetes. Built for LLMs, embeddings, and speech-to-text.

AI Audio & Speech 0

Implementation of MusicLM, Google's new SOTA model for music generation using attention networks, in Pytorch.

AI Audio & Speech 0
github.com

An open-source framework by NVIDIA for building speech AI systems, including automatic speech recognition and text-to-speech. opensource.

AI Audio & Speech 0
github.com

Open Source framework for voice and multimodal conversational AI.

AI Assistants & Chatbots 0
github.com

A "machine learning framework to automate text-and voice-based conversations.".

Models & Machine Learning 0
github.com

An Optimized Speech-to-Text Pipeline for the Whisper Model.

AI Audio & Speech 0