An offline speech recognition toolkit with C++ support, designed for low-resource devices and multiple languages.
Directory
Search results
Published directory entries matching your search.
A simple and efficient end-to-end Automatic Speech Recognition (ASR) system from Facebook AI Research.
Open-source project that uses AI for speech, voice, transcription, music, or other audio workflows.
USE BAZEL VERSION=5.0.0./bazelisk-linux-amd64 build wavegru mod -c opt --copt=-march=native.
A Whisper CLI client compatible with the original OpenAI client, using CTranslate2 for faster inference. opensource.
Port of OpenAI's Whisper model in C/C++. opensource.
WolframTones uses AI for speech, voice, transcription, music, or other audio workflows.
Open-source project that uses AI for speech, voice, transcription, music, or other audio workflows.
Zyphra uses AI for speech, voice, transcription, music, or other audio workflows.
Manage fleets of local and cloud agents from one surface. Plan, delegate, review, and ship without leaving your editor.
✨ Fully autonomous AI Agent that can perform complicated tasks and projects using terminal, browser, and editor.
Plugin for CMS Adobe Experience Manager (AEM) or Composum Pages helping the editor to create / edit / translate texts.
An alternative to Supabase for AI Code editors and Vibe Coding tools.
Visual pipeline editor and workflow orchestrator with an easy to use UI and based on Kubernetes.
AI studio for visual content. Create marketing content, designs, and influencers. Image and video generation with an intuitive AI editor.
Outline, draft, and revise your novel in one place, with editorial feedback and manuscript formatting built in. Free to start.
Discord invite for joining a related community, support server, updates, or project discussion.
Turn live audio into stunning visuals.
Creates or edits images with generative models and visual controls.
Helps find, analyze or synthesize information for research and knowledge discovery.
Curated list of top Hugging Face models for NLP, vision, and audio tasks with demos and benchmarks.
Curate and annotate vision, audio, and LLM datasets, track experiments, and manage models on a single platform.
Discord invite for joining a related community, support server, updates, or project discussion.
A large Dataset of synchronised Audio, LyrIcs and vocal notes.