Space emoji (emoji-only character allowed).
Directory
Search results
Published directory entries matching your search.
VoiceSphere AI: Streamline document handling for PDFs, DOCs, PPTs, videos, texts. Fast, precise answers. #VoiceSphereAI #AIChat #DocumentManagement.
An offline speech recognition toolkit with C++ support, designed for low-resource devices and multiple languages.
A simple and efficient end-to-end Automatic Speech Recognition (ASR) system from Facebook AI Research.
USE BAZEL VERSION=5.0.0./bazelisk-linux-amd64 build wavegru mod -c opt --copt=-march=native.
A Whisper CLI client compatible with the original OpenAI client, using CTranslate2 for faster inference. opensource.
Port of OpenAI's Whisper model in C/C++. opensource.
Turn text, scripts, and blog posts into videos with 2,000+ AI voices in 80+ languages. Free AI video generator - no camera, no editing skills needed.
Voice-first AI Assistant for online meetings that can actively participate and solve tasks live during the meeting.
Fine-tuned on AMD MI300X for the AMD Developer Hackathon 2026 (Fine-Tuning Track).
Voice AI that turns calls into outcomes - and gets sharper with each one. Our own model, built for the phone. Sub-400ms response. Live in days.
Open Source framework for voice and multimodal conversational AI.
The AI content team for solo founders. Ships articles in your voice that rank on Google and get cited by ChatGPT, Perplexity, and Gemini. From $29/mo.
A "machine learning framework to automate text-and voice-based conversations.".
Never miss another call with our AI-powered call answering service. 10x better than voicemail. 10x cheaper than a traditional phone answering service.
Enterprise-grade Voice AI simulation SDK for scenario-driven stress testing of multimodal and agentic systems.
Create AI generated videos from text with the most advanced AI avatars and voiceovers in 160+ languages. Try our free AI video generator now!
↗ external - Open-source dictation that types where you talk.
Translate texts & full document files instantly. Accurate translations for individuals and Teams. Millions translate with DeepL every day.
An open source Go transpiler for machine learning models.
Implementation of image to image (pix2pix) translation from the paper by isola et al.[DEEP LEARNING].
Muse: Text-To-Image Generation via Masked Generative Transformers.
Creates or edits images with generative models and visual controls.
AIFreeVideo uses AI to create, edit, transform, or animate video content.