Free and open source face recognition with deep neural networks.
Directory
Search results
Published directory entries matching your search.
On-device voice dictation for macOS — transcribes a 5-minute clip in 2.8 s; noise-robust, paste at cursor. 99 languages, ~8 MB, no telemetry. MIT.
Open-source project that uses AI for speech, voice, transcription, music, or other audio workflows.
Otter AI Meeting Agent supports real-time transcription, live chat, automated summaries, insights, and action items.
Pack Generator uses AI for speech, voice, transcription, music, or other audio workflows.
Paper2Audio uses AI for speech, voice, transcription, music, or other audio workflows.
Parse indexes AI recommendations so brands know where they stand.
A complete object-oriented environment for machine learning in Matlab.
A Python library for implementing a Recommender System.
Works with speech, voice, music or other audio using machine-learning models.
Read AI uses AI for speech, voice, transcription, music, or other audio workflows.
ReadSpeaker uses AI for speech, voice, transcription, music, or other audio workflows.
ReadWise uses AI for speech, voice, transcription, music, or other audio workflows.
Works with speech, voice, music or other audio using machine-learning models.
A C library for product recommendations/suggestions using collaborative filtering (CF).
Remove Vocals uses AI for speech, voice, transcription, music, or other audio workflows.
Replica Studios uses AI for speech, voice, transcription, music, or other audio workflows.
Resemble AI uses AI for speech, voice, transcription, music, or other audio workflows.
Professional AI voice generator for real production workflows, with free testing and flexible integration for studios and media teams.
Rev AI, part of the Rev family, is a developer-first API that delivers industry- accuracy and fast performance at global scale. Click to learn more.
Works with speech, voice, music or other audio using machine-learning models.
Open-source project that uses AI for speech, voice, transcription, music, or other audio workflows.
An open-source tool for recording screen and audio activity with AI-powered search, automations, and support for local LLMs. opensource.
An Optimized Speech-to-Text Pipeline for the Whisper Model.