Directory

Search results

Published directory entries matching your search.

Search results

58 listings
github.com

Open-source project that uses AI for speech, voice, transcription, music, or other audio workflows.

AI Audio & Speech 0
github.com

On-device voice dictation for macOS — transcribes a 5-minute clip in 2.8 s; noise-robust, paste at cursor. 99 languages, ~8 MB, no telemetry. MIT.

AI Audio & Speech 0

Open-source project that uses AI for speech, voice, transcription, music, or other audio workflows.

AI Audio & Speech 0
github.com

Open-source project that uses AI for speech, voice, transcription, music, or other audio workflows.

AI Audio & Speech 0
github.com

Open-source project that uses AI for speech, voice, transcription, music, or other audio workflows.

AI Audio & Speech 0
github.com

Open-source project that uses AI for speech, voice, transcription, music, or other audio workflows.

AI Audio & Speech 0
github.com

Open-source project that uses AI for speech, voice, transcription, music, or other audio workflows.

AI Audio & Speech 0

Accelerates transcription with the combination of OpenAI's Whisper Large v2, HF Transformers, Optimum, and flash attention.

AI Audio & Speech 0
github.com

Open-source project that uses AI for speech, voice, transcription, music, or other audio workflows.

AI Audio & Speech 0
github.com

Open-source project that uses AI for speech, voice, transcription, music, or other audio workflows.

AI Audio & Speech 0
github.com

Open-source project that uses AI for speech, voice, transcription, music, or other audio workflows.

AI Audio & Speech 0
github.com

Open-source project that uses AI for speech, voice, transcription, music, or other audio workflows.

AI Audio & Speech 0
github.com

A Whisper CLI client compatible with the original OpenAI client, using CTranslate2 for faster inference. opensource.

AI Audio & Speech 0
github.com

Port of OpenAI's Whisper model in C/C++. opensource.

AI Audio & Speech 0
github.com

Open-source project that uses AI for speech, voice, transcription, music, or other audio workflows.

AI Audio & Speech 0
github.com

Face recognition library that recognizes and manipulates faces from Python or from the command line.

Models & Machine Learning 0
github.com

The Gesture Recognition Toolkit (GRT) is a cross-platform, open-source, C++ machine learning library designed for real-time gesture recognition.

Models & Machine Learning 0

This package contains the matlab implementation of the algorithms described in the book Pattern Recognition and Machine Learning by C. Bishop.

Models & Machine Learning 0
github.com

Browser extension that helps solve audio CAPTCHA challenges using speech recognition.

Internet Utilities 0
github.com

🔥🔥🔥自定义Android相机(仿抖音 TikTok),其中功能包括视频人脸识别贴纸,美颜,分段录制,视频裁剪,视频帧处理,获取视频关键帧,视频旋转,添加滤镜,添加水印,合成Gif到视频,文字转视频,图片转视频,音视频合成,音频变声处理,SoundTouch,Fmod音频处理。 Android camera(imitation Tik Tok), which includes video editor,audio editor,video face recognition stickers, segment recording,video cropping, video frame processing, get the first video frame, key frame, video rotation, add filter Mirror ,add watermark ,add gif to video, add text to video, picture to video, audio and video synthesis, audio change processing

TikTok 0
github.com

Named-entity recognition using neural networks providing state-of-the-art-results.

AI Image Generation 0