Search results

Published directory entries matching your search.

Filters
Clear filters

Search results

53 listings
Preview unavailable

Awesome Vehicle Security

Awesome Vehicle Security is a code project with source, releases, documentation, or setup notes.

Surveillance Awareness GitHub repo
Preview unavailable

face recognition

Face recognition library that recognizes and manipulates faces from Python or from the command line.

Models & Machine Learning GitHub repo
Preview unavailable

FunASR

Speech recognition toolkit with streaming ASR, VAD, punctuation, speaker diarization, and OpenAI-compatible serving for voice AI applications.

AI Audio & Speech GitHub repo
Preview unavailable

grt

The Gesture Recognition Toolkit (GRT) is a cross-platform, open-source, C++ machine learning library designed for real-time gesture recognition.

Models & Machine Learning GitHub repo
Preview unavailable

NeMo

An open-source framework by NVIDIA for building speech AI systems, including automatic speech recognition and text-to-speech. opensource.

AI Audio & Speech GitHub repo
Preview unavailable

Pattern Recognition and Machine Learning

This package contains the matlab implementation of the algorithms described in the book Pattern Recognition and Machine Learning by C. Bishop.

Models & Machine Learning GitHub repo
Preview unavailable

Vosk

An offline speech recognition toolkit with C++ support, designed for low-resource devices and multiple languages.

AI Audio & Speech GitHub repo
Preview unavailable

wav2letter

A simple and efficient end-to-end Automatic Speech Recognition (ASR) system from Facebook AI Research.

AI Audio & Speech GitHub repo
Preview unavailable

AndroidCamera

🔥🔥🔥自定义Android相机(仿抖音 TikTok),其中功能包括视频人脸识别贴纸,美颜,分段录制,视频裁剪,视频帧处理,获取视频关键帧,视频旋转,添加滤镜,添加水印,合成Gif到视频,文字转视频,图片转视频,音视频合成,音频变声处理,SoundTouch,Fmod音频处理。 Android camera(imitation Tik Tok), which includes video editor,audio editor,video face recognition stickers, segment recording,video cropping, video frame processing, get the first video frame, key frame, video rotation, add filter Mirror ,add watermark ,add gif to video, add text to video, picture to video, audio and video synthesis, audio change processing

TikTok GitHub repo
Preview unavailable

Audiblez

Open-source project that uses AI for speech, voice, transcription, music, or other audio workflows.

AI Audio & Speech GitHub repo
Preview unavailable

Audio-WebUI

Open-source project that uses AI for speech, voice, transcription, music, or other audio workflows.

AI Audio & Speech GitHub repo
Preview unavailable

Awesome AI Music

You can also find more comprehensive list on and There's an AI AI Voice Cloning list.

AI Audio & Speech GitHub repo
Preview unavailable

Bark

Open-source project that uses AI for speech, voice, transcription, music, or other audio workflows.

AI Audio & Speech GitHub repo
Preview unavailable

Buster

Browser extension that helps solve audio CAPTCHA challenges using speech recognition.

Internet Utilities GitHub repo
Preview unavailable

Chatterbox

Open-source project that uses AI for speech, voice, transcription, music, or other audio workflows.

AI Audio & Speech GitHub repo
Preview unavailable

Ebook2Audiobook

Open-source project that uses AI for speech, voice, transcription, music, or other audio workflows.

AI Audio & Speech GitHub repo
Preview unavailable

EmotiVoice

Open-source project that uses AI for speech, voice, transcription, music, or other audio workflows.

AI Audio & Speech GitHub repo
Preview unavailable

EspNet

Open-source project that uses AI for speech, voice, transcription, music, or other audio workflows.

AI Audio & Speech GitHub repo
Preview unavailable

Faster Whisper

Fast inference engine for whisper in C++ using CTranslate2.

AI Audio & Speech GitHub repo
Preview unavailable

ggerganov/whisper.cpp

Port of OpenAI's Whisper model in C/C++. It can be executed locally.

AI Audio & Speech GitHub repo
Preview unavailable

GPT-SoVITS

Open-source project that uses AI for speech, voice, transcription, music, or other audio workflows.

AI Audio & Speech GitHub repo
Preview unavailable

Kaldi

Open-source project that uses AI for speech, voice, transcription, music, or other audio workflows.

AI Audio & Speech GitHub repo
Preview unavailable

KittenTTS

Open-source project that uses AI for speech, voice, transcription, music, or other audio workflows.

AI Audio & Speech GitHub repo