Fast inference engine for whisper in C++ using CTranslate2.
Directory
Search results
Published directory entries matching your search.
FlexClip uses AI to create, edit, transform, or animate video content.
Flixier uses AI to create, edit, transform, or animate video content.
Open-source project that uses AI to create, edit, transform, or animate video content.
結婚式スピーチ・弔辞・年賀状・退職挨拶・お礼状・お詫び文など、冠婚葬祭と日常の手紙・挨拶文をAIが作成します。シチュエーションを選んで項目を埋めるだけで、そのまま使える文面が完成。82種類のツールを登録不要・完全無料で今すぐ使えます。.
Speech recognition toolkit with streaming ASR, VAD, punctuation, speaker diarization, and OpenAI-compatible serving for voice AI applications.
Gen-2 by Runway uses AI to create, edit, transform, or animate video content.
Creates, rewrites or summarizes written content with language models.
Genmo uses AI to create, edit, transform, or animate video content.
Port of OpenAI's Whisper model in C/C++. It can be executed locally.
GlitchAI uses AI to create, edit, transform, or animate video content.
Google Flow uses AI to create, edit, transform, or animate video content.
Implementation of model parallel autoregressive transformers on GPUs, based on the DeepSpeed library.
Framework for building applications with LLMs and Transformers (e.g. agents, semantic search, question-answering).
HeyGen uses AI to create, edit, transform, or animate video content.
Distribution, publishing, funding, marketing, and a hands-on team for independent artists. We.
High-quality text-to-speech and voice recognition.
A Java port of SciPy's signal processing module, offering filters, transformations, and other scientific computing utilities.
Kapwing AI Video Editor uses AI to create, edit, transform, or animate video content.
Simple API for Neural Network. Better for image processing with CPU/GPU + Transfer Learning.
A library of statistical distribution sampling and transducing functions.
Klipy uses AI to create, edit, transform, or animate video content.
Deploy and scale machine learning models on Kubernetes. Built for LLMs, embeddings, and speech-to-text.
Works with speech, voice, music or other audio using machine-learning models.