A collection of algorithms for image processing in Python.
Directory
Search results
Published directory entries matching your search.
Produces high quality object masks from input prompts such as points or boxes, and it can be used to generate masks for all objects in an image.
A C++ library for unsupervised text tokenization and detokenization, widely used in modern NLP models.
An Optimized Speech-to-Text Pipeline for the Whisper Model.
A library for processing Chinese text.
A compendium of information regarding Stable Diffusion (SD).
Creates or edits images with generative models and visual controls.
Open-source project that provides machine-learning models, research, training resources, or evaluation tools.
Open-source project that provides machine-learning models, research, training resources, or evaluation tools.
Text preprocessing package for use in NLP tasks.
Creates, rewrites or summarizes written content with language models.
Open-source project that provides machine-learning models, research, training resources, or evaluation tools.
Automatically create masks for Stable Diffusion inpainting using natural language.
Generate TikTok Text-to-Speech voices in your browser
Text Retrieval and Annotation Toolkit, definitely the most comprehensive toolkit I’ve encountered so far for Ruby.
Creates or edits images with generative models and visual controls.
Creates or edits images with generative models and visual controls.
Source 2 Viewer is an all-in-one tool to browse VPK archives, view, extract, and decompile Source 2 assets, including maps, models, materials, textures, sounds, and more.
Real-time inference for Stable Diffusion - 0.88s latency. Covers AITemplate, nvFuser, TensorRT, FlashAttention. (Archived).
Creates or edits images with generative models and visual controls.
Automatic "Differentiation" via Text, using large language models to backpropagate textual gradients.
Python-zpar - Python bindings for, a statistical part-of-speech-tagger, constituency parser, and dependency parser for English.