Course materials and notes for Stanford class CS231n: Deep Learning for Computer Vision.
Directory
Search results
Published directory entries matching your search.
Curate and annotate vision, audio, and LLM datasets, track experiments, and manage models on a single platform.
Host inference APIs, bulk inference and fine tune text, vision, audio and multi-modal models.
Greg (Grzegorz) Surma - Portfolio; Machine Learning, Computer Vision, Self-Driving Cars, AI.
Open source computer vision API based on open source models.
Research data portal for exploring art history, music, theatre and media studies records, including image.
Specialized AI vision for extracting engineering specs from PDF/JPG to Excel.
Browse research datasets and datasets related to basic in biology and medicine.
TwelveLabs delivers enterprise video AI powered by multimodal intelligence. Search, analyze, and understand video across vision, audio, and language.
VLFeat is an open and portable library of computer vision algorithms, which has a Matlab toolbox.