Source code and experiments results for 2018 Data Science Bowl.
Directory
Search results
Published directory entries matching your search.
Peer-to-peer network of data owners and data scientists who can collectively train AI models using PySyft.
schema-org helps create or test structured data, schema markup, and rich-result eligibility.
Synthetic tabular data generation using GANs, Diffusion Models, and LLMs with adversarial filtering and privacy metrics.
Research archive containing dark-web authorship verification datasets and baseline models.
automates gathering website profiling data into a CSV from the "BuiltWith" or "Wappalyzer" API for tech stack information, technographic data, website
GitHub awesome list collecting OSINT tools, datasets, training links, and research resources.
50 AI agent skills for affiliate marketing. Research trending content, write data-backed posts, generate infographics, build landing pages, deploy — full flywheel with social intelligence. Works with Claude Code, Pi, ChatGPT, Gemini, Cursor, Windsurf, any AI.
🤖 Curated AI OSINT resources — Google dorks, Shodan queries, GitHub dorks, and techniques to discover exposed LLM endpoints, leaked AI API keys, misconfigured vector databases, and unprotected AI agents
A comprehensive set of fairness metrics for datasets and machine learning models.
Easy way to turn any app into searchable data for LLMs.
Code and documentation to train Stanford's Alpaca models, and generate the data.
Ambrosia helps you clean up your LLM datasets using other LLMs.
Lightweight analytics reporting and publishing tool for Digital Analytics Program's Google Analytics 360 data.
Lightweight, Portable, Flexible Distributed/Mobile Deep Learning with Dynamic, Mutation-aware Dataflow Dep Scheduler.
Platform for Production Data Science.
Scripts to generate a dataset with static frames from the Arcade Learning Environment.
Automated machine learning for image, text, tabular, time-series, and multi-modal data.
Open-source project that provides AI-assisted learning, courses, tutorials, or study support.
AI Native database for embedding vectors.
) - A fast web client boilerplate written in C# / Blazor, that uses an in-browser SQLite database.
) - An efficient, scalable, and deduplicated local blob storage that maintains metadata in SQLite. Fully compatible with Blossom, it gives developers a reliable database option for building their own Blossom servers.
Latest Papers and Datasets on Multimodal Large Language Models, and Their Evaluation.
Open-source project that provides machine-learning models, research, training resources, or evaluation tools.