Kuwala is the no-code data platform for BI analysts and engineers enabling you to build powerful analytics workflows. We are set out to bring state-of-the-art data engineering tools you love, such as Airbyte, dbt, or Great Expectations together in one intuitive interface built with React Flow. In addition we provide third-party data into data science models and products with a focus on geospatial data. Currently, the following data connectors are available worldwide: a) High-resolution demograph
Directory
Search results
Published directory entries matching your search.
Python Data Science Handbook is a code project with source, releases, documentation, or setup notes.
Data Science Ipython Notebooks is a GitHub repository or organization with source code, releases, documentation, or project resources.
OSSU Data Science is a GitHub repository or organization with source code, releases, documentation, or project resources.
A low-level Linear Regression Engine utilizing the Ordinary Least Squares (OLS) method and QR decomposition.
A repository of useful data science prompts for ChatGPT.
Template repository for data science lifecycle project.
Machine learning development environment for data science and AI/ML engineering teams.
Computer Science Lecture Links is a GitHub repository or organization with source code, releases, documentation, or project resources.
Math and Science Lectures is a GitHub repository or organization with source code, releases, documentation, or project resources.
Cleansing, pre-processing, feature engineering, exploratory data analysis and easy ML with PySpark backend.
Spark is a fast and general engine for large-scale data processing.
Curated list of awesome vector search framework/engine, library, cloud service and research papers to vector similarity search.
Open-source command-line tool that queries multiple onion search engines and exports results.
A lightweight, GPU accelerated, SQL engine for Python. Built on RAPIDS cuDF.
ML powered analytics engine for outlier/anomaly detection and root cause analysis.
Storage layer that brings scalable, ACID transactions to Apache Spark and other engines.
Code for Data Science at Olin College, Spring 2014.
Microsoft are pleased to offer a 10-week, 20-lesson curriculum all about Data Science.
Source code and experiments results for 2018 Data Science Bowl.
A Playwright-based Node.js tool that bypasses search engine anti-scraping mechanisms to execute Google searches. Local alternative to SERP APIs with MCP server integration.
nuxt-seo helps track or improve visibility in AI search, answer engines, and generative results.
A lightweight set of tools for loading and sharing data in data science projects.
MeTA: ModErn Text Analysis is a C++ Data Sciences Toolkit that facilitates mining big text data.