Template repository for data science lifecycle project.
Directory
Search results
Published directory entries matching your search.
Computer Science Lecture Links is a GitHub repository or organization with source code, releases, documentation, or project resources.
Math and Science Lectures is a GitHub repository or organization with source code, releases, documentation, or project resources.
Kuwala is the no-code data platform for BI analysts and engineers enabling you to build powerful analytics workflows. We are set out to bring state-of-the-art data engineering tools you love, such as Airbyte, dbt, or Great Expectations together in one intuitive interface built with React Flow. In addition we provide third-party data into data science models and products with a focus on geospatial data. Currently, the following data connectors are available worldwide: a) High-resolution demograph
Code for Data Science at Olin College, Spring 2014.
Microsoft are pleased to offer a 10-week, 20-lesson curriculum all about Data Science.
Source code and experiments results for 2018 Data Science Bowl.
A lightweight set of tools for loading and sharing data in data science projects.
MeTA: ModErn Text Analysis is a C++ Data Sciences Toolkit that facilitates mining big text data.
Open-source project that applies AI to data analysis, extraction, visualization, or business intelligence.
GDM Science Skills to speed up agentic scientific workflows with better grounding and higher token efficiency. Integrate insights from AlphaGenome, AFDB, UniProt and 30+ other databases and tools.
Reproducible data setup for reproducible science.
Curated GitHub list of public datasets across science, government, finance, and research topics.
Open-source DMS (data management system) for powering data hubs and data portals.
A federated, open-source data catalog for all your big data and small data.
Open-source project that applies AI to data analysis, extraction, visualization, or business intelligence.
Data Version Control - Git for Data & Models - ML Experiments Management.
Open-source project that applies AI to data analysis, extraction, visualization, or business intelligence.
Pandas API on Apache Spark. Makes data scientists more productive when interacting with big data.
Fast and easy data exploration by automating the visualization and data analysis process.
CLI tool that allows you to build data profiles and write assertion tests for easily evaluating and tracking your data's reliability over time.
Platform for Production Data Science.
Machine learning development environment for data science and AI/ML engineering teams.
Aims at simplifying the Data Science experience of deploying Kubeflow Pipelines workflows.