The AI Scientist: Towards Fully Automated Open-Ended Scientific.
Directory
Search results
Published directory entries matching your search.
Platform for Production Data Science.
Automated machine learning toolkit and a drop-in replacement for a scikit-learn estimator.
Curated GitHub list of public datasets across science, government, finance, and research topics.
BigScience Large Open-science Open-access Multilingual Language Model.
Reproducible data setup for reproducible science.
Open-source project that provides machine-learning models, research, training resources, or evaluation tools.
Machine learning development environment for data science and AI/ML engineering teams.
Set of tools for creating and testing machine learning features, with a scikit-learn compatible API.
A lightweight set of tools for loading and sharing data in data science projects.
Open-source project that provides machine-learning models, research, training resources, or evaluation tools.
A Java port of SciPy's signal processing module, offering filters, transformations, and other scientific computing utilities.
Aims at simplifying the Data Science experience of deploying Kubeflow Pipelines workflows.
An unsupervised machine learning extension library for NetworkX with a Scikit-Learn like API.
Open source MLOps project that eases model handoffs between data scientist and DevOps.
Knowledge sharing platform for data scientists and other technical professions.
Pandas API on Apache Spark. Makes data scientists more productive when interacting with big data.
Kuwala is the no-code data platform for BI analysts and engineers enabling you to build powerful analytics workflows. We are set out to bring state-of-the-art data engineering tools you love, such as Airbyte, dbt, or Great Expectations together in one intuitive interface built with React Flow. In addition we provide third-party data into data science models and products with a focus on geospatial data. Currently, the following data connectors are available worldwide: a) High-resolution demograph
A graph sampling extension library for NetworkX with a Scikit-Learn like API.
MeTA: ModErn Text Analysis is a C++ Data Sciences Toolkit that facilitates mining big text data.
All-in-one web-based IDE specialized for machine learning and data science.
500+ ML/AI interview Q&A with runnable code — covers ML fundamentals, deep learning, NLP, PyTorch, scikit-learn pipelines, and system design.
A high performance, memory efficient, maximally parallelized ensemble learning, integrated with scikit-learn.
Generic mechanism for data scientists to build, run, and monitor ML tasks and pipelines.