pile.eleuther.ai
The Pile is a 825 GiB diverse, open source language modelling data set that consists of 22 smaller, high-quality datasets combined together.
Directory
Published directory entries matching your search.
The Pile is a 825 GiB diverse, open source language modelling data set that consists of 22 smaller, high-quality datasets combined together.
Discover datasets around the world!
Enriches training datasets with features from public and community shared data sources.
Vaex is a Python library that allows you to visualize large datasets and calculate statistics at high speeds.
VBench supports machine learning models, deployment, inspection, datasets, or AI development workflows.
What Models? supports machine learning models, deployment, inspection, datasets, or AI development workflows.