Python library for data-centric AI and machine learning with messy, real-world data and labels.
Directory
Search results
Published directory entries matching your search.
DataRobot delivers the industry- AI applications and platform that maximize impact and minimize risk for your business.
Eurybia monitors data and model drift over time and securizes model deployment with data validation.
Synthetic tabular data generation using GANs, Diffusion Models, and LLMs with adversarial filtering and privacy metrics.
Terloka Data Insight Tool is an interactive, AI-powered data exploration and visualization tool for analytics.
This AI Data Analyst chatbot generates SQL code using AI, like ChatGPT for SQL Databases. Connect and chat with database in ChatGPT.
The primary interface for interacting with OpenAI models, supporting chat, coding, data analysis, research, and agent workflows.
Databerry provides conversational AI for questions, tasks, and general assistance.
Ego-Exo4D: a foundational dataset by Meta for research on video learning and multimodal perception.
Interact your data and environment using the local GPT, no data leaks, 100% privately, 100% security.
Turn entire websites into LLM-ready markdown or structured data. Scrape, crawl and extract with a single API.
AI Data Laundering - Waxy.org supports AI-assisted research, knowledge retrieval, summarization, or source analysis.
Data Scientist Agent that helps you:.
AI for Database supports AI agents, automated workflows, orchestration, or delegated tasks.
Data School provides AI-assisted learning, courses, tutorials, or study support.
Databricks’ Dolly, a large language model trained on the Databricks Machine Learning Platform.
Integrate with providers using LangChain Python.
Provides a central interface to connect your LLM's with external data.
LlamaIndex is a data framework for your LLM applications.
A data framework for building LLM applications over external data.
An open dataset with 30 trillion tokens for training Large Language Models.
Curate and annotate vision, audio, and LLM datasets, track experiments, and manage models on a single platform.
Code and documentation to train Stanford's Alpaca models, and generate the data.
Latest Papers and Datasets on Multimodal Large Language Models, and Their Evaluation.