Design and Deploy Large Language Model Apps.
Directory
Search results
Published directory entries matching your search.
Eden AI provides machine-learning models, research, training resources, or evaluation tools.
Everything AI/ML supports machine learning models, deployment, inspection, datasets, or AI development workflows.
Falcon LLM is a generative large language model (LLM) that helps advance applications and use cases to future-proof our world.
Feast is an end-to-end open source feature store for machine learning. It allows teams to define, manage, discover, and serve features.
Feature store for machine learning.
Open-source project that provides machine-learning models, research, training resources, or evaluation tools.
Feature Stores for ML provides machine-learning models, research, training resources, or evaluation tools.
Open-source project that provides machine-learning models, research, training resources, or evaluation tools.
PyTorch Lightning extension that accelerates and enhances foundation model experimentation with flexible fine-tuning schedules.
Running large language models on a single GPU for throughput-oriented scenarios. (Archived).
Drag & drop UI to build your customized LLM flow using LangchainJS.
Library for high performance deep learning inference on NVIDIA GPUs.
Host inference APIs, bulk inference and fine tune text, vision, audio and multi-modal models.
Testing framework dedicated to ML models, from tabular to LLMs. Detect risks of biases, performance issues and errors in 4 lines of code.
Cloud-Native LLM Routing Engine. Improve LLM app resilience and speed.
Open Bilingual Pre-Trained Model (ICLR 2023).
Open Bilingual Pre-Trained Model, quantization of ChatGLM-130B, can run on consumer-level GPUs.
Google tool for building, testing, and experimenting with Gemini prompts, models, and API workflows.
Google tool for building, testing, and experimenting with Gemini prompts, models, and API workflows.
gotoHuman provides machine-learning models, research, training resources, or evaluation tools.
Implementation of model parallel autoregressive transformers on GPUs, based on the DeepSpeed library.
Creating semantic cache to store responses from LLM queries.
GPU cluster manager for running and managing LLMs.