Library for high performance deep learning inference on NVIDIA GPUs.
Directory
Search results
Published directory entries matching your search.
Open-source self-hostable end-to-end LLMOps platform unifying tracing, evals, simulations, datasets, gateway, and guardrails.
Production-grade SDK for observability, automated evaluations and prompt management with sub-100ms guardrails for LLM/agent workflows.
Host inference APIs, bulk inference and fine tune text, vision, audio and multi-modal models.
Browse research datasets and datasets related to early modern history.
Testing framework dedicated to ML models, from tabular to LLMs. Detect risks of biases, performance issues and errors in 4 lines of code.
Cloud-Native LLM Routing Engine. Improve LLM app resilience and speed.
Open Bilingual Pre-Trained Model (ICLR 2023).
Open Bilingual Pre-Trained Model, quantization of ChatGLM-130B, can run on consumer-level GPUs.
Google tool for building, testing, and experimenting with Gemini prompts, models, and API workflows.
Google tool for building, testing, and experimenting with Gemini prompts, models, and API workflows.
gotoHuman provides machine-learning models, research, training resources, or evaluation tools.
Implementation of model parallel autoregressive transformers on GPUs, based on the DeepSpeed library.
Creating semantic cache to store responses from LLM queries.
Real-time GPU cloud price comparison across 30+ providers.
GPU cluster manager for running and managing LLMs.
Create customizable UI components around your models.
GraphPipe supports deploying, serving, monitoring, or operating AI and machine-learning systems.
Groq supports deploying, serving, monitoring, or operating AI and machine-learning systems.
Guild AI supports deploying, serving, monitoring, or operating AI and machine-learning systems.
A curated list of Large Language Model.
Helicone AI supports deploying, serving, monitoring, or operating AI and machine-learning systems.
Hermes Agent helps build, test, automate, or manage AI agents, prompts, models, and API workflows.
HF Learn supports machine learning models, deployment, inspection, datasets, or AI development workflows.