Open-source project that provides machine-learning models, research, training resources, or evaluation tools.
Directory
Search results
Published directory entries matching your search.
Running large language models on a single GPU for throughput-oriented scenarios. (Archived).
Drag & drop UI to build your customized LLM flow using LangchainJS.
Library for high performance deep learning inference on NVIDIA GPUs.
Open-source self-hostable end-to-end LLMOps platform unifying tracing, evals, simulations, datasets, gateway, and guardrails.
Production-grade SDK for observability, automated evaluations and prompt management with sub-100ms guardrails for LLM/agent workflows.
Testing framework dedicated to ML models, from tabular to LLMs. Detect risks of biases, performance issues and errors in 4 lines of code.
Cloud-Native LLM Routing Engine. Improve LLM app resilience and speed.
Creating semantic cache to store responses from LLM queries.
GPU cluster manager for running and managing LLMs.
Create customizable UI components around your models.
Platform for deploying your Machine Learning to production.
Code for hyperparameter tuning/optimization of machine learning and deep learning algorithms.
Open-source project that supports deploying, serving, monitoring, or operating AI and machine-learning systems.
Open-source project that supports deploying, serving, monitoring, or operating AI and machine-learning systems.
Kubernetes-based system for hyperparameter tuning and neural architecture search.
Kubernetes custom resource definition for serving ML models on arbitrary frameworks.
Open-source project that supports deploying, serving, monitoring, or operating AI and machine-learning systems.
Open-source project that supports deploying, serving, monitoring, or operating AI and machine-learning systems.
Standardized Serverless ML Inference Platform on Kubernetes.
Open-source all-in-one platform for engineering AI products. Traces, Evals, Datasets, Labels.
FastAPI framework to build production-grade LLM applications.
Developer-friendly, serverless vector database for AI applications. Easily add long-term memory to your LLM apps!
Serverless LLM apps on Production with Jina AI Cloud (Archived).