GPU cluster manager for running and managing LLMs.
Directory
Search results
Published directory entries matching your search.
Create customizable UI components around your models.
A curated list of Large Language Model.
Python materials for the online course on diffusion models by @huggingface.
Platform for deploying your Machine Learning to production.
A service for deployment Apache Spark MLLib machine learning models as realtime, batch or reactive web services.
Code for hyperparameter tuning/optimization of machine learning and deep learning algorithms.
Open-source project that supports deploying, serving, monitoring, or operating AI and machine-learning systems.
Open-source project that supports deploying, serving, monitoring, or operating AI and machine-learning systems.
Kubernetes-based system for hyperparameter tuning and neural architecture search.
A python package that integrates an LLM copilot inside the keras model development workflow.
Kubernetes custom resource definition for serving ML models on arbitrary frameworks.
Open-source project that supports deploying, serving, monitoring, or operating AI and machine-learning systems.
Open-source project that supports deploying, serving, monitoring, or operating AI and machine-learning systems.
Standardized Serverless ML Inference Platform on Kubernetes.
Open-source all-in-one platform for engineering AI products. Traces, Evals, Datasets, Labels.
FastAPI framework to build production-grade LLM applications.
Developer-friendly, serverless vector database for AI applications. Easily add long-term memory to your LLM apps!
Serverless LLM apps on Production with Jina AI Cloud (Archived).
Open-source project that provides machine-learning models, research, training resources, or evaluation tools.
LLM Ops platform with Analytics, Monitoring, Evaluations and an LLM Optimization Studio powered by DSPy.
Open-source project that supports deploying, serving, monitoring, or operating AI and machine-learning systems.
Neural network inference from the command line, implemented in CHICKEN Scheme.
A lightweight, portable pure C99 onnx inference engine for embedded devices with hardware acceleration support.