Creating semantic cache to store responses from LLM queries.
Directory
Search results
Published directory entries matching your search.
GPU cluster manager for running and managing LLMs.
Create customizable UI components around your models.
A curated list of Large Language Model.
Python materials for the online course on diffusion models by @huggingface.
Platform for deploying your Machine Learning to production.
A service for deployment Apache Spark MLLib machine learning models as realtime, batch or reactive web services.
Code for hyperparameter tuning/optimization of machine learning and deep learning algorithms.
Open-source project that supports deploying, serving, monitoring, or operating AI and machine-learning systems.
Open-source project that supports deploying, serving, monitoring, or operating AI and machine-learning systems.
Kubernetes-based system for hyperparameter tuning and neural architecture search.
A python package that integrates an LLM copilot inside the keras model development workflow.
Kubernetes custom resource definition for serving ML models on arbitrary frameworks.
Open-source project that supports deploying, serving, monitoring, or operating AI and machine-learning systems.
Open-source project that supports deploying, serving, monitoring, or operating AI and machine-learning systems.
Standardized Serverless ML Inference Platform on Kubernetes.
Open-source all-in-one platform for engineering AI products. Traces, Evals, Datasets, Labels.
FastAPI framework to build production-grade LLM applications.
Developer-friendly, serverless vector database for AI applications. Easily add long-term memory to your LLM apps!
Serverless LLM apps on Production with Jina AI Cloud (Archived).
Open-source project that provides machine-learning models, research, training resources, or evaluation tools.
LLM Ops platform with Analytics, Monitoring, Evaluations and an LLM Optimization Studio powered by DSPy.
Open-source project that supports deploying, serving, monitoring, or operating AI and machine-learning systems.
Neural network inference from the command line, implemented in CHICKEN Scheme.