Open-source project that supports deploying, serving, monitoring, or operating AI and machine-learning systems.
Directory
Search results
Published directory entries matching your search.
Machine Learning Operations - An awesome list of references for MLOps.
Store, search, organize and make machine-learned inferences over big data at serving time.
Provides an optimized cloud and edge inferencing solution.
Lets you create apps for your ML projects with deceptively simple Python scripts.
An MLOps/LLMOps platform for model building, evaluation, and fine-tuning.
Python-free Rust inference server with OpenAI API compatibility and hot model swapping.
MLOps framework to package, deploy, monitor and manage thousands of production machine learning models.
Provides containers to encapsulate and deploy EdgeML pipelines and applications.
OpenAI-compatible LLM inference server for Apple Silicon using MLX. 2-4x faster than Ollama with tool calling and prompt caching.
The open source solution for monitoring your AI models in production.
Open-source project that supports deploying, serving, monitoring, or operating AI and machine-learning systems.
Open-source project that supports deploying, serving, monitoring, or operating AI and machine-learning systems.
Turns your ML code into microservices with web API, interactive GUI, and more.
Open-source GenAI and LLM observability platform native to OpenTelemetry with traces and metrics. opensource.
Open-source project that supports deploying, serving, monitoring, or operating AI and machine-learning systems.
Open-source project that supports deploying, serving, monitoring, or operating AI and machine-learning systems.
Open-source project that supports deploying, serving, monitoring, or operating AI and machine-learning systems.
A framework providing the right abstractions to ease research, development, and deployment of your ML pipelines.
Open-source project that supports deploying, serving, monitoring, or operating AI and machine-learning systems.
Ncnn is a high-performance neural network inference framework optimized for the mobile platform.
A lightweight, portable pure C99 onnx inference engine for embedded devices with hardware acceleration support.
Open-source project that supports deploying, serving, monitoring, or operating AI and machine-learning systems.
LLM Ops platform with Analytics, Monitoring, Evaluations and an LLM Optimization Studio powered by DSPy.