OpenAI-compatible LLM inference server for Apple Silicon using MLX. 2-4x faster than Ollama with tool calling and prompt caching.
Directory
Search results
Published directory entries matching your search.
The open source solution for monitoring your AI models in production.
Open-source project that supports deploying, serving, monitoring, or operating AI and machine-learning systems.
Open-source project that supports deploying, serving, monitoring, or operating AI and machine-learning systems.
Open-source vector similarity search for Postgres.
Vector database plugin for Postgres, written in Rust, specifically designed for LLM.
Turns your ML code into microservices with web API, interactive GUI, and more.
Open-source GenAI and LLM observability platform native to OpenTelemetry with traces and metrics. opensource.
Open-source project that supports deploying, serving, monitoring, or operating AI and machine-learning systems.
Open-source project that supports deploying, serving, monitoring, or operating AI and machine-learning systems.
Open-source project that supports deploying, serving, monitoring, or operating AI and machine-learning systems.
Neural networks framework in pure C: training and inference, no dependencies.
An Easy-to-Use and High-Performance AI deployment framework.
A framework providing the right abstractions to ease research, development, and deployment of your ML pipelines.
Easy-to-use library to boost AI inference.
Open-source project that supports deploying, serving, monitoring, or operating AI and machine-learning systems.
Ncnn is a high-performance neural network inference framework optimized for the mobile platform.
Proof-of-concept OpenAI Gym environment for Neural Architecture Search (NAS).
Machine learning model serving framework with dynamic batching and pipelined stages, provides an easy-to-use Python interface.
Open source MLOps platform that helps you collaborate, reproduce and share your ML work.
Version and deploy your ML models following GitOps principles.
A lightweight, portable pure C99 onnx inference engine for embedded devices with hardware acceleration support.
Neural network inference from the command line, implemented in CHICKEN Scheme.
Open-source project that supports deploying, serving, monitoring, or operating AI and machine-learning systems.