Open-source project that supports deploying, serving, monitoring, or operating AI and machine-learning systems.
Directory
Search results
Published directory entries matching your search.
Easy-to-use library to boost AI inference.
A framework providing the right abstractions to ease research, development, and deployment of your ML pipelines.
An Easy-to-Use and High-Performance AI deployment framework.
Neural networks framework in pure C: training and inference, no dependencies.
Serve Llama 2 and other large language models locally from command line or through a browser interface.
Open-source project that supports deploying, serving, monitoring, or operating AI and machine-learning systems.
Fujitsu Research's post-training quantization pipeline for LLMs (QEP, AutoBit, JointQ, rotation) with vLLM plugin (arXiv:2603.28845).
Open-source project that supports deploying, serving, monitoring, or operating AI and machine-learning systems.
Open-source project that supports deploying, serving, monitoring, or operating AI and machine-learning systems.
Open-source GenAI and LLM observability platform native to OpenTelemetry with traces and metrics. opensource.
Jan - Run LLMs like Mistral or Llama2 locally and offline on your computer, or connect to remote AI APIs.
Turns your ML code into microservices with web API, interactive GUI, and more.
Vector database plugin for Postgres, written in Rust, specifically designed for LLM.
Open-source vector similarity search for Postgres.
Use AutoML to do model compression.
A high-speed inference engine for deploying LLMs locally.
Open-source tool to simplify the process of creating and managing LLM workflows and prompts as a self-hosted solution.
Open-source project that provides machine-learning models, research, training resources, or evaluation tools.
Open-source project that supports deploying, serving, monitoring, or operating AI and machine-learning systems.
Open-source project that supports deploying, serving, monitoring, or operating AI and machine-learning systems.
Open-source project that provides machine-learning models, research, training resources, or evaluation tools.
The open source solution for monitoring your AI models in production.
OpenAI-compatible LLM inference server for Apple Silicon using MLX. 2-4x faster than Ollama with tool calling and prompt caching.