Serve Llama 2 and other large language models locally from command line or through a browser interface.
Directory
Search results
Published directory entries matching your search.
Open-source project that supports deploying, serving, monitoring, or operating AI and machine-learning systems.
Fujitsu Research's post-training quantization pipeline for LLMs (QEP, AutoBit, JointQ, rotation) with vLLM plugin (arXiv:2603.28845).
Open-source project that supports deploying, serving, monitoring, or operating AI and machine-learning systems.
Open-source project that supports deploying, serving, monitoring, or operating AI and machine-learning systems.
Open-source GenAI and LLM observability platform native to OpenTelemetry with traces and metrics. opensource.
Open platform for operating large language models (LLMs) in production. Fine-tune, serve, deploy, and monitor any LLMs with ease.
Vector database plugin for Postgres, written in Rust, specifically designed for LLM.
Open-source vector similarity search for Postgres.
Write maintainable, production-ready pipelines. Develop locally, deploy to the cloud.
A high-speed inference engine for deploying LLMs locally.
Open-source tool to simplify the process of creating and managing LLM workflows and prompts as a self-hosted solution.
Open-source project that provides machine-learning models, research, training resources, or evaluation tools.
Open-source project that supports deploying, serving, monitoring, or operating AI and machine-learning systems.
Deep learning framework to train, deploy, and ship AI products Lightning fast.
Open-source project that supports deploying, serving, monitoring, or operating AI and machine-learning systems.
Open-source project that provides machine-learning models, research, training resources, or evaluation tools.
OpenAI-compatible LLM inference server for Apple Silicon using MLX. 2-4x faster than Ollama with tool calling and prompt caching.
Open-source project that provides machine-learning models, research, training resources, or evaluation tools.
Examples showing how to use the OpenAI vision API to run inference on images, video files and webcam streams.
Open-source project that provides machine-learning models, research, training resources, or evaluation tools.
Provides containers to encapsulate and deploy EdgeML pipelines and applications.
A CLI utility to train and deploy ML/DL models on AWS SageMaker.
MLOps framework to package, deploy, monitor and manage thousands of production machine learning models.