Vector database plugin for Postgres, written in Rust, specifically designed for LLM.
Directory
Search results
Published directory entries matching your search.
Open-source vector similarity search for Postgres.
Use AutoML to do model compression.
A high-speed inference engine for deploying LLMs locally.
Open-source tool to simplify the process of creating and managing LLM workflows and prompts as a self-hosted solution.
Open-source project that provides machine-learning models, research, training resources, or evaluation tools.
Open-source project that supports deploying, serving, monitoring, or operating AI and machine-learning systems.
Open-source project that supports deploying, serving, monitoring, or operating AI and machine-learning systems.
Open-source project that provides machine-learning models, research, training resources, or evaluation tools.
The open source solution for monitoring your AI models in production.
OpenAI-compatible LLM inference server for Apple Silicon using MLX. 2-4x faster than Ollama with tool calling and prompt caching.
Open-source project that provides machine-learning models, research, training resources, or evaluation tools.
Examples showing how to use the OpenAI vision API to run inference on images, video files and webcam streams.
Open-source project that provides machine-learning models, research, training resources, or evaluation tools.
Provides containers to encapsulate and deploy EdgeML pipelines and applications.
MLOps framework to package, deploy, monitor and manage thousands of production machine learning models.
Open-source project that provides machine-learning models, research, training resources, or evaluation tools.
Open-source project that provides machine-learning models, research, training resources, or evaluation tools.
AI-generated visualization prototyping and editing platform, support 2D, 3D models, combined with LLM(Large Language Model) for quick editing.
Lets you create apps for your ML projects with deceptively simple Python scripts.
Inference engine for TensorRT on Nvidia GPUs.
Inference for text-embedding models.
Flexible, high-performance serving system for machine learning models.
Uniform deep learning inference framework for mobile, desktop and server.