Turns your ML code into microservices with web API, interactive GUI, and more.
Directory
Search results
Published directory entries matching your search.
Platform and SDK for AI Engineers providing tools for LLM evaluation, observability, and a version-controlled enhanced prompt playground.
Vector database plugin for Postgres, written in Rust, specifically designed for LLM.
Open-source vector similarity search for Postgres.
MLOps and LLMOps, from the trenches.
Democratize and productionize Gen AI across your entire org with Portkey.
Quix is the agentic AI platform for hardware engineering — turn test-rig and sensor data into real-time insight.
OpenAI-compatible LLM inference server for Apple Silicon using MLX. 2-4x faster than Ollama with tool calling and prompt caching.
Examples showing how to use the OpenAI vision API to run inference on images, video files and webcam streams.
Provides containers to encapsulate and deploy EdgeML pipelines and applications.
MLOps framework to package, deploy, monitor and manage thousands of production machine learning models.
Python-free Rust inference server with OpenAI API compatibility and hot model swapping.
An MLOps/LLMOps platform for model building, evaluation, and fine-tuning.
Lets you create apps for your ML projects with deceptively simple Python scripts.
TensorZero builds open-source tools for production-grade LLM applications: LLM gateway, observability, optimization, evaluations, and experimentation.
Inference for text-embedding models.
A flexible and easy to use tool for serving PyTorch models.
Provides an optimized cloud and edge inferencing solution.
Package, configure, and iterate on a model with Truss at whatever level of control your model needs.
Highly Scalable Distributed Vector Search Engine.
Scalable, fast, and disk-friendly vector search in Postgres, the successor of pgvecto.rs.
Python vector database you just need - no more, no less.
Gemini Enterprise Agent Platform (formerly Vertex AI) is a comprehensive platform for developers to build, scale, govern and optimize agents.
Store, search, organize and make machine-learned inferences over big data at serving time.