Lets you create apps for your ML projects with deceptively simple Python scripts.
Directory
Search results
Published directory entries matching your search.
MLOps framework to package, deploy, monitor and manage thousands of production machine learning models.
Examples showing how to use the OpenAI vision API to run inference on images, video files and webcam streams.
OpenAI-compatible LLM inference server for Apple Silicon using MLX. 2-4x faster than Ollama with tool calling and prompt caching.
Neural networks framework in pure C: training and inference, no dependencies.
An Easy-to-Use and High-Performance AI deployment framework.
A framework providing the right abstractions to ease research, development, and deployment of your ML pipelines.
Ncnn is a high-performance neural network inference framework optimized for the mobile platform.
Private AI for individuals, teams, and organizations. Workspaces, agents, models, and knowledge—one connected system on your terms.
Machine learning model serving framework with dynamic batching and pipelined stages, provides an easy-to-use Python interface.
Helps teams build, deploy, observe or operate machine-learning systems.
Open-source all-in-one platform for engineering AI products. Traces, Evals, Datasets, Labels.
Kubernetes custom resource definition for serving ML models on arbitrary frameworks.
Respan unifies LLM observability, evals, prompt optimization, and an AI gateway so teams can ship reliable AI applications.
Kubernetes-based system for hyperparameter tuning and neural architecture search.
Helps teams build, deploy, observe or operate machine-learning systems.
Build, deploy, and scale production ML systems with Hopsworks. The Feature Store and MLOps platform for real-time AI, trusted by teams.
Curated list of awesome vector search framework/engine, library, cloud service and research papers to vector similarity search.
Helps teams build, deploy, observe or operate machine-learning systems.
Run sandboxes, task queues, and custom model inference with ultrafast boot times, instant autoscaling, and a developer experience that just works.
Inference hosting for AI teams who ship fast and scale faster.
Amadeus Code supports deploying, serving, monitoring, or operating AI and machine-learning systems.
Helps teams build, deploy, observe or operate machine-learning systems.
A temporal extension of PyTorch Geometric for dynamic graph representation learning.