Write maintainable, production-ready pipelines. Develop locally, deploy to the cloud.
Directory
Search results
Published directory entries matching your search.
A high-speed inference engine for deploying LLMs locally.
Open-source tool to simplify the process of creating and managing LLM workflows and prompts as a self-hosted solution.
Open-source project that provides machine-learning models, research, training resources, or evaluation tools.
Open-source project that supports deploying, serving, monitoring, or operating AI and machine-learning systems.
Deep learning framework to train, deploy, and ship AI products Lightning fast.
Open-source project that supports deploying, serving, monitoring, or operating AI and machine-learning systems.
Open-source project that provides machine-learning models, research, training resources, or evaluation tools.
The open source solution for monitoring your AI models in production.
OpenAI-compatible LLM inference server for Apple Silicon using MLX. 2-4x faster than Ollama with tool calling and prompt caching.
Open-source project that provides machine-learning models, research, training resources, or evaluation tools.
Examples showing how to use the OpenAI vision API to run inference on images, video files and webcam streams.
Open-source project that provides machine-learning models, research, training resources, or evaluation tools.
Provides containers to encapsulate and deploy EdgeML pipelines and applications.
A CLI utility to train and deploy ML/DL models on AWS SageMaker.
MLOps framework to package, deploy, monitor and manage thousands of production machine learning models.
Open-source project that provides machine-learning models, research, training resources, or evaluation tools.
Python-free Rust inference server with OpenAI API compatibility and hot model swapping.
Open-source project that provides machine-learning models, research, training resources, or evaluation tools.
Open-source DevOps agent to help you secure, deploy, and maintain production-ready infrastructure.
An MLOps/LLMOps platform for model building, evaluation, and fine-tuning.
Lets you create apps for your ML projects with deceptively simple Python scripts.
Inference engine for TensorRT on Nvidia GPUs.
Inference for text-embedding models.