OpenAI-compatible LLM inference server for Apple Silicon using MLX. 2-4x faster than Ollama with tool calling and prompt caching.
Directory
Search results
Published directory entries matching your search.
Open-source project that provides machine-learning models, research, training resources, or evaluation tools.
Examples showing how to use the OpenAI vision API to run inference on images, video files and webcam streams.
Open-source project that provides machine-learning models, research, training resources, or evaluation tools.
Create playable games with AI prompts. Describe your idea or choose a template, then build and deploy your game. No coding required.
Provides containers to encapsulate and deploy EdgeML pipelines and applications.
RunThisLLM provides machine-learning models, research, training resources, or evaluation tools.
A CLI utility to train and deploy ML/DL models on AWS SageMaker.
MLOps framework to package, deploy, monitor and manage thousands of production machine learning models.
Open-source project that provides machine-learning models, research, training resources, or evaluation tools.
Python-free Rust inference server with OpenAI API compatibility and hot model swapping.
Open-source project that provides machine-learning models, research, training resources, or evaluation tools.
SinglebaseCloud supports deploying, serving, monitoring, or operating AI and machine-learning systems.
SocialBu supports deploying, serving, monitoring, or operating AI and machine-learning systems.
Open-source DevOps agent to help you secure, deploy, and maintain production-ready infrastructure.
Stan supports deploying, serving, monitoring, or operating AI and machine-learning systems.
An MLOps/LLMOps platform for model building, evaluation, and fine-tuning.
Lets you create apps for your ML projects with deceptively simple Python scripts.
Superwise supports deploying, serving, monitoring, or operating AI and machine-learning systems.
systemprompt.io supports deploying, serving, monitoring, or operating AI and machine-learning systems.
TeamoRouter provides machine-learning models, research, training resources, or evaluation tools.
Tecton supports deploying, serving, monitoring, or operating AI and machine-learning systems.
Inference engine for TensorRT on Nvidia GPUs.
TensorZero builds open-source tools for production-grade LLM applications: LLM gateway, observability, optimization, evaluations, and experimentation.