The open source solution for monitoring your AI models in production.
Directory
Search results
Published directory entries matching your search.
OpenAI-compatible LLM inference server for Apple Silicon using MLX. 2-4x faster than Ollama with tool calling and prompt caching.
Real-Time Latent Consistency Model Text to Image creates or edits images with generative AI and text-based controls.
Real-Time Latent Consistency Model Text-to-Image-Lora-SD1.5 creates or edits images with generative AI and text-based controls.
Multi-model simultaneous generation from a single prompt, fully unrestricted and packed with the latest greatest AI models.
Examples showing how to use the OpenAI vision API to run inference on images, video files and webcam streams.
Provides containers to encapsulate and deploy EdgeML pipelines and applications.
Explore leaderboards with expert-driven LLM benchmarks and updated AI model rankings across coding, reasoning and more.
MLOps framework to package, deploy, monitor and manage thousands of production machine learning models.
SinglebaseCloud supports deploying, serving, monitoring, or operating AI and machine-learning systems.
SocialBu supports deploying, serving, monitoring, or operating AI and machine-learning systems.
AI-generated visualization prototyping and editing platform, support 2D, 3D models, combined with LLM(Large Language Model) for quick editing.
Stan supports deploying, serving, monitoring, or operating AI and machine-learning systems.
Examples for ICASSP2024 paper “StemGen: A music generation model that listens”.
Lets you create apps for your ML projects with deceptively simple Python scripts.
Superwise supports deploying, serving, monitoring, or operating AI and machine-learning systems.
systemprompt.io supports deploying, serving, monitoring, or operating AI and machine-learning systems.
Tecton supports deploying, serving, monitoring, or operating AI and machine-learning systems.
Run GFPGAN created by tencentarc, the #1 AI model for Practical Face Restoration. Restore old photos or AI generated faces with GFPGAN.
Inference engine for TensorRT on Nvidia GPUs.
Creates or edits images with generative models and visual controls.
Inference for text-embedding models.
Flexible, high-performance serving system for machine learning models.
The FLUX.1 family of models – Replicate creates or edits images with generative AI and text-based controls.