OpenAI-compatible LLM inference server for Apple Silicon using MLX. 2-4x faster than Ollama with tool calling and prompt caching.
Directory
Search results
Published directory entries matching your search.
Examples showing how to use the OpenAI vision API to run inference on images, video files and webcam streams.
Provides containers to encapsulate and deploy EdgeML pipelines and applications.
MLOps framework to package, deploy, monitor and manage thousands of production machine learning models.
seo-os helps monitor local SEO visibility, map rankings, business listings, and local search performance.
Python-free Rust inference server with OpenAI API compatibility and hot model swapping.
An MLOps/LLMOps platform for model building, evaluation, and fine-tuning.
Lets you create apps for your ML projects with deceptively simple Python scripts.
Multi-monitor wallpaper manager for spanning, cropping, and managing desktop backgrounds.
Inference for text-embedding models.
A flexible and easy to use tool for serving PyTorch models.
Provides an optimized cloud and edge inferencing solution.
Highly Scalable Distributed Vector Search Engine.
Scalable, fast, and disk-friendly vector search in Postgres, the successor of pgvecto.rs.
Python vector database you just need - no more, no less.
Store, search, organize and make machine-learned inferences over big data at serving time.
Machine Learning Operations - An awesome list of references for MLOps.
VSCode Solidity Auditor provides smart contract auditing, blockchain security reviews, or Web3 risk assessment.