Jan - Run LLMs like Mistral or Llama2 locally and offline on your computer, or connect to remote AI APIs.
Directory
Search results
Published directory entries matching your search.
Turns your ML code into microservices with web API, interactive GUI, and more.
Vector database plugin for Postgres, written in Rust, specifically designed for LLM.
Open-source vector similarity search for Postgres.
A high-speed inference engine for deploying LLMs locally.
Open-source tool to simplify the process of creating and managing LLM workflows and prompts as a self-hosted solution.
Open-source project that provides machine-learning models, research, training resources, or evaluation tools.
Open-source project that supports deploying, serving, monitoring, or operating AI and machine-learning systems.
Open-source project that supports deploying, serving, monitoring, or operating AI and machine-learning systems.
Open-source project that provides machine-learning models, research, training resources, or evaluation tools.
The open source solution for monitoring your AI models in production.
OpenAI-compatible LLM inference server for Apple Silicon using MLX. 2-4x faster than Ollama with tool calling and prompt caching.
Open-source project that provides machine-learning models, research, training resources, or evaluation tools.
Examples showing how to use the OpenAI vision API to run inference on images, video files and webcam streams.
Open-source project that provides machine-learning models, research, training resources, or evaluation tools.
Provides containers to encapsulate and deploy EdgeML pipelines and applications.
Open-source project that provides machine-learning models, research, training resources, or evaluation tools.
Python-free Rust inference server with OpenAI API compatibility and hot model swapping.
Open-source project that provides machine-learning models, research, training resources, or evaluation tools.
Lets you create apps for your ML projects with deceptively simple Python scripts.
Suite of tools that users, both novice and advanced, can use to optimize machine learning models for deployment and execution.
Inference engine for TensorRT on Nvidia GPUs.
Inference for text-embedding models.
Large Language Model Text Generation Inference.