Go binding for MXNet c predict api to do inference with a pre-trained model.
Directory
Search results
Published directory entries matching your search.
Showcase streaming text generation using the @huggingface/inference JS lib.
Standardized Serverless ML Inference Platform on Kubernetes.
Neural network inference from the command line, implemented in CHICKEN Scheme.
A lightweight, portable pure C99 onnx inference engine for embedded devices with hardware acceleration support.
Algorithms for learning and inference with discrete probabilistic models.
Extensible Toolkit for Finetuning and Inference of Large Foundation Models.
MindSpore is a new open source deep learning training/inference framework that could be used for mobile, edge and cloud scenarios.
Ncnn is a high-performance neural network inference framework optimized for the mobile platform.
Easy-to-use library to boost AI inference.
Neural networks framework in pure C: training and inference, no dependencies.
NeMo Parakeet ASR Models attain strong speech recognition accuracy while being efficient for inference. Available in CTC and RNN-Transducer variants.
A high-speed inference engine for deploying LLMs locally.
Python Library for Probabalistic Programming (Bayesian Inference and Machine Learning).
OpenAI-compatible LLM inference server for Apple Silicon using MLX. 2-4x faster than Ollama with tool calling and prompt caching.
Examples showing how to use the OpenAI vision API to run inference on images, video files and webcam streams.
Python-free Rust inference server with OpenAI API compatibility and hot model swapping.
AI-powered social listening tool that tracks brand mentions across Google, Instagram, Facebook, TikTok, Reddit, and X (Twitter). Uses HasData SERP API + LLM inference to extract structured insights, generate charts, export CSV reports, and send Telegram notifications on a 24-hour automated cycle.
Inference engine for TensorRT on Nvidia GPUs.
Uniform deep learning inference framework for mobile, desktop and server.
Provides an optimized cloud and edge inferencing solution.
Store, search, organize and make machine-learned inferences over big data at serving time.
An inference and serving engine for language models.
A Whisper CLI client compatible with the original OpenAI client, using CTranslate2 for faster inference. opensource.