Latest Papers and Datasets on Multimodal Large Language Models, and Their Evaluation.
Directory
Search results
Published directory entries matching your search.
Track, log, visualize and evaluate your LLM prompts and prompt chains.
A python library for accurate and scalable fuzzy matching, record deduplication and entity-resolution.
Deploy and scale machine learning models on Kubernetes. Built for LLMs, embeddings, and speech-to-text.
A Clojure library of optimisation and control theory tools and convenience functions based on Neanderthal.
C, C++, and Python tools for named entity recognition and relation extraction.
OpenAI compatible API for LLMs and embeddings (LLaMA, Vicuna, ChatGLM and many others).
CEA-List's CAD framework for designing and simulating Deep Neural Network, and building full DNN-based applications on embedded platforms.
Evals is a framework for evaluating LLMs and LLM systems, and an open-source registry of benchmarks.
Open-source tool to simplify the process of creating and managing LLM workflows and prompts as a self-hosted solution.
Fast and simple framework for building and running distributed applications.
An LLM-powered repository agent designed to assist developers and teams in generating documentation and understanding repositories quickly.
Open-source SDK for running LLMs and multimodal models on-device across iOS, Android, and cross-platform apps.
Simple and efficient library to minimize expensive and noisy black-box functions.
An open-source tool for recording screen and audio activity with AI-powered search, automations, and support for local LLMs. opensource.
Synthetic tabular data generation using GANs, Diffusion Models, and LLMs with adversarial filtering and privacy metrics.
A comprehensive testing and evaluation framework for voice agents across language models, prompts, and agent personas.
Open source framework for debugging LLM agents and RAG pipelines with a 16-mode ProblemMap and practical triage checklists.
Build and control your personal LLMs with fast and efficient fine-tuning.
A simple way to train and use PyTorch models with multi-GPU, TPU, mixed-precision.
Open-source project that supports deploying, serving, monitoring, or operating AI and machine-learning systems.
Open-source project that supports deploying, serving, monitoring, or operating AI and machine-learning systems.
Open-source Python library enabling ML model inspection and interpretation.
An open source Python library focused on outlier, adversarial and drift detection.