Uncover LLM evaluation's importance and explore methods for assessing its performance and impact across industries.
Directory
Search results
Published directory entries matching your search.
Collection of patterns for experimenting with agents, llm pipelines, and ChainOfThoughtStrategy.
Open source generative AI development platform for building AI agents, LLM orchestration, and more.
Production-grade SDK for observability, automated evaluations and prompt management with sub-100ms guardrails for LLM/agent workflows.
Cloud-Native LLM Routing Engine. Improve LLM app resilience and speed.
An LLM by xAI with open source and open weights. opensource.
A curated list of Large Language Model.
Run LLM backends, APIs, frontends, and services with one command.
Indic LLM Arena provides machine-learning models, research, training resources, or evaluation tools.
Respan unifies LLM observability, evals, prompt optimization, and an AI gateway so teams can ship reliable AI applications.
Developer-friendly, serverless vector database for AI applications. Easily add long-term memory to your LLM apps!
A Challenging, Contamination-Free LLM Benchmark.
LLM provides machine-learning models, research, training resources, or evaluation tools.
LLM App is a Python library that helps you build real-time AI-powered data pipelines with few lines of code.
No-code batch compute platform for LLM evaluation and tuning workloads.
LLM Explorer provides machine-learning models, research, training resources, or evaluation tools.
LLM Model VRAM Calculator provides machine-learning models, research, training resources, or evaluation tools.
LLM Papers provides machine-learning models, research, training resources, or evaluation tools.
LLM Pricing provides machine-learning models, research, training resources, or evaluation tools.
LLM Resources Hub provides machine-learning models, research, training resources, or evaluation tools.
No-code platform to build generative AI apps, chatbots and agents with your data.
LLM Stats provides machine-learning models, research, training resources, or evaluation tools.
LLMFlows is a framework for building simple, explicit, and transparent LLM applications such as chatbots, question-answering systems, and agents.
LLMs Bullshit Benchmark provides machine-learning models, research, training resources, or evaluation tools.