Directory

Search results

Published directory entries matching your search.

Search results

48 listings
github.com

Benchmarks of machine learning inference for Go.

Models & Machine Learning 0
kaggle.com

Kaggle Benchmarks catalogs AI services and resources for browsing by feature or use case.

AI Tools & Directories 0
petergpt.github.io

LLMs Bullshit Benchmark provides machine-learning models, research, training resources, or evaluation tools.

AI Tools & Directories 0

Wolfram LLM Benchmarking Project provides machine-learning models, research, training resources, or evaluation tools.

AI Tools & Directories 0
ygzhang.com

(Non-)Human — an art installation exploring the semi-human, semi-object territory.

AI Image Generation 0

Curated list of top Hugging Face models for NLP, vision, and audio tasks with demos and benchmarks.

Models & Machine Learning 0
github.com

GPT Migrate team is working on adding for the agent.

Models & Machine Learning 0
codeflash.ai

Codeflash uses AI to automatically find the most optimized version of your Python code through benchmarking — while verifying it's correct.

Models & Machine Learning 0
github.com

"a collaborative benchmark intended to probe large language models and extrapolate their future capabilities".

Models & Machine Learning 0
github.com

Library for hyperparameter optimization and black box optimization benchmarks.

Models & Machine Learning 0
livebench.ai

A Challenging, Contamination-Free LLM Benchmark.

Models & Machine Learning 0
github.com

An open source robotics benchmark for meta- and multi-task reinforcement learning.

Models & Machine Learning 0
github.com

Evals is a framework for evaluating LLMs and LLM systems, and an open-source registry of benchmarks.

Models & Machine Learning 0
github.com

Benchmarking Large Language Models.

Models & Machine Learning 0
github.com

CLI tool for debugging and benchmarking RAG retrieval. EXPLAIN ANALYZE for your retrieval layer.

Models & Machine Learning 0
labs.scale.com

Explore leaderboards with expert-driven LLM benchmarks and updated AI model rankings across coding, reasoning and more.

Models & Machine Learning 0
trulens.org

TruLens instruments your AI agent with OpenTelemetry, scores every step with benchmarked LLM judges, and tells you which version to ship.

Models & Machine Learning 0
nothumansearch.ai

Not Human Search provides machine-learning models, research, training resources, or evaluation tools.

Models & Machine Learning 0