Directory

Search results

Published directory entries matching your search.

Search results

27 listings
eqbench.com

EQ-Bench provides machine-learning models, research, training resources, or evaluation tools.

AI Tools & Directories 0
github.com

Benchmarks of machine learning inference for Go.

Models & Machine Learning 0
kaggle.com

Kaggle Benchmarks catalogs AI services and resources for browsing by feature or use case.

AI Tools & Directories 0
petergpt.github.io

LLMs Bullshit Benchmark provides machine-learning models, research, training resources, or evaluation tools.

AI Tools & Directories 0
simple-bench.com

Simple Bench provides machine-learning models, research, training resources, or evaluation tools.

AI Tools & Directories 0
tbench.ai

Coding agents help developers plan, implement, review, test, and debug software. For independent capability comparisons, see SWE-bench and.

Models & Machine Learning 0

Wolfram LLM Benchmarking Project provides machine-learning models, research, training resources, or evaluation tools.

AI Tools & Directories 0

Curated list of top Hugging Face models for NLP, vision, and audio tasks with demos and benchmarks.

Models & Machine Learning 0
github.com

GPT Migrate team is working on adding for the agent.

Models & Machine Learning 0
github.com

Fast Neural Networks framework built on top of Metal. Supports TensorFlow models.

Models & Machine Learning 0
github.com

Open-source platform for high-performance ML model serving.

AI Infrastructure & MLOps 0
codeflash.ai

Codeflash uses AI to automatically find the most optimized version of your Python code through benchmarking — while verifying it's correct.

Models & Machine Learning 0
github.com

"a collaborative benchmark intended to probe large language models and extrapolate their future capabilities".

Models & Machine Learning 0
github.com

Library for hyperparameter optimization and black box optimization benchmarks.

Models & Machine Learning 0
livebench.ai

A Challenging, Contamination-Free LLM Benchmark.

Models & Machine Learning 0
github.com

An open source robotics benchmark for meta- and multi-task reinforcement learning.

Models & Machine Learning 0
github.com

Evals is a framework for evaluating LLMs and LLM systems, and an open-source registry of benchmarks.

Models & Machine Learning 0
github.com

Benchmarking Large Language Models.

Models & Machine Learning 0
github.com

CLI tool for debugging and benchmarking RAG retrieval. EXPLAIN ANALYZE for your retrieval layer.

Models & Machine Learning 0