Search results

Published directory entries matching your search.

Filters
Clear filters

Search results

28 listings
Preview unavailable

Language Benchmarks

Language Benchmarks is a developer resource for coding, documentation, communities, security research, or software tools.

Programming Languages Website
Preview unavailable

Benchmarks Game

Benchmarks Game is a developer resource for coding, documentation, communities, security research, or software tools.

Programming Languages Website
Preview unavailable

Cybenetics PSU Benchmarks

Cybenetics PSU Benchmarks is a curated directory or link collection for discovering useful websites, tools, and resources.

Shopping & Consumer Website
Preview unavailable

Human Benchmark

Human Benchmark is a miscellaneous web resource for discovery, directories, archives, web culture, or useful tools.

Fun & Interesting Website
Preview unavailable

Open Benchmarking

Open Benchmarking is a miscellaneous web resource for discovery, directories, archives, web culture, or useful tools.

Shopping & Consumer Website
Preview unavailable

UNIGINE Benchmarks

UNIGINE Benchmarks is a system utility for Windows, hardware checks, cleanup, process control, or desktop customization.

Hardware & Diagnostics Website
Preview unavailable

go-ml-benchmarks

Benchmarks of machine learning inference for Go.

Models & Machine Learning GitHub repo
Preview unavailable

Kaggle Benchmarks

Kaggle Benchmarks catalogs AI services and resources for browsing by feature or use case.

AI Tools & Directories Website
Preview unavailable

LLMs Bullshit Benchmark

LLMs Bullshit Benchmark provides machine-learning models, research, training resources, or evaluation tools.

AI Tools & Directories Website
Preview unavailable

Wolfram LLM Benchmarking Project

Wolfram LLM Benchmarking Project provides machine-learning models, research, training resources, or evaluation tools.

AI Tools & Directories Website
Preview unavailable

Awesome Hugging Face Models

Curated list of top Hugging Face models for NLP, vision, and audio tasks with demos and benchmarks.

Models & Machine Learning GitHub repo
Preview unavailable

benchmarks

GPT Migrate team is working on adding for the agent.

Models & Machine Learning GitHub repo
Preview unavailable

Codeflash

Codeflash uses AI to automatically find the most optimized version of your Python code through benchmarking — while verifying it's correct.

Models & Machine Learning Website
Preview unavailable

google/BIG-bench

"a collaborative benchmark intended to probe large language models and extrapolate their future capabilities".

Models & Machine Learning GitHub repo
Preview unavailable

HPOlib2

Library for hyperparameter optimization and black box optimization benchmarks.

Models & Machine Learning GitHub repo
Preview unavailable

LiveBench

A Challenging, Contamination-Free LLM Benchmark.

Models & Machine Learning Website
Preview unavailable

metaworld

An open source robotics benchmark for meta- and multi-task reinforcement learning.

Models & Machine Learning GitHub repo
Preview unavailable

openai/evals

Evals is a framework for evaluating LLMs and LLM systems, and an open-source registry of benchmarks.

Models & Machine Learning GitHub repo
Preview unavailable

qcri/LLMeBench

Benchmarking Large Language Models.

Models & Machine Learning GitHub repo
Preview unavailable

RagTune

CLI tool for debugging and benchmarking RAG retrieval. EXPLAIN ANALYZE for your retrieval layer.

Models & Machine Learning GitHub repo