Benchmarks of machine learning inference for Go.
Directory
Search results
Published directory entries matching your search.
Curated list of top Hugging Face models for NLP, vision, and audio tasks with demos and benchmarks.
GPT Migrate team is working on adding for the agent.
"a collaborative benchmark intended to probe large language models and extrapolate their future capabilities".
Library for hyperparameter optimization and black box optimization benchmarks.
An open source robotics benchmark for meta- and multi-task reinforcement learning.
Evals is a framework for evaluating LLMs and LLM systems, and an open-source registry of benchmarks.
Benchmarking Large Language Models.
CLI tool for debugging and benchmarking RAG retrieval. EXPLAIN ANALYZE for your retrieval layer.
Here, CLI tools, libraries, Add-ons, Reports, Benchmarks and Sample Scripts for taking advantage of Google Apps Script which are publishing in my blog, Gists and GitHub are summarized.
Research archive containing dark-web authorship verification datasets and baseline models.