benchmarks
Tracked open-source repos tagged benchmarks, sorted by stars.
Related topics
Topics that frequently appear alongside benchmarks on the same repo.
Recent risers
Repos created in the last 90 days, tagged benchmarks.
- #1
A curated, non-BS library of the best resources for building and evaluating AI agents — papers, blogs, talks, tools, benchmarks. Maintained by BenchFlow.
★ 865
- #1★ 3,871+0Star change over the last 7 days
- #2
Prime number projects in 100+ programming languages, to compare their speed - and their programmer's cleverness
★ 3,019+3Star change over the last 7 days - #3
Some benchmarks of different languages
★ 2,923+1Star change over the last 7 days - #4
VPS Fusion Monster Server Test GO Version Aiming to be the most comprehensive server testing project, implemented in Go with zero environment dependencies. VPS融合怪服务器测评项目 GO版本 尽量成为最全能的服务器测评项目,使用 Go 实现,无需任何环境依赖。
★ 2,329+13Star change over the last 7 days - #5★ 2,088+0Star change over the last 7 days
- #6
C++14 concurrent lock-free low-latency queue.
★ 1,892+0Star change over the last 7 days - #7
🏃♂️🏃♀️🏃 JS minification benchmarks: babel-minify, esbuild, terser, uglify-js, swc, google closure compiler, tdewolff/minify, oxc-minify
★ 1,619+0Star change over the last 7 days - #8
Open Source Continuous Inference Benchmark Research Platform — Kimi K3 2.8T, MiniMax M3, DeepSeekv4, GLM5 - GB200 NVL72 vs MI355X vs B200 vs GB300 NVL72 & soon™ TPUv6e/v7/Trainium2/3 | 开源持续推理基准研究平台 — Kimi K2.7-Code、MiniMax M3、DeepSeekv4、GLM5 - GB200 NVL72 vs MI355X vs B200 vs GB300 NVL72,即将推出™ TPUv6e/v7/Trainium2/3
★ 1,613+33Star change over the last 7 days - #9
Repo for AI Agents The Definitive Guide
★ 1,567+220Star change over the last 7 days - #10★ 1,279+2Star change over the last 7 days
- #11
Classifies GPUs based on their 3D rendering benchmark score allowing the developer to provide sensible default settings for graphically intensive applications.
★ 1,211+0Star change over the last 7 days - #12★ 1,155+7Star change over the last 7 days
- #13
Jeff Geerling's SBC review data - Raspberry Pi, Radxa, Orange Pi, etc.
★ 994+4Star change over the last 7 days - #14
A curated, non-BS library of the best resources for building and evaluating AI agents — papers, blogs, talks, tools, benchmarks. Maintained by BenchFlow.
★ 865+17Star change over the last 7 days - #15
📊 Benchmark Comparison of Packages with Runtime Validation and TypeScript Support
★ 830+3Star change over the last 7 days - #16
A curated list of 3D Vision papers relating to Robotics domain in the era of large models i.e. LLMs/VLMs, inspired by awesome-computer-vision, including papers, codes, and related websites
★ 821+1Star change over the last 7 days - #17
Yet another implementation of computer language benchmarks game
★ 800-1Star change over the last 7 days - #18★ 782+2Star change over the last 7 days
- #19
XBOW Validation Benchmarks
★ 699+8Star change over the last 7 days - #20★ 633+0Star change over the last 7 days
- #21★ 631+0Star change over the last 7 days
- #22
[NeurIPS D&B '25] The one-stop repository for LLM unlearning
★ 594+6Star change over the last 7 days - #23★ 559+0Star change over the last 7 days
- #24
CodaLab Competitions
★ 538-1Star change over the last 7 days - #25
An in-the-wild benchmark for AI agents in the production harness.
★ 516+1Star change over the last 7 days - #26
TUI and CLI for browsing models.dev, benchmarks, coding agents, and statuses for AI providers.
★ 505+3Star change over the last 7 days