Skip to main content
buildradar
Sign in
Topic · benchmarks

benchmarks

Tracked open-source repos tagged benchmarks, sorted by stars.

Repos
26
Total stars
34,310
Avg. stars
1,320
Share
0.00%

Topics that frequently appear alongside benchmarks on the same repo.

Recent risers

Repos created in the last 90 days, tagged benchmarks.

  • awesome-evals@benchflow-ai

    A curated, non-BS library of the best resources for building and evaluating AI agents — papers, blogs, talks, tools, benchmarks. Maintained by BenchFlow.

    865
  • JCTools@JCTools
    3,871+0Star change over the last 7 days
  • Primes@PlummersSoftwareLLC

    Prime number projects in 100+ programming languages, to compare their speed - and their programmer's cleverness

    3,019+3Star change over the last 7 days
  • benchmarks@kostya

    Some benchmarks of different languages

    2,923+1Star change over the last 7 days
  • ecs@oneclickvirt

    VPS Fusion Monster Server Test GO Version Aiming to be the most comprehensive server testing project, implemented in Go with zero environment dependencies. VPS融合怪服务器测评项目 GO版本 尽量成为最全能的服务器测评项目,使用 Go 实现,无需任何环境依赖。

    2,329+13Star change over the last 7 days
  • avalanche@ContinualAI

    Avalanche: an End-to-End Library for Continual Learning based on PyTorch.

    2,088+0Star change over the last 7 days
  • atomic_queue@max0x7ba

    C++14 concurrent lock-free low-latency queue.

    1,892+0Star change over the last 7 days
  • minification-benchmarks@privatenumber

    🏃‍♂️🏃‍♀️🏃 JS minification benchmarks: babel-minify, esbuild, terser, uglify-js, swc, google closure compiler, tdewolff/minify, oxc-minify

    1,619+0Star change over the last 7 days
  • InferenceX@SemiAnalysisAI

    Open Source Continuous Inference Benchmark Research Platform — Kimi K3 2.8T, MiniMax M3, DeepSeekv4, GLM5 - GB200 NVL72 vs MI355X vs B200 vs GB300 NVL72 & soon™ TPUv6e/v7/Trainium2/3 | 开源持续推理基准研究平台 — Kimi K2.7-Code、MiniMax M3、DeepSeekv4、GLM5 - GB200 NVL72 vs MI355X vs B200 vs GB300 NVL72,即将推出™ TPUv6e/v7/Trainium2/3

    1,613+33Star change over the last 7 days
  • ai-agents-the-definitive-guide@Nicolepcx

    Repo for AI Agents The Definitive Guide

    1,567+220Star change over the last 7 days
  • TDC@mims-harvard

    Therapeutics Commons (TDC): Multimodal Foundation for Therapeutic Science

    1,279+2Star change over the last 7 days
  • detect-gpu@pmndrs

    Classifies GPUs based on their 3D rendering benchmark score allowing the developer to provide sensible default settings for graphically intensive applications.

    1,211+0Star change over the last 7 days
  • Gym@NVIDIA-NeMo

    Evaluate and improve models and agents using environments

    1,155+7Star change over the last 7 days
  • sbc-reviews@geerlingguy

    Jeff Geerling's SBC review data - Raspberry Pi, Radxa, Orange Pi, etc.

    994+4Star change over the last 7 days
  • awesome-evals@benchflow-ai

    A curated, non-BS library of the best resources for building and evaluating AI agents — papers, blogs, talks, tools, benchmarks. Maintained by BenchFlow.

    865+17Star change over the last 7 days
  • 📊 Benchmark Comparison of Packages with Runtime Validation and TypeScript Support

    830+3Star change over the last 7 days
  • A curated list of 3D Vision papers relating to Robotics domain in the era of large models i.e. LLMs/VLMs, inspired by awesome-computer-vision, including papers, codes, and related websites

    821+1Star change over the last 7 days
  • Yet another implementation of computer language benchmarks game

    800-1Star change over the last 7 days
  • HammerDB@TPC-Council

    HammerDB: The industry standard open-source database benchmark

    782+2Star change over the last 7 days
  • validation-benchmarks@xbow-engineering

    XBOW Validation Benchmarks

    699+8Star change over the last 7 days
  • memo_wise@panorama-ed

    The wise choice for Ruby memoization

    633+0Star change over the last 7 days
  • robohive@vikashplus

    A unified framework for robot learning

    631+0Star change over the last 7 days
  • open-unlearning@locuslab

    [NeurIPS D&B '25] The one-stop repository for LLM unlearning

    594+6Star change over the last 7 days
  • langtest@PacificAI

    Deliver safe & effective language models

    559+0Star change over the last 7 days
  • CodaLab Competitions

    538-1Star change over the last 7 days
  • WildClawBench@InternLM

    An in-the-wild benchmark for AI agents in the production harness.

    516+1Star change over the last 7 days
  • models@reyamira

    TUI and CLI for browsing models.dev, benchmarks, coding agents, and statuses for AI providers.

    505+3Star change over the last 7 days
← Back to topics