Skip to main content
buildradar
Sign in
Topic · benchmark

benchmark

Tracked open-source repos tagged benchmark, sorted by stars.

158 repos
  • llm-colosseum@OpenGenerativeAI

    Benchmark LLMs by fighting in Street Fighter 3! The new way to evaluate the quality of an LLM

    1,483+0Star change over the last 7 days
  • pytest-benchmark@ionelmc

    pytest fixture for benchmarking code

    1,448+0Star change over the last 7 days
  • divan@nvzqz

    Fast and simple benchmarking for Rust projects

    1,444+1Star change over the last 7 days
  • MedMNIST@MedMNIST

    [pip install medmnist] 18x Standardized Datasets for 2D and 3D Biomedical Image Classification

    1,399+0Star change over the last 7 days
  • Fast Compiler for C# Expression Trees and the lightweight LightExpression alternative. Diagnostic and code generation tools for the expressions.

    1,375+1Star change over the last 7 days
  • smac@oxwhirl

    SMAC: The StarCraft Multi-Agent Challenge

    1,367+1Star change over the last 7 days
  • SLM-Lab@kengz

    Modular Deep Reinforcement Learning framework in PyTorch. Companion library of the book "Foundations of Deep Reinforcement Learning".

    1,362+0Star change over the last 7 days
  • Latest Advances on System-2 Reasoning

    1,353+0Star change over the last 7 days
  • kg-gen@stair-lab

    [NeurIPS '25] Knowledge Graph Generation from Any Text

    1,266+6Star change over the last 7 days
  • github-action-benchmark@benchmark-action

    GitHub Action for continuous benchmarking to keep performance

    1,251+0Star change over the last 7 days
  • boomer@myzhan

    A better load generator for locust, written in golang.

    1,237+0Star change over the last 7 days
  • LongBench@THUDM

    LongBench v2 and LongBench (ACL 25'&24')

    1,232+1Star change over the last 7 days
  • KernelBench@ScalingIntelligence

    KernelBench: Can LLMs Write GPU Kernels? - Benchmark + Toolkit with Torch -> CUDA (+ more DSLs)

    1,227+7Star change over the last 7 days
  • PDEBench@pdebench

    PDEBench: An Extensive Benchmark for Scientific Machine Learning

    1,192+0Star change over the last 7 days
  • bench-scripts@haydenjames

    A compilation of Linux server benchmarking scripts.

    1,190+0Star change over the last 7 days
  • flow@flow-project

    Computational framework for reinforcement learning in traffic control

    1,188+0Star change over the last 7 days
  • VectorDBBench@zilliztech

    Benchmark for vector databases.

    1,173+2Star change over the last 7 days
  • omnisafe@PKU-Alignment

    JMLR: OmniSafe is an infrastructural framework for accelerating SafeRL research.

    1,150+1Star change over the last 7 days
  • OpenSTL@chengtan9907

    OpenSTL: A Comprehensive Benchmark of Spatio-Temporal Predictive Learning

    1,143+5Star change over the last 7 days
  • primesieve@kimwalisch

    🚀 Fast prime number generator

    1,114+1Star change over the last 7 days
  • ClickBench@ClickHouse

    ClickBench: a Benchmark For Analytical Databases

    1,100+2Star change over the last 7 days
  • lzbench@inikep

    lzbench is an in-memory benchmark of open-source compressors

    1,084+4Star change over the last 7 days
  • NoSQL Redis and Memcache traffic generation and benchmarking tool.

    1,053+5Star change over the last 7 days
  • opencv_zoo@opencv

    Model Zoo For OpenCV DNN and Benchmarks.

    1,046+8Star change over the last 7 days
  • benchmark@pytorch

    TorchBench is a collection of open source benchmarks used to evaluate PyTorch performance.

    1,042-1Star change over the last 7 days
  • pyperformance@python

    Python Performance Benchmark Suite

    1,029+1Star change over the last 7 days
  • GLM-5.3-Flash × J-Space capability realization — benchmark presentation of the J-Space Cognition Suite

    1,026-11Star change over the last 7 days
  • ADBench@Minqi824

    Official Implement of "ADBench: Anomaly Detection Benchmark", NeurIPS 2022.

    1,023+2Star change over the last 7 days
  • asv@airspeed-velocity

    Airspeed Velocity: A simple Python benchmarking tool with web-based reporting

    1,018+4Star change over the last 7 days
  • moses@molecularsets

    Molecular Sets (MOSES): A Benchmarking Platform for Molecular Generation Models

    988+0Star change over the last 7 days
← Back to topics