benchmark
Tracked open-source repos tagged benchmark, sorted by stars.
- #31★ 2,800+12Star change over the last 7 days
- #32
Java web common vulnerabilities and security code which is base on springboot and spring security
★ 2,675+1Star change over the last 7 days - #33
CPU-X is a Free software that gathers information on CPU, motherboard and more
★ 2,647+0Star change over the last 7 days - #34
Tsung is a high-performance benchmark framework for various protocols including HTTP, XMPP, LDAP, etc.
★ 2,629+0Star change over the last 7 days - #35
A harness optimized to smaller LLMs
★ 2,528+16Star change over the last 7 days - #36★ 2,525+4Star change over the last 7 days
- #37
[ECCV2024] Video Foundation Models & Data for Multimodal Understanding
★ 2,374+5Star change over the last 7 days - #38★ 2,362+0Star change over the last 7 days
- #39
VPS Fusion Monster Server Test GO Version Aiming to be the most comprehensive server testing project, implemented in Go with zero environment dependencies. VPS融合怪服务器测评项目 GO版本 尽量成为最全能的服务器测评项目,使用 Go 实现,无需任何环境依赖。
★ 2,329+13Star change over the last 7 days - #40
A Heterogeneous Benchmark for Information Retrieval. Easy to use, evaluate your models across 15+ diverse IR datasets.
★ 2,283+5Star change over the last 7 days - #41★ 2,264+14Star change over the last 7 days
- #42
Cista is a simple, high-performance, zero-copy C++ serialization & reflection library.
★ 2,247+1Star change over the last 7 days - #43
📰 Must-read papers and blogs on LLM based Long Context Modeling 🔥
★ 2,164-1Star change over the last 7 days - #44
:zap: Go web framework benchmark
★ 2,135+1Star change over the last 7 days - #45
An objective comparison of multiple frameworks that allow us to "transform" our web apps to desktop applications.
★ 1,989+0Star change over the last 7 days - #46★ 1,987+0Star change over the last 7 days
- #47★ 1,973+2Star change over the last 7 days
- #48
τ-Bench: A Benchmark for Tool-Agent-User Interaction in Real-World Domains
★ 1,938+30Star change over the last 7 days - #49
Playing around "Less Slow" coding practices in C++ 20, C, CUDA, PTX, & Assembly, from numerics & SIMD to coroutines, ranges, exception handling, networking and user-space IO
★ 1,923-1Star change over the last 7 days - #50★ 1,806+1Star change over the last 7 days
- #51★ 1,771-1Star change over the last 7 days
- #52★ 1,757+7Star change over the last 7 days
- #53
SkillsBench evaluates how well skills work and how effective agents are at using them.
★ 1,742+11Star change over the last 7 days - #54
Simple, fast, accurate single-header microbenchmarking functionality for C++11/14/17/20
★ 1,733+1Star change over the last 7 days - #55
Quickly find bottlenecks in Rust - one profiler for CPU, memory, SQL, HTTP, I/O and async code.
★ 1,707+7Star change over the last 7 days - #56
BEHAVIOR-1K: a platform for accelerating Embodied AI research. Join our Discord for support: https://discord.gg/bccR5vGFEx
★ 1,677+8Star change over the last 7 days - #57★ 1,624+2Star change over the last 7 days
- #58
The official GitHub page for the survey paper "A Survey on Evaluation of Large Language Models".
★ 1,608-1Star change over the last 7 days - #59
ADR secures enterprise AI agents through observability, security benchmarking, and threat detection. Deployed at Uber.
★ 1,527+21Star change over the last 7 days - #60
:bar_chart: Benchmark multiple object trackers (MOT) in Python
★ 1,487+0Star change over the last 7 days