Collection of Suffix Array Construction Algorithms (SACAs)
Quick and easy resource usage monitoring and benchmarking for any command's CPU, memory, disk usage and runtime.
The official python toolkit for running experiments and evaluate performance on VideoCube benchmark @TPAMI2023
Generate markdown comparison tables from `cargo-criterion` JSON output
I/O benchmark for different image processing python libraries.
BeHonest: Benchmarking Honesty in Large Language Models
Critical difference diagrams with Python and Tikz
Lua <-> C++ bindings libraries benchmark
:hammer: :wrench: Test Driven Development :repeat: with Golang :hamster:
The Stanford Word Substitution (Swords) Benchmark
ansible-vault CLI reimplemented in go
An unified framework of quality enhancement approaches for compressed images based on PyTorch.
Just a small test to see which language is better for extending python when using lists of lists
[IROS2021] NYU-VPR: Long-Term Visual Place Recognition Benchmark with View Direction and Data Anonymization Influences
Heterogenous, Task- and Domain-Specific Benchmark for Unsupervised Sentence Embeddings used in the TSDAE paper: https://arxiv.org/abs/2104.06979.
A Comprehensive and Versatile Open-Source Federated Learning Framework
[NeurIPS DBT 2021] HPO-B
LLM Divergent Thinking Creativity Benchmark. LLMs generate 25 unique words that start with a given letter with no connections to each other or to 50 i...
Official Repo for the paper: VCR: Visual Caption Restoration. Check arxiv.org/pdf/2406.06462 for details.
Longitudinal Evaluation of LLMs via Data Compression
A benchmark for standalone WebAssembly
Code and data for "Living in the Moment: Can Large Language Models Grasp Co-Temporal Reasoning?" (ACL 2024)
C# ECS Benchmarks
Spurious Features Everywhere - Large-Scale Detection of Harmful Spurious Features in ImageNet
🔥 Collection of useful javascript snippets with automated benchmarks
R package for benchmarking single cell analysis methods
The dataset and source code for our paper: "Did You Ask a Good Question? A Cross-Domain Question IntentionClassification Benchmark for Text-to-SQL"
Arline Benchmarks platform allows to benchmark various algorithms for quantum circuit mapping/compression against each other on a list of predefined h...
Linköping GraphQL Benchmark (LinGBM)
Benchee (Elixir benchmarking) integration for Livebook
A modular benchmarking library with V8 warmup and cpu/ram denoising for the most accurate and consistent results.
[NAACL 2025 🔥] CAMEL-Bench is an Arabic benchmark for evaluating multimodal models across eight domains with 29,000 questions.
[ACL 2024 Main] NewsBench: A Systematic Evaluation Framework for Assessing Editorial Capabilities of Large Language Models in Chinese Journalism
phpbenchmarks.com kit to add your benchmark.
corebench - run your benchmarks against high performance computing servers with many CPU cores
A C++ Thread Pool Colosseum
A benchmark library for Dynamic Algorithm Configuration.
PHP User Agent Parser Benchmarks
a benchmark for compile-time and/or runtime Nim 🏆
Tool to easily benchmark QML/QtQuick (or your own QML components) performance on different hardware.
The code accompaniment for the CoRL 2020 paper: A User's Guide to Calibrating Robotics Simulators (https://arxiv.org/abs/2011.08985), from NVIDIA Rese...
📚Examples
SERAB: a multi-lingual benchmark for speech emotion recognition
Benchmarks comparing ESNext features to their ES5 and various pre-processor equivalents
glxgears demo/benchmark for windows - compiled in Visual Studio 2017
Micro-Runner, a CLI playground for benchmarking your JavaScript code
Jakarta EE 5/6/7/8/10 WildFly/JBoss EAP Clustering Benchmark Application
JavaScript test-runners benchmark
A benchmark program for dgraph.