bashkit
Benches

Latest benchmark snapshot

Static aggregate generated from repository result artifacts. Use the linked files for raw measurements and full eval traces.

Latest reports

Open Markdown reports

Runtime snapshot

Latest benchmark categories

Browse benchmark runs
CategoryCasesLast run
stringsString expansion, pattern handling, and text manipulation.80.018 msbash median: 7.36 ms
arithmeticInteger math, substitutions, and expression-heavy shell snippets.60.023 msbash median: 5.213 ms
startupSmall commands where interpreter startup dominates runtime.40.027 msbash median: 4.411 ms
pipesPipeline construction, streaming, and command chaining.60.03 msbash median: 9.414 ms
variablesVariable assignment, lookup, expansion, and environment handling.80.032 msbash median: 6.57 ms
controlConditionals, loops, case statements, and branching scripts.90.042 msbash median: 5.255 ms
toolsBuiltin and external-tool style command workloads.210.042 msbash median: 7.603 ms
ioFile reads, writes, redirects, and filesystem-facing commands.60.042 msbash median: 4.396 ms
Eval pressure

Lowest eval categories

Browse eval runs
Latest LLM eval93%54/58 tasks
CategoryPassedPass rate
system_info1/250%tasks passed
file_operations3/466.7%tasks passed
scripting5/768.6%tasks passed
json_processing8/8100%tasks passed
data_transformation6/6100%tasks passed
complex_tasks6/6100%tasks passed
text_processing6/6100%tasks passed
pipelines5/5100%tasks passed