bashkit
Benches

Latest benchmark snapshot

Static aggregate generated from repository result artifacts. Use the linked files for raw measurements and full eval traces.

Latest reports

Open Markdown reports

Runtime snapshot

Latest benchmark categories

Browse benchmark runs
CategoryCasesLast run
startupSmall commands where interpreter startup dominates runtime.40.017 msbash median: 5.408 ms
arraysIndexed array reads, writes, expansion, and iteration.60.019 msbash median: 6.865 ms
stringsString expansion, pattern handling, and text manipulation.80.02 msbash median: 6.807 ms
pipesPipeline construction, streaming, and command chaining.60.021 msbash median: 7.346 ms
controlConditionals, loops, case statements, and branching scripts.90.026 msbash median: 8.239 ms
subshellCommand substitution and nested shell execution paths.60.028 msbash median: 7.704 ms
ioFile reads, writes, redirects, and filesystem-facing commands.60.036 msbash median: 9.014 ms
arithmeticInteger math, substitutions, and expression-heavy shell snippets.60.049 msbash median: 8.676 ms
Eval pressure

Lowest eval categories

Browse eval runs
Latest LLM eval93%54/58 tasks
CategoryPassedPass rate
system_info1/250%tasks passed
file_operations3/466.7%tasks passed
scripting5/768.6%tasks passed
json_processing8/8100%tasks passed
data_transformation6/6100%tasks passed
complex_tasks6/6100%tasks passed
text_processing6/6100%tasks passed
pipelines5/5100%tasks passed