shimbabomb
shimbabomb-ai-benchmark
ShimbaBomb AI Benchmark
A GSM8K-style benchmark dataset for evaluating AI models on ShimbaBomb — an English-like scripting language that compiles to native C.
Overview
This dataset contains 81 problems with chain-of-thought reasoning for training and evaluating AI models on ShimbaBomb code generation, understanding, debugging, and explanation.
Each problem has:
question: A natural language description of a programming task or question about SB code
answer:… See the full description on the dataset page: https://huggingface.co/datasets/shimbaaa/shimbabomb-ai-benchmark.shimbabomb-benchmark
ShimbaBomb Interpreter Extreme Stress Benchmark
Overview
Benchmark results from extreme stress testing of the ShimbaBomb (SB) v1.11.0 interpreter — an English-like scripting language that compiles to native C.
The benchmark suite spawns CPU_Logical_Cores * 2 (or higher) threads to saturate the interpreter engine, running diverse SB scripts simultaneously for 30-second sustained windows per phase.
System
Parameter
Value
Platform
Windows… See the full description on the dataset page: https://huggingface.co/datasets/shimbaaa/shimbabomb-benchmark.
