datasets
Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.
local-code-arena-mbpp-deepseek-coder_6.7b
Local Code Arena Telemetry: MBPP Benchmark on DeepSeek Coder 6.7B
This repository hosts the raw evaluation metrics, execution telemetry logs, and structural syntax outputs captured from running the Mostly Basic Python Problems (MBPP) benchmark against the DeepSeek Coder 6.7B parameter model.
This specific run catalogs mid-tier parameter dynamics for legacy code specialists, providing an anchor point to evaluate generational alignment improvements in newer architectures.… See the full description on the dataset page: https://huggingface.co/datasets/ShahzebKhoso/local-code-arena-mbpp-deepseek-coder_6.7b.local-code-arena-mbpp-deepseek-coder_1.3b
Local Code Arena Telemetry: MBPP Benchmark on DeepSeek Coder 1.3B
This repository hosts the raw evaluation metrics, execution telemetry logs, and structural syntax outputs captured from running the Mostly Basic Python Problems (MBPP) benchmark against the ultra-lightweight DeepSeek Coder 1.3B model.
This specific run establishes the absolute maximum throughput envelope of our local hardware setup while tracking the accuracy trade-offs of legacy, lightweight code specialists.… See the full description on the dataset page: https://huggingface.co/datasets/ShahzebKhoso/local-code-arena-mbpp-deepseek-coder_1.3b.
