datasets
Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.
Phi4-Mini-P2T-4B-TestingTesting Results for USS-Inferprise/Phi4-Mini-Prose2Tags-4B (https://huggingface.co/USS-Inferprise/Phi4-Mini-Prose2Tags-4B)
local-code-arena-mbpp-phi4-mini
Local Code Arena Telemetry: MBPP Benchmark on Phi-4 Mini
This repository hosts the raw evaluation metrics, execution telemetry logs, and structural syntax outputs captured from running the Mostly Basic Python Problems (MBPP) benchmark against Microsoft's Phi-4 Mini (3.8B) model.
This run establishes a vital cross-vendor reference point, documenting how high-density synthetic reasoning filtration scales relative to dedicated code-only specialists on local consumer hardware.… See the full description on the dataset page: https://huggingface.co/datasets/ShahzebKhoso/local-code-arena-mbpp-phi4-mini.
