datasets
Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.
hads-physical-commonsense
HADS – Human Action and Decision Sense (حَدس)
📄 Paper: HADS: A Large-Scale Parallel Benchmark for Physical Commonsense Reasoning — IEEE Access, 2026 (doi:10.1109/ACCESS.2026.3705337)
HADS is a large-scale Arabic parallel adaptation of the English
PIQA benchmark for physical commonsense reasoning.
The name derives from the Arabic word حَدس (hads), meaning physical intuition
or gut sense — the tacit embodied knowledge the benchmark measures.
Dataset summary… See the full description on the dataset page: https://huggingface.co/datasets/IWAN/hads-physical-commonsense.shastraai__Shastra-LLAMA2-Math-Commonsense-SFT-details
Dataset Card for Evaluation run of shastraai/Shastra-LLAMA2-Math-Commonsense-SFT
Dataset automatically created during the evaluation run of model shastraai/Shastra-LLAMA2-Math-Commonsense-SFT
The dataset is composed of 38 configuration(s), each one corresponding to one of the evaluated task.
The dataset has been created from 1 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard/shastraai__Shastra-LLAMA2-Math-Commonsense-SFT-details.
