datasets
Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.
lm-eval-results-shyamieee-Padma-SLM-7b-v1.0-private
Dataset Card for Evaluation run of shyamieee/Padma-SLM-7b-v1.0
Dataset automatically created during the evaluation run of model shyamieee/Padma-SLM-7b-v1.0
The dataset is composed of 62 configuration(s), each one corresponding to one of the evaluated task.
The dataset has been created from 2 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always pointing to the latest results.
An… See the full description on the dataset page: https://huggingface.co/datasets/nyu-dice-lab/lm-eval-results-shyamieee-Padma-SLM-7b-v1.0-private.lm-eval-results-shyamieee-Padma-SLM-7b-v3.0-private
Dataset Card for Evaluation run of shyamieee/Padma-SLM-7b-v3.0
Dataset automatically created during the evaluation run of model shyamieee/Padma-SLM-7b-v3.0
The dataset is composed of 62 configuration(s), each one corresponding to one of the evaluated task.
The dataset has been created from 2 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always pointing to the latest results.
An… See the full description on the dataset page: https://huggingface.co/datasets/nyu-dice-lab/lm-eval-results-shyamieee-Padma-SLM-7b-v3.0-private.pad-auto-solver-reviewed
PAD Reviewed Dataset
Canonical reviewed PAD board/orb artifacts for dw-indie/pad-auto-solver-reviewed. This repository
contains immutable reviewed package revisions and does not contain raw captures,
training runs, checkpoints, or model binaries.
Packages exported: 28
Active catalog datasets: 14
Catalog schema: 3
Layout
packages/<dataset_id>.tar: deterministic self-contained reviewed package
catalog.json: active revision heads and coverage summary… See the full description on the dataset page: https://huggingface.co/datasets/dw-indie/pad-auto-solver-reviewed.nfqa-multilingual-dataset
NFQA Multilingual Dataset
A large-scale multilingual dataset for Non-Factoid Question Answering (NFQA) classification, covering 49 languages and 8 question categories.
Dataset Statistics
Split
Examples
Train
28,653
Validation
3,539
Test
3,671
Total (Balanced)
35,863
Full Dataset (High Quality)
63,647
Dataset Composition
Languages (49 total)
Arabic (ar), Azerbaijani (az), Bulgarian (bg), Bengali (bn), Catalan (ca)… See the full description on the dataset page: https://huggingface.co/datasets/PaDaS-Lab/nfqa-multilingual-dataset.PADBen
PADBen: Paraphrase and AI-Generated Text Detection Benchmark
📊 Dataset Overview
PADBen is a comprehensive benchmark for evaluating AI-generated text detection methods, specifically designed to test detection capabilities across various paraphrasing scenarios and attack vectors. For detailed implementation of how this dataset is generated/curated, please see https://github.com/JonathanZha47/PadBen-Paraphrase-Attack-Benchmark.
Total Dataset Size: 486,990 samples across 46… See the full description on the dataset page: https://huggingface.co/datasets/JonathanZha/PADBen.swemera-10tasks-pyconfHere’s your Markdown text, organized for clear readability:
Task Description
Instances
Instance ID
Short Title
reframe-0
Performance threshold goes to -inf when it should be zero.
pyflakes-1
Walrus operator + annotation can cause F821
sqlglot-2
MySQL dialect fails to parse PRIMARY KEY USING BTREE syntax
matchms-3
matchms fails when reading spectra where abundance is in scientific notation #809
guarddog-4
Add Mach-O magic bytes to bundled binary detector… See the full description on the dataset page: https://huggingface.co/datasets/padamenko/swemera-10tasks-pyconf.PADBen-Task1
PADBen Task 1: Paraphrase Source Attribution (Binary Classification)
📋 Dataset Summary
PADBen Task 1 is a binary classification dataset for distinguishing between human-authored and LLM-generated paraphrases. This task evaluates whether AI detectors can identify the source of paraphrased text without additional context.
Key Features
Task Type: Binary text classification
Total Samples: 16,233 sentences
Train Split: 12,986 samples (80%)
Test Split: 3,247… See the full description on the dataset page: https://huggingface.co/datasets/JonathanZha/PADBen-Task1.phyworld-data-pad_featuresshyamieee__Padma-v7.0-details
Dataset Card for Evaluation run of shyamieee/Padma-v7.0
Dataset automatically created during the evaluation run of model shyamieee/Padma-v7.0
The dataset is composed of 38 configuration(s), each one corresponding to one of the evaluated task.
The dataset has been created from 1 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always pointing to the latest results.
An additional… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard/shyamieee__Padma-v7.0-details.pad-pid-geospatial
Geospatial data-use annotations in World Bank PADs and PIDs
Focus-review annotations joined to Project Appraisal Documents (PADs) and Project Information Documents (PIDs) by canonical document id. This export preserves the complete reviewed passage rows and all focus labels—not only spatially linked mentions—alongside document/project metadata and origin fields.
What's included
3,804 passage records across 1,591 documents; exact metadata types are PAD/PID.
6,503… See the full description on the dataset page: https://huggingface.co/datasets/rafmacalaba/pad-pid-geospatial.
