datasets
Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.
Multi_News_fact_checking_claims
Dataset Card for "v2"
More Information needed
task966_ruletaker_fact_checking_based_on_given_context
Dataset Card for Natural Instructions (https://github.com/allenai/natural-instructions) Task: task966_ruletaker_fact_checking_based_on_given_context
Additional Information
Citation Information
The following paper introduces the corpus in detail. If you use the corpus in published work, please cite it:
@misc{wang2022supernaturalinstructionsgeneralizationdeclarativeinstructions,
title={Super-NaturalInstructions: Generalization via Declarative Instructions on… See the full description on the dataset page: https://huggingface.co/datasets/Lots-of-LoRAs/task966_ruletaker_fact_checking_based_on_given_context.checking
HumaniBench: A Human-Centric Visual QA Dataset
HumaniBench is a dataset for evaluating visual question answering models on tasks that involve human-centered attributes such as gender, age, and occupation.
Each data point includes:
ID: Unique identifier
Attribute: A social attribute (e.g., gender, race)
Question: A visual question related to the image
Answer: The ground-truth answer
image: Embedded image in base64 or file format for visual preview
Example Entry
{… See the full description on the dataset page: https://huggingface.co/datasets/shainaraza/checking.vietnamese-fact-checking-verifier-data
Vietnamese fact-checking verifier data
Leakage-aware document-level 80/10/10 split derived from
Loctran123/vietnamese-fact-checking-claims at revision 63963b973af864ecc20313636b1786c31bbb4a41.
Input is (evidence_text, claim) and labels are SUPPORTED, REFUTED, and
NOT_ENOUGH_INFO. Exact duplicate claims are retained only once.
portuguese-fact-checking
Portuguese Automated Fact-Checking
Fake.BR
COVID19.BR
MuMiN-PT
Info (fake/true)
🖥️
💬
X
Domain
General
Health
"General" (Health)
Year
2016–2018
2020
2020–2022
Approach [1]
bottom-up
bottom-up
top-down
Size
3580/3580
848/1139
1339/65
% URL
1.0%/0.7%
28.9%/56.9%
0.3%/0.0%
Avg. # words
181.4/183.1
167.7/111.1
18.9/16.9
Corpora characteristics after cleaning. Top-down starts with fact-checked claims; bottom-up seeks for new misinformation in posts.… See the full description on the dataset page: https://huggingface.co/datasets/ju-resplande/portuguese-fact-checking.vietnamese-fact-checking-verifier-data-v3-1
Vietnamese fact-checking verifier data
Leakage-aware document-level 80/10/10 split derived from
aiMy144/vietnamese-fact-checking-claims-v3-1 at revision e20c1eddcbb4a5de862dbc6965eabee43c9766da.
Input is (evidence_text, claim) and labels are SUPPORTED, REFUTED, and
NOT_ENOUGH_INFO. Exact duplicate claims are retained only once.
condition-checking-dataset
Condition Checking Dataset
This dataset contains condition checking conversations for robotics applications, with embedded base64 images from multiple camera viewpoints.
Dataset Structure
Data Fields
id: Unique identifier for each sample
images: Dictionary containing base64-encoded images from multiple camera viewpoints
conversations: List of conversation turns (human question + assistant answer)
Camera Viewpoints
The dataset includes images from 5… See the full description on the dataset page: https://huggingface.co/datasets/jeffshen4011/condition-checking-dataset.checking-test
Dataset Card for "checking-test"
More Information needed
rlvr_task966_ruletaker_fact_checking_based_on_given_contextflan_combined_task966_ruletaker_fact_checking_based_on_given_contextultra-feedback_checkingqwen3_0.6b-rlvr_task966_ruletaker_fact_checking_based_on_given_contextcheckingcuda-stack-casssubset-checkingdZhabDtALkOxnJYluQUrCBvRrGd2_checking_bugchecking_metric_trackingD-EVAL__standard_eval_v3__checking_evals-eval_0
