nguyenkhanh87/ViLegalQA-Synthetic-Curation
ViLegalQA Synthetic Curation Dataset summary This repository releases the synthetic Vietnamese legal QA research artifacts produced in the accompanying study. The primary resource contains 10,095 synthetic QA items spanning true/false, multiple-choice, and open-ended tasks. It is accompanied by the final curation/quality annotations used in the study, plus aggregated labels for 600 items from the five-expert human calibration panel. Manuscript: Human-Calibrated… See the full description on the dataset page: https://huggingface.co/datasets/nguyenkhanh87/ViLegalQA-Synthetic-Curation.
Release v1.2.0 exact Step 09 replay metadata
Release v1.1.0 statistical replay artifacts
Refresh release manifest and integrity checksums
Standardize dataset configs and add human-calibration sampling metadata
Add sampling design fields to human calibration
Update README.md
Release synthetic legal QA dataset and curation annotations
Release synthetic legal QA dataset and curation annotations
Release synthetic legal QA dataset and curation annotations
initial commit
