CoolFace
Datasetpublic

nguyenkhanh87/ViLegalQA-Synthetic-Curation

ViLegalQA Synthetic Curation Dataset summary This repository releases the synthetic Vietnamese legal QA research artifacts produced in the accompanying study. The primary resource contains 10,095 synthetic QA items spanning true/false, multiple-choice, and open-ended tasks. It is accompanied by the final curation/quality annotations used in the study, plus aggregated labels for 600 items from the five-expert human calibration panel. Manuscript: Human-Calibrated… See the full description on the dataset page: https://huggingface.co/datasets/nguyenkhanh87/ViLegalQA-Synthetic-Curation.

sourceHugging Faceupdated 23d agoView on Hugging Face
0likes119downloads
10 commits on main
c84378923d ago

Release v1.2.0 exact Step 09 replay metadata

nguyenkhanh87
3c9642823d ago

Release v1.1.0 statistical replay artifacts

nguyenkhanh87
3f4e1ad23d ago

Refresh release manifest and integrity checksums

nguyenkhanh87
22cf0f923d ago

Standardize dataset configs and add human-calibration sampling metadata

nguyenkhanh87
31d0da324d ago

Add sampling design fields to human calibration

nguyenkhanh87
0e9c10c1mo ago

Update README.md

nguyenkhanh87
a66d5a81mo ago

Release synthetic legal QA dataset and curation annotations

nguyenkhanh87
5b6b39b1mo ago

Release synthetic legal QA dataset and curation annotations

nguyenkhanh87
ee779a01mo ago

Release synthetic legal QA dataset and curation annotations

nguyenkhanh87
28468a71mo ago

initial commit

nguyenkhanh87