gplsi/uji_parallel_va_en
UJI_PARALLEL_VA_EN Dataset Dataset Summary UJI_PARALLEL_VA_EN is a parallel dataset for machine translation between Valencian (VA) and English (EN).It consists of aligned sentence pairs along with the source file from which each pair was extracted.The dataset is intended for research in machine translation, cross-lingual NLP, and linguistic analysis. Dataset Structure Each row in the dataset includes the following fields: VA: A sentence in… See the full description on the dataset page: https://huggingface.co/datasets/gplsi/uji_parallel_va_en.
Update README.md
Update README.md
Delete train
Upload data/train/va-en-sentence.length.ner.curated.jsonl with huggingface_hub
Upload data/train/va-en-paragraph.length.curated.jsonl with huggingface_hub
Upload data/train/va-en-documents.curated.jsonl with huggingface_hub
Update README.md
Update README.md
Delete va-en-paragraphs.jsonl
Delete train.jsonl
Upload 2 files
Upload va-en-paragraphs.jsonl
Upload train.jsonl
Update README.md
initial commit
