BSC-LT/Legal_Catalan_Spanish_Parallel_Corpus
Dataset Card for Legal Catalan-Spanish Parallel Corpus Dataset Summary The Legal Catalan-Spanish Parallel Corpus is a multilingual dataset of authentic parallel text in Catalan and Spanish from the legal domain, drawn from official Catalan public institution documents. It comprises three sub-corpora organized at different textual granularities: sentence, paragraph, and document level. This multi-level structure enables research and training that goes beyond… See the full description on the dataset page: https://huggingface.co/datasets/BSC-LT/Legal_Catalan_Spanish_Parallel_Corpus.
Conversations for this repository live on Hugging Face.
CoolFace shows imported repositories read-only. Posting into someone else’s repository from here would need an authorised integration and the account holder’s consent, so the link goes to the source instead.
Open discussions on Hugging Face