nihedb/EUR-Lex-Triples
EUR-Lex-Triples: A Legal Relation Extraction Dataset from European Legislation EUR-Lex-Sum dataset Aumiller, 2022 annotated with triples Relation Extraction Baselines Code/RE-Baselines contains the code used to run the RE baselines : Fine-Tuning and Inference. Results of baseline models for Relation Extraction are : Model Precision Recall F1-Score Legal-Bert 0.64 0.59 0.60 Bert 0.58 0.52 0.54 Rebel-Large 0.88 0.75 0.80 Mistral 7b zero-Shot 0.38… See the full description on the dataset page: https://huggingface.co/datasets/nihedb/EUR-Lex-Triples.
EUR-Lex-Triples: A Legal Relation Extraction Dataset from European Legislation
EUR-Lex-Sum dataset Aumiller, 2022 annotated with triples
Dataset Description
- EUR-Lex-Triples consists on 1504 annotated documents. All Documents come from the english part of EUR-Lex-Sum Dataset.
- ``
Filtered_Annotated_Documents`` contains json files containing for each document its summary, the annotated paragraphs, and for each paragraph the triples derived from the annotations.
Relation Extraction Baselines
- ``
Code/RE-Baselines`` contains the code used to run the RE baselines : Fine-Tuning and Inference. - Results of baseline models for Relation Extraction are :
Citation
EUR-Lex-Triples: A Legal Relation Extraction Dataset from European Legislation. Paper accepted at TPDL 2025.
Licence
Copyright for the editorial content of EUR-Lex website, the summaries of EU legislation and the consolidated texts owned by the EU, are licensed under the Creative Commons Attribution 4.0 International licence, i.e., CC BY 4.0 as mentioned on the official EUR-Lex website. Any data artifacts remain licensed under the CC BY 4.0 license.
