datasets
Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.
biographical
Biographical Relation Extraction Dataset
Welcome to the repository of datasets tailored for biographical relation extraction, crafted utilizing Guided Distant Supervision (GDS). Explore datasets available in both English and German, which facilitate extensive research in relation extraction from biographical data. Below you can find an overview of the datasets currently available, as well as the relations that are in each set. Please note there are different sets for each language… See the full description on the dataset page: https://huggingface.co/datasets/plumaj/biographical.biographical
Biographical Dataset for Relation Extraction (RE)
Overview
This dataset is a reconstructed version of the Biographical Dataset, specifically designed for relation extraction (RE) tasks. It serves as a valuable resource for digital humanities (DH) and historical research, enabling the study of relationships within biographical data. The dataset is generated by automatically aligning sentences from Wikipedia articles with structured data sourced from platforms like… See the full description on the dataset page: https://huggingface.co/datasets/Despina/biographical.biographical_knowledge_graph_n200_m3_seed42biographical_knowledge_graph_n20_m3_seed42
Synthetic Knowledge Graph Dataset
Dataset Description
This is a synthetic dataset for knowledge graph relation extraction, generated from a scale-free knowledge graph.
Dataset Summary
Total samples: 70
Total triples: 70
Relations: 14
Supported Tasks
Relation extraction
Knowledge graph construction
Natural language understanding
Dataset Structure
Each sample contains:
sentence: Natural language sentence
subject: Subject entity… See the full description on the dataset page: https://huggingface.co/datasets/r-takahashi/biographical_knowledge_graph_n20_m3_seed42.biographical_knowledge_graph_n100_m3_seed42
Synthetic Knowledge Graph Dataset
Dataset Description
This is a synthetic dataset for knowledge graph relation extraction, generated from a scale-free knowledge graph.
Dataset Summary
Total samples: 263
Total triples: 263
Relations: 14
Supported Tasks
Relation extraction
Knowledge graph construction
Natural language understanding
Dataset Structure
Each sample contains:
sentence: Natural language sentence
subject: Subject entity… See the full description on the dataset page: https://huggingface.co/datasets/r-takahashi/biographical_knowledge_graph_n100_m3_seed42.biographical_knowledge_graph_n50_m3_seed42
Synthetic Knowledge Graph Dataset
Dataset Description
This is a synthetic dataset for knowledge graph relation extraction, generated from a scale-free knowledge graph.
Dataset Summary
Total samples: 160
Total triples: 160
Relations: 14
Supported Tasks
Relation extraction
Knowledge graph construction
Natural language understanding
Dataset Structure
Each sample contains:
sentence: Natural language sentence
subject: Subject entity… See the full description on the dataset page: https://huggingface.co/datasets/r-takahashi/biographical_knowledge_graph_n50_m3_seed42.
