datasets
Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.
Akka_Finetuning_Llama3.2T2G-1k-Llama3.2-3B
T2G
Overview
T2G is a synthetic data consisting of text-graph pairs designed to finetune LLMs on information extraction tasks, specifically text-to-graph conversion.
Dataset Structure
The dataset is organized into the following main components:
Train Set: 800 instances for training models.
Validation Set: 100 for validating model performance.
Test Set: 100 instances for final evaluation.
Data Fields
Each instance in the dataset contains the… See the full description on the dataset page: https://huggingface.co/datasets/ESITime/T2G-1k-Llama3.2-3B.llama3.2training_datasetllama-3.2-train-validategfr_rules_formatted_llama3.2_datasetLlama-3.2_CQAllama-3.2-comparison-resultssample_form_data_for_llama3.2_3bllama3.2-3b-instructllama3.2_plutoedu
