datasets
Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.
chat_structured_extraction
Lead Extraction Dataset
Dataset Description
This dataset contains structured extraction examples for lead information from conversational input in Spanish.
Dataset Structure
Format: JSONL (JSON Lines)
Total Examples: 120
Splits:
Train: 90 examples
Dev: 10 examples
Test: 20 examples
Schema
Each row in the dataset follows the schema defined in schemas/lead_extraction_row_1.0.0.json.
Task
Extract structured lead information from user… See the full description on the dataset page: https://huggingface.co/datasets/mauroibz/chat_structured_extraction.structured-generation-information-extraction-vlms-openbmb-RLAIF-V-Datasetstructured-generation-information-extraction-vlms-openbmb-RLAIF-V-Datasetstructured-generation-information-extraction-vlms-openbmb-RLAIF-V-Datasetstructured-generation-information-extraction-vlms-openbmb-RLAIF-V-DatasetStructured_Information_Extraction
