datasets
Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.
sea-product-attribute-extraction-sample
SEA Multilingual Product Attribute Extraction Sample
This public sample contains 1,000 synthetic, AI-generated marketplace-style
records for product attribute extraction and catalog normalization.
Languages
English
Chinese
Malay
Indonesian
Formats
CSV
JSONL
Intended Use
Use this sample for inspection, evaluation, catalog normalization prototypes,
search-filtering experiments, and multilingual data-quality testing.… See the full description on the dataset page: https://huggingface.co/datasets/nwchang/sea-product-attribute-extraction-sample.AttributeExtractionForMistral7B
