datasets
Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.
Pattern-Recognition
Pattern Completion Dataset
A 30 GB synthetic dataset of numeric sequence‑completion prompts and their next values, designed to teach large language models how to recognize and extrapolate patterns.
Each row contains a prompt (the sequence with a ? indicating the missing next element) and a completion (the correct next number).
Dataset Structure
Format: CSV (no header row)
Columns:
prompt – "Find the next number in the sequence: a,b,c,... ,?"
completion – the… See the full description on the dataset page: https://huggingface.co/datasets/Corpus-NZ/Pattern-Recognition.Pattern-Recognition
Pattern Completion Dataset
A 30 GB synthetic dataset of numeric sequence‑completion prompts and their next values, designed to teach large language models how to recognize and extrapolate patterns.
Each row contains a prompt (the sequence with a ? indicating the missing next element) and a completion (the correct next number).
Dataset Structure
Format: CSV (no header row)
Columns:
prompt – "Find the next number in the sequence: a,b,c,... ,?"
completion – the… See the full description on the dataset page: https://huggingface.co/datasets/Gugu8/Pattern-Recognition.law_entity_recognition
Dataset Card for Dataset Name
The dataset transforms complex legal passages into structured outputs, detailing entities, their interrelationships, and claims, providing a foundation for a legal knowledge graph to facilitate advanced analysis and applications.
Dataset Details
Dataset Description
The dataset in question is a specialized collection designed for legal text analysis, where each input is a passage of legal text—ranging from case law to statutory… See the full description on the dataset page: https://huggingface.co/datasets/rubenamtz0/law_entity_recognition.
