datasets
Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.
inkslop-results
InkSlop Benchmark Results
Model evaluation results for the InkSlop Benchmark - a vibe-coded benchmark for spatial reasoning with digital ink.
Collection: InkSlop Benchmark
Contents
This dataset contains inference results and evaluation metrics for multiple VLMs across all InkSlop tasks:
overlap_easy / overlap_hard - Overlapped handwriting recognition
autocomplete_easy / autocomplete_hard - Handwriting autocompletion
derender_easy / derender_hard - Ink derendering (image… See the full description on the dataset page: https://huggingface.co/datasets/amaksay/inkslop-results.inkslop-autocomplete-hard
InkSlop Autocomplete Hard
Part of the InkSlop Benchmark a vibe-coded benchmark for spatial reasoning with digital ink.
Collection: InkSlop Benchmark
Task
Handwriting Autocompletion: Given a partial handwritten input, generate the completion as digital ink. This "hard" variant contains human-collected handwriting samples.
Data Format
This dataset contains two top-level directories:
original/ # Raw collected data
└── samples/
└──… See the full description on the dataset page: https://huggingface.co/datasets/amaksay/inkslop-autocomplete-hard.
