histogram
visual_qa_histograms
What Lies Beneath: A Call for Distribution-based Visual Question & Answer Datasets
Publication: JCDL 2025 Website and on arXiv
GitHub Repo: ReadingTimeMachine/LLM_VQA_JCDL2025
This is a histogram-based dataset for visual question and answer (VQA) with humans and large language/multimodal models (LMMs).
Data contains synthetically generated single-panel histograms images, data used to create histograms, bounding box data for titles, axis and tick labels, and… See the full description on the dataset page: https://huggingface.co/datasets/ReadingTimeMachine/visual_qa_histograms.Amp-Phase-2D-Histogram-DatasetThis dataset contains a comprehensive collection of 2D amplitude-phase histograms derived from raw IEEE 802.11 (Wi-Fi) IQ signal captures. It is intended to support research in Automatic Modulation Classification (AMC) under realistic and challenging indoor wireless propagation environments.
The dataset is particularly useful for:
modulation recognition and classification
machine learning and deep learning for RF sensing
channel robustness studies
SNR-aware signal analysis
indoor Wi-Fi signal… See the full description on the dataset page: https://huggingface.co/datasets/Raluca13/Amp-Phase-2D-Histogram-Dataset.histogram-comparisons-small-v1This is a small subset of the huge histogram-comparisons-v1 dataset with 3M rows.
This dataset contains 150000 items in total. There are 3 curriculums each containing 50000 items.
Each item is a markdown document.
Each item contains between 2 and 6 image comparisons, with a Summary at the bottom.
The images are between 3x3 and 14x14.
The markdown document contains a ## Response, that separates the prompt from the answer.
The structure of the markdown document with 3 comparisons: A, B, C.
#… See the full description on the dataset page: https://huggingface.co/datasets/neoneye/histogram-comparisons-small-v1.histogram-comparisons-v1If you want a small subset of this dataset, there is histogram-comparisons-small-v1 with 150k rows.
This dataset contains 3000000 items in total. There are 3 curriculums each containing 1000000 items.
Each item is a markdown document.
Each item contains between 2 and 6 image comparisons, with a Summary at the bottom.
The images are between 3x3 and 14x14.
The markdown document contains a ## Response, that separates the prompt from the answer.
The structure of the markdown document with 3… See the full description on the dataset page: https://huggingface.co/datasets/neoneye/histogram-comparisons-v1.simon-arc-histogram-v9
Version 1
The counters are in the range 1-20.
Version 2
The counters are in the range 1-50.
Version 3
The counters are in the range 1-100.
Version 4
The counters are in the range 1-200.
Histogram.remove_other_colors() added.
Version 5
I forgot to update the range of the counters when doing comparisons.
Now the counters are in the range 1-100.
Version 6
The counters are in the range 1-200.
Version 7
The counters are in… See the full description on the dataset page: https://huggingface.co/datasets/neoneye/simon-arc-histogram-v9.simon-arc-histogram-v8
Version 1
The counters are in the range 1-20.
Version 2
The counters are in the range 1-50.
Version 3
The counters are in the range 1-100.
Version 4
The counters are in the range 1-200.
Histogram.remove_other_colors() added.
Version 5
I forgot to update the range of the counters when doing comparisons.
Now the counters are in the range 1-100.
Version 6
The counters are in the range 1-200.
Version 7
The counters are in… See the full description on the dataset page: https://huggingface.co/datasets/neoneye/simon-arc-histogram-v8.
