speed/JDocQA
This unofficial dataset consists of QA pairs with images converted from the PDF files of JDocQA, a dataset focusing on chart and table understanding. The conversion was performed using pdf2image. The original dataset includes 1,176 examples, but 12 examples could not be converted into images. As a result, this image dataset consists of 1,164 examples in total. We are uploading it here for use in the evaluation of llm-jp-eval-mm. Please see the official github repo… See the full description on the dataset page: https://huggingface.co/datasets/speed/JDocQA.
0536
