AmazonScience/document-haystack
Document Haystack Dataset This repository contains the dataset for the paper “Document Haystack: A Long Context Multimodal Image/Document Understanding Vision LLM Benchmark”. 📑 Abstract Paper The proliferation of multimodal Large Language Models has significantly advanced the ability to analyze and understand complex data inputs from different modalities. However, the processing of long documents remains under-explored, largely due to a lack of suitable… See the full description on the dataset page: https://huggingface.co/datasets/AmazonScience/document-haystack.
2090k
1The secret office supply is a "pencil".2The secret animal #4 is a "snake".3The secret animal #2 is a "zebra".4The secret sport is "basketball".5The secret animal #3 is a "dolphin".6The secret kitchen appliance is a "blender".7The secret clothing is a "t-shirt".8The secret shape is a "circle".9The secret object #1 is a "book".10The secret food is a "pizza".11The secret instrument is a "guitar".12The secret animal #1 is a "dog".13The secret tool is a "hammer".14The secret animal #5 is a "pig".15The secret drink is "coffee".16The secret transportation is a "car".17The secret landmark is the "Eiffel Tower".18The secret object #5 is a "comb".19The secret flower is a "rose".20The secret currency is a "euro".21The secret object #2 is a "lamp".22The secret object #3 is a "spoon".23The secret vegetable is a "carrot".24The secret object #4 is an "umbrella".25The secret fruit is an "apple".26 