CoolFace
20 results

rlm

RL-MIND /XHRBench XHRBench Ultra-High-Resolution Remote Sensing Understanding and Reasoning 🤗 Hugging Face · 🤖 ModelScope · 📄 Paper · 💻 Code English | 中文 📚 Introduction XHRBench evaluates fine-grained perception and complex reasoning in multimodal large language models using ultra-high-resolution remote-sensing imagery. This repository retains the name XHRBench and belongs to the same RSHR benchmark project as RSHR-Bench, with a… See the full description on the dataset page: https://huggingface.co/datasets/RL-MIND/XHRBench.imageimage-text-to-text1K<n<10K7 likes6.6k downloads9d agoHugging FaceRL-MIND /HARD-VQA HARD-VQA Ultra-High-Resolution Aerial VQA with Original Images Embedded per Question 🤗 Hugging Face · 🟣 ModelScope · 📊 Statistics English | 中文:Hugging Face · ModelScope 📚 Introduction HARD-VQA packages ultra-high-resolution aerial visual question answering data as self-contained Parquet shards. Each row is one multiple-choice question, with all required original JPEG bytes embedded in its ordered images… See the full description on the dataset page: https://huggingface.co/datasets/RL-MIND/HARD-VQA.imagevisual-question-answering1K<n<10K0 likes926 downloads9d agoHugging FaceRL-MIND /RSHR-Bench RSHR-Bench Ultra-High-Resolution Remote Sensing Understanding and Reasoning 🤗 Hugging Face · 🤖 ModelScope · 📄 Paper · 💻 Code English | 中文 📚 Introduction RSHR-Bench evaluates ultra-high-resolution remote-sensing visual understanding and reasoning in multimodal large language models across single-image, multi-image, and multi-turn settings. This release embeds original-resolution images directly in Parquet shards for use… See the full description on the dataset page: https://huggingface.co/datasets/RL-MIND/RSHR-Bench.imageimage-text-to-text1K<n<10K5 likes762 downloads9d agoHugging FaceRL-MIND /NJU-HARD-Tracking NJU-HARD-Tracking Multi-Object Tracking across 122 MP UAV Image Sequences 🤗 Hugging Face · 🟣 ModelScope · 📊 Statistics: HF / MS English | 中文: Hugging Face · ModelScope 🌍 Overview NJU-HARD-Tracking provides the multi-object-tracking release of HARD, with full-resolution frames, temporal ordering, and the original instance annotations. It supports studying how detection and association behave across wide-area aerial… See the full description on the dataset page: https://huggingface.co/datasets/RL-MIND/NJU-HARD-Tracking.imageother1K<n<10K0 likes684 downloads7d agoHugging FaceRL-MIND /NJU-HARD-Detection NJU-HARD-Detection Object Detection in 122 MP Wide-Area UAV Imagery 🤗 Hugging Face · 🟣 ModelScope · 📊 Statistics: HF / MS English | 中文: Hugging Face · ModelScope 🌍 Overview NJU-HARD-Detection provides the object-detection release of HARD, with full-resolution frames and per-frame pedestrian and vehicle boxes. Detection training and evaluation use the categories and boxes; the retained source identity fields are… See the full description on the dataset page: https://huggingface.co/datasets/RL-MIND/NJU-HARD-Detection.imageobject-detection1K<n<10K0 likes682 downloads7d agoHugging FaceRL-MIND /NJU-HARD NJU-HARD 🤗 Hugging Face · 🟣 ModelScope · 📊 Statistics English | 中文:Hugging Face · ModelScope 📚 Introduction NJU-HARD is the deduplicated, full-resolution release of the HARD visual question answering data. It contains 1,563 valid VQA records across 8 task types, using 854 original aerial images. Only images referenced by these valid questions are included. Every original JPEG is embedded once in a native… See the full description on the dataset page: https://huggingface.co/datasets/RL-MIND/NJU-HARD.imagevisual-question-answering1K<n<10K0 likes570 downloads9d agoHugging Face