datasets
Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.
MathVerse-lmmseval
Dataset Card for MathVerse
This is the version for lmms-eval. This shares the same data with the official dataset.
Dataset Description
Paper Information
Dataset Examples
Leaderboard
Citation
Dataset Description
The capabilities of Multi-modal Large Language Models (MLLMs) in visual math problem-solving remain insufficiently evaluated and understood. We investigate current benchmarks to incorporate excessive visual content within textual questions, which potentially… See the full description on the dataset page: https://huggingface.co/datasets/CaraJ/MathVerse-lmmseval.MMSearch
MMSearch 🔥: Benchmarking the Potential of Large Models as Multi-modal Search Engines
Official repository for the paper "MMSearch: Benchmarking the Potential of Large Models as Multi-modal Search Engines".
🌟 For more details, please refer to the project page with dataset exploration and visualization tools: https://mmsearch.github.io/.
[🌐 Webpage] [📖 Paper] [🤗 Huggingface Dataset] [🏆 Leaderboard] [🔍 Visualization]
💥 News
[2024.09.25] 🌟 The evaluation code now… See the full description on the dataset page: https://huggingface.co/datasets/CaraJ/MMSearch.CapSpeech-MCQ
CapSpeech-MCQ Dataset
Dataset Description
This dataset contains multiple-choice questions (MCQs) and detail questions generated from the CapTTS-SFT dataset. The questions are designed for training and evaluating models on speech-related attributes and caption understanding.
Dataset Structure
Splits
{chr(10).join([f'- {split.title()}: {split.title()} data' for split in available_splits])}
Total Rows: {total_rows:,}
Question Types… See the full description on the dataset page: https://huggingface.co/datasets/carankt/CapSpeech-MCQ.
