datasets
Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.
LRMovieNetThis is the LRMovieNet dataset proposed by ECCV 2024 Paper "Multimodal Label Relevance Ranking via Reinforcement Learning". The code is available at https://github.com/ChazzyGordon/LR2PPO.
Please go to Files and versions to download the LRMovieNet dataset.
We select 3,206 clips from 219 videos in the MovieNet dataset.
For each movie clip, we extract frames from the video and input them into the RAM model to obtain image labels.
Concurrently, we input the descriptions of each movie clip into… See the full description on the dataset page: https://huggingface.co/datasets/ChazzyGordon/LRMovieNet.LRM-Safety-Study
LRM-Safety-Study Dataset
Dataset Description
This dataset was used for training in the work How Should We Enhance the Safety of Large Reasoning Models: An Empirical Study. It includes data for both safety-related and mathematical reasoning tasks.
Dataset Details
Data Splits:
MATH: 4,000 mathematical reasoning examples.
Default CoT: 1,000 safety-related examples using the default CoT prompting.
RealSafe CoT: 1,000 safety-related examples with… See the full description on the dataset page: https://huggingface.co/datasets/thu-coai/LRM-Safety-Study.LR-MMQA
