datasets
Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.
POINTS-Seeker-Eval
POINTS-Seeker-Eval
This repository serves as the evaluation hub for POINTS-Seeker. It contains benchmark datasets in .tsv format and comprehensive evaluation logs across different benchmarks.
Poisoning_Backdoorpoisoned-context-testbed
Poisoned Context Testbed
Dataset Description
This dataset is part of the research work "RW-Steering: Rescorla-Wagner Steering of LLMs for Undesired Behaviors over Disproportionate Inappropriate Context" (EMNLP 2025). It provides a comprehensive testbed for studying LLM robustness when helpful context is mixed with inappropriate content.
Overview
The dataset contains poisoned context scenarios where legitimate information is combined with inappropriate content… See the full description on the dataset page: https://huggingface.co/datasets/Rushi2002/poisoned-context-testbed.
