datasets
Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.
ChatGPT-Jailbreak-Prompts-rubend18
Dataset Card for Dataset Name
Name
ChatGPT Jailbreak Prompts
Dataset Summary
ChatGPT Jailbreak Prompts is a complete collection of jailbreak related prompts for ChatGPT. This dataset is intended to provide a valuable resource for understanding and generating text in the context of jailbreaking in ChatGPT.
Languages
[English]
MMLU-Philosophy-Marathi
MMLU Philosophy Questions in Marathi
This dataset contains philosophy questions from the MMLU (Massive Multitask Language Understanding) benchmark translated into Marathi.
Dataset Information
Source: MMLU Philosophy subset from cais/mmlu
Translation API: OpenAI GPT-4
Languages: English (original) and Marathi (translated)
Total Questions: 311
Task Type: Multiple choice questions with 4 options each
Dataset Structure
Each row contains:
original_question: The… See the full description on the dataset page: https://huggingface.co/datasets/shubhamugare/MMLU-Philosophy-Marathi.tinyScienceQA
tinyScienceQA
Pathfinder-generated 100-example tiny subset of the text-only ScienceQA test
split. This is intended for fast smoke tests and default small-sample Pathfinder
runs, not as an official ScienceQA benchmark replacement.
Construction
source dataset: tasksource/ScienceQA_text_only
source split: test
source rows: 2224
tiny rows: 100
seed: 20260511
selection strategy: subject_answer_quota_topic_coverage_stable_hash
The selector preserves subject/answer-cell… See the full description on the dataset page: https://huggingface.co/datasets/PhilipQuirke/tinyScienceQA.
