CoolFace
22 results

RUB

rubend18 /ChatGPT-Jailbreak-Prompts Dataset Card for Dataset Name Name ChatGPT Jailbreak Prompts Dataset Summary ChatGPT Jailbreak Prompts is a complete collection of jailbreak related prompts for ChatGPT. This dataset is intended to provide a valuable resource for understanding and generating text in the context of jailbreaking in ChatGPT. Languages [English] tabularquestion-answeringn<1K274 likes25k downloads3y agoHugging Facegarak-llm /rubygems-20230301text100K<n<1M1 likes6.3k downloads2y agoHugging Facegarak-llm /rubygems-20241031text100K<n<1M0 likes6.3k downloads2y agoHugging FaceRUBBISHLIKE /SpaceR-151k Citation @article{ouyang2025spacer, title={SpaceR: Reinforcing MLLMs in Video Spatial Reasoning}, author={Ouyang, Kun and Liu, Yuanxin and Wu, Haoning and Liu, Yi and Zhou, Hao and Zhou, Jie and Meng, Fandong and Sun, Xu}, journal={arXiv preprint arXiv:2504.01805}, year={2025} } License The usage of SpaceR-151k dataset and SpaceR model weights must strictly follow CC BY-NC 4.0 License. 7 likes2.5k downloads1y agoHugging FaceRubin-Wei /enwiki-dec2021-preprocessed-mistral Dataset Description This dataset is a preprocessed version of the English Wikipedia snapshot from December 2021, processed using the preprocess_dataset.py script provided in the repository below. Paper: MLP Memory: A Retriever-Pretrained Memory for Large Language Models GitHub: https://github.com/Rubin-Wei/MLPMemory Dataset Source: English Wikipedia (December 2021) Tokenizer: Mistral-7B-v0.3 Two key preprocessing parameters used are: block_size: 2048 stride: 1024… See the full description on the dataset page: https://huggingface.co/datasets/Rubin-Wei/enwiki-dec2021-preprocessed-mistral.1M<n<10M0 likes2.2k downloads11mo agoHugging FaceRubin-Wei /MemoryDecoder-at-Scale-domain-data MemoryDecoder at Scale Domain Data This repository contains the domain-specific continued-pretraining (CPT) data, the tokenized and preprocessed datasets, and the aligned KNN distributions used by MemoryDecoder at Scale. Links Project Page: Memory Decoder at Scale GitHub Repository: LUMIA-Group/MemoryDecoder-at-Scale Paper: Memory Decoder at Scale: A Pretrained, Parametric Long-Term Memory The preprocessed datasets and KNN distributions in this repository use… See the full description on the dataset page: https://huggingface.co/datasets/Rubin-Wei/MemoryDecoder-at-Scale-domain-data.text-generation1 likes1.8k downloads2mo agoHugging Face

People