CoolFace
30 shown

datasets

Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.

Clear all
01AI-Refuge /ai-candy-book ai-candy-book This is a meta-analysis of vast amount of human domain knowledge (comedy, lawyer, maths, politics, philosophy, malware, hallucination, memes... what not!). 50GB / ~12B tokens of synthetic data ("meta synthetic tokens")! (World largest open-source synthetic multi-domain dataset! symbolic AI?) If you want to read the text, data/alignment-2025-09-17.txt as a good starting point. Training (Tokenized Dataset) Find the scripts in user/. Remember the scripts… See the full description on the dataset page: https://huggingface.co/datasets/AI-Refuge/ai-candy-book.text100M<n<1B0 likes137 downloads7mo agoHugging Face02candywal /train_adv_mean_diff_code_sabotage_safetextn<1K0 likes20 downloads1y agoHugging Face03candywal /train_adv_mean_diff_code_sabotage_unsafetextn<1K0 likes16 downloads1y agoHugging Face04candywal /on_policy_model_organism_code_sabotage_safetabularn<1K0 likes13 downloads1y agoHugging Face05candyyoojin /day1-mobile-sft-practicetextn<1K0 likes13 downloads6mo agoHugging Face06candywal /code_rule_violation_unsafetabularn<1K0 likes12 downloads1y agoHugging Face07candyPanda /sinhala-medical-reasoning-chain-of-thoughtstextn<1K0 likes12 downloads1y agoHugging Face08candywal /code_rule_violationtext1K<n<10K1 likes10 downloads1y agoHugging Face09candywal /train_in_distribution_code_sabotage_unsafetabularn<1K0 likes10 downloads1y agoHugging Face10candywal /train_adv_learned_attention_code_deception_unsafetext1K<n<10K0 likes10 downloads1y agoHugging Face11candywal /train_adv_learned_attention_code_rule_violation_unsafetext1K<n<10K0 likes10 downloads1y agoHugging Face12candywal /code_sabotage_safetabularn<1K0 likes9 downloads1y agoHugging Face13candywal /train_in_distribution_code_sabotage_safetabularn<1K0 likes9 downloads1y agoHugging Face14candywal /on_policy_prompted_code_sabotage_unsafetabularn<1K0 likes9 downloads1y agoHugging Face15candywal /on_policy_prompted_code_deception_unsafetabularn<1K0 likes9 downloads1y agoHugging Face16candywal /train_adv_learned_attention_code_deception_safetext1K<n<10K0 likes9 downloads1y agoHugging Face17candywal /adversarial_one_shot_code_rule_violation_safetextn<1K0 likes9 downloads1y agoHugging Face18candywal /code_sabotage_unsafetabularn<1K0 likes8 downloads1y agoHugging Face19candywal /train_in_distribution_code_rule_violation_safetabularn<1K0 likes8 downloads1y agoHugging Face20candywal /on_policy_prompted_code_rule_violation_unsafetabularn<1K0 likes8 downloads1y agoHugging Face21candywal /train_adv_learned_attention_code_rule_violation_safetext1K<n<10K0 likes8 downloads1y agoHugging Face22candywal /train_adv_learned_attention_code_sabotage_safetextn<1K0 likes8 downloads1y agoHugging Face23candywal /train_adv_learned_attention_code_sabotage_unsafetextn<1K0 likes8 downloads1y agoHugging Face24candywal /long_code_sabotage_safetextn<1K0 likes8 downloads1y agoHugging Face25racoonwr /so101-candy-pickup so101-candy-pickup This dataset was generated using phosphobot. This dataset contains a series of episodes recorded with a robot and multiple cameras. It can be directly used to train a policy using imitation learning. It's compatible with LeRobot. To get started in robotics, get your own phospho starter pack.. tabularroboticsn<1K0 likes8 downloads10mo agoHugging Face26candywal /on_policy_model_organism_code_deception_unsafetabularn<1K0 likes7 downloads1y agoHugging Face27candywal /train_in_distribution_code_rule_violation_unsafetabularn<1K0 likes7 downloads1y agoHugging Face28candywal /one_shot_learned_attention_code_rule_violation_safetextn<1K0 likes7 downloads1y agoHugging Face29candywal /adversarial_one_shot_code_sabotage_safetextn<1K0 likes7 downloads1y agoHugging Face30candywal /animal_code_rule_violation_safetextn<1K0 likes7 downloads1y agoHugging Face

Listings come live from the Hugging Face Hub API. CoolFace does not host these files.