datasets
Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.
trilingual_fraud_consumer_protection_v2
This dataset is a remastered version of this dataset prepared using Adaption's Adaptive Data platform.
Punjab Fraud Intent Benchmark (Trilingual – Punjabi/Hindi/English)
Core Thesis
Most fraud datasets focus on obvious scams.
But in real-world communication, the harder problem is different:
messages that look almost identical but have completely different intent
This dataset focuses on that missing middle:
intent boundaries in Punjabi–Hindi–English… See the full description on the dataset page: https://huggingface.co/datasets/karanverma19/trilingual_fraud_consumer_protection_v2.robocalls_consumer_protections
Robocall Harassment Dataset Simulator
Description
Simulates a dataset to educate on safe customer phone services, raise awareness against AI/robocall harassment by corporations, and help optimize compliant calling practices. Based on adult dataset stats with added legal/compliance features.
Features
Customer demographics (age, education, marital, occupation, etc.)
Economic indicators (CPI, CCI, irate, employment)
Call details (robot/human calls, 800 number)
Location/legal… See the full description on the dataset page: https://huggingface.co/datasets/supersam7/robocalls_consumer_protections.
