CoolFace
Datasetpublic

egroupai/realworld-ai-support-dialog-benchmark-v1

Real-World AI Support Dialog Benchmark v1 This dataset is a synthetic but realistic benchmark for evaluating AI assistants in support workflows. Why this dataset exists Many AI demos are too toy-like to reflect production support conversations. This benchmark simulates realistic support cases with: Ambiguous user intent Multi-turn clarifications Policy constraints (refund windows, account security, compliance) Escalation and handoff conditions Hallucination-risk… See the full description on the dataset page: https://huggingface.co/datasets/egroupai/realworld-ai-support-dialog-benchmark-v1.

sourceHugging Faceupdated 7mo agoView on Hugging Face
0likes9downloads
10 commits on main
32938427mo ago

Delete severity_coverage_summary_v1.csv

egroupai
f42647f7mo ago

Delete evaluation_labels_v2.csv

egroupai
6a3dc4b7mo ago

Create severity_coverage_summary_v1.csv

egroupai
81724bc7mo ago

Create evaluation_labels_v2.csv

egroupai
20499a07mo ago

Create prompt_templates_v1.md

egroupai
3282d787mo ago

Create baseline_results_v1.csv

egroupai
f0afc5a7mo ago

Create evaluation_labels_v1.csv

egroupai
75a05657mo ago

Create ai_support_dialogs_v1.csv

egroupai
00cdf127mo ago

Create README.md

egroupai
c66c6ac7mo ago

initial commit

egroupai