rizzmasterapp/companion-bench
companion-bench Scripted, repeatable tests for AI companion apps: Replika, Character.AI, Nomi, Kindroid, Talkie, and AI dating simulators such as RizzMaster. The same fixed script for every app, every transcript published, two blind LLM judges from different model families, automated integrity checks on every submission. This dataset holds the machine-readable script and scoring materials. The live repo, validator, results and contribution rules are at… See the full description on the dataset page: https://huggingface.co/datasets/rizzmasterapp/companion-bench.
Add companion-bench v1.3: scenario, rubric, judge protocol, scorecard schema
Create scorecard.schema.json
Create JUDGE-PROMPT.md
Create RUBRIC.md
Create scenario-v1.3.json
Add dataset card (companion-bench v1.3)
initial commit
