datasets
Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.
multi-ai-interpretive-responses
Multi-AI Interpretive Responses
arena.ai のサイドバイサイド / ダイレクトバトルで行った
日本語チャットセッションのアーカイブ。
同じ問い(哲学・倫理・サブカル・メタ認知ネタ)に対する複数 LLM の
解釈・応答差を観察するためのデータセット。
「楽しい を教えるお仕事ならしまーす」
倫理と哲学だけメタ超級。バシャール可。仏陀可。ウィトゲン可。サブカル可。
ファイル
ファイル
説明
data.jsonl
1 行 = 1 セッション。HF Datasets Viewer はこれを読みます。
timeline.md
人間用:会話開始時刻順の年表(タイトル・モデル数・所要時間付き)。
filename_map.csv
新ファイル名 ↔ 元日本語タイトルの対応表。
files/chat_XXXX.json
arena.ai 由来の生 JSON(元構造そのまま)。連番は会話開始時刻順。
スキーマ(data.jsonl の… See the full description on the dataset page: https://huggingface.co/datasets/TonbokiriRaikiriMuramasa/multi-ai-interpretive-responses.mental_health_counseling_responses
Dataset Card for Mental Health Counseling Responses
This dataset contains responses to questions from mental health counseling sessions.
The responses are rated by LLMs using the dimensions: empathy, appropriateness, and relevance.
A detailed explanation of the rating process can be found in this blog post.
For a detailed analysis of LLM-generated responses and their comparison to human responses, refer to this blog post.
The original data with the human responses can be found here.… See the full description on the dataset page: https://huggingface.co/datasets/tcabanski/mental_health_counseling_responses.qudrat-student-responses
Qudrat Student Response Dataset
A dataset of 197 multiple-choice questions from the Saudi General Aptitude Test (Qudrat / اختبار القدرات العامة) verbal section, with real student response distributions.
Dataset Description
Each row contains a question with four answer choices, the correct answer, and the percentage of students who selected each option. This enables analysis of student error patterns, question difficulty, and comparison with LLM answer distributions.… See the full description on the dataset page: https://huggingface.co/datasets/hassanalsawadi/qudrat-student-responses.
