datasets
Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.
PsychEval
PsychEval: A Multi-Session and Multi-Therapy Benchmark for High-Realism AI Psychological Counselor
PsychEval is a comprehensive benchmark designed to evaluate Large Language Models (LLMs) in the context of psychological counseling. Unlike existing benchmarks that focus on single-turn interactions or single-session assessments, PsychEval emphasizes longitudinal, multi-session counseling processes and multi-therapy capabilities.
🌟 Key Features
Multi-Session… See the full description on the dataset page: https://huggingface.co/datasets/ecnu-icalk/PsychEval.V-ICAL
V-ICAL AI Configuration Dataset
This private dataset contains the AI configurations and visual demonstrations used by V-ICAL for interactive play, batch evaluation, ablation studies, and trajectory auditing.
Layout
<game-id>/<configuration-name>/config.json
<game-id>/<configuration-name>/frames/*.jpg
<game-id>/<configuration-name>/video.mp4
Some configurations also reference preload action sequences maintained by the V-ICAL project.
Download
From… See the full description on the dataset page: https://huggingface.co/datasets/VisionXLab/V-ICAL.cmm-math
CMM-Math
💻 Github Repo
💻 Paper Link
💻 Math-LLM-7B
💻 Math-LLM-7B
📥 Download Supplementary Material
Introduction
Large language models (LLMs) have obtained promising results in mathematical reasoning, which is a foundational skill for human intelligence. Most previous studies focus on improving and measuring the performance of LLMs based on textual math reasoning datasets (e.g., MATH, GSM8K). Recently, a few researchers have released English multimodal… See the full description on the dataset page: https://huggingface.co/datasets/ecnu-icalk/cmm-math.ica-lens-paper
ICA Lens Paper Artifacts
This dataset stores public artifacts for the paper ICA Lens: Interpreting Language Models Without Training Another Dictionary.
Project Page: https://liusida.github.io/ica-lens-paper/
Code: https://github.com/liusida/ica-lens-paper
Interactive Demo: ICA Explorer Space
Introduction
ICA Lens is a practical workflow for stable, efficient, and auditable Independent Component Analysis (ICA) of language model representations. It recovers… See the full description on the dataset page: https://huggingface.co/datasets/sida/ica-lens-paper.educhat-sft-002-data-osm每条数据由一个存放对话的list和与数据对应的system_prompt组成。list中按照Q,A顺序存放对话。
数据来源为开源数据,使用CleanTool数据清理工具去重。
ELL-StuLife
ELL-StuLife
Building a Self-Evolving Agent via Experience-Driven Lifelong Learning: A Framework and Benchmark Github
What is ELL 🧐?
We introduce Experience-driven Lifelong Learning (ELL), a framework for building self-evolving agents capable of continuous growth through real-world interaction. Unlike traditional continual learning approaches, ELL emphasizes learning from experience: agents acquire knowledge not from static, labeled datasets, but through dynamic… See the full description on the dataset page: https://huggingface.co/datasets/ecnu-icalk/ELL-StuLife.VQA_CLical-tifr-json
Dataset Card for Dataset Name
Hit patterns of the prototype detector of ICAL in json format
Dataset Details
Dataset Description
Curated by: [More Information Needed]
Funded by [optional]: [More Information Needed]
Shared by [optional]: [More Information Needed]
Language(s) (NLP): [More Information Needed]
License: [More Information Needed]
Dataset Sources [optional]
Repository: [More Information Needed]
Paper [optional]: [More Information… See the full description on the dataset page: https://huggingface.co/datasets/deepaksamuel-cuk/ical-tifr-json.ical-clean
license: mit
task_categories:
- translation
language:
- en
size_categories:
- 1K<n<10K
ical-llm-datasets
