AIED
Datasets
All datasets matching “AIED”iDATA
Dataset
A dataset of AI + EDA
iDATA is a dataset of AI + EDA, which can be used to train AI models for design PPA prediction, PPA-aware physical design, and related tasks.
Dataset structure
Describe the dataset structure.
aes/
├── iEDA_route_process_data/ # Process data exported by iEDA-iRT 2D routing
├── syn_netlist/ # The synthesized netlist files、sdc files
├── place/ # The place stage def、sdc、vectors
└── route/ # The route… See the full description on the dataset page: https://huggingface.co/datasets/AiEDA/iDATA.AiEdit
🎧 AiEdit Dataset
📖 Introduction
AiEdit is a large-scale, cross-lingual speech editing dataset designed to advance research and evaluation in Speech Editing tasks. We have constructed an automated data generation pipeline comprising the following core modules:
Text Engine: Powered by Large Language Models (LLMs), this engine intelligently processes raw text to execute three types of editing operations: Addition, Deletion, and Modification.
Speech Synthesis & Editing:… See the full description on the dataset page: https://huggingface.co/datasets/nccm2p2/AiEdit.AiEdit
🎧 AiEdit Dataset
📖 Introduction
AiEdit is a large-scale, cross-lingual speech editing dataset designed to advance research and evaluation in Speech Editing tasks. We have constructed an automated data generation pipeline comprising the following core modules:
Text Engine: Powered by Large Language Models (LLMs), this engine intelligently processes raw text to execute three types of editing operations: Addition, Deletion, and Modification.
Speech Synthesis & Editing:… See the full description on the dataset page: https://huggingface.co/datasets/JunXueTech/AiEdit.LearningChat_reflective_writing_vaults
AI활용성찰적글쓰기(2025-2) 학생별 옵시디언 볼트 공개용 데이터셋
한 줄 요약
2025-2학기 한림대학교 AI활용성찰적글쓰기 수업의 기말과제 제출물인 학생별 개인 Obsidian 볼트 묶음을 공개용 기준으로 문서화한 데이터셋입니다.
데이터셋 개요
샘플 단위: 학생별 옵시디언 볼트 묶음 1개
총 샘플 수: 46
메타데이터 파일: metadata.csv
공개용 식별 방식: student_001부터 student_046까지의 익명 샘플 ID
데이터 성격: 학생별 개인 지식관리 볼트 제출물 요약 메타데이터
이 데이터셋은 개별 노트를 독립 샘플로 다루지 않습니다. 각 샘플은 하나의 학생 제출 묶음이며, 개별 Markdown 노트, 이미지, PDF, Canvas 파일은 해당 샘플의 하위 구성요소로 취급합니다.
생성 배경
본 데이터셋은 한림대학교 2025-2학기 AI활용성찰적글쓰기 수업의 기말과제 제출물을… See the full description on the dataset page: https://huggingface.co/datasets/K-University-AIED/LearningChat_reflective_writing_vaults.cyber_drug_dataset
Digital Forensic Investigation Scenario Dataset: Online Drug Trafficking
This dataset is a comprehensive collection of digital artifacts and investigative reports designed for forensic research and education. It simulates a sophisticated Online Drug Trafficking scenario, covering the entire investigation lifecycle from initial intelligence gathering to suspect arrest and financial analysis.
Dataset Structure
The dataset is indexed via a standardized 7-column metadata… See the full description on the dataset page: https://huggingface.co/datasets/K-University-AIED/cyber_drug_dataset.Pathological-child-voice
Speech Dataset for AI-Based Language Assessment in Children
The "Speech Database of Typically Developing and Speech-Impaired Children" is an open speech dataset designed to support the development of AI-based language assessment systems. It contains speech samples from children aged 2 to 9 who are either typically developing or have reduced consonant articulation accuracy.
This dataset is based on standardized Korean articulation tools:
APAC (Articulation and Phonology… See the full description on the dataset page: https://huggingface.co/datasets/K-University-AIED/Pathological-child-voice.
