kiara
Datasets
All datasets matching “kiara”farsi-asr-unified-cleaned
🎧 Farsi ASR Unified Dataset (Parquet Sharded Edition)
Overview
The Farsi ASR Unified Dataset is a large-scale, high-quality, and fully standardized collection of Persian (Farsi) speech-to-text data — designed specifically for modern machine learning and ASR (Automatic Speech Recognition) workflows.
This dataset consolidates audio–text pairs from multiple open sources, applies a rigorous cleaning and normalization pipeline, and stores everything efficiently in Parquet… See the full description on the dataset page: https://huggingface.co/datasets/kiarashQ/farsi-asr-unified-cleaned.kiarashQ-aug1NightFogPedestrianDataset_KiaRayPOV
Night Fog Pedestrian Dataset (Kia Ray POV)
이 데이터셋은 유니티 퍼셉션(Unity Perception)을 활용하여 생성된 **합성 데이터셋(Synthetic Dataset)**입니다. 자율주행 모델이 안개 낀 야간 환경에서 보행자를 얼마나 정확하게 탐지하는지 테스트하기 위해 제작되었습니다.
1. 데이터 개요
시점 (POV): 기아 레이(Kia Ray) 차량의 블랙박스 위치 (지면으로부터 약 1.4m 높이)
환경 조건: 야간 (Night), 안개 (Foggy/Low Visibility)
클래스: Pedestrian (단일 클래스)
데이터 포맷: YOLOv8 (images/labels)
해상도: 1101 x 514 pixels
2. 데이터셋 구조
.
├── dataset.yaml # YOLOv8 설정 파일
├── images/ # .jpg 이미지 파일… See the full description on the dataset page: https://huggingface.co/datasets/JTSGRIT/NightFogPedestrianDataset_KiaRayPOV.GhostWordcrmsc-envsGPTMicro-Nanowire-Sintering
GPTMicro — Nanowire Sintering & Symbolic Regression Dataset
Curated data for data-driven discovery of governing equations in nanowire
sintering. It pairs raw molecular-dynamics (MD) trajectories with the ML-ready
train/validation/test splits used to learn closed-form models for the sintering
dynamics (change in flattening ddelta and rotation dtheta) and for two
effective material properties (effective diffusion coefficient D_eff and
effective relaxation/viscosity coefficient… See the full description on the dataset page: https://huggingface.co/datasets/Kiarash99/GPTMicro-Nanowire-Sintering.
