CoolFace
20 results

caa

CaasiHUANG /InstructTTSEval InstructTTSEval InstructTTSEval is a comprehensive benchmark designed to evaluate Text-to-Speech (TTS) systems' ability to follow complex natural-language style instructions. The dataset provides a hierarchical evaluation framework with three progressively challenging tasks that test both low-level acoustic control and high-level style generalization capabilities. Github Repository: https://github.com/KexinHUANG19/InstructTTSEval Paper: InstructTTSEval: Benchmarking Complex… See the full description on the dataset page: https://huggingface.co/datasets/CaasiHUANG/InstructTTSEval.audiotext-to-speech1K<n<10K18 likes439 downloads1y agoHugging Faceanon-caa-neurips /CAA Clinical Agent Annotator (CAA) A clinician-in-the-loop benchmark for long-horizon medical LLM agents: 333 KG-grounded clinical tasks, multi-gate evaluation across diagnosis, required tool use, parameterised actions, and must-ask history-taking topics, plus the full harbor evaluation harness so you can re-run every number locally. This repository bundles three artifacts: Task corpus — 333 clinician-approved tasks (and the 290-task authoring set the case study trains on).… See the full description on the dataset page: https://huggingface.co/datasets/anon-caa-neurips/CAA.textquestion-answeringn<1K0 likes153 downloads5mo agoHugging Faceaarajbhattarai /CA-AIN-V2 Nepali Source-Grounded Instruction Dataset Synthetic Nepali instruction-tuning data generated with NVIDIA NeMo Data Designer from authoritative Nepali documents (agriculture manuals, legal texts). Answers are grounded strictly in the source; unanswerable questions get an explicit refusal. Records use chat messages format plus metadata and per-record quality_scores (grounding / correctness / naturalness, 1-5, LLM-as-judge). One data/train-<shard>.jsonl per source document; shards… See the full description on the dataset page: https://huggingface.co/datasets/aarajbhattarai/CA-AIN-V2.texttext-generationn<1K0 likes69 downloads21d agoHugging Faceakiliaiafrica /caava-grpo-training-data0 likes68 downloads9mo agoHugging Facelldbrett /archaeological-sites-caa2025 Archaeological Site Dataset (CAA UK 2025) Dataset Summary This dataset provides a comprehensive multi-channel remote sensing dataset for training machine learning models to detect archaeological sites. The dataset combines Sentinel-2 satellite imagery, FABDEM elevation data, and derived spectral indices to create 11-channel representations of 1×1 km grid cells at 10m resolution. Key Features: Multi-modal data: 6 spectral bands + 3 spectral indices + 2 terrain features… See the full description on the dataset page: https://huggingface.co/datasets/lldbrett/archaeological-sites-caa2025.geospatialimage-classification1K<n<10K0 likes39 downloads9mo agoHugging FaceethanCSL /eval_steering_caa_high_2This dataset was created using LeRobot. Dataset Structure meta/info.json: { "codebase_version": "v3.0", "robot_type": "koch_follower", "total_episodes": 8, "total_frames": 1898, "total_tasks": 1, "chunks_size": 1000, "data_files_size_in_mb": 100, "video_files_size_in_mb": 200, "fps": 30, "splits": { "train": "0:8" }, "data_path": "data/chunk-{chunk_index:03d}/file-{file_index:03d}.parquet", "video_path":… See the full description on the dataset page: https://huggingface.co/datasets/ethanCSL/eval_steering_caa_high_2.tabularrobotics1K<n<10K0 likes38 downloads2mo agoHugging Face