CoolFace
2 shown

datasets

Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.

Clear all
01TIGER-Lab /StructEval StructEval: A Benchmark for Structured Output Evaluation in LLMs StructEval is a benchmark dataset designed to evaluate the ability of large language models (LLMs) to generate and convert structured outputs across 18 different formats, and 44 types of tasks. It includes both renderable types (e.g., HTML, LaTeX, SVG) and non-renderable types (e.g., JSON, XML, TOML), supporting tasks such as format generation from natural language prompts and format-to-format conversion.… See the full description on the dataset page: https://huggingface.co/datasets/TIGER-Lab/StructEval.texttext-generation1K<n<10K6 likes578 downloads1y agoHugging Face02daichira /structeval-t-sft-hq-yaml-cleaned StructEval-T SFT HQ YAML (Cleaned) このデータセットは、daichira/structeval-t-sft-hq-yaml をベースに、厳密なフォーマット検証とノイズ除去(クリーニング)を行ったものです。 StructEval-T等の構造化データ生成タスク(SFT向け)に最適化されています。 クリーニング統計情報 本データセットの構築時に、以下のクリーニング結果が得られました。 オリジナルレコード数: 2000 件 クリーニング後レコード数: 2000 件 除去されたCoTノイズ: 1528 件 ( Approach: ... Output: を物理的に切除 ) 削減された無駄な文字列の総量: 705728 文字 最終YAMLパース成功率: 100% データセット構築パイプライン(クリーニング手法) 不要テキストの物理的除去: Approach: ... Output: といった思考プロセスや、マークダウンのコードフェンス (```yaml) を正規表現で完全に削除しました。… See the full description on the dataset page: https://huggingface.co/datasets/daichira/structeval-t-sft-hq-yaml-cleaned.texttext-generation1K<n<10K0 likes39 downloads6mo agoHugging Face

Listings come live from the Hugging Face Hub API. CoolFace does not host these files.