datasets
Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.
communication-adaptation-sft-100k
Communication Adaptation SFT (100K)
100,000 ShareGPT conversations demonstrating skilled communication style adaptation across 15 task types. Each example shows how to take the same underlying content and adjust register, technical depth, length, and framing for different audiences and purposes.
Motivation
Communication adaptation is a core professional skill that LLMs often handle clumsily. Common failures:
Technical monologue: explaining cloud storage to a… See the full description on the dataset page: https://huggingface.co/datasets/stindardlogic/communication-adaptation-sft-100k.alia_gva_communications
📘 ALIA_GVA_Communications Dataset
The ALIA_GVA_Communications dataset is a multilingual resource designed for text generation.
The dataset consists of textual documents formatted in Markdown (.md), each provided as structured JSONL entries.
Each entry includes information about the text's language, format, text, source, and metadata.
🧾 Column Descriptions
Field
Type
Description
format
string
Indicates the text format. All entries use "md" (Markdown).… See the full description on the dataset page: https://huggingface.co/datasets/gplsi/alia_gva_communications.communication_datasetll-v4-5-12-hard-question-communication-20260622english-soft-skills-communication-30Communication-Clarity-Helper
