CoolFace
11 shown

datasets

Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.

Clear all
01LLM-OS-Models /LFM2.5-KO-SFT-Stage2-Diverse-KoSWE-Reasoning-LFMChat-Raw LFM2.5-KO-SFT-Stage2-Diverse-KoSWE-Reasoning-LFMChat-Raw Stage2 raw LFM chat JSONL shards: Korean domain, behavior, SWE/coding, reasoning, finance, legal, Text2SQL. This dataset is part of the LFM2.5-8B-A1B-KO-SFT / Agentic SFT workflow. Main SFT model: https://huggingface.co/LLM-OS-Models/LFM2.5-8B-A1B-KO-SFT CPT base model: https://huggingface.co/LLM-OS-Models/LFM2.5-8B-A1B-KO-CPT-FULL Agentic follow-up model: https://huggingface.co/LLM-OS-Models/LFM2.5-8B-A1B-KO-Agentic-SFT… See the full description on the dataset page: https://huggingface.co/datasets/LLM-OS-Models/LFM2.5-KO-SFT-Stage2-Diverse-KoSWE-Reasoning-LFMChat-Raw.0 likes161 downloads3mo agoHugging Face02LLM-OS-Models /LFM2.5-KO-SFT-Stage0-Legal-LFMChat-8K LFM2.5-KO-SFT-Stage0-Legal-LFMChat-8K Stage0 Korean legal warmup, LFM tokenizer, response-only SFT arrays. This dataset is part of the LFM2.5-8B-A1B-KO-SFT / Agentic SFT workflow. Main SFT model: https://huggingface.co/LLM-OS-Models/LFM2.5-8B-A1B-KO-SFT CPT base model: https://huggingface.co/LLM-OS-Models/LFM2.5-8B-A1B-KO-CPT-FULL Agentic follow-up model: https://huggingface.co/LLM-OS-Models/LFM2.5-8B-A1B-KO-Agentic-SFT SFT GitHub: https://github.com/gyunggyung/LFM25-KO-SFT CPT… See the full description on the dataset page: https://huggingface.co/datasets/LLM-OS-Models/LFM2.5-KO-SFT-Stage0-Legal-LFMChat-8K.0 likes52 downloads3mo agoHugging Face03LLM-OS-Models /LFM2.5-KO-SFT-Stage2-Diverse-KoSWE-Reasoning-LFMChat-4K LFM2.5-KO-SFT-Stage2-Diverse-KoSWE-Reasoning-LFMChat-4K Stage2 diverse Korean/SWE/reasoning prepared SFT arrays, LFM tokenizer. This dataset is part of the LFM2.5-8B-A1B-KO-SFT / Agentic SFT workflow. Main SFT model: https://huggingface.co/LLM-OS-Models/LFM2.5-8B-A1B-KO-SFT CPT base model: https://huggingface.co/LLM-OS-Models/LFM2.5-8B-A1B-KO-CPT-FULL Agentic follow-up model: https://huggingface.co/LLM-OS-Models/LFM2.5-8B-A1B-KO-Agentic-SFT SFT GitHub:… See the full description on the dataset page: https://huggingface.co/datasets/LLM-OS-Models/LFM2.5-KO-SFT-Stage2-Diverse-KoSWE-Reasoning-LFMChat-4K.0 likes52 downloads3mo agoHugging Face04LLM-OS-Models /LFM2.5-KO-SFT-Stage1-Legal-Terminal-LFMChat-8K LFM2.5-KO-SFT-Stage1-Legal-Terminal-LFMChat-8K Stage1 8k Korean legal/terminal/tool-use prepared SFT arrays. This dataset is part of the LFM2.5-8B-A1B-KO-SFT / Agentic SFT workflow. Main SFT model: https://huggingface.co/LLM-OS-Models/LFM2.5-8B-A1B-KO-SFT CPT base model: https://huggingface.co/LLM-OS-Models/LFM2.5-8B-A1B-KO-CPT-FULL Agentic follow-up model: https://huggingface.co/LLM-OS-Models/LFM2.5-8B-A1B-KO-Agentic-SFT SFT GitHub: https://github.com/gyunggyung/LFM25-KO-SFT… See the full description on the dataset page: https://huggingface.co/datasets/LLM-OS-Models/LFM2.5-KO-SFT-Stage1-Legal-Terminal-LFMChat-8K.0 likes49 downloads3mo agoHugging Face05LLM-OS-Models /LFM2.5-KO-SFT-Stage0B-Finance-Text2SQL-LFMChat-4K LFM2.5-KO-SFT-Stage0B-Finance-Text2SQL-LFMChat-4K Stage0b finance/Text2SQL/legal fast mix, LFM tokenizer. This dataset is part of the LFM2.5-8B-A1B-KO-SFT / Agentic SFT workflow. Main SFT model: https://huggingface.co/LLM-OS-Models/LFM2.5-8B-A1B-KO-SFT CPT base model: https://huggingface.co/LLM-OS-Models/LFM2.5-8B-A1B-KO-CPT-FULL Agentic follow-up model: https://huggingface.co/LLM-OS-Models/LFM2.5-8B-A1B-KO-Agentic-SFT SFT GitHub: https://github.com/gyunggyung/LFM25-KO-SFT CPT… See the full description on the dataset page: https://huggingface.co/datasets/LLM-OS-Models/LFM2.5-KO-SFT-Stage0B-Finance-Text2SQL-LFMChat-4K.0 likes43 downloads3mo agoHugging Face06LLM-OS-Models /LFM2.5-KO-SFT-Stage1-Finance-Text2SQL-LFMChat-4K LFM2.5-KO-SFT-Stage1-Finance-Text2SQL-LFMChat-4K Stage1 4k finance/Text2SQL prepared SFT arrays. This dataset is part of the LFM2.5-8B-A1B-KO-SFT / Agentic SFT workflow. Main SFT model: https://huggingface.co/LLM-OS-Models/LFM2.5-8B-A1B-KO-SFT CPT base model: https://huggingface.co/LLM-OS-Models/LFM2.5-8B-A1B-KO-CPT-FULL Agentic follow-up model: https://huggingface.co/LLM-OS-Models/LFM2.5-8B-A1B-KO-Agentic-SFT SFT GitHub: https://github.com/gyunggyung/LFM25-KO-SFT CPT GitHub:… See the full description on the dataset page: https://huggingface.co/datasets/LLM-OS-Models/LFM2.5-KO-SFT-Stage1-Finance-Text2SQL-LFMChat-4K.0 likes40 downloads3mo agoHugging Face07LLM-OS-Models /LFM2.5-KO-SFT-Stage2-Plus-KoTSQA-LFMChat-4K LFM2.5-KO-SFT-Stage2-Plus-KoTSQA-LFMChat-4K Stage2 diverse prepared arrays plus KoTSQA train supplement. This dataset is part of the LFM2.5-8B-A1B-KO-SFT / Agentic SFT workflow. Main SFT model: https://huggingface.co/LLM-OS-Models/LFM2.5-8B-A1B-KO-SFT CPT base model: https://huggingface.co/LLM-OS-Models/LFM2.5-8B-A1B-KO-CPT-FULL Agentic follow-up model: https://huggingface.co/LLM-OS-Models/LFM2.5-8B-A1B-KO-Agentic-SFT SFT GitHub: https://github.com/gyunggyung/LFM25-KO-SFT CPT… See the full description on the dataset page: https://huggingface.co/datasets/LLM-OS-Models/LFM2.5-KO-SFT-Stage2-Plus-KoTSQA-LFMChat-4K.0 likes35 downloads3mo agoHugging Face08gyung /toolbench-lfm-chatml ToolBench ChatML Dataset Train: 187,494 examples Eval: 762 examples Source: ToolBench (toolllama_G123_dfs) Format: ChatML messages text100K<n<1M3 likes33 downloads7mo agoHugging Face09LLM-OS-Models /LFM2.5-KO-SFT-Stage2-KoTSQA-Train-LFMChat-Raw LFM2.5-KO-SFT-Stage2-KoTSQA-Train-LFMChat-Raw KoTSQA v2 train split converted to LFM chat JSONL. Test split is held out for evaluation. This dataset is part of the LFM2.5-8B-A1B-KO-SFT / Agentic SFT workflow. Main SFT model: https://huggingface.co/LLM-OS-Models/LFM2.5-8B-A1B-KO-SFT CPT base model: https://huggingface.co/LLM-OS-Models/LFM2.5-8B-A1B-KO-CPT-FULL Agentic follow-up model: https://huggingface.co/LLM-OS-Models/LFM2.5-8B-A1B-KO-Agentic-SFT SFT GitHub:… See the full description on the dataset page: https://huggingface.co/datasets/LLM-OS-Models/LFM2.5-KO-SFT-Stage2-KoTSQA-Train-LFMChat-Raw.0 likes33 downloads3mo agoHugging Face10LLM-OS-Models /LFM2.5-KO-Agentic-Fable-Grounded-LFMChat-Raw LFM2.5-KO-Agentic-Fable-Grounded-LFMChat-Raw Fable5/Helio Korean agentic traces and local grounded document/log examples converted to LFM chat JSONL. This dataset is part of the LFM2.5-8B-A1B-KO-SFT / Agentic SFT workflow. Main SFT model: https://huggingface.co/LLM-OS-Models/LFM2.5-8B-A1B-KO-SFT CPT base model: https://huggingface.co/LLM-OS-Models/LFM2.5-8B-A1B-KO-CPT-FULL Agentic follow-up model: https://huggingface.co/LLM-OS-Models/LFM2.5-8B-A1B-KO-Agentic-SFT SFT GitHub:… See the full description on the dataset page: https://huggingface.co/datasets/LLM-OS-Models/LFM2.5-KO-Agentic-Fable-Grounded-LFMChat-Raw.0 likes28 downloads3mo agoHugging Face11LLM-OS-Models /LFM2.5-KO-Agentic-Fable-Grounded-LFMChat-8K LFM2.5-KO-Agentic-Fable-Grounded-LFMChat-8K Agentic/Fable grounded 8k prepared response-only SFT arrays. This dataset is part of the LFM2.5-8B-A1B-KO-SFT / Agentic SFT workflow. Main SFT model: https://huggingface.co/LLM-OS-Models/LFM2.5-8B-A1B-KO-SFT CPT base model: https://huggingface.co/LLM-OS-Models/LFM2.5-8B-A1B-KO-CPT-FULL Agentic follow-up model: https://huggingface.co/LLM-OS-Models/LFM2.5-8B-A1B-KO-Agentic-SFT SFT GitHub: https://github.com/gyunggyung/LFM25-KO-SFT CPT… See the full description on the dataset page: https://huggingface.co/datasets/LLM-OS-Models/LFM2.5-KO-Agentic-Fable-Grounded-LFMChat-8K.0 likes17 downloads3mo agoHugging Face

Listings come live from the Hugging Face Hub API. CoolFace does not host these files.