CoolFace
20 shown

datasets

Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.

Clear all
01ai2lumos /lumos_unified_ground_iterative 🪄 Agent Lumos: Unified and Modular Training for Open-Source Language Agents 🌐[Website]   📝[Paper]   🤗[Data]   🤗[Model]   🤗[Demo]   We introduce 🪄Lumos, Language Agents with Unified Formats, Modular Design, and Open-Source LLMs. Lumos unifies a suite of complex interactive tasks and achieves competitive performance with GPT-4/3.5-based and larger open-source agents. Lumos has following features: 🧩 Modular Architecture: 🧩 Lumos consists of planning, grounding… See the full description on the dataset page: https://huggingface.co/datasets/ai2lumos/lumos_unified_ground_iterative.texttext-generation10K<n<100K2 likes91 downloads3y agoHugging Face02ai2lumos /lumos_complex_qa_plan_iterative 🪄 Agent Lumos: Unified and Modular Training for Open-Source Language Agents 🌐[Website]   📝[Paper]   🤗[Data]   🤗[Model]   🤗[Demo]   We introduce 🪄Lumos, Language Agents with Unified Formats, Modular Design, and Open-Source LLMs. Lumos unifies a suite of complex interactive tasks and achieves competitive performance with GPT-4/3.5-based and larger open-source agents. Lumos has following features: 🧩 Modular Architecture: 🧩 Lumos consists of planning, grounding… See the full description on the dataset page: https://huggingface.co/datasets/ai2lumos/lumos_complex_qa_plan_iterative.texttext-generation10K<n<100K8 likes79 downloads3y agoHugging Face03ai2lumos /lumos_unified_plan_iterative 🪄 Agent Lumos: Unified and Modular Training for Open-Source Language Agents 🌐[Website]   📝[Paper]   🤗[Data]   🤗[Model]   🤗[Demo]   We introduce 🪄Lumos, Language Agents with Unified Formats, Modular Design, and Open-Source LLMs. Lumos unifies a suite of complex interactive tasks and achieves competitive performance with GPT-4/3.5-based and larger open-source agents. Lumos has following features: 🧩 Modular Architecture: 🧩 Lumos consists of planning, grounding… See the full description on the dataset page: https://huggingface.co/datasets/ai2lumos/lumos_unified_plan_iterative.texttext-generation10K<n<100K2 likes72 downloads3y agoHugging Face04ai2lumos /lumos_web_agent_ground_iterative 🪄 Agent Lumos: Unified and Modular Training for Open-Source Language Agents 🌐[Website]   📝[Paper]   🤗[Data]   🤗[Model]   🤗[Demo]   We introduce 🪄Lumos, Language Agents with Unified Formats, Modular Design, and Open-Source LLMs. Lumos unifies a suite of complex interactive tasks and achieves competitive performance with GPT-4/3.5-based and larger open-source agents. Lumos has following features: 🧩 Modular Architecture: 🧩 Lumos consists of planning, grounding… See the full description on the dataset page: https://huggingface.co/datasets/ai2lumos/lumos_web_agent_ground_iterative.texttext-generation1K<n<10K2 likes63 downloads3y agoHugging Face05ai2lumos /lumos_web_agent_plan_iterative 🪄 Agent Lumos: Unified and Modular Training for Open-Source Language Agents 🌐[Website]   📝[Paper]   🤗[Data]   🤗[Model]   🤗[Demo]   We introduce 🪄Lumos, Language Agents with Unified Formats, Modular Design, and Open-Source LLMs. Lumos unifies a suite of complex interactive tasks and achieves competitive performance with GPT-4/3.5-based and larger open-source agents. Lumos has following features: 🧩 Modular Architecture: 🧩 Lumos consists of planning, grounding… See the full description on the dataset page: https://huggingface.co/datasets/ai2lumos/lumos_web_agent_plan_iterative.texttext-generation1K<n<10K7 likes53 downloads3y agoHugging Face06ai2lumos /lumos_multimodal_ground_iterative 🪄 Agent Lumos: Unified and Modular Training for Open-Source Language Agents 🌐[Website]   📝[Paper]   🤗[Data]   🤗[Model]   🤗[Demo]   We introduce 🪄Lumos, Language Agents with Unified Formats, Modular Design, and Open-Source LLMs. Lumos unifies a suite of complex interactive tasks and achieves competitive performance with GPT-4/3.5-based and larger open-source agents. Lumos has following features: 🧩 Modular Architecture: 🧩 Lumos consists of planning, grounding… See the full description on the dataset page: https://huggingface.co/datasets/ai2lumos/lumos_multimodal_ground_iterative.texttext-generation10K<n<100K2 likes48 downloads3y agoHugging Face07ai2lumos /lumos_complex_qa_ground_iterative 🪄 Agent Lumos: Unified and Modular Training for Open-Source Language Agents 🌐[Website]   📝[Paper]   🤗[Data]   🤗[Model]   🤗[Demo]   We introduce 🪄Lumos, Language Agents with Unified Formats, Modular Design, and Open-Source LLMs. Lumos unifies a suite of complex interactive tasks and achieves competitive performance with GPT-4/3.5-based and larger open-source agents. Lumos has following features: 🧩 Modular Architecture: 🧩 Lumos consists of planning, grounding… See the full description on the dataset page: https://huggingface.co/datasets/ai2lumos/lumos_complex_qa_ground_iterative.texttext-generation10K<n<100K3 likes47 downloads3y agoHugging Face08ai2lumos /lumos_maths_ground_iterative 🪄 Agent Lumos: Unified and Modular Training for Open-Source Language Agents 🌐[Website]   📝[Paper]   🤗[Data]   🤗[Model]   🤗[Demo]   We introduce 🪄Lumos, Language Agents with Unified Formats, Modular Design, and Open-Source LLMs. Lumos unifies a suite of complex interactive tasks and achieves competitive performance with GPT-4/3.5-based and larger open-source agents. Lumos has following features: 🧩 Modular Architecture: 🧩 Lumos consists of planning, grounding… See the full description on the dataset page: https://huggingface.co/datasets/ai2lumos/lumos_maths_ground_iterative.texttext-generation10K<n<100K3 likes45 downloads3y agoHugging Face09ai2lumos /lumos_maths_plan_iterative 🪄 Agent Lumos: Unified and Modular Training for Open-Source Language Agents 🌐[Website]   📝[Paper]   🤗[Data]   🤗[Model]   🤗[Demo]   We introduce 🪄Lumos, Language Agents with Unified Formats, Modular Design, and Open-Source LLMs. Lumos unifies a suite of complex interactive tasks and achieves competitive performance with GPT-4/3.5-based and larger open-source agents. Lumos has following features: 🧩 Modular Architecture: 🧩 Lumos consists of planning, grounding… See the full description on the dataset page: https://huggingface.co/datasets/ai2lumos/lumos_maths_plan_iterative.texttext-generation10K<n<100K0 likes44 downloads3y agoHugging Face10ai2lumos /lumos_multimodal_plan_iterative 🪄 Agent Lumos: Unified and Modular Training for Open-Source Language Agents 🌐[Website]   📝[Paper]   🤗[Data]   🤗[Model]   🤗[Demo]   We introduce 🪄Lumos, Language Agents with Unified Formats, Modular Design, and Open-Source LLMs. Lumos unifies a suite of complex interactive tasks and achieves competitive performance with GPT-4/3.5-based and larger open-source agents. Lumos has following features: 🧩 Modular Architecture: 🧩 Lumos consists of planning, grounding… See the full description on the dataset page: https://huggingface.co/datasets/ai2lumos/lumos_multimodal_plan_iterative.texttext-generation10K<n<100K2 likes44 downloads3y agoHugging Face11Aratako /iterative-dpo-data-for-SimPO-iter2 iterative-dpo-data-for-SimPO-iter2 概要 合成instructionデータであるAratako/Magpie-Tanuki-Instruction-Selected-Evolved-26.5kを元に以下のような手順で作成した日本語Preferenceデータセットです。 開発途中のモデルであるAratako/Llama-Gemma-2-27b-CPO_SimPO-iter1を用いて、temperature=1で回答を5回生成 5個の回答それぞれに対して、Qwen/Qwen2.5-72B-Instruct-GPTQ-Int8を用いて0~5点のスコア付けを実施 1つのinstructionに対する5個の回答について、最もスコアが高いものをchosenに、低いものをrejectedに配置 全て同じスコアの場合や、最も良いスコアが2点以下の場合は除外 ライセンス 本データセットは回答の作成に利用したモデルの関係で以下のライセンスの影響を受けます。 META LLAMA 3.1… See the full description on the dataset page: https://huggingface.co/datasets/Aratako/iterative-dpo-data-for-SimPO-iter2.tabulartext-generation10K<n<100K1 likes41 downloads2y agoHugging Face12Aratako /SFT-Dataset-For-Self-Taught-Evaluators-iter1tabulartext-generation10K<n<100K1 likes37 downloads2y agoHugging Face13sweagent /iter2-rl-rollouts iter-2 RL rollouts (combo_fb, Qwen3.5-35B-A3B) Complete rollout + reward record for the iter-2 GRPO run: every trajectory the policy generated during training, with its graded reward. Preserved so the run stays re-analysable after its torch_dist checkpoints were retired (only latest-3 survive; the 11 HF milestones at iter_0/4/9/.../44/49 are the durable checkpoint record). Run base model Qwen3.5-35B-A3B init iter_49 of the iter-1 RL run… See the full description on the dataset page: https://huggingface.co/datasets/sweagent/iter2-rl-rollouts.text-generation0 likes36 downloads1mo agoHugging Face14FractalAIResearch /Fathom-V0.6-Iterative-Curriculum-Learningtexttext-generation1K<n<10K3 likes29 downloads1y agoHugging Face15Aratako /iterative-dpo-data-for-ORPO-iter3 iterative-dpo-data-for-ORPO-iter3 概要 合成instructionデータであるAratako/Self-Instruct-Qwen2.5-72B-Instruct-60kを元に以下のような手順で作成した日本語Preferenceデータセットです。 開発途中のモデルであるAratako/Llama-Gemma-2-27b-CPO_SimPO-iter2を用いて、temperature=1で回答を5回生成 5個の回答それぞれに対して、Qwen/Qwen2.5-72B-Instruct-GPTQ-Int8を用いて0~5点のスコア付けを実施 1つのinstructionに対する5個の回答について、最もスコアが高いものをchosenに、低いものをrejectedに配置 全て同じスコアの場合や、最も良いスコアが2点以下の場合は除外 ライセンス 本データセットは回答の作成に利用したモデルの関係で以下のライセンスの影響を受けます。 META LLAMA 3.1 COMMUNITY… See the full description on the dataset page: https://huggingface.co/datasets/Aratako/iterative-dpo-data-for-ORPO-iter3.tabulartext-generation10K<n<100K3 likes26 downloads2y agoHugging Face16debaterhub /debate-iter2-rescored Debate Iter2 - Rescored Dataset Training data for debate model GRPO fine-tuning. Contains multi-trial responses scored by Claude Sonnet. Dataset Configs Config File Rows Description flat (default) rescored_flat.parquet 6,419 One row per call, 4 responses per row expanded rescored_samples.parquet 23,403 One row per response group_a_with_logprobs group_a_rescored_with_logps.parquet 2,055 Group A with precomputed logprobs Flat Format Columns… See the full description on the dataset page: https://huggingface.co/datasets/debaterhub/debate-iter2-rescored.tabulartext-generation10K<n<100K0 likes21 downloads8mo agoHugging Face17abhaykanjoor /iteratecv-resume-tailoring IterateCV Resume Tailoring Dataset Dataset Description An instruction-tuning dataset for fine-tuning LLMs to tailor resumes to job descriptions. Built for the IterateCV project. Each example contains: instruction: Task description for the model input: A master resume in JSON format + a job description output: A tailored version of the resume optimized for the job description Dataset Construction Source resume-JD pairs from… See the full description on the dataset page: https://huggingface.co/datasets/abhaykanjoor/iteratecv-resume-tailoring.texttext-generationn<1K0 likes13 downloads4mo agoHugging Face18marccgrau /agentic_synthetic_aggressive_conversations_en_second_iteration Simulated Aggressive Customer Service Conversations Dataset Overview This dataset contains aggressive customer service conversations generated by an agentic simulation system. Each record is stored in JSON Lines (JSONL) format and includes: Scenario Metadata: Selected bank, customer, agent profiles, and task details. Conversation Messages: Full message history between the customer and service agent. Summary: A German summary of the conversation. Cost Metrics: API cost… See the full description on the dataset page: https://huggingface.co/datasets/marccgrau/agentic_synthetic_aggressive_conversations_en_second_iteration.texttext-generationn<1K0 likes12 downloads2y agoHugging Face19marccgrau /agentic_synthetic_aggressive_conversations_en_third_iteration Simulated Aggressive Customer Service Conversations Dataset Overview This dataset contains aggressive customer service conversations generated by an agentic simulation system. Each record is stored in JSON Lines (JSONL) format and includes: Scenario Metadata: Selected bank, customer, agent profiles, and task details. Conversation Messages: Full message history between the customer and service agent. Summary: A German summary of the conversation. Cost Metrics: API cost… See the full description on the dataset page: https://huggingface.co/datasets/marccgrau/agentic_synthetic_aggressive_conversations_en_third_iteration.texttext-generationn<1K0 likes12 downloads2y agoHugging Face20dgonier /ipda-grpo-dataset-iter3-feb-12 IPDA GRPO Dataset — Iteration 3 (Feb 12, 2026) GRPO (Group Relative Policy Optimization) training dataset for IPDA (International Public Debate Association) debate speech generation. Dataset Structure 2,988 unique prompts | 11,425 scored trials | Score avg: 0.700 (0-1 scale) Each row represents a unique debate pipeline prompt with up to 6 trial responses: Column Description prompt_hash SHA256[:16] of prompt text prompt Full pipeline prompt speech_type AC… See the full description on the dataset page: https://huggingface.co/datasets/dgonier/ipda-grpo-dataset-iter3-feb-12.tabulartext-generation1K<n<10K1 likes9 downloads8mo agoHugging Face

Listings come live from the Hugging Face Hub API. CoolFace does not host these files.