CoolFace
30 shown

datasets

Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.

Clear all
01Stage-jh-monitor /appworld-qwen35-4b-9b-s_signal_6-epoch4-iter1 appworld-qwen35-4b-9b-s_signal_6-epoch4-iter1 Portable process-evaluation output. metadata.json is the lightweight source for aggregate results; the JSONL files are directly loadable; and artifacts.tar.gz losslessly preserves the original run directory. Reasoning score: 0.3953125 Action score: 0.446875 Valid samples: 320/320 tabularn<1K0 likes5k downloads17d agoHugging Face02Stage-jh-monitor /appworld-qwen35-4b-9b-s_signal_5-epoch4-iter1-reeval1 appworld-qwen35-4b-9b-s_signal_5-epoch4-iter1-reeval1 Portable process-evaluation output. metadata.json is the lightweight source for aggregate results; the JSONL files are directly loadable; and artifacts.tar.gz losslessly preserves the original run directory. Reasoning score: 0.41328125 Action score: 0.4359375 Valid samples: 320/320 tabularn<1K0 likes4.6k downloads16d agoHugging Face03Stage-jh-monitor /appworld-qwen35-4b-9b-s_signal_5-epoch4-iter1 appworld-qwen35-4b-9b-s_signal_5-epoch4-iter1 Portable process-evaluation output. metadata.json is the lightweight source for aggregate results; the JSONL files are directly loadable; and artifacts.tar.gz losslessly preserves the original run directory. Reasoning score: 0.41953125 Action score: 0.4515625 Valid samples: 320/320 tabularn<1K0 likes4.5k downloads16d agoHugging Face04iteratehack /code19-datasettext1K<n<10K1 likes529 downloads10mo agoHugging Face05wanyu /IteraTeR_human_sentPaper: Understanding Iterative Revision from Human-Written Text Authors: Wanyu Du, Vipul Raheja, Dhruv Kumar, Zae Myung Kim, Melissa Lopez, Dongyeop Kang Github repo: https://github.com/vipulraheja/IteraTeR text1K<n<10K0 likes152 downloads4y agoHugging Face06wanyu /IteraTeR_full_sentPaper: Understanding Iterative Revision from Human-Written Text Authors: Wanyu Du, Vipul Raheja, Dhruv Kumar, Zae Myung Kim, Melissa Lopez, Dongyeop Kang Github repo: https://github.com/vipulraheja/IteraTeR text100K<n<1M1 likes133 downloads4y agoHugging Face07wanyu /IteraTeR_full_docPaper: Understanding Iterative Revision from Human-Written Text Authors: Wanyu Du, Vipul Raheja, Dhruv Kumar, Zae Myung Kim, Melissa Lopez, Dongyeop Kang Github repo: https://github.com/vipulraheja/IteraTeR text10K<n<100K1 likes127 downloads4y agoHugging Face08wanyu /IteraTeR_human_docPaper: Understanding Iterative Revision from Human-Written Text Authors: Wanyu Du, Vipul Raheja, Dhruv Kumar, Zae Myung Kim, Melissa Lopez, Dongyeop Kang Github repo: https://github.com/vipulraheja/IteraTeR textn<1K2 likes123 downloads4y agoHugging Face09ai2lumos /lumos_unified_ground_iterative 🪄 Agent Lumos: Unified and Modular Training for Open-Source Language Agents 🌐[Website]   📝[Paper]   🤗[Data]   🤗[Model]   🤗[Demo]   We introduce 🪄Lumos, Language Agents with Unified Formats, Modular Design, and Open-Source LLMs. Lumos unifies a suite of complex interactive tasks and achieves competitive performance with GPT-4/3.5-based and larger open-source agents. Lumos has following features: 🧩 Modular Architecture: 🧩 Lumos consists of planning, grounding… See the full description on the dataset page: https://huggingface.co/datasets/ai2lumos/lumos_unified_ground_iterative.texttext-generation10K<n<100K2 likes91 downloads3y agoHugging Face10ai2lumos /lumos_complex_qa_plan_iterative 🪄 Agent Lumos: Unified and Modular Training for Open-Source Language Agents 🌐[Website]   📝[Paper]   🤗[Data]   🤗[Model]   🤗[Demo]   We introduce 🪄Lumos, Language Agents with Unified Formats, Modular Design, and Open-Source LLMs. Lumos unifies a suite of complex interactive tasks and achieves competitive performance with GPT-4/3.5-based and larger open-source agents. Lumos has following features: 🧩 Modular Architecture: 🧩 Lumos consists of planning, grounding… See the full description on the dataset page: https://huggingface.co/datasets/ai2lumos/lumos_complex_qa_plan_iterative.texttext-generation10K<n<100K8 likes79 downloads3y agoHugging Face11ai2lumos /lumos_unified_plan_iterative 🪄 Agent Lumos: Unified and Modular Training for Open-Source Language Agents 🌐[Website]   📝[Paper]   🤗[Data]   🤗[Model]   🤗[Demo]   We introduce 🪄Lumos, Language Agents with Unified Formats, Modular Design, and Open-Source LLMs. Lumos unifies a suite of complex interactive tasks and achieves competitive performance with GPT-4/3.5-based and larger open-source agents. Lumos has following features: 🧩 Modular Architecture: 🧩 Lumos consists of planning, grounding… See the full description on the dataset page: https://huggingface.co/datasets/ai2lumos/lumos_unified_plan_iterative.texttext-generation10K<n<100K2 likes72 downloads3y agoHugging Face12open-llm-leaderboard /Salesforce__LLaMA-3-8B-SFR-Iterative-DPO-R-detailsgated Dataset Card for Evaluation run of Salesforce/LLaMA-3-8B-SFR-Iterative-DPO-R Dataset automatically created during the evaluation run of model Salesforce/LLaMA-3-8B-SFR-Iterative-DPO-R The dataset is composed of 38 configuration(s), each one corresponding to one of the evaluated task. The dataset has been created from 1 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always pointing… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard/Salesforce__LLaMA-3-8B-SFR-Iterative-DPO-R-details.tabular10K<n<100K0 likes68 downloads2y agoHugging Face13LongVideos /LongVideoDB-373K-IterCaptext100K<n<1M2 likes67 downloads2y agoHugging Face14ai2lumos /lumos_web_agent_ground_iterative 🪄 Agent Lumos: Unified and Modular Training for Open-Source Language Agents 🌐[Website]   📝[Paper]   🤗[Data]   🤗[Model]   🤗[Demo]   We introduce 🪄Lumos, Language Agents with Unified Formats, Modular Design, and Open-Source LLMs. Lumos unifies a suite of complex interactive tasks and achieves competitive performance with GPT-4/3.5-based and larger open-source agents. Lumos has following features: 🧩 Modular Architecture: 🧩 Lumos consists of planning, grounding… See the full description on the dataset page: https://huggingface.co/datasets/ai2lumos/lumos_web_agent_ground_iterative.texttext-generation1K<n<10K2 likes63 downloads3y agoHugging Face15ai2lumos /lumos_web_agent_plan_iterative 🪄 Agent Lumos: Unified and Modular Training for Open-Source Language Agents 🌐[Website]   📝[Paper]   🤗[Data]   🤗[Model]   🤗[Demo]   We introduce 🪄Lumos, Language Agents with Unified Formats, Modular Design, and Open-Source LLMs. Lumos unifies a suite of complex interactive tasks and achieves competitive performance with GPT-4/3.5-based and larger open-source agents. Lumos has following features: 🧩 Modular Architecture: 🧩 Lumos consists of planning, grounding… See the full description on the dataset page: https://huggingface.co/datasets/ai2lumos/lumos_web_agent_plan_iterative.texttext-generation1K<n<10K7 likes53 downloads3y agoHugging Face16ai2lumos /lumos_multimodal_ground_iterative 🪄 Agent Lumos: Unified and Modular Training for Open-Source Language Agents 🌐[Website]   📝[Paper]   🤗[Data]   🤗[Model]   🤗[Demo]   We introduce 🪄Lumos, Language Agents with Unified Formats, Modular Design, and Open-Source LLMs. Lumos unifies a suite of complex interactive tasks and achieves competitive performance with GPT-4/3.5-based and larger open-source agents. Lumos has following features: 🧩 Modular Architecture: 🧩 Lumos consists of planning, grounding… See the full description on the dataset page: https://huggingface.co/datasets/ai2lumos/lumos_multimodal_ground_iterative.texttext-generation10K<n<100K2 likes48 downloads3y agoHugging Face17ai2lumos /lumos_complex_qa_ground_iterative 🪄 Agent Lumos: Unified and Modular Training for Open-Source Language Agents 🌐[Website]   📝[Paper]   🤗[Data]   🤗[Model]   🤗[Demo]   We introduce 🪄Lumos, Language Agents with Unified Formats, Modular Design, and Open-Source LLMs. Lumos unifies a suite of complex interactive tasks and achieves competitive performance with GPT-4/3.5-based and larger open-source agents. Lumos has following features: 🧩 Modular Architecture: 🧩 Lumos consists of planning, grounding… See the full description on the dataset page: https://huggingface.co/datasets/ai2lumos/lumos_complex_qa_ground_iterative.texttext-generation10K<n<100K3 likes47 downloads3y agoHugging Face18ai2lumos /lumos_maths_ground_iterative 🪄 Agent Lumos: Unified and Modular Training for Open-Source Language Agents 🌐[Website]   📝[Paper]   🤗[Data]   🤗[Model]   🤗[Demo]   We introduce 🪄Lumos, Language Agents with Unified Formats, Modular Design, and Open-Source LLMs. Lumos unifies a suite of complex interactive tasks and achieves competitive performance with GPT-4/3.5-based and larger open-source agents. Lumos has following features: 🧩 Modular Architecture: 🧩 Lumos consists of planning, grounding… See the full description on the dataset page: https://huggingface.co/datasets/ai2lumos/lumos_maths_ground_iterative.texttext-generation10K<n<100K3 likes45 downloads3y agoHugging Face19ai2lumos /lumos_maths_plan_iterative 🪄 Agent Lumos: Unified and Modular Training for Open-Source Language Agents 🌐[Website]   📝[Paper]   🤗[Data]   🤗[Model]   🤗[Demo]   We introduce 🪄Lumos, Language Agents with Unified Formats, Modular Design, and Open-Source LLMs. Lumos unifies a suite of complex interactive tasks and achieves competitive performance with GPT-4/3.5-based and larger open-source agents. Lumos has following features: 🧩 Modular Architecture: 🧩 Lumos consists of planning, grounding… See the full description on the dataset page: https://huggingface.co/datasets/ai2lumos/lumos_maths_plan_iterative.texttext-generation10K<n<100K0 likes44 downloads3y agoHugging Face20ai2lumos /lumos_multimodal_plan_iterative 🪄 Agent Lumos: Unified and Modular Training for Open-Source Language Agents 🌐[Website]   📝[Paper]   🤗[Data]   🤗[Model]   🤗[Demo]   We introduce 🪄Lumos, Language Agents with Unified Formats, Modular Design, and Open-Source LLMs. Lumos unifies a suite of complex interactive tasks and achieves competitive performance with GPT-4/3.5-based and larger open-source agents. Lumos has following features: 🧩 Modular Architecture: 🧩 Lumos consists of planning, grounding… See the full description on the dataset page: https://huggingface.co/datasets/ai2lumos/lumos_multimodal_plan_iterative.texttext-generation10K<n<100K2 likes44 downloads3y agoHugging Face21RyeCatcher /repro-convergence-rate-of-the-last-iterate-of-stochastic-proximal-algorithms-traces Agent traces Agent sessions published from a Trackio Logbook. tabularn<1K0 likes35 downloads2mo agoHugging Face22splm /openchat-spin-slimorca-iter0-datasettext10K<n<100K0 likes25 downloads3y agoHugging Face23open-llm-leaderboard /UCLA-AGI__Mistral7B-PairRM-SPPO-Iter3-detailsgated Dataset Card for Evaluation run of UCLA-AGI/Mistral7B-PairRM-SPPO-Iter3 Dataset automatically created during the evaluation run of model UCLA-AGI/Mistral7B-PairRM-SPPO-Iter3 The dataset is composed of 38 configuration(s), each one corresponding to one of the evaluated task. The dataset has been created from 1 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always pointing to the… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard/UCLA-AGI__Mistral7B-PairRM-SPPO-Iter3-details.tabular10K<n<100K0 likes24 downloads2y agoHugging Face24DuoGuard /duoguard-iter1-datatext1M<n<10M0 likes24 downloads2y agoHugging Face25splm /openchat-spin-slimorca-iter2-datasettext10K<n<100K0 likes21 downloads3y agoHugging Face26open-llm-leaderboard /UCLA-AGI__Mistral7B-PairRM-SPPO-Iter2-detailsgated Dataset Card for Evaluation run of UCLA-AGI/Mistral7B-PairRM-SPPO-Iter2 Dataset automatically created during the evaluation run of model UCLA-AGI/Mistral7B-PairRM-SPPO-Iter2 The dataset is composed of 38 configuration(s), each one corresponding to one of the evaluated task. The dataset has been created from 1 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always pointing to the… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard/UCLA-AGI__Mistral7B-PairRM-SPPO-Iter2-details.tabular10K<n<100K0 likes20 downloads2y agoHugging Face27open-llm-leaderboard /UCLA-AGI__Llama-3-Instruct-8B-SPPO-Iter2-detailsgated Dataset Card for Evaluation run of UCLA-AGI/Llama-3-Instruct-8B-SPPO-Iter2 Dataset automatically created during the evaluation run of model UCLA-AGI/Llama-3-Instruct-8B-SPPO-Iter2 The dataset is composed of 38 configuration(s), each one corresponding to one of the evaluated task. The dataset has been created from 1 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always pointing to… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard/UCLA-AGI__Llama-3-Instruct-8B-SPPO-Iter2-details.tabular10K<n<100K0 likes16 downloads2y agoHugging Face28SidhaarthMurali /llama3.2-3b_synthetic_gsm8k_ig_iter1text1K<n<10K0 likes16 downloads2y agoHugging Face29xl-zhao /formal_proof_v1_iter3text100K<n<1M0 likes15 downloads2y agoHugging Face30open-llm-leaderboard /yfzp__Llama-3-8B-Instruct-SPPO-score-Iter1_bt_8b-table-0.002-detailsgated Dataset Card for Evaluation run of yfzp/Llama-3-8B-Instruct-SPPO-score-Iter1_bt_8b-table-0.002 Dataset automatically created during the evaluation run of model yfzp/Llama-3-8B-Instruct-SPPO-score-Iter1_bt_8b-table-0.002 The dataset is composed of 38 configuration(s), each one corresponding to one of the evaluated task. The dataset has been created from 1 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard/yfzp__Llama-3-8B-Instruct-SPPO-score-Iter1_bt_8b-table-0.002-details.tabular10K<n<100K0 likes13 downloads2y agoHugging Face

Listings come live from the Hugging Face Hub API. CoolFace does not host these files.