CoolFace
30 shown

datasets

Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.

Clear all
01Salesforce /xlam-function-calling-60kgated APIGen Function-Calling Datasets Paper | Website | Models This repo contains 60,000 data collected by APIGen, an automated data generation pipeline designed to produce verifiable high-quality datasets for function-calling applications. Each data in our dataset is verified through three hierarchical stages: format checking, actual function executions, and semantic verification, ensuring its reliability and correctness. We conducted human evaluation over 600 sampled data points… See the full description on the dataset page: https://huggingface.co/datasets/Salesforce/xlam-function-calling-60k.textquestion-answering10K<n<100K719 likes36k downloads2y agoHugging Face02Salesforce /APIGen-MT-5k Summary APIGen-MT is an automated agentic data generation pipeline designed to synthesize verifiable, high-quality, realistic datasets for agentic applications This dataset was released as part of APIGen-MT: Agentic PIpeline for Multi-Turn Data Generation via Simulated Agent-Human Interplay Code: https://github.com/apigen-mt/apigen-mt.github.io The repo contains 5000 multi-turn trajectories collected by APIGen-MT This dataset is a subset of the data used to train the xLAM-2 model… See the full description on the dataset page: https://huggingface.co/datasets/Salesforce/APIGen-MT-5k.textquestion-answering1K<n<10K115 likes4.6k downloads1y agoHugging Face03Salesforce /CRMArenaPro Dataset Card for CRMArena-Pro Dataset Description Paper Information Citation Dataset Description CRMArena-Pro is a benchmark for evaluating LLM agents' ability to perform real-world work tasks in realistic environment. It expands on CRMArena with nineteen expert-validated tasks across sales, service, and "configure, price, and quote" (CPQ) processes, for both Business-to-Business (B2B) and Business-to-Customer (B2C) scenarios. CRMArena-Pro distinctively incorporates… See the full description on the dataset page: https://huggingface.co/datasets/Salesforce/CRMArenaPro.text1K<n<10K18 likes2k downloads1y agoHugging Face04SamuelChien821 /salesbench-100 SalesBench-100 SalesBench-100 is a synthetic long-horizon sales-agent benchmark with 100 original workflows across Salesforce, HubSpot, Gong, and a seeded evidence room. Each task begins with a high-level employee request and has its own authored causal rule and provider transition. Identity, operating facts, authority, governed policy, live-system indexes, and exceptions are separated so no mounted business asset publishes a selected option or precomputed change. Every task has… See the full description on the dataset page: https://huggingface.co/datasets/SamuelChien821/salesbench-100.documenttext-generationn<1K1 likes1.8k downloads24d agoHugging Face05Salesforce /CRMArena Dataset Card for CRMArena Dataset Description Paper Information Citation Dataset Description CRMArena is a benchmark for evaluating LLM agents' ability to perform real-world work tasks in realistic environment. This benchmark is introduced in the paper "CRMArena: Understanding the Capacity of LLM Agents to Perform Professional CRM Tasks in Realistic Environments". We include 16 commonly-used industrial objects (e.g., account, order, knowledge article, case) with… See the full description on the dataset page: https://huggingface.co/datasets/Salesforce/CRMArena.text1K<n<10K8 likes1.1k downloads1y agoHugging Face06ameer4wisam /iraqi-arabic-sales-dialogue-dataset Iraqi Arabic Sales Dialogue Dataset A large synthetic dataset of Iraqi (Baghdadi-based) Arabic dialogue, centered on retail sales, haggling, and everyday conversation. النسخة العربية متوفرة بالكامل بالأسفل — Arabic version available in full below. What this is 210,832 template-generated conversations, of which 171,601 (81%) are exact-unique message sequences, spanning 20 topical categories in colloquial Iraqi Arabic. The core of the dataset (10 categories) is… See the full description on the dataset page: https://huggingface.co/datasets/ameer4wisam/iraqi-arabic-sales-dialogue-dataset.texttext-generation100K<n<1M0 likes703 downloads2mo agoHugging Face07Salesforce /Hard2Verify Hard2Verify: A Step-Level Verification Benchmark for Open-Ended Frontier Math Large language model (LLM)-based reasoning systems have recently achieved gold medal-level performance in the IMO 2025 competition, writing mathematical proofs where, to receive full credit, each step must be not only correct but also sufficiently supported. To train LLM-based reasoners in such challenging, open-ended settings, strong verifiers capable of catching step-level mistakes are necessary… See the full description on the dataset page: https://huggingface.co/datasets/Salesforce/Hard2Verify.textquestion-answeringn<1K7 likes264 downloads11mo agoHugging Face08elg4 /Gym_Salesman_Dataset Gym Salesman Dataset 11,997 synthetic gym-membership sales conversations, each labelled SUCCESS or FAILURE. 🔗 Project Links | Live App — practice against an AI customer | Hugging Face Space | | Telegram Bot — practice on the go | @ido_salescoach_bot | | Dataset — 11,997 labelled conversations | elg4/Gym_Salesman_Dataset | | Data Generation — how the data was built | notebook | | Recommendation — the embedding retriever | notebook | Every conversation is a… See the full description on the dataset page: https://huggingface.co/datasets/elg4/Gym_Salesman_Dataset.tabulartext-classification10K<n<100K1 likes199 downloads1mo agoHugging Face09Salesforce /SCOPE-Persona SCOPE Personas (Nemotron Augmentation) This dataset contains synthetic persona profiles constructed from socio-psychological framework (SCOPE) [https://arxiv.org/pdf/2601.07110], designed to better support LLM simulation usecases in social and behavioral science. It is intended to be used alongside Nemotron-Persona [https://huggingface.co/datasets/nvidia/Nemotron-Personas-USA]. Personas are grounded in a 141-item sociopsychological questionnaire spanning eight facets. You can… See the full description on the dataset page: https://huggingface.co/datasets/Salesforce/SCOPE-Persona.text100K<n<1M2 likes189 downloads5mo agoHugging Face10Thomasgudan /kapibala-sales-dialogues Kapibala Sales Dialogues A sales-conversation dataset with outcome, conversation-level and sentence-level labels 🤗 Hugging Face · Annotation details · 中文 630 synthetic sales conversations (11,688 messages, Chinese and English, five domains) between an LLM-simulated customer and an AI salesperson. Every conversation carries three layers of labels, each produced by a single method across the whole dataset: L1 — outcome. Did the customer buy, agree to a next step, stay undecided… See the full description on the dataset page: https://huggingface.co/datasets/Thomasgudan/kapibala-sales-dialogues.tabulartext-classification10K<n<100K2 likes186 downloads6d agoHugging Face11Salesforce /PROVE Trust but Verify: Programmatic VLM Evaluation in the Wild Viraj Prabhu, Senthil Purushwalkam, An Yan, Caiming Xiong, Ran Xu Explorer | Paper | Quickstart Vision-Language Models (VLMs) often generate plausible but incorrect responses to visual queries. However, reliably quantifying the effect of such hallucinations in free-form responses to open-ended queries is challenging as it requires visually verifying each claim within the response. We propose Programmatic VLM… See the full description on the dataset page: https://huggingface.co/datasets/Salesforce/PROVE.image1K<n<10K5 likes139 downloads2y agoHugging Face12Salesforce /summedits Factual Consistency in Summarization Can you tell which edits of summaries are consistent, and which are inconsistent? SummEdits Benchmark (Section 6-7) We release the 6,348 samples of data for the 10 domains in the SummEdits. Each sample has entries for: domain: out of the 10 domains in SummEdits, id: a unique ID for the sample, doc: the input document, summary: the summary that is either consistent or inconsistent with the facts in the document, label: 1 if the… See the full description on the dataset page: https://huggingface.co/datasets/Salesforce/summedits.texttext-classification1K<n<10K11 likes123 downloads5mo agoHugging Face135CD-AI /Vietnamese-Salesforce-xlam-function-calling-60k-gg-translatedtextquestion-answering10K<n<100K8 likes104 downloads2y agoHugging Face14Salesforce /vibepass VIBEPASS: Can Vibe Coders Really Pass the Vibe Check? Authors: Srijan Bansal, Jiao Fangkai, Yilun Zhou, Austin Xu, Shafiq Joty, Semih Yavuz TL;DR: As LLMs shift programming toward human-guided "vibe coding", agentic tools increasingly rely on models to self-diagnose and repair their own subtle faults—a capability central to autonomous software engineering yet never systematically evaluated. VIBEPASS presents the first empirical benchmark that decomposes fault-targeted reasoning into… See the full description on the dataset page: https://huggingface.co/datasets/Salesforce/vibepass.texttext-generationn<1K2 likes82 downloads6mo agoHugging Face15TianfuXinqu /northwind_sales_anonymized_2023 Sales Transactions (Anonymized) Anonymized sales transaction records generated internally by Northwind Analytics. No external source. License: MIT. textn<1K0 likes82 downloads1mo agoHugging Face16Salesforce /CogAlign Dataset Card for CogAlign Dataset Description Citation Dataset Description CogAlign is a post-training strategy for Vision Language Models (VLMs) aimed at enhancing their visual arithmetic capabilities. This repository presents the training data for CogAlign, a synthetic dataset containing 64,000 examples designed to facilitate this post-training process. CogAlign is inspired by Piaget's theory of cognitive development and focuses on improving a VLM's understanding of… See the full description on the dataset page: https://huggingface.co/datasets/Salesforce/CogAlign.imagevisual-question-answering10K<n<100K6 likes73 downloads1y agoHugging Face17waiszer /real_estate_sales 房地产销冠话术 - 多轮对话 texttext-generation1K<n<10K2 likes72 downloads1y agoHugging Face18ebowwa /people-profiles-io-salestextn<1K3 likes60 downloads2y agoHugging Face19ebowwa /human-biases-sales-marketing-iotextn<1K2 likes59 downloads2y agoHugging Face20open-llm-leaderboard /Salesforce__LLaMA-3-8B-SFR-Iterative-DPO-R-detailsgated Dataset Card for Evaluation run of Salesforce/LLaMA-3-8B-SFR-Iterative-DPO-R Dataset automatically created during the evaluation run of model Salesforce/LLaMA-3-8B-SFR-Iterative-DPO-R The dataset is composed of 38 configuration(s), each one corresponding to one of the evaluated task. The dataset has been created from 1 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always pointing… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard/Salesforce__LLaMA-3-8B-SFR-Iterative-DPO-R-details.tabular10K<n<100K0 likes56 downloads2y agoHugging Face21kapibala-ai /kapibala-sales-dialogues Kapibala Sales Dialogues A sales-conversation dataset with outcome, conversation-level and sentence-level labels 🤗 Hugging Face · Annotation details · 中文 630 synthetic sales conversations (11,688 messages, Chinese and English, five domains) between an LLM-simulated customer and an AI salesperson. Every conversation carries three layers of labels, each produced by a single method across the whole dataset: L1 — outcome. Did the customer buy, agree to a next step, stay undecided… See the full description on the dataset page: https://huggingface.co/datasets/kapibala-ai/kapibala-sales-dialogues.tabulartext-classification10K<n<100K0 likes55 downloads4d agoHugging Face22referencesource /vehicle-trade-in-sales-tax-credit-by-state Vehicle trade-in sales tax credit rules by state -- full price taxed, or only the difference Canonical, always-current version: https://referencesource.org/vehicle-trade-in-sales-tax-credit-by-state/ Machine-readable: https://referencesource.org/vehicle-trade-in-sales-tax-credit-by-state/data.json — this mirror is a point-in-time copy. Last verified: 2026-08-20 Stale after: 2027-08-20 (past this date, prefer the canonical copy — it re-verifies on a cadence this snapshot does… See the full description on the dataset page: https://huggingface.co/datasets/referencesource/vehicle-trade-in-sales-tax-credit-by-state.textn<1K0 likes50 downloads28d agoHugging Face23mkly /crypto-sales-question-answersA dataset consisting of questions, answers, and cryptocurrency descriptions textquestion-answeringn<1K3 likes43 downloads3y agoHugging Face24Salesforce /shared-imagination Dataset Card for Shared Imagination This dataset contains the problems used in the paper Shared Dataset Description This dataset contains the questions generated for the investigations described in the TMLR paper Shared Imagination: LLMs Hallucinate Alike. If you want to use this dataset to assess new models, please use the default config (i.e., datasets.load_dataset('Salesforce/shared-imagination')). This config contains questions for which the four candidate choices… See the full description on the dataset page: https://huggingface.co/datasets/Salesforce/shared-imagination.tabularmultiple-choice10K<n<100K0 likes43 downloads1y agoHugging Face25Salesforce /InstruSumgated InstruSum This is the dataset corresponding to our paper "Benchmarking Generation and Evaluation Capabilities of Large Language Models for Instruction Controllable Summarization". dataset The dataset subset contains 100 human-written data examples by us. Each example contains an article, a summary instruction, a LLM-generated summary, and a hybrid LLM-human summary. human_eval This subset contains human evaluation results for the 100 examples in the dataset… See the full description on the dataset page: https://huggingface.co/datasets/Salesforce/InstruSum.tabularn<1K6 likes39 downloads2y agoHugging Face26TianfuXinqu /filesystem_terminal_huggingface_5022_js2why_sales Product Usage Telemetry 2024 Product usage telemetry events logged by the web and mobile clients during 2024. Columns event_id user_id product_id event_type event_time textn<1K0 likes38 downloads1mo agoHugging Face27Salesforce /summexecedit Factual Consistency in Summarization Evaluate your model's ability to detect and explain the factual inconsistency in summaries. This repo contains the benchmark from our paper "SummExecEdit: A Factual Consistency Benchmark in Summarization with Executable Edits". SummExecEdit Benchmark This benchmark is built over our previous benchmark - SummEdits. Consistent summaries are used from SummEdits. New inconsistent and challenging summaries are generated using executable… See the full description on the dataset page: https://huggingface.co/datasets/Salesforce/summexecedit.texttext-classification1K<n<10K1 likes36 downloads2y agoHugging Face28stindardlogic /sales-methodology-sft-100k Sales Methodology SFT (100K) 100,000 ShareGPT conversations demonstrating expert-level B2B sales execution across discovery, objection handling, enterprise pricing, prospecting, account management, and sales leadership. Motivation Enterprise sales is one of the highest-leverage skills in business — great salespeople and sales leaders drive disproportionate revenue. Models commonly fail at sales tasks by: Generic frameworks without execution detail: Describing… See the full description on the dataset page: https://huggingface.co/datasets/stindardlogic/sales-methodology-sft-100k.texttext-generation100K<n<1M0 likes33 downloads2mo agoHugging Face29ChaosAIVision /Vietnamese-Salesforce-xlam-function-calling-60k-gg-translatedtextquestion-answering10K<n<100K0 likes32 downloads9mo agoHugging Face30miulab /SalesBot2.0text1K<n<10K2 likes31 downloads3y agoHugging Face

Listings come live from the Hugging Face Hub API. CoolFace does not host these files.