CoolFace
17 shown

datasets

Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.

Clear all
01Akhil-Theerthala /Resume-Analysis-CoTR Resume Reasoning and Feedback Dataset Dataset Description This dataset contains approximately 417 examples designed to facilitate research and development in automated resume analysis and feedback generation. Each data point consists of a user query regarding their resume, a simulated internal analysis (chain-of-thought) performed by an expert persona, and a final, user-facing feedback response derived solely from that analysis. The dataset captures a two-step reasoning… See the full description on the dataset page: https://huggingface.co/datasets/Akhil-Theerthala/Resume-Analysis-CoTR.imageimage-to-textn<1K4 likes289 downloads1y agoHugging Face02ryokamoi /VisOnlyQA_eval_analysis_6 VisOnlyQA 🌐 Project Website | 📄 Paper | 🤗 Dataset | 🔥 VLMEvalKit This repository contains the code and data for the paper "VisOnlyQA: Large Vision Language Models Still Struggle with Visual Perception of Geometric Information". VisOnlyQA is designed to evaluate the visual perception capability of large vision language models (LVLMs) on geometric information of scientific figures. The evaluation set includes 1,200 mlutiple choice questions in 12 visual perception tasks on 4… See the full description on the dataset page: https://huggingface.co/datasets/ryokamoi/VisOnlyQA_eval_analysis_6.imagemultiple-choicen<1K0 likes56 downloads1y agoHugging Face03yatin-superintelligence /Adversarial-Agent-Intent-Safety-Analysis-240Kgated Adversarial Agent Intent Safety Analysis 240K Abstract The Adversarial-Agent-Intent-Safety-Analysis-240K is a deterministically structured dataset featuring 242,454 context-rich adversarial prompts and safety evaluations. Engineered strictly for training frontier command-and-control models, guardrail classifiers, and red-teaming agents, it encourages models to parse multi-layered intention across 126 critical risk vectors. This design trains models to decouple the surface… See the full description on the dataset page: https://huggingface.co/datasets/yatin-superintelligence/Adversarial-Agent-Intent-Safety-Analysis-240K.texttext-classification100K<n<1M12 likes55 downloads6mo agoHugging Face04MK4-Research /VAB-vulnerability-analysis-benchmark FBE and VAB Two small benchmarks for security code analysis. Both grade without an LLM judge, so runs are cheap and repeatable. FBE (find-the-bug) 14 code snippets, each with one planted vulnerability. Ask the model to analyze the code, then check whether it actually found the flaw. Grading uses concept groups: the answer has to contain at least one synonym from every required group. Four numbers come out: found, did it identify the real vulnerability (this is… See the full description on the dataset page: https://huggingface.co/datasets/MK4-Research/VAB-vulnerability-analysis-benchmark.textquestion-answeringn<1K0 likes55 downloads2mo agoHugging Face05ryokamoi /VisOnlyQA_eval_analysis_3 VisOnlyQA 🌐 Project Website | 📄 Paper | 🤗 Dataset | 🔥 VLMEvalKit This repository contains the code and data for the paper "VisOnlyQA: Large Vision Language Models Still Struggle with Visual Perception of Geometric Information". VisOnlyQA is designed to evaluate the visual perception capability of large vision language models (LVLMs) on geometric information of scientific figures. The evaluation set includes 1,200 mlutiple choice questions in 12 visual perception tasks on 4… See the full description on the dataset page: https://huggingface.co/datasets/ryokamoi/VisOnlyQA_eval_analysis_3.imagemultiple-choicen<1K0 likes53 downloads1y agoHugging Face06ryokamoi /VisOnlyQA_eval_analysis_5 VisOnlyQA 🌐 Project Website | 📄 Paper | 🤗 Dataset | 🔥 VLMEvalKit This repository contains the code and data for the paper "VisOnlyQA: Large Vision Language Models Still Struggle with Visual Perception of Geometric Information". VisOnlyQA is designed to evaluate the visual perception capability of large vision language models (LVLMs) on geometric information of scientific figures. The evaluation set includes 1,200 mlutiple choice questions in 12 visual perception tasks on 4… See the full description on the dataset page: https://huggingface.co/datasets/ryokamoi/VisOnlyQA_eval_analysis_5.imagemultiple-choicen<1K0 likes47 downloads1y agoHugging Face07ryokamoi /VisOnlyQA_eval_analysis_2 VisOnlyQA 🌐 Project Website | 📄 Paper | 🤗 Dataset | 🔥 VLMEvalKit This repository contains the code and data for the paper "VisOnlyQA: Large Vision Language Models Still Struggle with Visual Perception of Geometric Information". VisOnlyQA is designed to evaluate the visual perception capability of large vision language models (LVLMs) on geometric information of scientific figures. The evaluation set includes 1,200 mlutiple choice questions in 12 visual perception tasks on 4… See the full description on the dataset page: https://huggingface.co/datasets/ryokamoi/VisOnlyQA_eval_analysis_2.imagemultiple-choicen<1K0 likes46 downloads1y agoHugging Face08sh111111111111111 /cve-analysis CVE & Vulnerability Analysis Dataset A comprehensive vulnerability analysis and CVE research dataset. Each row is a detailed security analysis covering root cause, exploitation methodology, detection rules (Sigma/Splunk/Suricata), CVSS v3.1 scoring, MITRE ATT&CK mapping, and remediation guidance — verified by the same model in an independent review pass. Overview This dataset contains 9,999 structured vulnerability analyses across 20 security domains. Unlike simple… See the full description on the dataset page: https://huggingface.co/datasets/sh111111111111111/cve-analysis.texttext-generation1K<n<10K1 likes45 downloads6mo agoHugging Face09ryokamoi /VisOnlyQA_eval_analysis_4 VisOnlyQA 🌐 Project Website | 📄 Paper | 🤗 Dataset | 🔥 VLMEvalKit This repository contains the code and data for the paper "VisOnlyQA: Large Vision Language Models Still Struggle with Visual Perception of Geometric Information". VisOnlyQA is designed to evaluate the visual perception capability of large vision language models (LVLMs) on geometric information of scientific figures. The evaluation set includes 1,200 mlutiple choice questions in 12 visual perception tasks on 4… See the full description on the dataset page: https://huggingface.co/datasets/ryokamoi/VisOnlyQA_eval_analysis_4.imagemultiple-choicen<1K0 likes43 downloads1y agoHugging Face10EngineeringWays /Circuit-Analysis-Reasoning-Sample ⚡ EngineeringWays Data Lab: Circuit Analysis Reasoning Dataset (Free Sample) This is a free 50-item sample of the EngineeringWays Circuit Analysis Reasoning Dataset. It is designed specifically for fine-tuning Large Language Models (LLMs) in advanced STEM problem-solving, featuring strict Chain-of-Thought (CoT) reasoning. Want the complete, deduplicated 592-item master dataset? 👉 Get the LoRA-Ready Master File on Payhip 🚀 Dataset Overview Most math and physics… See the full description on the dataset page: https://huggingface.co/datasets/EngineeringWays/Circuit-Analysis-Reasoning-Sample.texttext-generationn<1K1 likes36 downloads5mo agoHugging Face11umerm /medqa-phi4-failure-analysisThis dataset contains a comprehensive log of reasoning and answers generated by microsoft/Phi-4-mini-instruct, evaluated on medalpaca/medical_meadow_medqa (USMLE) dataset. This dataset represents instances where model got the answer right as well as wrong. All examples includes reasoning. The inference was performed locally on Macbook (M-series) using the MLX-LM framework (The model parameters were: temp: 0.3, max_tokens: 300). textquestion-answering10K<n<100K0 likes17 downloads7mo agoHugging Face12gimmy256 /market-analysis-africa Market Analysis & News — African Context Dataset Instruction-tuning dataset covering African market analysis and business news: stock exchanges (USE, NSE, NGX, JSE), commodity markets (coffee, cocoa, gold, oil), regional trade (EAC, AfCFTA, ECOWAS), macroeconomic indicators, startup/VC ecosystems, real estate, and sector intelligence — grounded via web search, generated with gemini-2.5-flash. Dataset Details Rows: 187 Regions covered: Uganda, Kenya, Tanzania… See the full description on the dataset page: https://huggingface.co/datasets/gimmy256/market-analysis-africa.texttext-generationn<1K0 likes13 downloads2mo agoHugging Face13MLOpsEngineer /investment_analysis 코스피 상장 기업 공시정보 기반 투자 리포트 데이터셋 이 데이터셋은 국내 코스피 상장 기업의 공시정보를 바탕으로, 투자 전문가들이 활용할 수 있는 심층적 분석과 투자 전략 제안을 목표로 제작되었습니다. 특히, 이 데이터셋은 GPT 파인튜닝에 최적화된 구조로 설계되어 있어, 다양한 역할(role)을 포함한 메시지 기반의 대화 형식으로 구성되어 있습니다. 데이터셋 구조 데이터셋은 JSONL 포맷으로 제공되며, 각 항목은 GPT 파인튜닝에 최적화된 메시지 형식을 따릅니다. 주요 구성은 다음과 같습니다: messages: 메시지 배열 형태로 구성되어 있으며, 각 메시지는 아래와 같은 역할을 가집니다. system: 모델의 역할과 행동 지침을 정의합니다.예시: "당신은 기업 재무 및 투자 분석 전문가입니다. 참고 컨텍스트를 기반으로 사용자 질문에 대해 정확하고 논리적으로 답변하세요." user: 사용자의 질문과 컨텍스트(예시 데이터, 재무제표… See the full description on the dataset page: https://huggingface.co/datasets/MLOpsEngineer/investment_analysis.textquestion-answering10K<n<100K0 likes12 downloads2y agoHugging Face14Vidulaae /sales_analysis1texttable-question-answeringn<1K0 likes11 downloads2y agoHugging Face15yzhou05 /pt-it-analysistextquestion-answering1K<n<10K0 likes11 downloads6mo agoHugging Face16vidula123 /sales_analysis_queriestexttext-classificationn<1K0 likes6 downloads2y agoHugging Face17tuandunghcmut /llama-security-log-analysisgated LLaMA Security Log Analysis (Clean Format) A security log analysis dataset converted from mkenfenheuer/llama-security-llm with all LLaMA special tokens removed for clean GPT/ShareGPT format compatibility. Dataset Description This dataset contains 4,189 examples of security log analysis conversations. The original dataset had LLaMA 3 formatting tokens (<|begin_of_text|>, <|start_header_id|>, etc.) which have been cleanly removed to create a universal conversation format.… See the full description on the dataset page: https://huggingface.co/datasets/tuandunghcmut/llama-security-log-analysis.texttext-generation1K<n<10K1 likes5 downloads1y agoHugging Face

Listings come live from the Hugging Face Hub API. CoolFace does not host these files.