CoolFace
30 shown

datasets

Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.

Clear all
01datasetmaster /resumes Dataset Card for Advanced Resume Parser & Job Matcher Resumes This dataset contains a merged collection of real and synthetic resume data in JSON format. The resumes have been normalized to a common schema to facilitate the development of NLP models for candidate-job matching in the technical recruitment domain. Dataset Details Dataset Description This dataset is a combined collection of real resumes and synthetically generated CVs. Curated by:… See the full description on the dataset page: https://huggingface.co/datasets/datasetmaster/resumes.texttoken-classification1K<n<10K18 likes1.9k downloads2y agoHugging Face02MinhND2301 /resumeDatasettextn<1K0 likes173 downloads2y agoHugging Face03hehhe89 /resumes Dataset Card for Advanced Resume Parser & Job Matcher Resumes This dataset contains a merged collection of real and synthetic resume data in JSON format. The resumes have been normalized to a common schema to facilitate the development of NLP models for candidate-job matching in the technical recruitment domain. Dataset Details Dataset Description This dataset is a combined collection of real resumes and synthetically generated CVs. Curated by: datasetmaster… See the full description on the dataset page: https://huggingface.co/datasets/hehhe89/resumes.texttoken-classification1K<n<10K0 likes88 downloads9mo agoHugging Face04ajaxdavis /spm_jsonresume_resumed spm Small Package Model is a method for creating micro llms trained to be an expert on a single software project. The goal is to generate fine tuned models that are so small they can be saved as a package, loaded as a dependency, and run locally. The advantage of this method is that the model can give accurate and up to date information on the particular code being run without needing external tools, it stays up to date with latest changes and understands the specific implementation… See the full description on the dataset page: https://huggingface.co/datasets/ajaxdavis/spm_jsonresume_resumed.textn<1K0 likes63 downloads1y agoHugging Face05Careerflow /ResumeExtractBench ResumeExtractBench ResumeExtractBench is a benchmark for schema-guided structured extraction from resume documents. Given a resume PDF and a JSON Schema, systems must return structured data covering personal details, work history, education, skills, and more. Dataset Size: 38 documents (handwritten + adversarial distractors) Schema Sections Scored: 9 (basics, experience, education, projects, summary, certifications, awards, volunteering, skills) Domains: 6 (engineering… See the full description on the dataset page: https://huggingface.co/datasets/Careerflow/ResumeExtractBench.documentdocument-question-answeringn<1K0 likes57 downloads10h agoHugging Face06Saba06huggingface /resume_dataset Dataset Card for Saba06huggingface/resume_dataset A collection of Resume Examples taken from livecareer.com for categorizing a given resume into any of the labels defined in the dataset. This dataset card aims to be a base template for new datasets. It has been generated using this raw template. Dataset Details Dataset Description About Dataset Context A collection of Resume Examples taken from livecareer.com for categorizing a given resume into any of… See the full description on the dataset page: https://huggingface.co/datasets/Saba06huggingface/resume_dataset.texttext-classificationn<1K2 likes47 downloads3y agoHugging Face07KSE-RESEARCH-Group /Work_UA_resumes WorkUA Resumes Dataset Dataset Summary This dataset contains 103,895 structured resume entries collected from publicly available candidate profiles on Work.ua, Ukraine's largest job platform. Resumes were scraped, parsed, cleaned, and deduplicated for research use. Scraping window: July 9 – August 22, 2025. Intended use: Resume parsing and information extraction Ukrainian-language NLP pipelines Vacancy–candidate matching Labor market and salary analysis Career… See the full description on the dataset page: https://huggingface.co/datasets/KSE-RESEARCH-Group/Work_UA_resumes.tabular100K<n<1M0 likes43 downloads6mo agoHugging Face08theonlymarjona /resume_dataset Dataset Card for Saba06huggingface/resume_dataset A collection of Resume Examples taken from livecareer.com for categorizing a given resume into any of the labels defined in the dataset. This dataset card aims to be a base template for new datasets. It has been generated using this raw template. Dataset Details Dataset Description About Dataset Context A collection of Resume Examples taken from livecareer.com for categorizing a given resume into any… See the full description on the dataset page: https://huggingface.co/datasets/theonlymarjona/resume_dataset.texttext-classificationn<1K0 likes42 downloads15d agoHugging Face09snehamulge2006 /resumes Dataset Card for Advanced Resume Parser & Job Matcher Resumes This dataset contains a merged collection of real and synthetic resume data in JSON format. The resumes have been normalized to a common schema to facilitate the development of NLP models for candidate-job matching in the technical recruitment domain. Dataset Details Dataset Description This dataset is a combined collection of real resumes and synthetically generated CVs. Curated by:… See the full description on the dataset page: https://huggingface.co/datasets/snehamulge2006/resumes.texttoken-classification1K<n<10K0 likes41 downloads17d agoHugging Face10strova-ai /resume-conversations-llm-training 📄 Resume Conversations for LLM Training High-quality conversational dataset for building AI that understands resumes, careers, and professional growth.Created and maintained by Syncora.ai. ✅ Overview This dataset provides resume-related conversations in a structured JSONL format, ideal for developers and AI practitioners working on chatbots, career advisory tools, or LLM fine-tuning. It includes realistic Q&A on career development, technology trends, and professional… See the full description on the dataset page: https://huggingface.co/datasets/strova-ai/resume-conversations-llm-training.texttext-generationn<1K3 likes37 downloads1y agoHugging Face11ssbML /resumes Dataset Card for Advanced Resume Parser & Job Matcher Resumes This dataset contains a merged collection of real and synthetic resume data in JSON format. The resumes have been normalized to a common schema to facilitate the development of NLP models for candidate-job matching in the technical recruitment domain. Dataset Details Dataset Description This dataset is a combined collection of real resumes and synthetically generated CVs. Curated by:… See the full description on the dataset page: https://huggingface.co/datasets/ssbML/resumes.texttoken-classification1K<n<10K0 likes33 downloads3mo agoHugging Face12alirezaaminzadeh /talentmatch-resume-samples TalentMatch Resume Samples Synthetic enterprise resumes and job descriptions with expert HR rankings for benchmark evaluation. Contents screenings.jsonl — model vs expert ranks per JD/resume pair manifest.json — corpus metadata benchmark_report.json — reproducible metrics Usage import json with open("screenings.jsonl") as f: for line in f: print(json.loads(line)) Built by Aria AI. tabularn<1K0 likes33 downloads2mo agoHugging Face13pandalla /datatager_llm_resume_scoring If you like our project, please give us a star ⭐ [GitHub | DataTager Home] Large Language Model Resume Scoring (LLM-RS) Task Dataset Prompt for Training When training your model with this dataset, prepend the following prompt to each input instance: 给定一个候选人的工作经历信息,你需要针对每个职位进行综合评分。每个工作经历包括职位名称、工作内容、技能需求等详细描述。根据职位的特性和需求,你应该为每个工作经历设计不同的评分标准。 针对每个工作经历,基于上述评分方面,给出一个具体的分数(1-10分)。每个评分方面的最高分为10分,确保评分具有差异性,反映出候选人在每个岗位上的表现强度和改进空间。 Description… See the full description on the dataset page: https://huggingface.co/datasets/pandalla/datatager_llm_resume_scoring.text1K<n<10K18 likes28 downloads2y agoHugging Face14weixu-zhang /visrl-14b-phase3e-step3814-resumetabularn<1K0 likes28 downloads4mo agoHugging Face15Sahiti99 /resumes Dataset Card for Advanced Resume Parser & Job Matcher Resumes This dataset contains a merged collection of real and synthetic resume data in JSON format. The resumes have been normalized to a common schema to facilitate the development of NLP models for candidate-job matching in the technical recruitment domain. Dataset Details Dataset Description This dataset is a combined collection of real resumes and synthetically generated CVs. Curated by:… See the full description on the dataset page: https://huggingface.co/datasets/Sahiti99/resumes.texttoken-classification1K<n<10K0 likes27 downloads1mo agoHugging Face162stacks /my-resume-v2 Andrew Stanley Resume Q&A A small, hand-curated chat-format dataset of 122 question/answer pairs covering the professional background, career history, technical skills, certifications, and military service of Andrew Stanley, CTO / Chief Innovation Officer at SMS Data Products Group (McLean, VA). The dataset is purpose-built for two things: A working demonstration of an end-to-end LLM fine-tuning workflow — source document → synthetic Q&A generation → QLoRA fine-tune → GGUF export →… See the full description on the dataset page: https://huggingface.co/datasets/2stacks/my-resume-v2.textquestion-answeringn<1K0 likes24 downloads5mo agoHugging Face17jbeiroa /resume-summarization-dataset Resume Summarization Dataset This dataset contains machine-generated summaries of 14,505 resumes using gpt-4o-mini. Each entry includes the original resume and a markdown-formatted summary divided into 5 sections. Structure Each row is a JSON object with: resume: The original resume text summary: The structured markdown summary input_tokens and output_tokens: (optional) token usage info License Some portions of this dataset are derived from public sources… See the full description on the dataset page: https://huggingface.co/datasets/jbeiroa/resume-summarization-dataset.tabularsummarization10K<n<100K1 likes23 downloads1y agoHugging Face18MikePfunk28 /resume-training-datasetgated Resume Training Dataset Dataset Summary This dataset contains 22,855 curated resume samples designed for training AI models on resume analysis, generation, and career development tasks. Each entry includes structured conversations between users seeking resume help and AI assistants providing feedback, making it ideal for training models to understand professional writing patterns, critique resumes, and suggest improvements. Dataset Details Supported… See the full description on the dataset page: https://huggingface.co/datasets/MikePfunk28/resume-training-dataset.textfeature-extraction10K<n<100K7 likes18 downloads1y agoHugging Face19hf-soo1001 /resume_sectionsThis dataset is mainly for creating NER model. These are the following sections: personal_info summary skills experience education certificates objective textn<1K0 likes16 downloads3mo agoHugging Face20uhfew /resumes Dataset Card for Advanced Resume Parser & Job Matcher Resumes This dataset contains a merged collection of real and synthetic resume data in JSON format. The resumes have been normalized to a common schema to facilitate the development of NLP models for candidate-job matching in the technical recruitment domain. Dataset Details Dataset Description This dataset is a combined collection of real resumes and synthetically generated CVs. Curated by:… See the full description on the dataset page: https://huggingface.co/datasets/uhfew/resumes.texttoken-classification1K<n<10K1 likes16 downloads2d agoHugging Face21kami-dayo /resumes Dataset Card for Advanced Resume Parser & Job Matcher Resumes This dataset contains a merged collection of real and synthetic resume data in JSON format. The resumes have been normalized to a common schema to facilitate the development of NLP models for candidate-job matching in the technical recruitment domain. Dataset Details Dataset Description This dataset is a combined collection of real resumes and synthetically generated CVs. Curated by: datasetmaster… See the full description on the dataset page: https://huggingface.co/datasets/kami-dayo/resumes.texttoken-classification1K<n<10K0 likes15 downloads9mo agoHugging Face22edwinbh /resume_wordtext1K<n<10K0 likes14 downloads2y agoHugging Face23Hatshe /resumes Dataset Card for Advanced Resume Parser & Job Matcher Resumes This dataset contains a merged collection of real and synthetic resume data in JSON format. The resumes have been normalized to a common schema to facilitate the development of NLP models for candidate-job matching in the technical recruitment domain. Dataset Details Dataset Description This dataset is a combined collection of real resumes and synthetically generated CVs. Curated by: datasetmaster… See the full description on the dataset page: https://huggingface.co/datasets/Hatshe/resumes.texttoken-classification1K<n<10K0 likes14 downloads4mo agoHugging Face24gkrishnan /Resume_Best_Practicestexttext-generationn<1K0 likes13 downloads3y agoHugging Face25server41k /resumetextn<1K0 likes13 downloads2y agoHugging Face26aicinema69 /Resume-Job-TextHello I hope you are doing well This dataset is in llama template format to tune. Feel free to use and Contirbute if possible texttext-generation1K<n<10K2 likes13 downloads2y agoHugging Face27Youssef-mohamed123 /resume_entitiestext1K<n<10K0 likes13 downloads1mo agoHugging Face28gunzzz24 /resume-qatextn<1K0 likes12 downloads2y agoHugging Face29Sakshivedi /synthetic-resume-datasettextn<1K1 likes11 downloads1y agoHugging Face30tholstholkappiyan /resumes Dataset Card for Advanced Resume Parser & Job Matcher Resumes This dataset contains a merged collection of real and synthetic resume data in JSON format. The resumes have been normalized to a common schema to facilitate the development of NLP models for candidate-job matching in the technical recruitment domain. Dataset Details Dataset Description This dataset is a combined collection of real resumes and synthetically generated CVs. Curated by:… See the full description on the dataset page: https://huggingface.co/datasets/tholstholkappiyan/resumes.texttoken-classification1K<n<10K0 likes11 downloads7mo agoHugging Face

Listings come live from the Hugging Face Hub API. CoolFace does not host these files.