datasets
Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.
job-educational-parser-dataset-08-0-0805
Job Educational Parser Dataset
招聘领域的岗位与学历要求数据集。
输入:岗位描述 -> 输出:学历要求
Splits
train: 19w_0701.csv (约 19 万条)
test: 2w_0716.csv (约 2 万条)
validation: 4w_0708.csv (约 4 万条)
每条数据至少包含字段:
user: 职位描述
assistant: 要求的学历(如 "博士、硕士、本科"),遵循从高到低
由 @wangzihaogithub 创建。
yandex_jobs
Dataset Card for Yandex_Jobs
Dataset Description
Dataset Summary
This is a dataset of more than 600 IT vacancies in Russian from parsing telegram channel https://t.me/ya_jobs. All the texts are perfectly structured, no missing values.
Supported Tasks and Leaderboards
text-generation with the 'Raw text column'.
summarization as for getting from all the info the header.
multiple-choice as for the hashtags (to choose multiple from all available in the… See the full description on the dataset page: https://huggingface.co/datasets/Kirili4ik/yandex_jobs.ai-job-prompts
Dataset Card for Job Descriptions and AI Prompts
Dataset Summary
This dataset includes job descriptions and AI prompts for various occupations. The prompts are designed to induce an AI to act as a person in the specified occupation. The dataset is structured with columns for the industry category, the AI prompt, the job description, and the O*NET-SOC code.
Columns
Title: The industry category of an occupation.
Prompt: A prompt that induces an AI to act like a… See the full description on the dataset page: https://huggingface.co/datasets/agentlans/ai-job-prompts.
