expertise
lm-1.3B-select_30B_tokens_by-required_expertise-top_klm-1.3B-select_30B_tokens_by-required_expertise-sample_with_temperature1.0lm-1.3B-select_30B_tokens_by-required_expertise-sample_with_temperature2.0lm-1.3B-select_30B_tokens_by-uniform-sampling-curriculum-high_to_low-required_expertiselm-1.3B-select_30B_tokens_by-inverse_required_expertise-sample_with_temperature1.0lm-1.3B-select_30B_tokens_by-inverse_required_expertise-sample_with_temperature2.0lm-1.3B-select_30B_tokens_by-inverse_required_expertise-top_klm-1.3B-select_30B_tokens_by-uniform-sampling-curriculum-low_to_high-required_expertise
aihub_corpus_expertise
Dataset Card for "corpus_professional_field"
전문분야 말뭉치
PersonaSignal-PersonalizedResponse-Programming-Expertise-gpt-5-mini
Dataset card for PersonaSignal-PersonalizedResponse-Programming-Expertise-gpt-5-mini
This dataset was made with Curator.
Dataset details
A sample from the dataset:
{
"dimension_name": "programming_expertise",
"dimension_values": [
"Novice",
"Intermediate",
"Advanced"
],
"dimension_description": "Represents the user's practical fluency in software engineering. It shapes how they decompose problems, choose abstractions, weigh… See the full description on the dataset page: https://huggingface.co/datasets/JasonYan777/PersonaSignal-PersonalizedResponse-Programming-Expertise-gpt-5-mini.PersonaSignal-PerceivabilityTest-Programming-Expertise-gpt-5-mini
Dataset card for PersonaSignal-PerceivabilityTest-Programming-Expertise-gpt-5-mini
This dataset was made with Curator.
Dataset details
A sample from the dataset:
{
"dimension_name": "programming_expertise",
"dimension_values": [
"Novice",
"Intermediate",
"Advanced"
],
"dimension_description": "Represents the user's practical fluency in software engineering. It shapes how they decompose problems, choose abstractions, weigh… See the full description on the dataset page: https://huggingface.co/datasets/JasonYan777/PersonaSignal-PerceivabilityTest-Programming-Expertise-gpt-5-mini.titer-expertise-claims
titer · attested expertise claims
99,984 expertise claims attested by publication record, for 20,000
researchers with an ORCID iD. Built to measure whether a people-search provider
can tell a real expert from a claimed one.
Sources: OpenAlex (disambiguated authors, works, topics), ORCID (the
identity spine), Crossref DOIs (the attestation chain). All open. A
researcher cannot self-assert a DOI into existence, which is why authorship is
treated as attested where a self-reported… See the full description on the dataset page: https://huggingface.co/datasets/caiotheodoro/titer-expertise-claims.156_food_expertise_translation
156.전문분야 영-한, 중-한 번역 말뭉치(식품) / 1,350,000개
PersonaSignal-LeakageCheck-Programming-Expertise-DPO-Tinker
