datasets
Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.
neat-evidence-c4b4b7
neat-evidence-c4b4b7
Synthetic products test data: 51 rows in data.csv.
All values are randomly generated fictional examples, not real observations, products, or user activity. Intended only for CSV loading and pipeline tests; not suitable for scientific or business conclusions. Columns are sampled independently and do not model real-world correlations.
Fields
sample_id: random identifier for this generated sample.
row_id: sequential row number starting at 1.… See the full description on the dataset page: https://huggingface.co/datasets/seongjin5777/neat-evidence-c4b4b7.just-potato-844c56
just-potato-844c56
Synthetic sensors test data: 59 rows in data.csv.
All values are randomly generated fictional examples, not real observations, products, or user activity. Intended only for CSV loading and pipeline tests; not suitable for scientific or business conclusions. Columns are sampled independently and do not model real-world correlations.
Fields
sample_id: random identifier for this generated sample.
row_id: sequential row number starting at 1.… See the full description on the dataset page: https://huggingface.co/datasets/seoyun74/just-potato-844c56.dead-outcome-feb7f5
dead-outcome-feb7f5
Synthetic sensors test data: 50 rows in data.csv.
All values are randomly generated fictional examples, not real observations, products, or user activity. Intended only for CSV loading and pipeline tests; not suitable for scientific or business conclusions. Columns are sampled independently and do not model real-world correlations.
Fields
sample_id: random identifier for this generated sample.
row_id: sequential row number starting at 1.… See the full description on the dataset page: https://huggingface.co/datasets/seongjin0459/dead-outcome-feb7f5.seoul-bike-rent-month
Seoul Bike Ttareungi Monthly Usage (서울 따릉이 월별 이용정보)
Monthly usage statistics for Seoul's public bike-sharing system "Ttareungi" (따릉이),
aggregated by station, subscription type, and age group.
Important: Text values (rent_type, station_name, age_group) are in Korean.
Column names are in English for global accessibility.
Dataset Summary
Records
~119,000
Features
11
Period
Multiple months (YYYYMM format)
Source
Seoul Open Data Plaza
Features… See the full description on the dataset page: https://huggingface.co/datasets/kpubdata/seoul-bike-rent-month.simple-club-523d12
simple-club-523d12
Synthetic products test data: 55 rows in data.csv.
All values are randomly generated fictional examples, not real observations, products, or user activity. Intended only for CSV loading and pipeline tests; not suitable for scientific or business conclusions. Columns are sampled independently and do not model real-world correlations.
Fields
sample_id: random identifier for this generated sample.
row_id: sequential row number starting at 1.… See the full description on the dataset page: https://huggingface.co/datasets/seongjin042/simple-club-523d12.dell-qa-en-to-ko-translated-by-ke-t5-base
Dell QA English to Korean Translation Dataset
Dataset Description
This dataset, dell-qa-en-to-ko-translated-by-ke-t5-base, is a Korean translation of the original English Dell QA dataset.
Source
The original dataset, dell_qa, is designed for question-answering tasks and contains questions and answers related to Dell technologies. This translated version extends the utility to Korean language tasks.
Dataset Structure
Data Fields
input… See the full description on the dataset page: https://huggingface.co/datasets/seongs/dell-qa-en-to-ko-translated-by-ke-t5-base.labor_market_context_datamx-wordpress-seo-health-benchmark
WordPress/PHP Technical SEO Health Benchmark — 4 Mexican SME Sites
A comparable technical-SEO health benchmark across 4 real, live production
websites in Mexico, spanning different stacks: WordPress (hand-coded theme),
WordPress (Astra + Elementor), WordPress (WooCommerce/Elementor), and a
CMS-free vanilla PHP + MySQL site. All four are scored with the same
deterministic rubric (Technical, On-Page, Speed, Headers → 0-100), so scores
are directly comparable across sites and… See the full description on the dataset page: https://huggingface.co/datasets/NODARISHUB/mx-wordpress-seo-health-benchmark.seo-descriptions-contextSEO-Data-Dive-Data-Challenge-2025
SEO Data Dive 2025: Discovering Patterns in Search (Data Challenge)
IMPORTANT NOTE: This is an exploratory analysis challenge, not a traditional machine learning competition.
Winners will be selected exclusively by an expert jury based on the quality of the submitted analyses. By participating, you explicitly agree to this evaluation process.
Welcome & The Mission
Welcome to the SEO Data Dive Data Challenge! Your mission, should you choose to accept it, is to decode the… See the full description on the dataset page: https://huggingface.co/datasets/searchstudies/SEO-Data-Dive-Data-Challenge-2025.seo-ai-scanner-benchmarks
SEO AI Visibility Scanner Benchmarks
Benchmark dataset of 20 brand visibility scan cases with individual scores for SEO signal, AI visibility, content signal, authority signal, gap, and coverage across 5 AI platforms — Google, ChatGPT, Gemini, Perplexity, and Microsoft Copilot.
Built by GetPR.Buzz.
Dataset Description
This dataset contains benchmark data for the SEO AI Visibility Scanner — a structured scanning framework that evaluates brand visibility across… See the full description on the dataset page: https://huggingface.co/datasets/getpr-buzz/seo-ai-scanner-benchmarks.short-seo-descriptionsseoul_museums서울시 박물관미술관 정보
서울특별시에 소재하는 박물관, 미술관 정보입니다(표준데이터)
시설명, 구분(공립,사립), 도로명주소, 위도, 경도, 전화번호, 운영기관명, 홈페이지 주소, 관람료, 운영시간 정보등을 제공합니다.
seo-report-template-benchmark-2026
SEO Report Template Benchmark 2026
An open, source-linked audit of 16 English-language public pages offering or documenting an SEO report template, dashboard, workbook, presentation deck, or reporting framework.
The dataset records access friction and ten reporting fields that were explicitly documented on each public page on July 10, 2026. It is a public-page audit, not a hands-on product ranking.
Dataset contents
Each row identifies one public resource and… See the full description on the dataset page: https://huggingface.co/datasets/demi-valerith/seo-report-template-benchmark-2026.CCPT_12.3KTitle-Keywords-SEO
Title and Keywords Dataset
The titles are taken from Medium articles dataset , and the keywords are extracted by our team at 🤖 https://exnrt.com
ecb-datasets
ECB Datasets: Cultural Bias Evaluation in Generative Image Models
Overview
This dataset contains human evaluation data for cultural bias analysis in image generation models, supporting the research paper "Exposing Blindspots: Cultural Bias Evaluation in Generative Image Models". ECB stands for "Evaluation Cultural Bias". The dataset includes prompts, generated images, and evaluation metrics across different countries and cultural contexts.
Dataset… See the full description on the dataset page: https://huggingface.co/datasets/seochan99/ecb-datasets.md_bbiyong
Dataset Card for Dataset Name
Dataset Details
Dataset Description
Curated by: [More Information Needed]
Funded by [optional]: [More Information Needed]
Shared by [optional]: [More Information Needed]
Language(s) (NLP): [More Information Needed]
License: [More Information Needed]
Dataset Sources [optional]
Repository: [More Information Needed]
Paper [optional]: [More Information Needed]
Demo [optional]: [More Information Needed]… See the full description on the dataset page: https://huggingface.co/datasets/Seokeunsoo/md_bbiyong.seo-descriptions-gpt4seo_urlsseoul-spot대한민국 서울시의 공원, 박물관과 미술관, 도서관의 이름과 고유 id를 맵핑한 데이터셋입니다.
seoul_store_infoTestabc
