CoolFace
Datasetpublic

samuelandaudreymedianetwork/top-100-travel-blogs-2010s-archive

Top 100 Travel Blogs 2010s Historical Archive Historical ranking archive — not a current ranking. This dataset preserves the Nomadic Samuel Top 100 Travel Blogs ranking from the early-to-mid 2010s as a structured historical archive. It includes the final Top 100 composite ranking, additional composite ranking rows from the source page, metric-specific ranking tables, blog entity records, methodology context, origin-story context, academic/research references, and public… See the full description on the dataset page: https://huggingface.co/datasets/samuelandaudreymedianetwork/top-100-travel-blogs-2010s-archive.

sourceHugging Facecc-by-nc-4.0updated 4mo agoView on Hugging Face
2likes48downloads
Dataset Card

Top 100 Travel Blogs 2010s Historical Archive

Historical ranking archive — not a current ranking.

This dataset preserves the Nomadic Samuel Top 100 Travel Blogs ranking from the early-to-mid 2010s as a structured historical archive. It includes the final Top 100 composite ranking, additional composite ranking rows from the source page, metric-specific ranking tables, blog entity records, methodology context, origin-story context, academic/research references, and public references showing how the list circulated through the travel-blogging ecosystem.

The rankings reflect the metrics, methodology, and web environment used at the time. They should not be interpreted as current standings or present-day evaluations of any travel blog.

Source page

  • —URL: https://nomadicsamuel.com/top100travelblogs
  • —Page title: Top 100 Travel Blogs by Domain Authority [List of the BEST TRAVEL blogs]
  • —Meta title: Top 100 Travel Blogs by Domain Authority: Best OG travel blogs list
  • —Meta description: Top 100 travel blogs and top 100 travel sites list ranking Domain Authority, Moz Page Authority, SEMRush and Alexa to rank OG blogs by score.

Why this dataset exists

The Top 100 Travel Blogs list was created by Nomadic Samuel as a data-driven alternative to subjective travel-blog rankings and informal popularity lists common in the early 2010s travel-blogging ecosystem. The source page explains the project through a Rotisserie Baseball / Moneyball analogy: instead of ranking travel blogs by personal preference, conference popularity, or industry relationships, the list attempted to compare blogs across multiple measurable web metrics.

The original ranking was compiled before modern AI tools or automated enrichment workflows. The source page describes a manual process of collecting data from services such as Alexa, Compete, Moz, SEMRush, and SimilarWeb for a broader tracked set of travel blogs. The resulting list is preserved here as a frozen snapshot of independent travel blogging during a period when long-form blogs, search traffic, link authority, and early creator-media networks played a central role in online travel publishing.

This origin context is included to explain why the ranking existed and how it should be interpreted. It does not make the list current, definitive, or universally authoritative. It documents one historical methodology and one creator’s attempt to measure a web-publishing ecosystem using the tools available at the time.

Dataset contents

Record typeCount
academic_reference2
archive_note1
blog_entity208
composite_ranking_entry78
methodology_note1
metric_ranking_entry1246
origin_story_note1
public_reference16
ranking_entry100
source_page1

Total records: 1654

Record types

ranking_entry

The 100 final ranked travel blogs from the source page’s card-grid Top 100 list. These records include final rank, blog name, URL/domain when available, and composite score points.

composite_ranking_entry

Additional composite ranking rows from the source page beyond the Top 100 card grid, including entries in the extended ranking table. These records are included to preserve more of the source page’s historical ranking universe, not to expand the official Top 100.

metric_ranking_entry

Metric-specific rows extracted from the source page’s historical ranking tables. Each row represents one blog’s position in one metric table.

Metric tableRecords
alexa188
compete192
domain_authority192
mozrank98
page_authority192
semrush192
similarweb192

The six primary composite metrics described on the source page are:

  • —Alexa Rank
  • —SEMRush
  • —Moz Domain Authority
  • —Moz Page Authority
  • —Compete
  • —SimilarWeb

The source page also includes a MozRank table. This dataset preserves MozRank as an auxiliary historical metric table because it appears in the source material, but the README and methodology fields distinguish it from the six primary composite metrics described in the methodology text.

blog_entity

A deduplicated blog-level record created from the final ranking, extended composite ranking, and metric tables. These records help users analyze which blogs appear across multiple metric tables, which blogs made the final Top 100, and which blogs were present in the broader tracked ecosystem.

academic_reference

Academic or research-facing references connected to the Top 100 Travel Blogs list or travel-blog ranking/list ecosystem. Some records include verification_status: needs_manual_review when a source is relevant but the exact citation relationship should be manually confirmed before being treated as a fully verified direct citation.

public_reference

Public references from blogger accolade pages, resource directories, interviews, media kits, and tertiary sources. These show how the Top 100 list circulated within the 2010s travel-blogging ecosystem.

origin_story_note

A structured historical-context record summarizing why the Top 100 list was created, how the source page frames the ranking, and how the dataset should be interpreted.

source_page, methodology_note, and archive_note

Context records describing the source URL, ranking method, limitations, and intended historical use.

Files

  • —top-100-travel-blogs-2010s-archive.jsonl — canonical structured dataset
  • —top-100-travel-blogs-2010s-archive.jsonl.gz — compressed JSONL
  • —top-100-travel-blogs-2010s-archive.csv — convenience CSV export of all records
  • —top-100-travel-blogs-2010s-archive.csv.gz — compressed CSV
  • —top-100-travel-blogs-2010s-archive-composite-ranking.csv — final and extended composite ranking convenience table
  • —top-100-travel-blogs-2010s-archive-metric-rankings.csv — metric-specific ranking convenience table
  • —top-100-travel-blogs-2010s-archive-blog-entities.csv — deduplicated blog entity convenience table
  • —top-100-travel-blogs-2010s-archive-references.csv — academic and public references convenience table
  • —llms.txt — short dataset guide
  • —llms-top-100-travel-blogs-2010s-archive.txt — full plain-text export
  • —DATA_DICTIONARY.md — field definitions
  • —SCHEMA.json — machine-readable schema
  • —CITATION.cff — citation metadata
  • —SHA256SUMS.txt — checksums

Reference strength values

Reference records use the field reference_strength:

  • —direct — the source directly references the Nomadic Samuel Top 100 Travel Blogs list
  • —indirect — the source references a related ranking status, accolade, or list use case
  • —tertiary — the source is a tertiary reference such as Wikipedia

Suggested uses

This dataset may be useful for:

  • —historical research into independent travel blogging
  • —creator-economy and web-publishing history
  • —tourism communication research
  • —SEO history and web-metric studies
  • —analysis of early travel-blog rankings
  • —retrieval experiments involving historical web entities
  • —comparing composite ranking systems with individual metric rankings
  • —studying how blogger accolades and resource pages circulated in the 2010s

Important limitations

This is a historical archive. The rankings are not current.

Some metrics used in the source page, including Alexa and Compete, are now defunct or historically frozen. Metric values should be interpreted only in the context of the source page and the period represented.

The source page states that 194 blogs were tracked. The extracted source material contains final ranking records, additional composite ranking rows, and metric tables with overlapping blog names and occasional inconsistencies typical of archived web content. This dataset preserves the source material as an historical record rather than attempting to modernize or correct the original ranking.

The dataset does not claim that the methodology is academically definitive. It documents one historically significant ranking method from the early travel-blogging ecosystem.

Citation

Samuel & Audrey Media Network. (2026). Top 100 Travel Blogs 2010s Historical Archive. Hugging Face. https://doi.org/10.57967/hf/8885

License

Creative Commons Attribution-NonCommercial 4.0 International (CC BY-NC 4.0).

For commercial licensing inquiries or expanded usage rights, contact nomadicsamuel@gmail.com.