datasets
Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.
conversation-calendar
Overview
This dataset was generated for evaluation purposes, focusing on compiling an individual’s information with questions and answers.
Use of ChatGPT-4o: Used for generating human-like text and structured calendar data.
Format: For easy processing, data is stored in JSON (calendar data) and TXT (conversation data) formats.
Persona: The dataset is centered around an individual named ‘Alex’.
Event Generation: Calendar events are designed to be realistic rather than randomly… See the full description on the dataset page: https://huggingface.co/datasets/asu-kim/conversation-calendar.tripsapien-festival-calendar-2026-2027
TripSapien World Festival Calendar 2026–2027
4,618 festivals and events across 2,299 cities worldwide with 2026–2027
dates — 3,872 with exact confirmed dates, 4,173 with a citable non-search
URL — plus event typing, published price signals, recurrence rules, and
per-row provenance.
Curated and continuously maintained as the festival layer of
TripSapien, the itinerary fact-checker that
validates every stop of a pasted travel plan against real opening hours,
closures, and booking… See the full description on the dataset page: https://huggingface.co/datasets/bingwow/tripsapien-festival-calendar-2026-2027.bn-calendar
🗓️ Bengali Calendar Dataset
Dataset Summary
The Bengali Calendar Dataset is a synthetic dataset that contains Bengali calendar-based system prompts, user queries, and assistant answers.It is designed for training and evaluating Bengali language models on reasoning about dates, weekdays, times, and ordinal formats.
Each example follows a structured format:
আজকের তারিখ: ১৯/১০/২০২৫
এখন সময়: রাত ১০:১৫
আজকের বার: রবিবার
and conversations like:
<|user|> আগামীকাল কোন বার?… See the full description on the dataset page: https://huggingface.co/datasets/rejauldu/bn-calendar.
