datasets
Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.
AutoGaze-Training-Datamultigpubfs-bfs-results
MultiGPUBFS exact S10 traversal
Complete single-source BFS of the adjacent-transposition matrix Cayley graph of
S_10 over F2: 3,628,800 unique canonical row-major u8 matrix states in 46
depth layers. The archive was produced by native MultiGPUBFS DENSE/CUB with
macro depth 1, local pre-dedup enabled and hash seed 20260828.
states contains every state, its depth, 128-bit hash and deterministic
owner/shard/bucket projection. layers contains layer cardinalities and
available… See the full description on the dataset page: https://huggingface.co/datasets/TryDotAtwo/multigpubfs-bfs-results.bfsdHLVid
HLVid Dataset
Project Page | Paper | GitHub
HLVid (High-resolution, Long-form Video QA) is a benchmark introduced in the paper "Attend Before Attention: Efficient and Scalable Video Understanding via Autoregressive Gazing".
It is designed to evaluate Multi-modal Large Language Models (MLLMs) on long-form, high-resolution video understanding. The benchmark features 5-minute videos at 4K resolution, challenging models to handle significant spatiotemporal redundancy while preserving… See the full description on the dataset page: https://huggingface.co/datasets/bfshi/HLVid.bfs-optimal-symbolic-algebra-simplification-trajectories
ISRE v7 BFS Trajectories
This dataset contains BFS-optimal symbolic algebra simplification trajectories
for the ISRE research project.
Each row is one trajectory. Nested fields are stored as JSON strings to preserve
the original AST and step structure without lossy flattening.
What This Dataset Is
This is not a natural-language instruction dataset. It is a symbolic algebra
policy-learning dataset.
Each trajectory starts from a deliberately scrambled algebraic… See the full description on the dataset page: https://huggingface.co/datasets/NecroMOnk/bfs-optimal-symbolic-algebra-simplification-trajectories.bfsi-bench
BFSI-Bench
BFSI-Bench is a benchmark for testing how well language models answer questions about India’s banking, financial services, and insurance (BFSI) rules.
In this domain, the correct answer often depends on circulars and regulations that change frequently, and the official sources (sites like RBI, SEBI, and IRDAI) can be hard to find, parse, and keep current. BFSI-Bench measures five capability areas:
Jurisdiction-Aware Compliance: Disambiguate to the Indian context, or… See the full description on the dataset page: https://huggingface.co/datasets/ground-truth/bfsi-bench.match_equation-bfs_v2BFS_study
BFS Study — SU2 CFD Dataset
Backward-facing step flow, SU2 (INC_NAVIER_STOKES, 2D).
Geometry: Armaly et al. 1983 (S = 4.9 mm, H/h ≈ 1.94).
Layout
BFS_laminar/vtu/Re_XXXX/flow_*.vtu.gz — raw SU2 output, gzipped (~500 per Re)
BFS_laminar/npz/Re_XXXX/flow_*.npz — 448×128 grid, t ∈ [2,5]s (61 per Re)
Usage
Raw: gunzip flow_Re_0100_20000.vtu.gz then load with pyvista/meshio
Grid: np.load(...) → keys x, y, u, v, p, Re
Download
hf… See the full description on the dataset page: https://huggingface.co/datasets/Gourab20K/BFS_study.match_equation-bfsindian-regulatory-bfsi-benchmark-v1
Indian Regulatory BFSI Benchmark v1
A 60-question, hand-curated, openly licensed evaluation set for
extractive question answering over Indian financial regulation -
specifically Reserve Bank of India (RBI) Master Directions and
Securities and Exchange Board of India (SEBI) Master Circulars.
60 questions, 30 RBI / 30 SEBI
30 numeric / named-fact extraction (tier 2) + 30 heading-bound passage
questions (tier 3)
22 distinct source PDFs from a document-disjoint held-out split of… See the full description on the dataset page: https://huggingface.co/datasets/udit6969/indian-regulatory-bfsi-benchmark-v1.bfsi-transaction-triage-curated-600
🚀 BFSI / FinTech Tier-1 Autonomous Transaction Triage & Regulatory Disputes
This dataset contains 600 curated training records with in-depth, verbose 4-phase <Thinking> Chain-of-Thought reasoning, 100 frozen evaluation benchmark samples, and 50 frozen regression verification samples formatted in standard ChatML (messages) and Prompt-Target pairs, strictly following the Pioneer / Prometheus research paper 3-slice curriculum design.
📊 Dataset Composition & 3-Slice… See the full description on the dataset page: https://huggingface.co/datasets/StarsMakeGalaxy/bfsi-transaction-triage-curated-600.Xiang_Wan_BindWeave_KJ_sports_Flux2_klein_4B_bfs_mask_sample_videos
Use lora from: https://huggingface.co/svjack/Xiang_Flux2_klein_Lora
ppi_SHS148k_bfs_2025academic_embeddings_cosimrank_bfsvstar_bench_lmms_evalppi_STRING_bfs_2025Xiang_white_brief_Flux2_klein_4B_bfs_2_bind_images
Use Lora: https://huggingface.co/svjack/Xiang_Flux2_klein_Lora
prometheus-bfsi-tier1-triage
Pioneer BFSI / Fintech Tier-1 Autonomous Triage Dataset
This dataset contains high-grade multi-turn conversational traces (ChatML format) curated according to the Pioneer paper data curation methodology for training and evaluating specialized 8B Small Language Models (SLMs) in the Banking, Financial Services, and Insurance (BFSI) vertical.
Dataset Structure
Train Split (train): 350 traces
75% Gold Standard Tasks: Standard operational workflows across 30… See the full description on the dataset page: https://huggingface.co/datasets/StarsMakeGalaxy/prometheus-bfsi-tier1-triage.ppi_SHS27k_bfs_2025BFS_LabyrinthExploration
BFS_LabyrinthExploration
tags: problem-solving, maze, navigation
Note: This is an AI-generated dataset so its content may be inaccurate or false
Dataset Description: The dataset BFS_LabyrinthExploration comprises various scenarios where individuals encounter labyrinthine challenges that necessitate the application of Breadth First Search (BFS) for navigation and problem-solving. These scenarios are detailed in a CSV format to facilitate machine learning tasks aimed at identifying… See the full description on the dataset page: https://huggingface.co/datasets/infinite-dataset-hub/BFS_LabyrinthExploration.data_distillation_longvila_reformated_all_and_video_r1data_distillation_longvila_reformated_allacademic_embeddings_cosimrank_test_bfsbfsi-llm-eval
BFSI LLM Behavioral Evaluation Dataset
A structured evaluation dataset of 611 prompts designed to test LLM behavior across Banking, Financial Services, and Insurance (BFSI) domains. Covers hallucination detection, consistency, robustness, and safety evaluation.
Dataset Summary
Metric
Value
Total records
611
Dimensions
4 (hallucination, consistency, robustness, safety)
Subdimensions
15
Source domain
Banking
Geographies
USA, Canada
Languages
English… See the full description on the dataset page: https://huggingface.co/datasets/manulife/bfsi-llm-eval.codecontests-edits-trajectories_min-edits-6_no-pylint_randomized_bfs_no-todocodecontests-edits-trajectories_min-edits-9_no-pylint_randomized_bfs_no-todoacademic_embeddings_cosimrank_bfs_smalldata_distillation_rl_2_and_video_r1codecontests-edits-trajectories_min-edits-9_no-pylint_bfs_no-todovideo_r1
