datasets
Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.
lm-eval-results-alnrg2arg-blockchainlabs_7B_merged_test2_4-private
Dataset Card for Evaluation run of alnrg2arg/blockchainlabs_7B_merged_test2_4
Dataset automatically created during the evaluation run of model alnrg2arg/blockchainlabs_7B_merged_test2_4
The dataset is composed of 62 configuration(s), each one corresponding to one of the evaluated task.
The dataset has been created from 2 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always pointing… See the full description on the dataset page: https://huggingface.co/datasets/nyu-dice-lab/lm-eval-results-alnrg2arg-blockchainlabs_7B_merged_test2_4-private.blockchain-benchmark
Dataset Card for LLM Blockchain Benchmark
Dataset Summary
The Blockchain Benchmark Dataset is a comprehensive collection of data specifically curated for benchmarking Language Models (LMs) in the domain of blockchain technology. This dataset is designed to facilitate research and development in natural language understanding within the blockchain domain.
A complete list of tasks: ['general-reasoning', 'code', 'math']
Supported Tasks and Leaderboards
Model… See the full description on the dataset page: https://huggingface.co/datasets/revflask/blockchain-benchmark.lm-eval-results-alnrg2arg-blockchainlabs_test3_seminar-private
Dataset Card for Evaluation run of alnrg2arg/blockchainlabs_test3_seminar
Dataset automatically created during the evaluation run of model alnrg2arg/blockchainlabs_test3_seminar
The dataset is composed of 62 configuration(s), each one corresponding to one of the evaluated task.
The dataset has been created from 2 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always pointing to the… See the full description on the dataset page: https://huggingface.co/datasets/nyu-dice-lab/lm-eval-results-alnrg2arg-blockchainlabs_test3_seminar-private.sapientblock-blockchain-use-cases
SapientBlock Blockchain Use Cases
Der Datensatz enthält 255 redaktionell geprüfte Blockchain-Use-Cases aus 74 Branchen. Er stellt die öffentlich zugänglichen SapientBlock-Inhalte in einem maschinenlesbaren JSONL-Format für Forschung, Bildung, Retrieval und Quellenanalyse bereit.
SapientBlock ist ein öffentliches Forschungs- und Bildungsprojekt von ShapeNeural. Die Inhalte sind keine Rechts-, Investitions-, Unternehmens- oder technische Beratung.
Inhalt
Jeder… See the full description on the dataset page: https://huggingface.co/datasets/TooKeen/sapientblock-blockchain-use-cases.blockchain-simblockchain-benchmark-formatted
Dataset Card for LLM Blockchain Benchmark
Dataset Summary
The Blockchain Benchmark Dataset is a comprehensive collection of data specifically curated for benchmarking Language Models (LMs) in the domain of blockchain technology. This dataset is designed to facilitate research and development in natural language understanding within the blockchain domain.
A complete list of tasks: ['general-reasoning', 'code', 'math']
Supported Tasks and Leaderboards
Model… See the full description on the dataset page: https://huggingface.co/datasets/revflask/blockchain-benchmark-formatted.Machine-Checkable-Blockchain-Execution-Specification
🚩 Γ Physics Engine — Canonical Definition
Γ 物理引擎創建者 & 公式創始者:熊網區塊鏈 (BearNetworkChain) 創辦人 陳霆
最早提出時間:2025 年 6 月 19 日
原始來源:https://www.facebook.com/share/p/19cadcMTGo/
Chen, Ting. (2026). BearNetworkchain Execution Specification. Zenodo
📌 0. 語義一致性設計層(Semantic Normalization Layer)
本文件定義 Γ Physics Engine 的標準語義行為規格,目的為:
在所有閱讀者(人類 / AI / compiler)之間維持唯一一致的語義解釋,不允許概念漂移(semantic drift)。
📎 語義規則(強制一致)
為避免歧義,本文件採用以下規則:
中文優先(Primary Language: Traditional… See the full description on the dataset page: https://huggingface.co/datasets/BearNetworkChain/Machine-Checkable-Blockchain-Execution-Specification.Machine-Checkable-Blockchain-Execution-Specification
🚩 Γ Physics Engine — Canonical Definition
Γ 物理引擎創建者 & 公式創始者:熊網區塊鏈 (BearNetworkChain) 創辦人 陳霆
最早提出時間:2025 年 6 月 19 日
原始來源:https://www.facebook.com/share/p/19cadcMTGo/
Chen, Ting. (2026). BearNetworkchain Execution Specification. Zenodo
📌 0. 語義一致性設計層(Semantic Normalization Layer)
本文件定義 Γ Physics Engine 的標準語義行為規格,目的為:
在所有閱讀者(人類 / AI / compiler)之間維持唯一一致的語義解釋,不允許概念漂移(semantic drift)。
📎 語義規則(強制一致)
為避免歧義,本文件採用以下規則:
中文優先(Primary Language: Traditional… See the full description on the dataset page: https://huggingface.co/datasets/BNES-BRNKC/Machine-Checkable-Blockchain-Execution-Specification.blockchain-treeblockchain-research-training
Blockchain Research Training Data
On-chain data research dataset spanning Bitcoin Ordinals, BRC-20 inscriptions, Arweave, and Solana content. Raw inscription metadata, content analysis, and training-ready text for blockchain AI research.
Dataset Summary
Metric
Value
Chains
Bitcoin, Arweave, Solana
Unified Records
135 (bitcoin: 127, arweave: 5, solana: 3)
Training Records
100 (text-based inscriptions)
Content Types
PNG, WebP, GIF, SVG, HTML, MP3, MP4, JSON… See the full description on the dataset page: https://huggingface.co/datasets/purplesquirrelnetworks/blockchain-research-training.Blockchain-OSINT-Reasoning-v1.2blockchain-datasets-for-QandArobotic_blockchain
Robotic Blockchain Dataset
Dataset ini berisi contoh dialog interaktif antara robot otonom dengan elemen blockchain dan DePIN (Decentralized Physical Infrastructure Network).
Fokus utama:
Execution task robot + claim reward via smart contract
Verifikasi data sensor/trajectories on-chain
Koordinasi swarm robot secara decentralized
Inspirasi dari proyek real seperti FrodoBots, Rice AI, NATIX, Peaq, IoTeX
Format: Chat messages (system-user-assistant) — cocok untuk fine-tune LLM jadi… See the full description on the dataset page: https://huggingface.co/datasets/zianrahmad/robotic_blockchain.arxiv_blockchain_crypto_papersCrypto_Blockchain_QA_SFTblockchain-tokenization-qaarxiv_blockchain_crypto_papers_semanticCryptoXChain_500K_Multi_Network_Blockchain_Transaction_Dataset
🔗 CryptoXChain-500K: Multi-Network Blockchain Transaction Dataset
Dataset Summary
A large-scale, multi-chain blockchain transaction dataset containing
500,000 real transactions sampled across 5 cryptocurrency networks:
Bitcoin (BTC), Bitcoin Cash (BCH), Dash (DASH), Dogecoin (DOGE), and
Ethereum Classic (ETC). Structured for immediate use in machine learning,
financial analytics, anomaly detection, graph neural networks, and
cross-chain comparative research.
Collected… See the full description on the dataset page: https://huggingface.co/datasets/Omarrran/CryptoXChain_500K_Multi_Network_Blockchain_Transaction_Dataset.blockchain_newsenglish-blockchain-basics-30fluaro-blockchainblockchain-tokenization-qa_datasblockchain-tokenization-datas_for_QAblockchain-modular-calculator-datasetsblockchain-datasets-modular-calculator
