datasets
Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.
tcga-brca-titan-idc-ilc
tcga-brca-titan-idc-ilc
1. Tổng quan
[CẦN ĐIỀN THỦ CÔNG: mục đích, ngữ cảnh tạo dataset]
Tổng số bản ghi (cộng tất cả manifest phát hiện được): 4228
Số manifest phát hiện được trong bộ nhớ: 3 (df, brca_df, full_df)
Repo HuggingFace: okbro1234/tcga-brca-titan-idc-ilc
2. Cấu trúc lưu trữ tại đích
/ # suy từ hàm `HfApi`
file.txt # suy từ hàm `HfApi`
lfs.bin # suy từ hàm `HfApi`
shard_{i}_of_5.bin # suy từ hàm `HfApi`
remote/file/path.h5 #… See the full description on the dataset page: https://huggingface.co/datasets/okbro1234/tcga-brca-titan-idc-ilc.titans_NPC
Titans - Pytorch
Unofficial implementation of Titans in Pytorch. Will also contain some explorations into architectures beyond their simple 1-4 layer MLP for the neural memory module, if it works well to any degree.
Paper review by Yannic
Quick Colab Run
Appreciation
Eryk for sharing his early experimental results with me, positive for 2 layer MLP
Install
$ pip install titans-pytorch
Usage
import torch
from titans_pytorch import… See the full description on the dataset page: https://huggingface.co/datasets/ChipYTY/titans_NPC.draco
DRACO: a Cross-Domain Benchmark for Deep Research Accuracy, Completeness, and Objectivity
The DRACO Benchmark consists of complex, open-ended research tasks with expert-curated rubrics for evaluating deep research systems. Tasks span 10 domains and require drawing on information sources from 40 countries. Each task is paired with a detailed, task-specific rubric featuring an average of ~40 evaluation criteria across four axes: factual accuracy, breadth and depth of analysis… See the full description on the dataset page: https://huggingface.co/datasets/titantv090/draco.titan-safety-check-catalog
Titan safety check catalog
This dataset packages Titan's shipped pre-trade safety-check catalog as small structured data:
gates.csv
gates.json
Related note: article-verifiable-receipts.md explains Titan's verifiable decision receipt boundary.
Each row includes:
order: cascade order from the shipped gate pipeline
gate_id: backend/API identifier
label: operator-facing display label
category: public Operator Guide group
applies_to: signal class from the shipped pipeline… See the full description on the dataset page: https://huggingface.co/datasets/favlo/titan-safety-check-catalog.bunnycore__Llama-3.1-8B-TitanFusion-Mix-details
Dataset Card for Evaluation run of bunnycore/Llama-3.1-8B-TitanFusion-Mix
Dataset automatically created during the evaluation run of model bunnycore/Llama-3.1-8B-TitanFusion-Mix
The dataset is composed of 38 configuration(s), each one corresponding to one of the evaluated task.
The dataset has been created from 1 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always pointing to the… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard/bunnycore__Llama-3.1-8B-TitanFusion-Mix-details.attack_on_titan_wikibunnycore__Gemma2-9B-TitanFusion-details
Dataset Card for Evaluation run of bunnycore/Gemma2-9B-TitanFusion
Dataset automatically created during the evaluation run of model bunnycore/Gemma2-9B-TitanFusion
The dataset is composed of 38 configuration(s), each one corresponding to one of the evaluated task.
The dataset has been created from 1 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always pointing to the latest… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard/bunnycore__Gemma2-9B-TitanFusion-details.adaption-marketing-optimized-neural-titans
Adaption Marketing Optimized Dataset - Neural Titans
Competition: Adaption AutoScientist Challenge ($50,000 Prize Pool)Track: MarketingTeam: Neural Titans (HackIndia)
Dataset Details
Metric
Value
Rows
5,000
Size
22.5 MB
Format
JSONL (instruction-tuning)
Pipeline Configuration
Recipes Applied
Deduplication - Removes duplicate and near-duplicate entries
Prompt Rephrasing - Diversifies prompt formulations for robust… See the full description on the dataset page: https://huggingface.co/datasets/rishini/adaption-marketing-optimized-neural-titans.test_deduplicated_datasetbunnycore__Phi-3.5-mini-TitanFusion-0.1-details
Dataset Card for Evaluation run of bunnycore/Phi-3.5-mini-TitanFusion-0.1
Dataset automatically created during the evaluation run of model bunnycore/Phi-3.5-mini-TitanFusion-0.1
The dataset is composed of 38 configuration(s), each one corresponding to one of the evaluated task.
The dataset has been created from 1 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always pointing to the… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard/bunnycore__Phi-3.5-mini-TitanFusion-0.1-details.jaspionjader__Kosmos-EVAA-v9-TitanFusion-Mix-8B-details
Dataset Card for Evaluation run of jaspionjader/Kosmos-EVAA-v9-TitanFusion-Mix-8B
Dataset automatically created during the evaluation run of model jaspionjader/Kosmos-EVAA-v9-TitanFusion-Mix-8B
The dataset is composed of 38 configuration(s), each one corresponding to one of the evaluated task.
The dataset has been created from 1 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard/jaspionjader__Kosmos-EVAA-v9-TitanFusion-Mix-8B-details.titan_poc_ds_1bunnycore__Llama-3.1-8B-TitanFusion-v3-details
Dataset Card for Evaluation run of bunnycore/Llama-3.1-8B-TitanFusion-v3
Dataset automatically created during the evaluation run of model bunnycore/Llama-3.1-8B-TitanFusion-v3
The dataset is composed of 38 configuration(s), each one corresponding to one of the evaluated task.
The dataset has been created from 1 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always pointing to the… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard/bunnycore__Llama-3.1-8B-TitanFusion-v3-details.Titan-L_QuickTrain-v4489,438 Instructions
Titan-configTitanbrainv1attack_on_titan_wiki_chinesetitanium27leaderboard-requests
