wuu
Datasets
All datasets matching “wuu”TraingTangTian
✨ Spatial-TTT: Streaming Visual-based Spatial Intelligence with Test-Time Training ✨
Fangfu Liu*,1,
Diankun Wu*,1,
Jiawei Chi*,1,
Yimo Cai1,
Yi-Hsin Hung1,
Xumin Yu2,
Hao Li3,
Han Hu2,
Yongming Rao†,2,
Yueqi Duan†,1
*Equal Contribution †Corresponding Author
1Tsinghua University 2Tencent Hunyuan 3NTU
Spatial-TTT: We propose… See the full description on the dataset page: https://huggingface.co/datasets/Wuuu3511/TraingTangTian.MMRC_Real_World_Conversation
MMRC - Multi-Modal Open-Ended Conversation Dataset
Overview:
MMRC is a benchmark dataset designed for evaluating Multi-Modal Large Language Models (MLLMs) in open-ended, multi-turn conversations. It provides diverse, real-world conversational data that integrates both textual and visual modalities, aiming to push the boundaries of MLLM performance in practical settings.
Dataset Details:
The MMRC dataset is composed of multi-turn conversations with integrated… See the full description on the dataset page: https://huggingface.co/datasets/WUUE/MMRC_Real_World_Conversation.bmvs_sparse_dtu2my-personal-codex-data
Coding Agent Conversation Logs
This is a performance art project. Anthropic built their models on the world's freely shared information, then introduced increasingly dystopian data policies to stop anyone else from doing the same with their data - pulling up the ladder behind them. DataClaw lets you throw the ladder back down. The dataset it produces is yours to share.
Exported with DataClaw.
Tag: dataclaw - Browse all DataClaw datasets
Stats
Metric
Value… See the full description on the dataset page: https://huggingface.co/datasets/wuuski/my-personal-codex-data.VeriOS-Bench
VeriOS-Bench
Paper | Code
Dataset Overview
This dataset is a Query-Driven Trustworthy OS Agent Dataset implemented as described in the paper:
VeriOS: Query-Driven Proactive Human-Agent-GUI Interaction for Trustworthy OS Agents
The VeriOS dataset is designed to train and evaluate operating system (OS) agents that automate tasks through on-device graphical user interfaces (GUIs). It focuses on enabling agents to decide when to query humans for more reliable task completion… See the full description on the dataset page: https://huggingface.co/datasets/wuuuuuz/VeriOS-Bench.GUI-CIDER
