datasets
Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.
trellis500k-sketchfab-archivestrellis500k-github-archives-10trellis500k-github-archives-9trellis500k-github-archives-5trellis500k-github-archives-4trellis500k-github-archives-8trellis500k-github-archives-6trellis500k-github-archives-3251103_soup_can_trellis_50_640_480_lighting_augmentedThis dataset was created using LeRobot.
Dataset Structure
meta/info.json:
{
"codebase_version": "v2.1",
"robot_type": "ffw_bg2",
"total_episodes": 300,
"total_frames": 49488,
"total_tasks": 1,
"total_videos": 900,
"total_chunks": 1,
"chunks_size": 1000,
"fps": 30,
"splits": {
"train": "0:300"},
"data_path": "data/chunk-{episode_chunk:03d}/episode_{episode_index:06d}.parquet",
"video_path":… See the full description on the dataset page: https://huggingface.co/datasets/kimyg119/251103_soup_can_trellis_50_640_480_lighting_augmented.trellismark-qwen3-4b
TrellisMark Qwen3-4B confirmation corpus
This is the frozen English confirmation corpus for
TrellisMark, an experimental
many-user AI-text watermark. It includes exact generated text and token IDs,
unwatermarked Qwen controls, public-key detector evidence, the public research
key, independent encoder vectors, and the result reports used for the
reader-facing curves. The standalone implementation, detector, and
reproduction instructions are in the
TrellisMark GitHub repository.… See the full description on the dataset page: https://huggingface.co/datasets/xlr8harder/trellismark-qwen3-4b.251105_soup_can_sim_50_lighting_augmented_trellisThis dataset was created using LeRobot.
Dataset Structure
meta/info.json:
{
"codebase_version": "v2.1",
"robot_type": "ffw_bg2",
"total_episodes": 300,
"total_frames": 115650,
"total_tasks": 1,
"total_videos": 900,
"total_chunks": 1,
"chunks_size": 1000,
"fps": 30,
"splits": {
"train": "0:300"},
"data_path": "data/chunk-{episode_chunk:03d}/episode_{episode_index:06d}.parquet",
"video_path":… See the full description on the dataset page: https://huggingface.co/datasets/kimyg119/251105_soup_can_sim_50_lighting_augmented_trellis.soup_can_trellis_50_640_480_lighting_augmentedThis dataset was created using LeRobot.
Dataset Structure
meta/info.json:
{
"codebase_version": "v2.1",
"robot_type": "ffw_bg2",
"total_episodes": 300,
"total_frames": 49488,
"total_tasks": 1,
"total_videos": 900,
"total_chunks": 1,
"chunks_size": 1000,
"fps": 30,
"splits": {
"train": "0:300"},
"data_path": "data/chunk-{episode_chunk:03d}/episode_{episode_index:06d}.parquet",
"video_path":… See the full description on the dataset page: https://huggingface.co/datasets/kimyg119/soup_can_trellis_50_640_480_lighting_augmented.trellismark-qwen3-4b-rephrasing
TrellisMark Qwen3-4B blind rephrasing corpus
This separate release contains the frozen detector-blind rephrasing experiment
for TrellisMark. Its 16,384
source documents are an exact document-ID-preserving subset of the
main TrellisMark Qwen3-4B confirmation corpus.
It publishes both rewriters' outcome records, retained rewrite text and token
IDs, aligned key-only and model-assisted evidence, and the reports behind the
article's countermeasure figures.
The rewriters received only… See the full description on the dataset page: https://huggingface.co/datasets/xlr8harder/trellismark-qwen3-4b-rephrasing.trellis-dual-contrast-flowedit-8gpu-ckpt-1-574
