datasets
Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.
ac-transit-apc
AC Transit Automatic Passenger Counter Records, 2019-2026
Stop-level boarding and alighting counts for the AC Transit bus network in
Alameda and Contra Costa counties, California, from January 2019 through
May 2026. The records come from the automatic passenger counters (APCs)
mounted at the doors of the buses: one row per stop event, with the number of
passengers who got on, the number who got off, and the load the bus left with.
89 monthly Parquet files, ~5.9 GB, partitioned… See the full description on the dataset page: https://huggingface.co/datasets/somemone/ac-transit-apc.koch_move_block_with_some_positionsThis dataset was created using LeRobot.
Dataset Structure
meta/info.json:
{
"codebase_version": "v2.0",
"robot_type": "koch",
"total_episodes": 60,
"total_frames": 22569,
"total_tasks": 1,
"total_videos": 120,
"total_chunks": 1,
"chunks_size": 1000,
"fps": 30,
"splits": {
"train": "0:60"
},
"data_path": "data/chunk-{episode_chunk:03d}/episode_{episode_index:06d}.parquet",
"video_path":… See the full description on the dataset page: https://huggingface.co/datasets/shin1107/koch_move_block_with_some_positions.something_something_v2_lerobotThis dataset was created using LeRobot.
Dataset Structure
meta/info.json:
{
"codebase_version": "v3.0",
"robot_type": "unknown",
"total_episodes": 220847,
"total_frames": 10008697,
"total_tasks": 124032,
"chunks_size": 1000,
"data_files_size_in_mb": 100,
"video_files_size_in_mb": 200,
"fps": 12,
"splits": {
"train": "0:220847"
},
"data_path": "data/chunk-{chunk_index:03d}/file-{file_index:03d}.parquet",
"video_path":… See the full description on the dataset page: https://huggingface.co/datasets/n3puiol/something_something_v2_lerobot.Q2CRBench-3
Q2CRBench-3
Q2CRBench-3 is a benchmark dataset designed to evaluate the performance of LLM in generating clinical recommendations. It is derived from the development records of three authoritative clinical guidelines: the 2020 EAN guideline for dementia, the 2021 ACR guideline for rheumatoid arthritis, and the 2024 KDIGO guideline for chronic kidney disease.
Due to copyright restrictions, we are unable to provide the screened records from the 2020 EAN Dementia and 2021 ACR RA… See the full description on the dataset page: https://huggingface.co/datasets/somewordstoolate/Q2CRBench-3.pickapic_v2_only_somesometimesanotion__Qwenvergence-14B-v12-Prose-DS-details
Dataset Card for Evaluation run of sometimesanotion/Qwenvergence-14B-v12-Prose-DS
Dataset automatically created during the evaluation run of model sometimesanotion/Qwenvergence-14B-v12-Prose-DS
The dataset is composed of 38 configuration(s), each one corresponding to one of the evaluated task.
The dataset has been created from 1 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard/sometimesanotion__Qwenvergence-14B-v12-Prose-DS-details.koch_move_block_with_some_shapesThis dataset was created using LeRobot.
Dataset Structure
meta/info.json:
{
"codebase_version": "v2.0",
"robot_type": "koch",
"total_episodes": 60,
"total_frames": 17980,
"total_tasks": 1,
"total_videos": 120,
"total_chunks": 1,
"chunks_size": 1000,
"fps": 30,
"splits": {
"train": "0:60"
},
"data_path": "data/chunk-{episode_chunk:03d}/episode_{episode_index:06d}.parquet",
"video_path":… See the full description on the dataset page: https://huggingface.co/datasets/shin1107/koch_move_block_with_some_shapes.sometimesanotion__Qwen-2.5-14B-Virmarckeoso-details
Dataset Card for Evaluation run of sometimesanotion/Qwen-2.5-14B-Virmarckeoso
Dataset automatically created during the evaluation run of model sometimesanotion/Qwen-2.5-14B-Virmarckeoso
The dataset is composed of 38 configuration(s), each one corresponding to one of the evaluated task.
The dataset has been created from 1 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always pointing… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard/sometimesanotion__Qwen-2.5-14B-Virmarckeoso-details.sometimesanotion__Qwenvergence-14B-v0.6-004-model_stock-details
Dataset Card for Evaluation run of sometimesanotion/Qwenvergence-14B-v0.6-004-model_stock
Dataset automatically created during the evaluation run of model sometimesanotion/Qwenvergence-14B-v0.6-004-model_stock
The dataset is composed of 38 configuration(s), each one corresponding to one of the evaluated task.
The dataset has been created from 1 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train"… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard/sometimesanotion__Qwenvergence-14B-v0.6-004-model_stock-details.sometimesanotion__Qwentinuum-14B-v7-details
Dataset Card for Evaluation run of sometimesanotion/Qwentinuum-14B-v7
Dataset automatically created during the evaluation run of model sometimesanotion/Qwentinuum-14B-v7
The dataset is composed of 38 configuration(s), each one corresponding to one of the evaluated task.
The dataset has been created from 1 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always pointing to the latest… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard/sometimesanotion__Qwentinuum-14B-v7-details.sometimesanotion__Qwen-14B-ProseStock-v4-details
Dataset Card for Evaluation run of sometimesanotion/Qwen-14B-ProseStock-v4
Dataset automatically created during the evaluation run of model sometimesanotion/Qwen-14B-ProseStock-v4
The dataset is composed of 38 configuration(s), each one corresponding to one of the evaluated task.
The dataset has been created from 1 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always pointing to… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard/sometimesanotion__Qwen-14B-ProseStock-v4-details.sometimesanotion__LamarckInfusion-14B-v2-details
Dataset Card for Evaluation run of sometimesanotion/LamarckInfusion-14B-v2
Dataset automatically created during the evaluation run of model sometimesanotion/LamarckInfusion-14B-v2
The dataset is composed of 38 configuration(s), each one corresponding to one of the evaluated task.
The dataset has been created from 1 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always pointing to… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard/sometimesanotion__LamarckInfusion-14B-v2-details.stack-overflow-datasetContraStylesRecapsometimesanotion__Qwen2.5-14B-Vimarckoso-v3-model_stock-details
Dataset Card for Evaluation run of sometimesanotion/Qwen2.5-14B-Vimarckoso-v3-model_stock
Dataset automatically created during the evaluation run of model sometimesanotion/Qwen2.5-14B-Vimarckoso-v3-model_stock
The dataset is composed of 38 configuration(s), each one corresponding to one of the evaluated task.
The dataset has been created from 1 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train"… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard/sometimesanotion__Qwen2.5-14B-Vimarckoso-v3-model_stock-details.sometimesanotion__Qwenvergence-14B-v13-Prose-DS-details
Dataset Card for Evaluation run of sometimesanotion/Qwenvergence-14B-v13-Prose-DS
Dataset automatically created during the evaluation run of model sometimesanotion/Qwenvergence-14B-v13-Prose-DS
The dataset is composed of 38 configuration(s), each one corresponding to one of the evaluated task.
The dataset has been created from 2 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard/sometimesanotion__Qwenvergence-14B-v13-Prose-DS-details.bi_so101_some_tasksThis dataset was created using LeRobot.
Dataset Structure
meta/info.json:
{
"codebase_version": "v3.0",
"fps": 60,
"features": {
"action": {
"dtype": "float32",
"names": [
"left_shoulder_pan.pos",
"left_shoulder_lift.pos",
"left_elbow_flex.pos",
"left_wrist_flex.pos",
"left_wrist_roll.pos",
"left_gripper.pos"… See the full description on the dataset page: https://huggingface.co/datasets/wuc1/bi_so101_some_tasks.eval_act_koch_move_block_with_some_shapesThis dataset was created using LeRobot.
Dataset Structure
meta/info.json:
{
"codebase_version": "v2.0",
"robot_type": "koch",
"total_episodes": 2,
"total_frames": 984,
"total_tasks": 1,
"total_videos": 4,
"total_chunks": 1,
"chunks_size": 1000,
"fps": 30,
"splits": {
"train": "0:2"
},
"data_path": "data/chunk-{episode_chunk:03d}/episode_{episode_index:06d}.parquet",
"video_path":… See the full description on the dataset page: https://huggingface.co/datasets/shin1107/eval_act_koch_move_block_with_some_shapes.some-task-582d4e
some-task-582d4e
Synthetic weather test data: 48 rows in data.csv.
All values are randomly generated fictional examples, not real observations, products, or user activity. Intended only for CSV loading and pipeline tests; not suitable for scientific or business conclusions. Columns are sampled independently and do not model real-world correlations.
Fields
sample_id: random identifier for this generated sample.
row_id: sequential row number starting at 1.… See the full description on the dataset page: https://huggingface.co/datasets/browntimothy/some-task-582d4e.some-rp-v2This is a filtered subset of lemonilia/Elliquiy-Role-Playing-Forums_2023-04 that's been converted into a typical turn-based conversation format (the added conversations column). These are all 2-user chats, and the first user was assigned to human and the second to gpt.
eval_act_koch_move_block_with_some_shapes_different_positionThis dataset was created using LeRobot.
Dataset Structure
meta/info.json:
{
"codebase_version": "v2.0",
"robot_type": "koch",
"total_episodes": 1,
"total_frames": 709,
"total_tasks": 1,
"total_videos": 2,
"total_chunks": 1,
"chunks_size": 1000,
"fps": 30,
"splits": {
"train": "0:1"
},
"data_path": "data/chunk-{episode_chunk:03d}/episode_{episode_index:06d}.parquet",
"video_path":… See the full description on the dataset page: https://huggingface.co/datasets/shin1107/eval_act_koch_move_block_with_some_shapes_different_position.pixelprose-flowerseval_act_koch_move_block_with_some_shapes2This dataset was created using LeRobot.
Dataset Structure
meta/info.json:
{
"codebase_version": "v2.0",
"robot_type": "koch",
"total_episodes": 8,
"total_frames": 4137,
"total_tasks": 1,
"total_videos": 16,
"total_chunks": 1,
"chunks_size": 1000,
"fps": 30,
"splits": {
"train": "0:8"
},
"data_path": "data/chunk-{episode_chunk:03d}/episode_{episode_index:06d}.parquet",
"video_path":… See the full description on the dataset page: https://huggingface.co/datasets/shin1107/eval_act_koch_move_block_with_some_shapes2.screwdriver_attach_panel_rs_080125_20_e5_some_splitThis dataset was created using LeRobot.
Dataset Structure
meta/info.json:
{
"codebase_version": "v3.0",
"robot_type": "koch_screwdriver_follower",
"total_episodes": 4,
"total_frames": 783,
"total_tasks": 1,
"chunks_size": 1000,
"data_files_size_in_mb": 100,
"video_files_size_in_mb": 500,
"fps": 30,
"splits": {
"train": "0:4"
},
"data_path": "data/chunk-{chunk_index:03d}/file-{file_index:03d}.parquet",
"video_path":… See the full description on the dataset page: https://huggingface.co/datasets/jackvial/screwdriver_attach_panel_rs_080125_20_e5_some_split.eval_act_koch_move_block_with_some_shapes_siriusThis dataset was created using LeRobot.
Dataset Structure
meta/info.json:
{
"codebase_version": "v2.0",
"robot_type": "koch",
"total_episodes": 8,
"total_frames": 4670,
"total_tasks": 1,
"total_videos": 16,
"total_chunks": 1,
"chunks_size": 1000,
"fps": 30,
"splits": {
"train": "0:8"
},
"data_path": "data/chunk-{episode_chunk:03d}/episode_{episode_index:06d}.parquet",
"video_path":… See the full description on the dataset page: https://huggingface.co/datasets/shin1107/eval_act_koch_move_block_with_some_shapes_sirius.something2This dataset was created using LeRobot.
Dataset Structure
meta/info.json:
{
"codebase_version": "v2.1",
"robot_type": "so101_follower",
"total_episodes": 1,
"total_frames": 896,
"total_tasks": 1,
"total_videos": 1,
"total_chunks": 1,
"chunks_size": 1000,
"fps": 30,
"splits": {
"train": "0:1"
},
"data_path": "data/chunk-{episode_chunk:03d}/episode_{episode_index:06d}.parquet",
"video_path":… See the full description on the dataset page: https://huggingface.co/datasets/sid003/something2.someone
someone* 订单筛选结果
来源:/home/GRQ/lw_9/ordermsg_logs_20260910。每单两份文件:<日期>_<order_id>.llm_record.jsonl(该单全部模型调用,按时间排序)、<日期>_<order_id>.app.log(该单全部运行日志行)。共 58 单。
订单
日期
产品
回复轮
detect 次
跨度(分)
日志行
20260910_someone1125_test1_2694
20260910
psyche-link
13
8
7.2
392
20260910_someone115_test1_6967
20260910
psyche-link
10
7
6.4
319
20260910_someone1186_test1_1252
20260910
psyche-link
12
7
7.3
356
20260910_someone1286_test1_562
20260910
psyche-link
8
0
6.1
151… See the full description on the dataset page: https://huggingface.co/datasets/HYGGEhygge/someone.sometimesanotion__IF-reasoning-experiment-80-details
Dataset Card for Evaluation run of sometimesanotion/IF-reasoning-experiment-80
Dataset automatically created during the evaluation run of model sometimesanotion/IF-reasoning-experiment-80
The dataset is composed of 38 configuration(s), each one corresponding to one of the evaluated task.
The dataset has been created from 1 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard/sometimesanotion__IF-reasoning-experiment-80-details.coyo-700m-qwen3vl-8b-capsafrica-burundi-uganda-and-burundi-climate-data-for-some-stations-eab2fac0
Uganda and Burundi. Climate data for some stations | Africa (Burundi official open data)
76 rows - 1 Africa country - not-applicable - Repackaged by Electric Sheep Africa
TL;DR
This dataset packages one official CSV resource from Burundi as
ML-ready Parquet. The source file is the provenance boundary; all usable
indicators or tabular columns from the resource stay together in this repo.
About the source
Source: Uganda and Burundi. Climate data… See the full description on the dataset page: https://huggingface.co/datasets/electricsheepafrica/africa-burundi-uganda-and-burundi-climate-data-for-some-stations-eab2fac0.
