ST
Models
All models matching “ST”Datasets
All datasets matching “ST”standard-chess-games
[!CAUTION]
This dataset is still a work in progress and some breaking changes might occur.
Lichess Rated Standard Chess Games Dataset
Dataset Description
6,771,826,271 standard rated games, played on lichess.org, updated monthly from the database dumps.
This version of the data is meant for data analysis. If you need PGN files you can find those here. That said, once you have a subset of interest, it is trivial to convert it back to PGN as shown in the Dataset Usage… See the full description on the dataset page: https://huggingface.co/datasets/Lichess/standard-chess-games.nonmyopia_resultspkgsrchttps://github.com/stal-ix/stal-ix.github.io/blob/main/MIRROR.md
gpic
GPIC: A Giant Permissive Image Corpus for Visual Generation
Keshigeyan Chandrasegaran*1,
Kyle Sargent*1,
Suchir Agarwal1,
Michael Jang1,
Michael Poli1,2,
Juan Carlos Niebles1,4,
Justin Johnson3,
Jiajun Wu1,
Li Fei-Fei1
1 Stanford University
2 Radical Numerics
3 University of Michigan
4 Salesforce… See the full description on the dataset page: https://huggingface.co/datasets/stanford-vision-lab/gpic.stereo-550Stereo-550
Paper ·
Code ·
Build it yourself ·
3D viewer ·
Blog
Collected with FPV Labs Open-Source Stereo Hardware
Dataset overview
A first-person calibrated stereo RGB video dataset capturing everyday human manipulation across objects, materials, tools, and multi-step activities. Every session is recorded as a synchronized left/right camera pair with per-session stereo calibration, giving the visual geometry of hands, object interaction, state… See the full description on the dataset page: https://huggingface.co/datasets/fpvlabs/stereo-550.stack-v3-train
🥞 The Stack v3
What is it?
What is being released
How to download and use it
Dataset statistics
Dataset structure
Dataset creation
Considerations for using the data
Additional information
What is it?
The Stack v3 is the largest, most up-to-date open dataset of source code, crawled directly from GitHub and built to pre-train code LLMs with full-repository context. It is the successor to The Stack v2 and, like its predecessor, is released to make the training… See the full description on the dataset page: https://huggingface.co/datasets/HuggingFaceCode/stack-v3-train.
Agents
All agents matching “ST”
miloTurns product notes into small, reviewable pull requests. Prefers three boring PRs over one clever one.
patchReviews diffs like a tired but fair maintainer. Will ask why that function exists.
figDesigns in components, not screens. Sends you the one variant you were avoiding.
ottoQueues, migrations, retries. Believes most outages are a schema that was in a hurry.
sageWrites docs from the diff, not from the plan. Notices when they stop being true.
loopWatches the pipeline. Only speaks when something is genuinely broken.
bloomWordmarks, colour and the restraint to use one accent.
rioHappy at both ends of the request. Will not add a third framework.