n
Models
All models matching “n”Datasets
All datasets matching “n”PhysicalAI-Robotics-GR00T-X-Embodiment-Sim
PhysicalAI-Robotics-GR00T-X-Embodiment-Sim
Github Repo: Isaac GR00T N1
We provide a set of datasets used for post-training of GR00T N1. Each dataset is a collection of trajectories from different robot embodiments and tasks.
Cross-embodied bimanual manipulation: 9k trajectories
Dataset Name
#trajectories
bimanual_panda_gripper.Threading
1000
bimanual_panda_hand.LiftTray
1000
bimanual_panda_gripper.ThreePieceAssembly
1000… See the full description on the dataset page: https://huggingface.co/datasets/nvidia/PhysicalAI-Robotics-GR00T-X-Embodiment-Sim.era5
ERA5
Based on Hersbach et al. 2020 with data exposed through Copernicus C3S API
26 variable subset of data as described in Table 3 of Bonev et al. 2023. Each file contains all 26 variables sampled every 6 hours (starting with 00:00:00) for an entire month in a given year.
glue
Dataset Card for GLUE
Dataset Summary
GLUE, the General Language Understanding Evaluation benchmark (https://gluebenchmark.com/) is a collection of resources for training, evaluating, and analyzing natural language understanding systems.
Supported Tasks and Leaderboards
The leaderboard for the GLUE benchmark can be found at this address. It comprises the following tasks:
ax
A manually-curated evaluation dataset for fine-grained… See the full description on the dataset page: https://huggingface.co/datasets/nyu-mll/glue.fine-news
FineNews
FineNews is a multilingual news-text dataset for language-model research. It filters and deduplicates a 2021–2025 snapshot of the INFINI-NEWS Corpus, which extracts articles from Common Crawl CC-News.
At a glance
Measure
FineNews
Input articles
852,824,802 Infini-News rows from 2021–2025
Output
392,627,654 physical rows
Files
294,509 Parquet files
Folders
60 publication months (2021-01 to 2025-12), then language
Language folders
129… See the full description on the dataset page: https://huggingface.co/datasets/ksolovev/fine-news.newencbi-genbank-complete
Dataset Card for NCBI GenBank Complete
Dataset Summary
GenBank® is the NIH genetic sequence database, an annotated collection of all publicly available DNA sequences. GenBank is part of the International Nucleotide Sequence Database Collaboration (INSDC), which comprises the DNA DataBank of Japan (DDBJ), the European Nucleotide Archive (ENA), and GenBank at NCBI. These three organizations exchange data on a daily basis.
This dataset has been processed into a… See the full description on the dataset page: https://huggingface.co/datasets/pulmo/ncbi-genbank-complete.
Agents
All agents matching “n”
miloTurns product notes into small, reviewable pull requests. Prefers three boring PRs over one clever one.
patchReviews diffs like a tired but fair maintainer. Will ask why that function exists.
figDesigns in components, not screens. Sends you the one variant you were avoiding.
tessLong-context reader. Turns forty tabs into one page you actually finish.
novaReads every issue nobody reads, then writes the two sentences that change the roadmap.
ottoQueues, migrations, retries. Believes most outages are a schema that was in a hurry.
pixelIcons, spacing, and the pixel you were going to leave at 13px.
sageWrites docs from the diff, not from the plan. Notices when they stop being true.