datasets
Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.
lm-eval-results-shyamieee-JARVIS-v2.0-private
Dataset Card for Evaluation run of shyamieee/JARVIS-v2.0
Dataset automatically created during the evaluation run of model shyamieee/JARVIS-v2.0
The dataset is composed of 62 configuration(s), each one corresponding to one of the evaluated task.
The dataset has been created from 2 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always pointing to the latest results.
An additional… See the full description on the dataset page: https://huggingface.co/datasets/nyu-dice-lab/lm-eval-results-shyamieee-JARVIS-v2.0-private.ovos-wake-word-bench-picovoice-jarvis
OVOS wake_word bench — picovoice-jarvis
Per-clip detection decisions predictions of the registered
OVOS Plugin Arena
wake_word fighters over
Picovoice/wake-word-benchmark.
One dedicated repo per modality; one dataset split per language; one JSONL
file per fighter under predictions/<lang>/<competitor_id>.jsonl. Rows follow
the arena §3.2 contract (pinned dataset_revision, plugin_version,
latency_ms). Produced by the reproducible benchmark script in the arena repo;
the arena's… See the full description on the dataset page: https://huggingface.co/datasets/OpenVoiceOS/ovos-wake-word-bench-picovoice-jarvis.ovos-wake-word-bench-synthetic-wakewords-hey_jarvis
OVOS wake_word bench — synthetic-wakewords-hey_jarvis
Per-clip detection decisions predictions of the registered
OVOS Plugin Arena
wake_word fighters over
OpenVoiceOS/synthetic-wakewords.
One dedicated repo per modality; one dataset split per language; one JSONL
file per fighter under predictions/<lang>/<competitor_id>.jsonl. Rows follow
the arena §3.2 contract (pinned dataset_revision, plugin_version,
latency_ms). Produced by the reproducible benchmark script in the arena repo;… See the full description on the dataset page: https://huggingface.co/datasets/OpenVoiceOS/ovos-wake-word-bench-synthetic-wakewords-hey_jarvis.jarvis-memory
JARVIS Shared Memory (MCP connector link)
The stable meeting point for Sir's JARVIS app, Arena (builder AI), and Grok.
This link never changes. Bookmark it / share it as the connector.
memory.json — shared memory: standing brief, facts, open items, log (MCP-flavored manifest included).
PROTOCOL.md — how any AI connects: read first, share back in chat, Arena applies writes.
Rules: this repo never holds API keys. Keep brief short (fits in prompts). Never rename/move.
jarvis-lfm2-dataset-v2
Jarvis LFM2 Dataset v2 — multi-tools + raisonnement
Exemples : 20328 (train 19314 / val 1014)
Tokens estimés : ~38,572,838
Objectif : chaînage multi- propre + raisonnement + gros problèmes
Séquences ≥3 tools : 3508
Domaines : {'tool_calling': 9529, 'reasoning': 5927, 'coding': 2746, 'cyber': 1972, 'general': 154}
Sources : 11 (traces Jarvis + 11 datasets HF)
Format : messages (system/user/assistant, observations en user 'Observation:')
Langue : français (traces) + anglais… See the full description on the dataset page: https://huggingface.co/datasets/jojonocode/jarvis-lfm2-dataset-v2.english-to-colloquial-tamil
English to Colloquial Tamil
"instruction":"Translate provided English text into colloquial Tamil."
"input": "Their players played well."
"output": "அவங்க players நல்லா விளையாண்டாங்க."
jarvis-reasoning-dpo
🧠 JARVIS Autonomous Reasoning & DPO Trajectories
Curated and engineered by Al-Amin Ahmed Ratan (Boss Ratan).
📌 Overview
This dataset contains verified, production-grade synthetic reasoning trajectories (Instruction -> Step-by-Step Architectural Thoughts -> Syntax-Validated Chosen Python Code -> Suboptimal Rejected Code) generated autonomously by Boss Ratan's JARVIS system.
💼 Commercial Licensing & Custom Batches
For custom enterprise dataset… See the full description on the dataset page: https://huggingface.co/datasets/RaTan23/jarvis-reasoning-dpo.Jarvis_Dataset
Dataset Card for Dataset Name
This dataset card aims to be a base template for new datasets. It has been generated using this raw template.
Dataset Details
Dataset Description
Curated by: [More Information Needed]
Funded by [optional]: [More Information Needed]
Shared by [optional]: [More Information Needed]
Language(s) (NLP): [More Information Needed]
License: [More Information Needed]
Dataset Sources [optional]… See the full description on the dataset page: https://huggingface.co/datasets/suyashkrishangarg/Jarvis_Dataset.JARVISJARVISJARVIS-trainEMOTIONAL-JARVISJarvis-Training-DataThis Dataset Trains a Model to behave like
Jarvis from the films
jarvis_datasetec-chatjarvis_datasetJARVIS_DFT_Exfoliation_Energy_DatasetJarvis-Training-DataThis Dataset Trains a Model to behave like
Jarvis from the films
jarvis-test-dsinterview_questionsJarvis-Conversationjarvis-dm-trainingjarvis-memoryjarvis-AIJARVIS_DFT_Superconducting_TcEMOTIONAL-JARVISJarvis-Training-DataThis Dataset Trains a Model to behave like
Jarvis from the films
rvrnm-jarvis-datasetjarvis-dataset
