autoscientist
chartwise-autoscientist-data
ChartWise AutoScientist
This is a deterministic synthetic source dataset generated for AutoScientist adaptation.
Task
Each row contains a chart image URL, a visual reasoning prompt, and an evidence-grounded completion.
The benchmark covers bar, line, grouped bar, stacked bar, and scatter plots.
Training Columns
image: chart image URL
prompt: question and answer format instruction
completion: direct answer and concise visual evidence
All other… See the full description on the dataset page: https://huggingface.co/datasets/doraking/chartwise-autoscientist-data.autoscientist-competition-datasetsfinmix-autoscientist-10k
FinMix AutoScientist 10k
A deterministic, upload-ready 10,000-row subset of
FinMix v1, created for
fast finance adaptation runs in the Adaption AutoScientist challenge.
Use with Adaption Adaptive Data
Import this Hugging Face dataset and map:
Prompt: prompt
Context: context
Completion: completion
Leave task_type, source, and group_key unmapped. They are retained for
provenance and auditing.
Fields
Field
Description
prompt
Financial… See the full description on the dataset page: https://huggingface.co/datasets/julian8897/finmix-autoscientist-10k.autoscientist-toolcaller-dataset
AutoScientist Tool-Calling Dataset
A curated function-calling / tool-use dataset for the Adaption AutoScientist Challenge. Its
distinguishing feature is a large slice of hard negatives and reliability-focused cases — where the
correct behavior is not a plain tool call.
Adaptive Data quality (real): on the fixed set (c4923b7f…, graded on 1,000 of 2,440 rows
under the free-tier cap) the platform reported 7.0 → 8.1, +15.7%, grade C → B — now confirmed by a
completed, uncapped run… See the full description on the dataset page: https://huggingface.co/datasets/pandeyankit84/autoscientist-toolcaller-dataset.autoscientist-healthcare-reasoning
🩺 Adapted Healthcare Clinical-Reasoning (AutoScientist)
Built with Adaptive Data by Adaption.
A grounded, safety-blueprinted clinical-reasoning dataset — and a rigorous,
fully-reproducible study of when data adaptation helps a small model, and when it doesn't.
📈 Adaptive Data quality
Before → After
Overall quality score
7.0 → 9.1 (+30%)
Quality grade
B → A
Completion quality
+37.9%
Message quality
+17.6%
Percentile vs. reference corpus
15.3 → 33.0… See the full description on the dataset page: https://huggingface.co/datasets/hetanshwaghela/autoscientist-healthcare-reasoning.gujarati-autoscientist-dataset
