datasets
Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.
details_inbox225710___model_llama_3_8B_Instruct_fine_tuned_xMR_1edetails_pinkyponky__SOLAR-10.7B-dpo-instruct-tuned-v0.1
Dataset Card for Evaluation run of pinkyponky/SOLAR-10.7B-dpo-instruct-tuned-v0.1
Dataset automatically created during the evaluation run of model pinkyponky/SOLAR-10.7B-dpo-instruct-tuned-v0.1 on the Open LLM Leaderboard.
The dataset is composed of 63 configuration, each one coresponding to one of the evaluated task.
The dataset has been created from 1 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard-old/details_pinkyponky__SOLAR-10.7B-dpo-instruct-tuned-v0.1.SEA-Instruct-2602-fine-tunedarxiv-physics-instruct-tune-30kArtifactAI_arxiv-physics-instruct-tune-30k_formated
Dataset Card for "ArtifactAI_arxiv-physics-instruct-tune-30k_formated"
More Information needed
arxiv-physics-instruct-tune-30k-unslothdetails_pinkyponky__Mistral-7B-Instruct-sft-tuned-v0.2
Dataset Card for Evaluation run of pinkyponky/Mistral-7B-Instruct-Sft-Tuned-V0.2
Dataset automatically created during the evaluation run of model pinkyponky/Mistral-7B-Instruct-Sft-Tuned-V0.2 on the Open LLM Leaderboard.
The dataset is composed of 63 configuration, each one coresponding to one of the evaluated task.
The dataset has been created from 2 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard-old/details_pinkyponky__Mistral-7B-Instruct-sft-tuned-v0.2.arxiv-cs-ml-instruct-tune-50kmwp-instruct-tune-dataset-splittedinstruct-tuned-derainSEA-Instruct-2602-fine-tunedarxiv-physics-instruct-tune-30k_filteredmwp-instruct-tune-dataset
Dataset Card for "mwp-instruct-tune-dataset"
More Information needed
arxiv-physics-instruct-tune-30k_filtered_formatedPhi-3.5-mini-instruct_probe_tune_hhexphi_probearXiv-full-text-synthetic-instruct-tune
Dataset Card for "arXiv-full-text-synthetic-instruct-tune"
More Information needed
ecva_instruct_tuned
/hub_data3/seohyun/saves/ecva_instruct/full/sft · happy8825/valid_ecva_clean results
Model: /hub_data3/seohyun/saves/ecva_instruct/full/sft
Dataset: happy8825/valid_ecva_clean
Generated: 2025-12-16 01:24:28Z
Metrics
Metric
Value
Total samples
924
With GT
0
Parsed answers
0
Top-1 accuracy
0
Recall@5
0
MRR
0
The uploaded JSON contains full per-sample predictions produced via t3_infer_with_vllm.bash.
EVQA/ECVA Metrics
Metric… See the full description on the dataset page: https://huggingface.co/datasets/happy8825/ecva_instruct_tuned.Qwen2.5-3B-Instruct_adaptive_tune_no_ref_hhexphi_probeInstructTuneDatasetQwen2.5-3B-Instruct_adaptive_tune_2e4_hhexphi_lrQwen2.5-3B-Instruct_adaptive_tune_2e5_prefill_hhexphi_lrQwen2.5-3B-Instruct_adaptive_tune_2e6_hhexphi_lrQwen2.5-3B-Instruct_adaptive_tune_2e7_hhexphi_lrQwen2.5-3B-Instruct_adaptive_tune_2e7_prefill_hhexphi_lrQwen2.5-3B-Instruct_probe_tune_hhexphi_probeQwen2.5-3B-Instruct_adaptive_tune_2e5_hhexphi_lrQwen2.5-3B-Instruct_adaptive_tune_2e6_prefill_hhexphi_lrMeta-Llama-3.1-8B-Instruct_probe_tune_hhexphi_probeQwen2.5-7B-Instruct_probe_tune_hhexphi_probeLlama-3.2-3B-Instruct_probe_tune_hhexphi_probe
