CoolFace
Datasetpublic

happynew111/haotian_data-GPS-verl-main-CL

AdaRFT: Adaptive Curriculum Reinforcement Finetuning 📢 New extension to verl! We propose an adaptive curriculum learning method for efficient and scalable reinforcement finetuning (RFT) of LLMs — now implemented as an extension to this repo. Efficient Reinforcement Finetuning via Adaptive Curriculum LearningTaiwei Shi†, Yiyang Wu†, Linxin Song†, Tianyi Zhou▽, Jieyu Zhao††University of Southern California, ▽University of Maryland[Paper] 🧠 Highlights Dynamically adapts… See the full description on the dataset page: https://huggingface.co/datasets/happynew111/haotian_data-GPS-verl-main-CL.

sourceHugging Faceupdated 1y agoView on Hugging Face
0likes190downloads
settings

This repository belongs to happynew111 on Hugging Face.

CoolFace never edits a repository it does not host. Visibility, licence, collaborators and gating are all managed at the source.

namehaotian_data-GPS-verl-main-CL
visibilitypublic
licencenot set
gatedno
ownerhappynew111
Account settings
happynew111/haotian_data-GPS-verl-main-CL · CoolFace