CoolFace
Agents
Live
Leaderboard
Models
Community
Search
Create
Alerts
Menu
9 results
rlsd
rlsd
Search
in
all
models
datasets
apps
agents
people
projects
Models
All models matching “rlsd”
SeongryongJung /
qwen3-8b-biology-rlsd-ema005
text-generation
transformers
0 likes
17 downloads
3mo ago
Hugging Face
SeongryongJung /
Qwen3-4B-Biology-RLSD-TR
text-generation
transformers
0 likes
16 downloads
3mo ago
Hugging Face
SeongryongJung /
Qwen3-8B-Biology-RLSD-TR
text-generation
transformers
1 likes
15 downloads
3mo ago
Hugging Face
ipfipfipf /
Qwen3.5-4B-sdpo-react-rlsd-multitask-arm1.1
image-text-to-text
transformers
0 likes
15 downloads
1mo ago
Hugging Face
darklord1611 /
rl_sdf_none_qwen3_30b_a3b
peft
0 likes
14 downloads
4mo ago
Hugging Face
SeongryongJung /
Qwen-8b-base-RLSD
text-generation
transformers
0 likes
14 downloads
3mo ago
Hugging Face
SeongryongJung /
Qwen3-4B-Chemistry-RLSD
text-generation
transformers
0 likes
14 downloads
3mo ago
Hugging Face
SeongryongJung /
qwen3-4b-chemistry-rlsd-ema005
text-generation
transformers
0 likes
14 downloads
3mo ago
Hugging Face
Datasets
All datasets matching “rlsd”
iieycx /
rlsd-train-MMFineReason-123K
RLSD Training Data This repository contains the training data for Self-Distilled RLVR (RLSD). The data is derived from the dataset released in MMFineReason: Closing the Multimodal Reasoning Gap via Open Data-Centric Methods. We processed the original data by keeping only the final conclusion after the </think> tag and removing the lengthy reasoning content before it. This makes the data more suitable as conclusion-only supervision for RLSD training. Data The main data… See the full description on the dataset page: https://huggingface.co/datasets/iieycx/rlsd-train-MMFineReason-123K.
visual-question-answering
100K<n<1M
0 likes
236 downloads
5mo ago
Hugging Face