Nguyen17/model_id_Dev_dataset_index_3_e10
010
TRL DDPO Model
This is a diffusion model that has been fine-tuned with reinforcement learning to guide the model outputs according to a value, function, or human feedback. The model can be used for image generation conditioned with text.
