CoolFace
Modelpublic

shrikant11/pokemon_text_to_image_2

sourceHugging Facecreativeml-openrail-mupdated 3y agoView on Hugging Face
1likes3downloads
Model Card

license: creativeml-openrail-m base_model: runwayml/stable-diffusion-v1-5 datasets:

  • —lambdalabs/pokemon-blip-captions tags:
  • —stable-diffusion
  • —stable-diffusion-diffusers
  • —text-to-image
  • —diffusers inference: true ---

Text-to-image finetuning - shrikant11/pokemontexttoimage2

This bla bla pipeline was finetuned from runwayml/stable-diffusion-v1-5 on the lambdalabs/pokemon-blip-captions dataset. Below are some example images generated with the finetuned pipeline using the following prompts: ['Pokemon with yellow eyes', 'Green colour pokemon', 'Blue colour pikacchu', 'Charlizzard', 'pikachu', 'dangerous looking pokemon']:

[image]

Pipeline usage

You can use the pipeline like so:

python
from diffusers import DiffusionPipeline
import torch

pipeline = DiffusionPipeline.from_pretrained("shrikant11/pokemon_text_to_image_2", torch_dtype=torch.float16)
prompt = "Pokemon with yellow eyes"
image = pipeline(prompt).images[0]
image.save("my_image.png")

Training info

These are the key hyperparameters used during training:

  • —Epochs: 1
  • —Learning rate: 1e-05
  • —Batch size: 1
  • —Gradient accumulation steps: 1
  • —Image resolution: 512
  • —Mixed-precision: None

More information on all the CLI arguments and the environment are available on your [wandb run page]().