CoolFace
Apppublic

BobSaget777/caveofwonders777cpu

sourceHugging Faceapache-2.0updated 11mo agoView on Hugging Face
0likes
App README

Wan2.2 Video Generation ๐ŸŽฅ

Generate high-quality videos from text prompts or images using the powerful Wan2.2-TI2V-5B model!

This Space provides an easy-to-use interface for creating videos with state-of-the-art AI technology.

Features โœจ

  • โ€”Text-to-Video: Generate videos from descriptive text prompts
  • โ€”Image-to-Video: Animate your images by adding an input image
  • โ€”High Quality: 720P resolution at 24fps
  • โ€”Customizable: Adjust resolution, number of frames, guidance scale, and more
  • โ€”Reproducible: Use seeds to recreate your favorite generations

Model Information ๐Ÿค–

Wan2.2-TI2V-5B is a unified text-to-video and image-to-video generation model with:

  • โ€”5 billion parameters optimized for consumer-grade GPUs
  • โ€”720P resolution support (1280x704 default)
  • โ€”24 fps smooth video output
  • โ€”Optimized duration: Default 3 seconds (optimized for Zero GPU limits)

The model uses a Mixture-of-Experts (MoE) architecture and delivers outstanding video generation quality, surpassing many commercial models.

How to Use ๐Ÿš€

Text-to-Video Generation

  1. 1.Enter your prompt describing the video you want to create
  2. 2.Adjust settings in "Advanced Settings" if desired
  3. 3.Click "Generate Video"
  4. 4.Wait for generation (typically 2-3 minutes on Zero GPU with default settings)

Image-to-Video Generation

  1. 1.Upload an input image
  2. 2.Enter a prompt describing how the image should animate
  3. 3.Click "Generate Video"
  4. 4.The output will maintain the aspect ratio of your input image
  5. 5.Generation takes 2-3 minutes with optimized settings

Advanced Settings โš™๏ธ

  • โ€”Width/Height: Video resolution (default: 1280x704)
  • โ€”Number of Frames: Longer videos need more frames (default: 73 frames โ‰ˆ 3 seconds, max: 145)
  • โ€”Inference Steps: More steps = better quality but slower (default: 35, optimized for speed)
  • โ€”Guidance Scale: How closely to follow the prompt (default: 5.0)
  • โ€”Seed: Set a specific seed for reproducible results

Note: Settings are optimized to complete within Zero GPU's 3-minute time limit for Pro users.

Tips for Best Results ๐Ÿ’ก

  1. 1.Detailed Prompts: Be specific about what you want to see
  2. 2.Good: "Two anthropomorphic cats in comfy boxing gear fight on stage with dramatic lighting"
  3. 3.Basic: "cats fighting"
  1. 1.Image-to-Video: Use clear, high-quality input images that match your prompt
  1. 1.Quality vs Speed (optimized for Zero GPU limits):
  2. 2.Fast: 25-30 steps (~2 minutes)
  3. 3.Balanced: 35 steps (default, ~2-3 minutes)
  4. 4.Higher Quality: 40-50 steps (~3+ minutes, may timeout)
  1. 1.Experiment: Try different guidance scales:
  2. 2.Lower (3-4): More creative, less literal
  3. 3.Default (5): Good balance
  4. 4.Higher (7-10): Strictly follows prompt

Example Prompts ๐Ÿ“

  • โ€”"Two anthropomorphic cats in comfy boxing gear fight on stage"
  • โ€”"A serene underwater scene with colorful coral reefs and tropical fish swimming gracefully"
  • โ€”"A bustling futuristic city at night with neon lights and flying cars"
  • โ€”"A peaceful mountain landscape with snow-capped peaks and a flowing river"
  • โ€”"An astronaut riding a horse through a nebula in deep space"
  • โ€”"A dragon flying over a medieval castle at sunset"

Technical Details ๐Ÿ”ง

  • โ€”Model: Wan-AI/Wan2.2-TI2V-5B-Diffusers
  • โ€”Framework: Hugging Face Diffusers
  • โ€”Backend: PyTorch with bfloat16 precision
  • โ€”GPU: Hugging Face Zero GPU (H200 with 70GB VRAM, automatically allocated)
  • โ€”GPU Duration: 180 seconds (3 minutes) for Pro users
  • โ€”Generation Time: ~2-3 minutes with optimized settings (73 frames, 35 steps)

Limitations โš ๏ธ

  • โ€”Generation requires compute time (2-3 minutes with default settings)
  • โ€”Zero GPU allocation is time-limited (3 minutes for Pro, 60 seconds for Free)
  • โ€”Videos longer than 6 seconds (145 frames) may timeout
  • โ€”Higher quality settings (50+ steps) may timeout on Zero GPU
  • โ€”Complex scenes with many objects may be challenging

Credits ๐Ÿ™

License ๐Ÿ“„

This Space uses the Wan2.2 model which is released under Apache 2.0 license.

Related Links ๐Ÿ”—


Note: This is a community-created Space for easy access to Wan2.2 video generation. Generation times may vary based on current GPU availability.