Search prompts for Stable Diffusion, ChatGPT & Midjourney

In this repository, we present Wan2.1, a comprehensive and open suite of video foundation models that pushes the boundaries of video generation. Wan2.1 offers these key features:

👍 SOTA Performance: Wan2.1 consistently outperforms existing open-source models and state-of-the-art commercial solutions across multiple benchmarks.
👍 Supports Consumer-grade GPUs: The T2V-1.3B model requires only 8.19 GB VRAM, making it compatible with almost all consumer-grade GPUs. It can generate a 5-second 480P video on an RTX 4090 in about 4 minutes (without optimization techniques like quantization). Its performance is even comparable to some closed-source models.
👍 Multiple Tasks: Wan2.1 excels in Text-to-Video, Image-to-Video, Video Editing, Text-to-Image, and Video-to-Audio, advancing the field of video generation.
👍 Visual Text Generation: Wan2.1 is the first video model capable of generating both Chinese and English text, featuring robust text generation that enhances its practical applications.
👍 Powerful Video VAE: Wan-VAE delivers exceptional efficiency and performance, encoding and decoding 1080P videos of any length while preserving temporal information, making it an ideal foundation for video and image generation.

This repository features our T2V-14B model, which establishes a new SOTA performance benchmark among both open-source and closed-source models. It demonstrates exceptional capabilities in generating high-quality visuals with significant motion dynamics. It is also the only video model capable of producing both Chinese and English text and supports video generation at both 480P and 720P resolutions.

What is Wan-AI 万相/ Wan2.1 Video Model (Safetensors) - Comfy&Kijai?

Wan-AI 万相/ Wan2.1 Video Model (Safetensors) - Comfy&Kijai is a highly specialized Image generation AI Model of type Safetensors / Checkpoint AI Model created by AI community user METAFILM_Ai. Derived from the powerful Stable Diffusion (Other) model, Wan-AI 万相/ Wan2.1 Video Model (Safetensors) - Comfy&Kijai has undergone an extensive fine-tuning process, leveraging the power of a dataset consisting of images generated by other AI models or user-contributed data. This fine-tuning process ensures that Wan-AI 万相/ Wan2.1 Video Model (Safetensors) - Comfy&Kijai is capable of generating images that are highly relevant to the specific use-cases it was designed for, such as base model, video, basemodel.

With a rating of 0 and over 0 ratings, Wan-AI 万相/ Wan2.1 Video Model (Safetensors) - Comfy&Kijai is a popular choice among users for generating high-quality images from text prompts.

Can I download Wan-AI 万相/ Wan2.1 Video Model (Safetensors) - Comfy&Kijai?

Yes! You can download the latest version of Wan-AI 万相/ Wan2.1 Video Model (Safetensors) - Comfy&Kijai from here.

How to use Wan-AI 万相/ Wan2.1 Video Model (Safetensors) - Comfy&Kijai?

To use Wan-AI 万相/ Wan2.1 Video Model (Safetensors) - Comfy&Kijai, download the model checkpoint file and set up an UI for running Stable Diffusion models (for example, AUTOMATIC1111). Then, provide the model with a detailed text prompt to generate an image. Experiment with different prompts and settings to achieve the desired results. If this sounds a bit complicated, check out our initial guide to Stable Diffusion – it might be of help. And if you really want to dive deep into AI image generation and understand how set up AUTOMATIC1111 to use Safetensors / Checkpoint AI Models like Wan-AI 万相/ Wan2.1 Video Model (Safetensors) - Comfy&Kijai, check out our crash course in AI image generation.

Download (10.3 GB) Download available on desktop only

You'll need to use a program like A1111 to run this – learn how in our crash course

Popularity

660 ~10

Info

Base model: Wan Video

Version ComfyORG wan_2.1_vae: 1 File

wanAIWan21VideoModelSafetensors_comfyorgWan21Vae.safetensors (236 MB)

To download these files, please visit this page from a desktop computer.

About this version: ComfyORG wan_2.1_vae

Comfy-Org/Wan_2.1_ComfyUI_repackaged

Wan 2.1 repackaged for ComfyUI use. For examples see: https://comfyanonymous.github.io/ComfyUI_examples/wan

---

Wan 2.1 Models

Wan 2.1 is a family of video models.

---

Files to Download

You will first need:

Text encoder and VAE:

umt5_xxl_fp8_e4m3fn_scaled.safetensors goes in: ComfyUI/models/text_encoders/

wan_2.1_vae.safetensors goes in: ComfyUI/models/vae/

---

Video Models

files go in: ComfyUI/models/diffusion_models/

These examples use the 16 bit files but you can use the fp8 ones instead if you don’t have enough memory.

---

Workflows

---

Text to Video

Workflow in Json format

files (put it in: ComfyUI/models/diffusion_models/). You can also use it with the 14B model.

---

Image to Video

Workflow in Json format

Note this example only generates 33 frames at 512x512 because I wanted it to be accessible, the model can do more than that. The 720p model is pretty good if you have the hardware/patience to run it.