DFloat11 / Wan2.2-T2V-A14B-DF11

huggingface.co
Total runs: 1.8K
24-hour runs: 24
7-day runs: 130
30-day runs: 405
Model's Last Updated: August 02 2025
text-to-video

Introduction of Wan2.2-T2V-A14B-DF11

Model Details of Wan2.2-T2V-A14B-DF11

DFloat11 Compressed Model: Wan-AI/Wan2.2-T2V-A14B

This is a DFloat11 losslessly compressed version of the original Wan-AI/Wan2.2-T2V-A14B model. It reduces model size by 32% compared to the original BFloat16 model, while maintaining bit-identical outputs and supporting efficient GPU inference .

🔥🔥🔥 Thanks to DFloat11 compression, Wan-AI/Wan2.2-T2V-A14B can now generate a 5-second 720P video on a single 24GB GPU, while maintaining full model quality. 🔥🔥🔥

📊 Performance Comparison
Model Model Size Peak GPU Memory (5-second 720P generation) Generation Time (A100 GPU)
Wan-AI/Wan2.2-T2V-A14B (BFloat16) ~56 GB O.O.M. -
Wan-AI/Wan2.2-T2V-A14B (DFloat11) 19.46 + 19.39 GB 41.06 GB 42 minutes
Wan-AI/Wan2.2-T2V-A14B (DFloat11 + CPU Offloading) 19.46 + 19.39 GB 22.49 GB 44 minutes
🔍 How It Works

We apply Huffman coding to the exponent bits of BFloat16 model weights, which are highly compressible. We leverage hardware-aware algorithmic designs to enable highly efficient, on-the-fly weight decompression directly on the GPU. Find out more in our research paper .

🔧 How to Use
  1. Install or upgrade the DFloat11 pip package (installs the CUDA kernel automatically; requires a CUDA-compatible GPU and PyTorch installed) :

    pip install -U dfloat11[cuda12]
    
  2. Install the latest diffusers package from source:

    pip install git+https://github.com/huggingface/diffusers
    
  3. Save the following code to a Python file t2v.py :

    import time
    import torch
    import argparse
    from diffusers import WanPipeline, AutoencoderKLWan
    from diffusers.utils import export_to_video
    from dfloat11 import DFloat11Model
    
    # Set up argument parser
    parser = argparse.ArgumentParser(description='Run Wan2.2 T2V model with custom parameters')
    parser.add_argument('--prompt', type=str, default="A serene koi pond at night, with glowing lanterns reflecting on the rippling water. Ethereal fireflies dance above as cherry blossoms gently fall, creating a dreamlike atmosphere.",
                        help='Text prompt for video generation')
    parser.add_argument('--negative_prompt', type=str, default="色调艳丽,过曝,静态,细节模糊不清,字幕,风格,作品,画作,画面,静止,整体发灰,最差质量,低质量,JPEG压缩残留,丑陋的,残缺的,多余的手指,画得不好的手部,画得不好的脸部,畸形的,毁容的,形态畸形的肢体,手指融合,静止不动的画面,杂乱的背景,三条腿,背景人很多,倒着走",
                        help='Negative prompt for video generation')
    parser.add_argument('--width', type=int, default=1280, help='Width of output video')
    parser.add_argument('--height', type=int, default=720, help='Height of output video')
    parser.add_argument('--num_frames', type=int, default=81, help='Number of frames to generate')
    parser.add_argument('--guidance_scale', type=float, default=4.0, help='Guidance scale for first stage')
    parser.add_argument('--guidance_scale_2', type=float, default=3.0, help='Guidance scale for second stage')
    parser.add_argument('--num_inference_steps', type=int, default=40, help='Number of inference steps')
    parser.add_argument('--cpu_offload', action='store_true', help='Enable CPU offloading')
    parser.add_argument('--output', type=str, default='t2v_out.mp4', help='Output video file path')
    parser.add_argument('--fps', type=int, default=16, help='FPS of output video')
    
    args = parser.parse_args()
    
    # Initialize models
    vae = AutoencoderKLWan.from_pretrained("Wan-AI/Wan2.2-T2V-A14B-Diffusers", subfolder="vae", torch_dtype=torch.float32)
    pipe = WanPipeline.from_pretrained("Wan-AI/Wan2.2-T2V-A14B-Diffusers", vae=vae, torch_dtype=torch.bfloat16)
    
    # Load DFloat11 models
    DFloat11Model.from_pretrained(
        "DFloat11/Wan2.2-T2V-A14B-DF11",
        device="cpu",
        cpu_offload=args.cpu_offload,
        bfloat16_model=pipe.transformer,
    )
    DFloat11Model.from_pretrained(
        "DFloat11/Wan2.2-T2V-A14B-2-DF11",
        device="cpu",
        cpu_offload=args.cpu_offload,
        bfloat16_model=pipe.transformer_2,
    )
    
    pipe.enable_model_cpu_offload()
    
    start_time = time.time()
    # Generate video
    output = pipe(
        prompt=args.prompt,
        negative_prompt=args.negative_prompt,
        height=args.height,
        width=args.width,
        num_frames=args.num_frames,
        guidance_scale=args.guidance_scale,
        guidance_scale_2=args.guidance_scale_2,
        num_inference_steps=args.num_inference_steps,
    ).frames[0]
    print(f"Time taken: {time.time() - start_time:.2f} seconds")
    
    export_to_video(output, args.output, fps=args.fps)
    
    # Print memory usage
    max_memory = torch.cuda.max_memory_allocated()
    print(f"Max memory: {max_memory / (1000 ** 3):.2f} GB")
    
  4. To run without CPU offloading (40GB VRAM required):

    python t2v.py
    

    To run with CPU offloading (22.5GB VRAM required):

    python t2v.py --cpu_offload
    
📄 Learn More

Runs of DFloat11 Wan2.2-T2V-A14B-DF11 on huggingface.co

1.8K
Total runs
24
24-hour runs
77
3-day runs
130
7-day runs
405
30-day runs

More Information About Wan2.2-T2V-A14B-DF11 huggingface.co Model

Wan2.2-T2V-A14B-DF11 huggingface.co

Wan2.2-T2V-A14B-DF11 huggingface.co is an AI model on huggingface.co that provides Wan2.2-T2V-A14B-DF11's model effect (), which can be used instantly with this DFloat11 Wan2.2-T2V-A14B-DF11 model. huggingface.co supports a free trial of the Wan2.2-T2V-A14B-DF11 model, and also provides paid use of the Wan2.2-T2V-A14B-DF11. Support call Wan2.2-T2V-A14B-DF11 model through api, including Node.js, Python, http.

Wan2.2-T2V-A14B-DF11 huggingface.co Url

https://huggingface.co/DFloat11/Wan2.2-T2V-A14B-DF11

DFloat11 Wan2.2-T2V-A14B-DF11 online free

Wan2.2-T2V-A14B-DF11 huggingface.co is an online trial and call api platform, which integrates Wan2.2-T2V-A14B-DF11's modeling effects, including api services, and provides a free online trial of Wan2.2-T2V-A14B-DF11, you can try Wan2.2-T2V-A14B-DF11 online for free by clicking the link below.

DFloat11 Wan2.2-T2V-A14B-DF11 online free url in huggingface.co:

https://huggingface.co/DFloat11/Wan2.2-T2V-A14B-DF11

Wan2.2-T2V-A14B-DF11 install

Wan2.2-T2V-A14B-DF11 is an open source model from GitHub that offers a free installation service, and any user can find Wan2.2-T2V-A14B-DF11 on GitHub to install. At the same time, huggingface.co provides the effect of Wan2.2-T2V-A14B-DF11 install, users can directly use Wan2.2-T2V-A14B-DF11 installed effect in huggingface.co for debugging and trial. It also supports api for free installation.

Wan2.2-T2V-A14B-DF11 install url in huggingface.co:

https://huggingface.co/DFloat11/Wan2.2-T2V-A14B-DF11

Url of Wan2.2-T2V-A14B-DF11

Wan2.2-T2V-A14B-DF11 huggingface.co Url

Provider of Wan2.2-T2V-A14B-DF11 huggingface.co

DFloat11
ORGANIZATIONS

Other API from DFloat11

huggingface.co

Total runs: 2.5K
Run Growth: 281
Growth Rate: 11.46%
Updated:Juni 26 2025