Alpha-channel video
According to the model card, Wan-Alpha generates transparent video by jointly learning RGB and alpha channels.
Open Source Model Profile · htdong
Wan-Alpha is a text-to-video model from htdong for RGBA generation with an alpha channel. According to the model card, it jointly learns RGB and alpha and adapts Wan2.1-T2V-14B weights.
Wan-Alpha is published by htdong as a text-to-video model. Tags identify rgba, transparency, and base_model Wan-AI/Wan2.1-T2V-14B with finetune lineage. According to the model card, version 1.0 open-sources Wan2.1-14B-T2V-adapted weights and inference code for transparent video.
According to the model card, Wan-Alpha generates transparent video by jointly learning RGB and alpha channels.
Tags identify base_model Wan-AI/Wan2.1-T2V-14B; according to the model card, v1.0 adapts those weights with a VAE encoding alpha into RGB latent space.
According to the model card, prompts support both Chinese and English input with explicit transparency and shot-type wording.
Captured metadata records an apache-2.0 license with text-to-video pipeline tagging.
Source: htdong/Wan-Alpha
Captured: Unknown. Processed: 2026-09-07T19:35:56.950040+00:00.
Wan-Alpha Wan-Alpha: High-Quality Text-to-Video Generation with Alpha Channel Qualitative results of video generation using Wan-Alpha . Our model successfully generates various scenes with accurate and clearly rendered transparency. Notably, it can synthesize diverse semi-transparent objects, glowing effects, and fine-grained details such as hair. Abstract RGBA video generation, which includes an alpha channel to represent transparency, is gaining increasing attention across a wide range of applications. However, existing methods often neglect visual quality, limiting their practical usability. In this paper, we propose Wan-Alpha, a…
F001F002F003F004F005F006F007F008F010F011F013