High-resolution synthesis claim
According to the model card, Sana can synthesize high-resolution images with text-image alignment at fast speed, including deployment on a laptop GPU.
Open Source Model Profile · Efficient-Large-Model
Sana_1600M_1024px is a Sana text-to-image release from Efficient-Large-Model. According to the model card, it targets 1024px-based generation with high-resolution synthesis.
Sana_1600M_1024px is published by Efficient-Large-Model as a text-to-image model. According to the model card, Sana is a Linear-Diffusion-Transformer framework developed by NVIDIA, and this checkpoint carries a 1,648M-parameter claim for 1024px-based generation. The captured record lists an Apache-2.0 license and the sana library.
According to the model card, Sana can synthesize high-resolution images with text-image alignment at fast speed, including deployment on a laptop GPU.
According to the model card, this checkpoint targets 1024px-based images with multi-scale height and width.
According to the model card, source code is available through the NVlabs Sana repository, with a research-oriented generative-models repository recommended.
Source: Efficient-Large-Model/Sana_1600M_1024px
Captured: Unknown. Processed: 2026-09-07T19:34:30.598713+00:00.
🐱 Sana Model Card Model We introduce Sana , a text-to-image framework that can efficiently generate images up to 4096 × 4096 resolution. Sana can synthesize high-resolution, high-quality images with strong text-image alignment at a remarkably fast speed, deployable on laptop GPU. Source code is available at https://github.com/NVlabs/Sana . Model Description Developed by: NVIDIA, Sana Model type: Linear-Diffusion-Transformer-based text-to-image generative model Model size: 1648M parameters Model resolution: This model is developed to generate 1024px based images with multi-scale heigh and width. License: Apache License 2.0 . Additio…
F001F002F003F004F005F007F008F009F010F011F013F015F016F017