Skip to content

EthenEthenEthen

Open Source Model Profile · Efficient-Large-Model

Sana_1600M_1024px

Sana_1600M_1024px is a Sana text-to-image release from Efficient-Large-Model. According to the model card, it targets 1024px-based generation with high-resolution synthesis.

Publisher
Efficient-Large-Model
Task
text-to-image
Model type
Unknown
License
apache-2.0
Library
sana
Publication status
Approved for indexing

Model overview

Sana_1600M_1024px is published by Efficient-Large-Model as a text-to-image model. According to the model card, Sana is a Linear-Diffusion-Transformer framework developed by NVIDIA, and this checkpoint carries a 1,648M-parameter claim for 1024px-based generation. The captured record lists an Apache-2.0 license and the sana library.

Recorded capabilities

High-resolution synthesis claim

According to the model card, Sana can synthesize high-resolution images with text-image alignment at fast speed, including deployment on a laptop GPU.

1024px operating point

According to the model card, this checkpoint targets 1024px-based images with multi-scale height and width.

Research source links

According to the model card, source code is available through the NVlabs Sana repository, with a research-oriented generative-models repository recommended.

Use cases in the source record

  • Text-to-image generation at the documented 1024px operating point, with framework-level support claimed up to 4096x4096.
  • Research uses the card scopes to artworks, design, and other artistic processes, excluding factual representations of people or events.

Limitations and unknowns

  • No parameter count was extracted from structured metadata; the 1,648M figure is a publisher model-card claim.
  • No evaluation results were extracted from this record.
  • The card says the model was not trained for factual representations of people or events, so such generation is out of scope.
  • Provider state is historical snapshot data, not independently refreshed current availability.

Source and provenance

Source: Efficient-Large-Model/Sana_1600M_1024px

Captured: Unknown. Processed: 2026-09-07T19:34:30.598713+00:00.

🐱 Sana Model Card Model We introduce Sana , a text-to-image framework that can efficiently generate images up to 4096 × 4096 resolution. Sana can synthesize high-resolution, high-quality images with strong text-image alignment at a remarkably fast speed, deployable on laptop GPU. Source code is available at https://github.com/NVlabs/Sana . Model Description Developed by: NVIDIA, Sana Model type: Linear-Diffusion-Transformer-based text-to-image generative model Model size: 1648M parameters Model resolution: This model is developed to generate 1024px based images with multi-scale heigh and width. License: Apache License 2.0 . Additio…

F001F002F003F004F005F007F008F009F010F011F013F015F016F017