{"id":1937,"date":"2026-04-15T14:10:18","date_gmt":"2026-04-15T14:10:18","guid":{"rendered":"https:\/\/wiro.ai\/blog\/?p=1937"},"modified":"2026-09-27T21:36:51","modified_gmt":"2026-09-27T21:36:51","slug":"sana-1600m-1024px-6-prompt-tests","status":"publish","type":"post","link":"https:\/\/wiro.ai\/blog\/sana-1600m-1024px-6-prompt-tests\/","title":{"rendered":"Sana 1600M (1024px): 6 Prompt Tests"},"content":{"rendered":"<h2>Sana 1600M (1024px): 6 Prompt Tests<\/h2>\n<p><strong>Sana 1600M (1024px)<\/strong> was tested here as a practical text-to-image model for concept work, stylized scenes, and photographic lighting studies. The point was not to hunt for a single hero image. The six prompts put pressure on different weak spots: small text, rigid style direction, blended composition, poster layout, exact object counts, and detailed natural texture. Those are the moments where a polished-looking image can still fail a production brief.<\/p>\n<p>The model page on Wiro offers several Sana 1600M checkpoints, including 1024px, 2K, and 4K variants. This post uses <a href=\"https:\/\/wiro.ai\/models\/wiro\/text-to-image-sana\">Efficient-Large-Model\/Sana_1600M_1024px_BF16_diffusers on Wiro<\/a>. Sana&#8217;s original technical report describes a deeply compressed autoencoder, linear attention, and an efficient sampling design aimed at high-resolution generation. The implementation and checkpoints are also documented in the <a href=\"https:\/\/github.com\/NVlabs\/Sana\" target=\"_blank\" rel=\"noopener\">official Sana GitHub repository<\/a> and the <a href=\"https:\/\/arxiv.org\/abs\/2410.10629\" target=\"_blank\" rel=\"noopener\">Sana research paper<\/a>.<\/p>\n<h2>What the Sana 1600M (1024px) test checked<\/h2>\n<ul>\n<li><strong>Checkpoint:<\/strong> Efficient-Large-Model\/Sana_1600M_1024px_BF16_diffusers<\/li>\n<li><strong>Canvas:<\/strong> 1024 x 1024 pixels<\/li>\n<li><strong>Samples:<\/strong> one output per prompt<\/li>\n<li><strong>Steps:<\/strong> 25 for the product test; 22 for tests two through six<\/li>\n<li><strong>Guidance scale:<\/strong> 3.5<\/li>\n<li><strong>Negative prompt:<\/strong> bad, ugly, low quality, watermark, blurry, deformed<\/li>\n<\/ul>\n<p>These settings matter. At 1024px, the model has enough room to show material detail and layout, but it still has to decide how to spend that detail. Lower guidance can leave more room for the model&#8217;s own visual choices; it also means exact wording and strict counts should be treated as instructions to test, not guarantees. The Wiro model documentation lists the available inputs and their defaults, but it does not publish a current per-output run time or cost for this checkpoint. No timing or cost figure is claimed here rather than turning an example task record into a price promise.<\/p>\n<nav><strong>In this post:<\/strong> <a href=\"#product\">product text<\/a> | <a href=\"#pixel-art\">pixel art<\/a> | <a href=\"#double-exposure\">double exposure<\/a> | <a href=\"#poster\">poster typography<\/a> | <a href=\"#counting\">counting<\/a> | <a href=\"#wildlife\">wildlife<\/a><\/nav>\n<h2 id=\"product\">Test 1: Studio product shot and a simple label<\/h2>\n<figure>\n  <img loading=\"lazy\" decoding=\"async\" width=\"1024\" height=\"1024\" class=\"wp-image-1930\" src=\"https:\/\/wiro.ai\/blog\/wp-content\/uploads\/2026\/03\/sana-1.png\" alt=\"Sana 1600M 1024px studio product test with a matte black water bottle\" srcset=\"https:\/\/wiro.ai\/blog\/wp-content\/uploads\/2026\/03\/sana-1.png 1024w, https:\/\/wiro.ai\/blog\/wp-content\/uploads\/2026\/03\/sana-1-510x510.webp 510w, https:\/\/wiro.ai\/blog\/wp-content\/uploads\/2026\/03\/sana-1-900x900.webp 900w, https:\/\/wiro.ai\/blog\/wp-content\/uploads\/2026\/03\/sana-1-150x150.webp 150w, https:\/\/wiro.ai\/blog\/wp-content\/uploads\/2026\/03\/sana-1-768x768.webp 768w\" sizes=\"auto, (max-width: 1024px) 100vw, 1024px\" \/><figcaption>Prompt: Studio product photo of a matte black reusable water bottle on a white seamless background. The bottle has a clean white label with the exact text &#8216;AURORA&#8217; on top line and &#8216;SPARKLING WATER&#8217; on second line. Sans serif, centered. Softbox reflections. 85mm lens, sharp, high detail.<\/figcaption><\/figure>\n<p>The output tests two separate jobs at once: product-lighting realism and exact label text. The bottle, pale background, and softbox reflection direction make the image read as a usable product concept. The label sits in a plausible place, which is useful when the goal is an art-directed comp. The fragile part is the lettering. The prompt asks for two exact lines, but generated letters can soften, substitute characters, or look almost correct at a glance. Pick Sana 1600M (1024px) for bottle shapes, material cues, and a first layout. Add final packaging copy in a design tool.<\/p>\n<h2 id=\"pixel-art\">Test 2: Pixel art scene for style control<\/h2>\n<figure>\n  <img loading=\"lazy\" decoding=\"async\" width=\"1024\" height=\"1024\" class=\"wp-image-1931\" src=\"https:\/\/wiro.ai\/blog\/wp-content\/uploads\/2026\/03\/sana-2.png\" alt=\"Pixel art seaside town at night with a lighthouse\" srcset=\"https:\/\/wiro.ai\/blog\/wp-content\/uploads\/2026\/03\/sana-2.png 1024w, https:\/\/wiro.ai\/blog\/wp-content\/uploads\/2026\/03\/sana-2-510x510.webp 510w, https:\/\/wiro.ai\/blog\/wp-content\/uploads\/2026\/03\/sana-2-900x900.webp 900w, https:\/\/wiro.ai\/blog\/wp-content\/uploads\/2026\/03\/sana-2-150x150.webp 150w, https:\/\/wiro.ai\/blog\/wp-content\/uploads\/2026\/03\/sana-2-768x768.webp 768w\" sizes=\"auto, (max-width: 1024px) 100vw, 1024px\" \/><figcaption>Prompt: Pixel art scene of a tiny seaside town at night. A lighthouse beam sweeps across the ocean. Small boats in the harbor. 16-bit era style. Crisp edges. Limited palette of navy, cyan, and warm yellow.<\/figcaption><\/figure>\n<p>This is a tighter test of style language. The output uses the requested night palette, small-town subject, lighthouse beam, and blocky visual vocabulary well enough to read as a 16-bit-inspired scene. It also shows why constraints help: the limited colors and simple shapes give the model fewer chances to drift into a painted illustration. The risk appears when a pixel-art brief asks for photographic lighting, tiny signage, or too many separate objects. Use this checkpoint for game mood boards, scene studies, and background ideas. For sprite sheets or grids that need consistent tiles, verify each cell by hand.<\/p>\n<h2 id=\"double-exposure\">Test 3: Double exposure portrait and blended composition<\/h2>\n<figure>\n  <img loading=\"lazy\" decoding=\"async\" width=\"1024\" height=\"1024\" class=\"wp-image-1932\" src=\"https:\/\/wiro.ai\/blog\/wp-content\/uploads\/2026\/03\/sana-3.png\" alt=\"Double exposure portrait with a city skyline inside a cyclist silhouette\" srcset=\"https:\/\/wiro.ai\/blog\/wp-content\/uploads\/2026\/03\/sana-3.png 1024w, https:\/\/wiro.ai\/blog\/wp-content\/uploads\/2026\/03\/sana-3-510x510.webp 510w, https:\/\/wiro.ai\/blog\/wp-content\/uploads\/2026\/03\/sana-3-900x900.webp 900w, https:\/\/wiro.ai\/blog\/wp-content\/uploads\/2026\/03\/sana-3-150x150.webp 150w, https:\/\/wiro.ai\/blog\/wp-content\/uploads\/2026\/03\/sana-3-768x768.webp 768w\" sizes=\"auto, (max-width: 1024px) 100vw, 1024px\" \/><figcaption>Prompt: Double exposure portrait of a cyclist wearing a helmet. Inside the silhouette, a modern city skyline at sunset. Teal and orange color grade. Soft film grain. Clean edges. High detail.<\/figcaption><\/figure>\n<p>Double exposure asks the model to preserve a readable outer silhouette while merging a second scene inside it. The cyclist outline and internal skyline make the intended idea legible, while the teal-and-orange treatment keeps both image layers in one visual family. That is a better result than a generic portrait because the prompt states the compositing logic clearly. Still, this is not a precise masking workflow. If a campaign needs a specific skyline to end exactly at a helmet edge, use Sana for a direction and finish the composite with layers and masks.<\/p>\n<h2 id=\"poster\">Test 4: Minimal poster and typography<\/h2>\n<figure>\n  <img loading=\"lazy\" decoding=\"async\" width=\"1024\" height=\"1024\" class=\"wp-image-1933\" src=\"https:\/\/wiro.ai\/blog\/wp-content\/uploads\/2026\/03\/sana-4.png\" alt=\"Minimal foggy forest poster generated in the Sana 1600M 1024px test\" srcset=\"https:\/\/wiro.ai\/blog\/wp-content\/uploads\/2026\/03\/sana-4.png 1024w, https:\/\/wiro.ai\/blog\/wp-content\/uploads\/2026\/03\/sana-4-510x510.webp 510w, https:\/\/wiro.ai\/blog\/wp-content\/uploads\/2026\/03\/sana-4-900x900.webp 900w, https:\/\/wiro.ai\/blog\/wp-content\/uploads\/2026\/03\/sana-4-150x150.webp 150w, https:\/\/wiro.ai\/blog\/wp-content\/uploads\/2026\/03\/sana-4-768x768.webp 768w\" sizes=\"auto, (max-width: 1024px) 100vw, 1024px\" \/><figcaption>Prompt: Minimalist movie poster. Background: foggy pine forest at dawn. Big title text reads &#8216;SANA&#8217;. Small tagline below reads &#8216;FAST HIGH RES&#8217;. Clean layout. Centered. Subtle paper texture.<\/figcaption><\/figure>\n<p>The foggy forest, centered structure, restrained palette, and paper-like mood show that the model understands poster direction. The hard requirement is the type. A short title and short tagline are easier than dense copy, yet image generation still cannot be trusted for release-ready letterforms. This output is useful as a background plate or art-direction reference, not a finished key art file. Readers comparing layout behavior may also want the <a href=\"https:\/\/wiro.ai\/blog\/gpt-image-1-5-5-prompts-for-clean-layouts\/\">GPT Image 1.5 layout test<\/a> and the <a href=\"https:\/\/wiro.ai\/blog\/sensenova-u1-8b-5-layout-tests-readable-text\/\">SenseNova readable-text test<\/a>.<\/p>\n<h2 id=\"counting\">Test 5: Exact counts in a breakfast flat lay<\/h2>\n<figure>\n  <img loading=\"lazy\" decoding=\"async\" width=\"1024\" height=\"1024\" class=\"wp-image-1934\" src=\"https:\/\/wiro.ai\/blog\/wp-content\/uploads\/2026\/03\/sana-5.png\" alt=\"Breakfast flat lay with a croissant, strawberries, blueberries, and almonds\" srcset=\"https:\/\/wiro.ai\/blog\/wp-content\/uploads\/2026\/03\/sana-5.png 1024w, https:\/\/wiro.ai\/blog\/wp-content\/uploads\/2026\/03\/sana-5-510x510.webp 510w, https:\/\/wiro.ai\/blog\/wp-content\/uploads\/2026\/03\/sana-5-900x900.webp 900w, https:\/\/wiro.ai\/blog\/wp-content\/uploads\/2026\/03\/sana-5-150x150.webp 150w, https:\/\/wiro.ai\/blog\/wp-content\/uploads\/2026\/03\/sana-5-768x768.webp 768w\" sizes=\"auto, (max-width: 1024px) 100vw, 1024px\" \/><figcaption>Prompt: Top-down flat lay breakfast photo on a pale stone table. Exactly 1 croissant, 2 strawberries, 3 blueberries, and 4 almonds. Arrange them in a spiral. Soft natural light. Sharp focus.<\/figcaption><\/figure>\n<p>The image looks like a coherent top-down food composition, which makes it a useful test of surface, light, and arrangement. Exact counting is the catch. Repeated small objects such as berries and almonds are easy for a generative model to duplicate, hide, or merge. A result can look balanced while missing the requested inventory. Keep the objects large, separated, and easy to inspect when count matters. For ecommerce or menu work, treat this as a sketch and validate every required object before it reaches a client.<\/p>\n<h2 id=\"wildlife\">Test 6: Cinematic wildlife texture and directed light<\/h2>\n<figure>\n  <img loading=\"lazy\" decoding=\"async\" width=\"1024\" height=\"1024\" class=\"wp-image-1935\" src=\"https:\/\/wiro.ai\/blog\/wp-content\/uploads\/2026\/03\/sana-6.png\" alt=\"Cinematic black panther in a rainforest at night\" srcset=\"https:\/\/wiro.ai\/blog\/wp-content\/uploads\/2026\/03\/sana-6.png 1024w, https:\/\/wiro.ai\/blog\/wp-content\/uploads\/2026\/03\/sana-6-510x510.webp 510w, https:\/\/wiro.ai\/blog\/wp-content\/uploads\/2026\/03\/sana-6-900x900.webp 900w, https:\/\/wiro.ai\/blog\/wp-content\/uploads\/2026\/03\/sana-6-150x150.webp 150w, https:\/\/wiro.ai\/blog\/wp-content\/uploads\/2026\/03\/sana-6-768x768.webp 768w\" sizes=\"auto, (max-width: 1024px) 100vw, 1024px\" \/><figcaption>Prompt: Cinematic photo of a black panther in a rainforest at night. Wet fur with small water droplets. Rim light from behind. Shallow depth of field. Ultra-detailed. Realistic.<\/figcaption><\/figure>\n<p>This is the cleanest fit for Sana 1600M (1024px). One dominant subject, a dark setting, rim light, and shallow depth of field give the model a clear hierarchy. The panther, wet-fur treatment, and softened rainforest background support the requested cinematic look without depending on fragile text or bookkeeping. Use this kind of prompt for atmospheric editorial art, thumbnail concepts, and early visual development. Check anatomy and paw detail at full resolution before choosing a final output.<\/p>\n<h2>What to choose Sana 1600M (1024px) for<\/h2>\n<table>\n<thead>\n<tr>\n<th>Need<\/th>\n<th>What the test suggests<\/th>\n<th>Practical choice<\/th>\n<\/tr>\n<\/thead>\n<tbody>\n<tr>\n<td>Fast 1024px visual direction<\/td>\n<td>Strong lighting, palette, and scene mood<\/td>\n<td>Choose Sana for concepts and art-direction references.<\/td>\n<\/tr>\n<tr>\n<td>Stylized scenes<\/td>\n<td>Pixel-art language and constrained palettes hold together well<\/td>\n<td>Use a short, specific style brief and inspect fine edges.<\/td>\n<\/tr>\n<tr>\n<td>Compositing ideas<\/td>\n<td>Double-exposure logic reads when subject and color direction are explicit<\/td>\n<td>Generate the concept, then build exact masks elsewhere.<\/td>\n<\/tr>\n<tr>\n<td>Finished typography or strict counts<\/td>\n<td>Text and repeated-object precision remain unreliable<\/td>\n<td>Use a model output as the visual base, then correct it manually.<\/td>\n<\/tr>\n<\/tbody>\n<\/table>\n<h2>Verdict<\/h2>\n<p>Sana 1600M (1024px) works best when visual intent matters more than literal compliance. It handles mood, light, subject hierarchy, and compact style briefs better than tasks that require perfect spelling or audited quantities. Start with a direct subject, a controlled palette, and one lighting idea. Keep a separate finishing step for text, logos, counts, and brand-critical edges. For another angle on choosing image generators for production work, see <a href=\"https:\/\/wiro.ai\/blog\/seedream-v5-pro-vs-nano-banana-pro\/\">Seedream V5 Pro vs Nano Banana Pro<\/a>.<\/p>\n<p><a href=\"https:\/\/wiro.ai\/models\/wiro\/text-to-image-sana\">Run Sana on Wiro<\/a> and use the same six prompt types as a quick acceptance test for a new visual brief.<\/p>\n","protected":false},"excerpt":{"rendered":"<p>Sana 1600M (1024px): 6 Prompt Tests Sana 1600M (1024px) was tested here as a practical text-to-image model for concept work, stylized scenes,&hellip;<\/p>\n","protected":false},"author":4,"featured_media":2038,"comment_status":"closed","ping_status":"closed","sticky":false,"template":"","format":"standard","meta":{"footnotes":""},"categories":[52],"tags":[182,81],"class_list":["post-1937","post","type-post","status-publish","format-standard","has-post-thumbnail","hentry","category-model-reviews","tag-sana","tag-text-to-image"],"_links":{"self":[{"href":"https:\/\/wiro.ai\/blog\/wp-json\/wp\/v2\/posts\/1937","targetHints":{"allow":["GET"]}}],"collection":[{"href":"https:\/\/wiro.ai\/blog\/wp-json\/wp\/v2\/posts"}],"about":[{"href":"https:\/\/wiro.ai\/blog\/wp-json\/wp\/v2\/types\/post"}],"author":[{"embeddable":true,"href":"https:\/\/wiro.ai\/blog\/wp-json\/wp\/v2\/users\/4"}],"replies":[{"embeddable":true,"href":"https:\/\/wiro.ai\/blog\/wp-json\/wp\/v2\/comments?post=1937"}],"version-history":[{"count":2,"href":"https:\/\/wiro.ai\/blog\/wp-json\/wp\/v2\/posts\/1937\/revisions"}],"predecessor-version":[{"id":4286,"href":"https:\/\/wiro.ai\/blog\/wp-json\/wp\/v2\/posts\/1937\/revisions\/4286"}],"wp:featuredmedia":[{"embeddable":true,"href":"https:\/\/wiro.ai\/blog\/wp-json\/wp\/v2\/media\/2038"}],"wp:attachment":[{"href":"https:\/\/wiro.ai\/blog\/wp-json\/wp\/v2\/media?parent=1937"}],"wp:term":[{"taxonomy":"category","embeddable":true,"href":"https:\/\/wiro.ai\/blog\/wp-json\/wp\/v2\/categories?post=1937"},{"taxonomy":"post_tag","embeddable":true,"href":"https:\/\/wiro.ai\/blog\/wp-json\/wp\/v2\/tags?post=1937"}],"curies":[{"name":"wp","href":"https:\/\/api.w.org\/{rel}","templated":true}]}}