Flux 3 Image release – maximum pixel‑level control and precise editing

Flux 3 Image release – maximum pixel‑level control and precise editing

Flux 3 Image introduces a layout‑driven workflow where every visual element is defined by a bounding box and a textual description, enabling pixel‑perfect composition and targeted edits while preserving the rest of the image.


What Flux 3 Image does differently

Flux 3 Image lets you control each pixel by describing elements inside explicit bounding boxes.

  • Users draw a box on a 0‑to‑1000 grid, assign an ID and a natural‑language description (e.g., dome_1[250,150,650,850] → “a massive, smooth parabolic dome of pale concrete”).
  • A single scene caption ties the elements together (e.g., “A glowing concrete dome rises from a twilight bay, a crowd on the beach before it.”).
  • The model renders the full canvas in native 4K resolution, preserving fine details such as textures, reflections, and small text.
  • After generation, any box can be edited, moved, or replaced without affecting untouched regions, supporting iterative, pixel‑perfect editing.

Core workflow

  1. Define aspect ratio – choose any shape; the canvas always spans a 0‑1000 coordinate system.
  2. Create element table – a JSON array where each entry contains id, bbox, and desc.
  3. Write a global caption – a one‑sentence description that references the IDs.
  4. Generate – Flux 3 Image fills each box with the described content.
  5. Edit – modify the JSON (or redraw boxes) and re‑run generation; unchanged boxes stay exactly the same.

“Drag out a box for every element that matters, then describe what goes in it.” – Black Forest Labs documentation


Example: "Le Festival du Soleil"

The page showcases a complex composition built from five boxes:

ID Bounding box Description
Fr_Text_1 [10,200,170,800] ""LE FESTIVAL DU SOL​EIL" written in a thin, elegant, serif typeface in a light cream color"
town_1 [280,700,420,1000] "faint lights and small buildings of a coastal town at the foot of the hills"
dome_1 [250,150,650,850] "a massive, smooth parabolic dome of pale concrete"
swimmers_1 [580,200,720,800] "dozens of small, silhouetted figures scattered in the dark water, wading"
crowd_1 [740,0,1000,1000] "a large crowd of people seated on the beach in casual, light‑colored summer attire"

The global caption reads: “A glowing concrete dome rises from a twilight bay, a crowd on the beach before it.” Flux 3 Image fills each region accordingly, producing a cohesive night‑scene illustration.


Precise editing use‑case

The platform supports pixel‑perfect edits such as recoloring a surfer’s wetsuit while keeping the surrounding wave, sky, and composition unchanged. The edit record shows the original IDs (surfer_wetsuit, surfboard) and the new descriptions (bright red neoprene wetsuit, bright red surfboard). All other elements (ocean_wave, cloudy_sky) remain locked.


Commercial availability

Flux 3 Image is offered under a commercial‑weights license for enterprises that need large‑scale generation. Companies can fine‑tune and self‑host the model on private infrastructure. Sales inquiries are handled via the Black Forest Labs contact form.


Community reaction on Hacker News

  • UX praise – Users highlighted the intuitive interface for placing elements:

    "The UX looks amazing and very steerable, congrats to the team for focusing on the interface." – @arnaudsm

  • Comparison to other tools – Commenters noted similarity to Ideogram V4’s JSON‑based layout but found Flux 3’s UI less cumbersome:

    "One of the things they seem to be emphasizing here is the UX around being able to place specific elements where you want them in an image. … It’s definitely a bit of a hassle. I'll probably be waiting until it goes open‑weight…" – @vunderba

  • Open‑weight expectations – Several users expressed anticipation for a future open‑weight release:

    "I think we are all waiting for the open weights or local model releases." – @vergessenmir

  • Practical success stories – Early adopters reported concrete gains in workflow speed:

    "I tried it for quick UI element replacements in a screenshot, and it nailed 3/4 elements I highlighted first attempt. … granular iteration, which is where it becomes useful in an industry‑wide manner." – @soundworlds

  • Limitations – Some users encountered edge cases where precise editing failed on small text or complex objects:

    "I tried precise editing with a photo of horse with fence in front of it. One shot didn't work…" – @pks016


Frequently asked questions (excerpt)

  • What is Flux 3 Image? – It is the image‑generation and editing component of Black Forest Labs’ multimodal Flux 3 family.
  • How do bounding boxes work? – Boxes are defined as [y_min, x_min, y_max, x_max] on a 0‑1000 grid; each box is paired with a textual description.
  • Do I have to draw every box myself? – No. An LLM can propose a layout automatically; you can then adjust any boxes you dislike.
  • Can I edit an existing image? – Yes. Re‑describe, move, or replace individual boxes; untouched boxes stay exactly as before.
  • What kinds of images benefit most? – Complex collages, editorial spreads, UI mockups, or any scene where precise spatial relationships matter.

Getting started

  1. Visit the Flux 3 Image Playground to draw boxes and enter captions.
  2. Use the API (documentation linked on the Black Forest Labs site) for programmatic generation and editing.
  3. For commercial use, contact sales to obtain a commercial‑weights license and self‑hosting terms.

Flux 3 Image represents a step forward in controllable generative AI, giving creators the ability to dictate exact composition while still benefiting from the model’s creative synthesis.

Sources

関連

  • プロジェクト
  • Dispatch
  • Dispatch
  • プロジェクト
  • プロジェクト