FLUX.1 Kontext Editing Guide: Instructions, Identity and Iterative Edits

How to write FLUX.1 Kontext edit instructions that change only what you ask for, keep a person recognisable and survive several rounds of edits, tested on real Pro and Max outputs.

FuserUpdated
Fuser canvas: a ceramicist photo feeds three FLUX.1 Kontext edits (yellow mug, orange apron, new sign text); the apron edit feeds a Kontext Max edit moving her to a market stall.

All guides · FLUX.1 Kontext in Fuser

Quick answer: to edit an image with FLUX.1 Kontext, give it the image and one plain instruction that names exactly what should change ("Change the mug color to mustard yellow"). The image supplies everything else. Use a specific verb for the part you want changed rather than "transform", say what must stay ("while maintaining the same facial features and hairstyle"), and put text you want replaced in quotes: Replace 'OPEN STUDIO' with 'CLOSED SUNDAY'. Those rules come from Black Forest Labs' own Kontext prompting notes; every example below is a real output we generated on 28 September 2026.

What FLUX.1 Kontext is, and where it sits now

FLUX.1 Kontext is Black Forest Labs' instruction-based image editor: one model that takes a text prompt plus a reference image and either makes a local edit or builds a new scene around what it sees (BFL overview). It comes in two hosted versions. [pro] is the fast default at $0.04 per image; [max] costs $0.08 and BFL positions it for maximum prompt adherence and typography (BFL model comparison, fal pricing).

BFL now labels Kontext a previous-generation model and recommends FLUX.2 for new editing work, citing multi-reference input and output up to 4 MP (BFL image editing docs). Fuser ships both. Kontext is still worth knowing: it is quick, cheap and predictable for single, targeted edits. For the newer path, see the FLUX.2 prompt guide.

How to word an instruction

BFL's guidance for Kontext is short, and our runs backed up each point:

  1. Say what changes, not what the picture is. The input image already carries the scene. "Change the mug color to mustard yellow" recoloured the mug and left the face, hands, sign and shelves as they were (hero image, top node).

  2. Use a specific verb for a specific part. BFL advises "change the clothes" over "transform the person". We tested both on the same seed: "Transform the person into a Viking" replaced our ceramicist with a bearded man. "Change her clothes to a Viking warrior's outfit with a fur cloak and leather armour, while maintaining the same facial features, glasses and hairstyle" kept her and changed only the outfit.

  3. Name what must stay. The preservation clause is part of the instruction, not a nice-to-have. When identity matters, list the features: face, glasses, hairstyle, the object in her hands.

  4. Quote text you want replaced. BFL's pattern is Replace '[original text]' with '[new text]' (BFL text editing).

  5. Be more explicit as the edit gets bigger. For several changes at once, BFL's advice is to add as many explicit details as possible. For a single colour swap, a short sentence is enough.

One vague verb swapped the whole person. Naming the clothes kept her. FLUX.1 Kontext [pro], seed 4217, 28 September 2026.

Keeping a person recognisable

Kontext is built for character consistency: BFL's paper reports better preservation of objects and characters across multiple turns than the editing systems it compared against (FLUX.1 Kontext paper). In practice identity holds best when the edit touches something other than the person. Our apron, mug, sign and outfit edits all kept the same face, glasses and grey curls.

Identity is most at risk when the whole scene changes. BFL's own consistency example uses scene-level phrasing ("She is now taking a selfie in the streets of Freiburg, it's a lovely day out", BFL image editing docs) and we got the best relocation results with that style too. Writing "She is now standing behind an outdoor market stall selling her mugs" on Kontext [max] produced a complete outdoor market scene with the same face, glasses, apron and teal mug. The same sentence on [pro] left her in the studio.

Relocating a subject was Pro's weak spot in our runs. Max handled the same instruction.

Pro held on to the original scene hard. "Put her on a beach at sunset" swapped the back wall for sea and sky but left the shelves, bench and pots in place. Adding "Replace the whole studio, including the shelves, the chalkboard sign, the workbench and the pottery" removed the sign and little else. If you need to move a subject somewhere new, switch the node to Max before you rewrite the prompt a fourth time.

Editing text in an image

Short replacements are where Kontext is strongest. Replace 'OPEN STUDIO' with 'CLOSED SUNDAY' came back spelled correctly, in the same chalk hand, with the arrow kept. A longer replacement, Replace 'OPEN STUDIO' with 'KILN FIRING - BACK AT 3PM', failed on both versions: Pro wrote "KILN FINING", and Max spelled every word right but dropped "BACK". Keep new text close in length to the text it replaces, and check spelling on every output before you use it.

A two-word swap came out clean; a six-word replacement slipped on both Pro and Max.

For text that has to move or resize, BFL documents a second technique: draw a brightly coloured box on the input image around the area to change and refer to it in the prompt. Kontext [pro] reads the box as an annotation and removes it from the output (BFL annotation boxes). In Fuser you can draw that box as a rectangle layer in the Compositor and export a PNG.

Iterative edits: chain or branch

Feeding each output into the next edit works, and identity survives it. Image quality is the thing to watch. We ran three Pro edits in sequence (orange apron, then a SALE sign, then evening light). By the third pass the image had heavy contrast, oversharpened edges and a blotchy wall, and the lighting change hardly showed. The same lighting prompt applied once to the original came back clean.

Each pass through the model adds contrast and sharpening. Branch from the cleanest image when you can.
  • Branch when edits are independent. Run each change from the original, as in the hero canvas where the mug, apron and sign edits all start from the same source.

  • Chain only when a step depends on the last one, and keep the chain short. Two or three steps is a sensible limit before you inspect the result closely.

  • Fix the seed while you iterate on wording. With the same seed, differences between two wordings come from the wording. The Kontext node in Fuser has a Seed field for this.

  • Combine small edits into one prompt instead of three passes when they don't conflict: "Change the apron to rust-orange linen and replace 'OPEN STUDIO' with 'SALE'."

FLUX.1 Kontext settings in Fuser

The FLUX.1 Kontext node picks the endpoint for you from two things: the Model dropdown (Pro by default, or Max) and how many images you connect.

  • No image: text-to-image, using the Aspect Ratio setting (1:1 by default).

  • One image: a single-image edit. The output follows the input's proportions, so the aspect ratio setting is not sent. BFL says Kontext matches the input dimensions as closely as it can (BFL parameters); our 1248 × 832 input came back at 1248 × 832.

  • Two to four images: the multi-reference endpoint, for combining elements from several pictures.

  • CFG Scale defaults to 3.5 and sets how closely the output follows the prompt. Seed and Output Format (JPEG or PNG) are also on the node.

  • Images from your Fuser library are downscaled to fit 1024 × 1024 before they are sent, which is close to Kontext's roughly 1 MP output anyway.

Because it is a node, the edit is one step in a larger graph. Generate a keyframe with GPT Image or FLUX.2, make targeted fixes with Kontext, then pass the result to an upscaler or an image-to-video model on the same canvas. To keep a character consistent across a whole series, see consistent characters with AI, and for a wider view of editors, the best AI image editing models.

FLUX.1 Kontext instruction cheat sheet.

What to write for common edits, based on BFL's guidance and our 28 September 2026 runs.

EditWriteResult in our test
Instructions
Recolour an object

Change the mug color to mustard yellow.

Mug recoloured, rest untouched (Pro)

Change clothing

Change her clothes to …, while maintaining the same facial features, glasses and hairstyle.

Outfit changed, identity kept (Pro)

Avoid

Transform the person into …

Replaced the person entirely (Pro)

Replace text

Replace 'OPEN STUDIO' with 'CLOSED SUNDAY'

Correct, same lettering style (Pro)

Long text

Keep replacements close to the original length.

Six words: errors on Pro and Max

Move the subject

She is now standing behind an outdoor market stall …

Full scene change on Max; Pro stayed in the studio

Several edits

Branch from the source, or combine them in one prompt.

Third chained pass visibly degraded

Questions, answered.

Name only what should change, with a specific verb for that part ("change the clothes", not "transform the person"). Add what must stay, such as "while maintaining the same facial features and hairstyle". The input image supplies everything else.

Edit something other than the person where possible, and list the features to preserve in the prompt. Avoid "transform" wording: in our test "Transform the person into a Viking" replaced the woman with a different man, while "Change her clothes …" kept her face.

Quote both strings: Replace 'old text' with 'new text'. Short swaps are reliable; in our test a six-word replacement came back with errors on both Pro and Max, so check spelling on every result.

Pro is the default; at API list prices it costs $0.04 per image and Max $0.08. Max is aimed at stronger prompt adherence and typography. In our runs Max handled moving the subject to a new location where Pro kept the original room.

Yes, and identity holds, but image quality drifts: by the third chained Pro edit our image was oversharpened and blotchy. Run independent edits from the original, or combine compatible changes into one prompt.

BFL now recommends FLUX.2 for new editing work, with more reference images and higher resolution. Kontext remains a quick, low-cost option for single targeted edits. Fuser offers both.

Edit, branch and compare on one canvas.

Run Kontext Pro and Max side by side from the same source and keep every version.

All articles