Generate
Tools
Templates
Learn
Use cases
Pricing
Sign in Start creating

FLUX.1 Kontext: Editing Images by Instruction

FLUX.1 Kontext is a family of AI image models from Black Forest Labs built for editing rather than pure generation. You give it a picture and a sentence - "make the jacket navy blue", "remove the sign", "same character, now on a beach" - and it redraws the image to match, with no masks, layers or selections. Below is what it does, where it struggles, how to write instructions it follows, and a free editor you can run right now on your own photo.

Made by
Black Forest Labs
Type
In-context image editing and generation
Input
An image + a plain-language instruction
Variants
[dev] open weights, [pro] and [max] via API
Masks needed
None - the model finds what you named
Try the workflow
Free editor on this page

Try instruction editing

Upload a photo, describe the change in one sentence. No masks, no layers.

Drop an image or browse

JPG, PNG or WebP, up to 16 MB

Edits on the free plan queue behind Pro and Max members, so expect at least 40 seconds. Upgrade for priority and no watermark.

FLUX.1 and FLUX.1 Kontext are products of Black Forest Labs. img.now is an independent tool and is not affiliated with or endorsed by Black Forest Labs. The editor on this page is our own instruction-based AI editor - it does the same job in the same way, so you can try the workflow without a GPU or an API key. Black Forest Labs' site is the source of truth for model releases, pricing and licence terms.

What FLUX.1 Kontext actually is

Most image models take text in and give an image back. Kontext takes text and an image in, which is what "in-context" means here: your photo is part of the prompt, not a starting point it is free to discard. That framing is the whole point. The model is asked to change the one thing you named and to leave the rest of the frame - the faces, the product, the layout, the lighting - recognisably intact.

In practice that turns editing into a conversation. You upload a photo, type a sentence, look at the result, then type the next sentence. There is no mask to paint, no layer stack, and no strength dial to guess at. Black Forest Labs, the German lab behind the FLUX family, released the first Kontext models in 2025 and positioned them squarely at this iterative editing loop rather than at one-shot art generation - although Kontext can generate from text alone too.

Instruction editing vs inpainting vs image-to-image

These three get mixed up constantly, and picking the wrong one is the usual reason an edit comes out badly. Instruction editing is the generalist: it is fastest to use and best when you can name the change in words. Inpainting is the specialist: you mark an exact region, so it is the safer choice when the edit has to stay inside a precise boundary - a logo on a shirt, one face in a crowd. Image-to-image is the transformer: it re-imagines the whole frame toward a prompt with a strength slider, which is what you want for a full style change and exactly what you do not want for a small fix.

Approach How you drive it Best for Watch out for
Instruction editing (Kontext-style) Upload + one sentence Named, specific changes: colour, object, background, weather, style Redraws the whole frame, so tiny details can shift
Inpainting / object removal Brush the exact area Erasing or replacing one region with a hard boundary You have to know and mark exactly where the edit goes
Image-to-image Upload + prompt + strength dial Restyling the entire picture while keeping the composition High strength reinvents the subject, not just the look

On img.now those three map to the AI Magic Edit tool (instruction), the object remover (masked) and image to image (whole-frame restyle).

Where it is strong

Keeping a subject consistent. The headline capability: run the same character or product through several edits - new background, new outfit, new angle - and it stays identifiably the same thing. That is what makes it useful for a product line or a set of marketing shots rather than one lucky image.

Local changes without a mask. Recolour an object, swap a background, change weather or time of day, add or delete an item, restyle clothing. You name the target in words and the model works out where it is.

Style transfer that survives. Push a photo into an illustration, a 3D render or a painting while the composition and the subject's identity hold.

Editing text that is already in the picture. Signage, labels and packaging copy can be swapped, which older editing models handled badly.

Where it struggles (worth knowing before you start)

Drift across many passes. Every edit is a redraw, so quality and likeness degrade slowly as you chain edits. Do the big structural change first, keep the chain short, and re-upload the original rather than the fifth generation.

Compound instructions. "Make the car red, add a dog, change the sky and remove the fence" usually lands one or two of the four. Split it up and run one sentence at a time.

Small typography. Text can be edited but fine print, exact fonts and brand marks are still unreliable - set real copy over the image in a design tool afterwards.

Precise geometry. Exact object counts, hands, reflections and engineering-accurate detail remain hard for every model in this class, Kontext included.

How to write an instruction it will follow

Name one change, name the thing precisely, and say what should stay. "Change the wall behind the sofa to warm cream, keep the sofa and the lamp as they are" beats "nicer living room" every time. Use the words a person would use to point at the object - "the red mug on the left", not "the object in the mid-ground".

Describe the destination, not the process. "Make it look like an overcast afternoon" works; "reduce the saturation and add a cool LUT" does not, because the model paints results, not adjustment layers. If a result strays, do not add more words - shorten the sentence and run it again, since each pass varies. And start from a clean, well-lit, reasonably large source: an edit cannot recover detail the original never had.

Change the car colour to matte black, keep everything else the same
Replace the cloudy sky with a clear blue sky
Remove the poster from the wall and leave the brickwork clean
Put the product on a marble countertop in soft daylight
Same character, now standing on a beach at sunset

Our prompt-writing guide goes deeper on phrasing, and the image-to-image guide covers when a whole-frame restyle beats a targeted edit.

The three Kontext variants

Black Forest Labs ships Kontext in tiers rather than as a single model. Which one a given site or app is using changes the price, the quality and - importantly - what you are allowed to do with the output.

Variant How you get it What to know
FLUX.1 Kontext [dev] Open weights, downloadable Runs locally on a capable GPU or through hosting providers and ComfyUI. Released under Black Forest Labs' non-commercial licence, so check the terms before using output in paid work.
FLUX.1 Kontext [pro] Hosted API / partner platforms The general-purpose commercial tier: faster, stronger instruction following than [dev], billed per image by Black Forest Labs or a reseller.
FLUX.1 Kontext [max] Hosted API / partner platforms The top tier, aimed at the hardest prompts and the best in-image typography. Costs more per image than [pro].

Model line-ups move fast. Check Black Forest Labs directly for the current variants, per-image pricing and licence wording before you build on any of them.

Editing by instruction, in three steps

1

Upload your image

Drop in the photo you want to change - a product shot, a portrait, a room or a screenshot.

2

Describe the change

One sentence: name the thing to change, and name what should stay exactly as it is.

3

Run it and download

Compare before and after with the slider, then save the result - or run another instruction on it.

FLUX.1 Kontext questions

What is FLUX.1 Kontext?
It is an AI image model family from Black Forest Labs built for editing in context: you supply an image plus a written instruction, and it returns the edited image. Unlike a plain text-to-image model, your picture is part of the prompt, so the model changes what you asked for and tries to leave everything else the same. It can also generate images from text alone.
How is Kontext different from FLUX.1 dev or schnell?
FLUX.1 dev and FLUX.1 schnell are text-to-image models: text in, brand-new image out. Kontext is the editing branch of the same family - it accepts an image alongside the text and is trained to make targeted changes to it while preserving the subject. Same lab, different job.
Do I need to mask the part I want changed?
No, and that is the main appeal. You describe the target in plain words - "the blue chair", "the sky", "the text on the sign" - and the model locates it. Masking still wins when an edit must stay inside an exact boundary, which is what a dedicated object remover is for.
Is FLUX.1 Kontext free to use?
The [dev] weights can be downloaded at no cost, but they carry a non-commercial licence and you need a capable GPU to run them. The [pro] and [max] tiers are paid, billed per image through Black Forest Labs' API or a partner platform. The editor on this page is free to try on the img.now free plan, with no GPU and no API key.
Does img.now run FLUX.1 Kontext?
img.now runs its own hosted image model and does not expose a model picker - the point is that you describe the edit and get a result without choosing plumbing. The editor here is our AI Magic Edit pipeline: the same upload-and-describe workflow Kontext popularised, done for you in the browser.
Can I use the edited images commercially?
Images you create on img.now are released under CC0, so commercial use is allowed with no attribution and no licence fee. If you run Kontext yourself, your rights depend on which variant and licence you used - the open [dev] weights are non-commercial, while the paid API tiers have their own terms.
Why did the edit change something I did not ask about?
Instruction editing regenerates the picture rather than patching pixels, so small drift is normal. Keep the instruction to one change, name what should stay put, and re-run it - each pass differs slightly. For an edit that must not touch a single pixel outside a region, use masked object removal instead.
Edit an image