Recraft's New Model Costs Five Cents and Finishes Before You Blink

Recraft released Recraft V4.1 Flash on 23 September 2026, and the two headline numbers are the whole pitch: roughly 1.3 to 1.5 seconds to generate a one-thousand-pixel image, and $0.007 per image.
That price is about a fifth of what standard Recraft V4.1 charges, and the company is calling Flash the fastest image model on the market. Recraft supports the claim with a test run on 21 September: one hundred prompts, each sent once to a spread of major image models, measuring the median and the slowest-few-percent response times. Flash came in fastest on both. The comparison set is instructive. Nano Banana 2 Lite landed around 4.1 seconds at the median, Ideogram v4 Instant around 10.2 seconds, and Seedream 5.0 Lite around 43.7 seconds. Recraft also notes it kept each provider's default settings, including whatever prompt enhancement they apply by default, because turning that off costs visible quality.
Seven-tenths of a cent was not the interesting number. The interesting number is what happens to a design team's behaviour when exploring an idea stops costing anything.
What Flash actually is
Flash is not a new model family. It is the speed and cost tier of the existing V4.1 line, sharing the same aspect ratios and overall look while trading resolution and deliberation for turnaround time.
Recraft is explicit about the trade being an attention problem rather than a throughput one. Ten seconds, in Recraft's framing, is long enough to forget the idea you had when you pressed generate. An image model that answers inside two seconds keeps you inside the thought instead of leaving you to reload it. That is a designer's argument, not an engineer's, and it is a real one. The most expensive part of ideation is rarely the render. It is the attention lost waiting for it.
What Flash gives you is narrow and clear. It produces roughly 1K raster images in the same nine aspect ratios as V4.1, runs up to six images per request, and accepts two colour controls: a palette for the image and a background colour. That is the whole surface area. It is available through Recraft's API and on OpenRouter at the same list price.
The features that got dropped
Here is where the release gets less comfortable. Flash supports generation only. It takes no image input, no custom styles, and no style references. Vector and SVG output are gone, and so is the 2K Pro tier.
The omission that matters most is style references, because that is the feature Recraft's business customers lean on hardest. Style references are how a team makes a set of images look like they came from one brand. Marketing assets, e-commerce hero shots, and poster series all depend on it. Strip it out and Flash stops being a cheaper Recraft. It becomes a different product wearing the Recraft name, one that can draft a direction but cannot hold it.
Recraft also published no quality comparison. The announcement claims the speed crown and states the price, and says nothing about how Flash images compare with V4.1 images or with any competitor. For a model whose entire proposition is a trade between speed and fidelity, the missing half of the trade is the fidelity. Buyers are being asked to accept speed on trust.
How the workflow is supposed to change
The intended pattern is easy to describe and, to be fair, easy to like. Run exploration on Flash at seven-tenths of a cent per image, generate a few dozen directions for well under a dollar, then take the concept you like and re-render it on V4.1 where style references work and the resolution is real.
That reshuffles the order of creative work. The old advice was to think clearly first and generate once. When a draft costs less than a cent and finishes in under two seconds, you can generate and think at the same time, which is closer to how designers actually iterate with a sketchbook than to how they use a render farm.
This is part of a broader shift that showed up across the model ecosystem in the same stretch of September. Pruna AI open-sourced few-step LoRA adapters that cut Qwen-Image-2.1 from 40 sampling steps down to 8 or 5, which the company clocks at up to 6.3 times faster. fal switched on Extend Video for its H3 Max model, letting users stretch a 15-second clip to 30 while holding the original look. The pattern repeats: generation is splitting into a cheap, fast draft pass and a slower, more controlled finish.
The honest read on Recraft V4.1 Flash is that Recraft shipped a very cheap ideation tool and announced it as a model release. That is not a dismissal. The price makes it trivial to find out whether you want it, and for teams that burn through options before choosing, a fast draft tier is a genuine upgrade to how the work gets organised. It just should not be dropped into a production pipeline and expected to hold a brand look. It cannot, and Recraft's release notes are the ones that say so.
Who should actually use it
The honest answer is narrower than the speed record suggests. Flash makes sense for a team that generates dozens of options before choosing one, or for a workflow that needs a quick visual answer to a narrow question, such as whether a layout can hold a headline at all. It makes less sense for anyone whose output has to look like it belongs to a brand, because the features that enforce brand consistency are the ones Recraft removed. A photographer testing lighting directions will get real value out of it. A studio delivering a finished campaign will not, at least not from Flash alone.

That gap is the part worth watching over the next quarter. If enough teams adopt the draft-on-Flash, finish-on-V4.1 pattern, Recraft has created a two-tier pricing model that competitors will copy. If teams find the output good enough to ship after a single render, then the entire premium tier starts to look overpriced, and Recraft will have undermined its own product with a faster version of it.
Where it fits
Flash is best understood as a response to a specific competitive reality. Image quality at the top end has compressed. The differences between the best models are real but narrow, and increasingly they are differences of control, editing and cost rather than raw beauty. When that happens, the model that wins an afternoon is often not the best one. It is the one that answers before you lose the thread.
Recraft is betting a meaningful slice of the market will pay almost nothing for a draft and then pay full price for the finish. Whether that slice is large enough to matter is the question the next few quarters will answer, one seven-tenths-of-a-cent generation at a time.
Related articles
The Video Model Leaderboard Nobody Markets: Where the Requests Actually Go
Every week brings a new video generation ranking, and almost all of them are built the same way.
Ai2 Open-Sourced the Training Stack, Not Another Set of Weights
Weights are easy to give away and hard to learn from.
Alibaba's Qwen3.8-Max Claims It Can Code by Itself for a Fortnight
Parameter counts stopped being news a while ago. Twelve days is the number worth examining.
PixelUMM Throws Out the Two Components Every Visual AI Model Depends On
Pixel-level diffusion has been proposed before and has repeatedly run into compute.