FLUX 3 Image
Thread
Loading the complete thread in the background. This saved snapshot is available now. Refresh
Unofficial Hacker News client; not affiliated with Y Combinator.
FLUX 3 Image
Loading the complete thread in the background. This saved snapshot is available now. Refresh
Unofficial Hacker News client; not affiliated with Y Combinator.
Grimblewald · · focus · HN ↗
reilly3000 · · focus · HN ↗
arnaudsm · · focus · HN ↗
Chats can be awful user interfaces.
neals · · focus · HN ↗
swiftcoder · · focus · HN ↗
htrp · · focus · HN ↗
What's new from the last post? GA?
zamadatix · · focus · HN ↗
> We will open up an early access phase for FLUX 3 Image in the following weeks.
Not sure if there was a separate post for early access or if they just skipped to this.
myself248 · · focus · HN ↗
pizzafeelsright · · focus · HN ↗
Flux is in the top 9000 of the most common words.
vunderba · · focus · HN ↗
Ideogram V4, an open-weight model released back in June can also do this [1], but you have to use a relatively cumbersome JSON structure to describe all the different bounding boxes. So it’s definitely a bit of a hassle.
I'll probably be waiting until it goes open-weight (hopefully soon) like they did with Flux.2 / Klein.
[1] - <a href="https://docs.ideogram.ai/using-ideogram/getting-started/prompting-guide/4.-json-prompting-ideogram-4.0" rel="nofollow">https://docs.ideogram.ai/using-ideogram/getting-started/prom...
kranke155 · · focus · HN ↗
CuriouslyC · · focus · HN ↗
hdjrudni · · focus · HN ↗
CuriouslyC · · focus · HN ↗
swiftcoder · · focus · HN ↗
It's one of the most uniquely hostile user experiences I've ever had the (dis)pleasure of working with
bavell · · focus · HN ↗
user43928 · · focus · HN ↗
However, today I don't see a reason to use ComfyUI at all.
For Qwen Image 2.1, I had Opus 5.5 create a backend outside of ComfyUI and it was able to make it take 20% less time with some optimizations.
If there was anything interesting in ComfyUI nodes I imagine I could just have the AI adopt the relevant code instead of dealing with ComfyUI or custom nodes.
orbital-decay · · focus · HN ↗
jarjoura · · focus · HN ↗
From where I'm sitting, it's just turning python functions into boxes and instead of write the function yourself, you drag from the output of one box to the input of another. For 2 or 3 boxes, this is cool, but I opened up a professional workflow and was taken into a view with 100s of boxes and wires all over the place. Uhh, ok?
For myself, I'd rather just create my own python environment, write some quick pytorch or mlx calls, wire up some cli to it and share that in a GitHub.
sorenjan · · focus · HN ↗
[0] <a href="https://github.com/saintbrodie/Orange" rel="nofollow">https://github.com/saintbrodie/Orange
sheepscreek · · focus · HN ↗
Disclaimer: never used it for actual generation so I don’t know if it’s doing anything special other than being a subgraph with a different name.
rf15 · · focus · HN ↗
vunderba · · focus · HN ↗
popalchemist · · focus · HN ↗
<a href="https://www.youtube.com/watch?v=2mecWZgbaEg" rel="nofollow">https://www.youtube.com/watch?v=2mecWZgbaEg
woadwarrior01 · · focus · HN ↗
popalchemist · · focus · HN ↗
NBJack · · focus · HN ↗
woadwarrior01 · · focus · HN ↗
NBJack · · focus · HN ↗
vergessenmir · · focus · HN ↗
pixelesque · · focus · HN ↗
The website mentions:
> FLUX 3 Image is available under a commercial weights license for companies running image generation at scale. Fine-tune and deploy it on your own infrastructure. Reach out to us to learn more.
vunderba · · focus · HN ↗
<a href="https://bfl.ai/legal/non-commercial-license-terms" rel="nofollow">https://bfl.ai/legal/non-commercial-license-terms
LordDragonfang · · focus · HN ↗
trentor · · focus · HN ↗
vunderba · · focus · HN ↗
"Open Weights version of FLUX 3 Image is launching in the coming weeks."
<a href="https://nitter.cf/bfl_ai/status/2105734605621825738" rel="nofollow">https://nitter.cf/bfl_ai/status/2105734605621825738
LordDragonfang · · focus · HN ↗
JimDabell · · focus · HN ↗
— <a href="https://x.com/bfl_ai/status/2105734605621825738" rel="nofollow">https://x.com/bfl_ai/status/2105734605621825738
KazaNLP · · focus · HN ↗
imgbenchdude · · focus · HN ↗
[dead]
assimpleaspossi · · focus · HN ↗
richardfulop · · focus · HN ↗
nazgulsenpai · · focus · HN ↗
assimpleaspossi · · focus · HN ↗
richardfulop · · focus · HN ↗
thorum · · focus · HN ↗
bonoboTP · · focus · HN ↗
9dev · · focus · HN ↗
Mashimo · · focus · HN ↗
How is this not clear?
armcat · · focus · HN ↗
popalchemist · · focus · HN ↗
Generate a sprite in an image editor, then use a video model to make the loop you want; then turn the resulting video back into individual sprite images.
armcat · · focus · HN ↗
popalchemist · · focus · HN ↗
armcat · · focus · HN ↗
I can still get absolutely insane results with MiniMax H3 - insane in the sense that it would not make sense at all and would make your head spin.
popalchemist · · focus · HN ↗
bobcatsmith · · focus · HN ↗
[dead]
popalchemist · · focus · HN ↗
Video Generation Models as World Models: Efficient Paradigms, Architectures and Algorithms
Clearly I'm the uneducated one here eh
yorwba · · focus · HN ↗
That paper sets up a task where generating a correct video requires correctly modeling physical laws. From the failure to always generate the correct video, they infer that the model has failed to correctly model the physical laws. The whole premise of the experiment is that learning visual statistics is equivalent to learning causal dynamics, such that failure at one implies failure at the other.
The main difference in applications is that the bar for entertainment is lower, so that even a very bad world model may be acceptable.
nicerice · · focus · HN ↗
<a href="https://media.discordapp.net/attachments/1401891025970008154/1552060356220686336/2511_vs_21.mp4" rel="nofollow">https://media.discordapp.net/attachments/1401891025970008154...
I don't know what went into making this, but their twitter is @araminta_k if you're curious
justinclift · · focus · HN ↗
* <a href="https://alvdansen.github.io/animating-on-twos/" rel="nofollow">https://alvdansen.github.io/animating-on-twos/
* <a href="https://github.com/alvdansen/animating-on-twos" rel="nofollow">https://github.com/alvdansen/animating-on-twos
* <a href="https://huggingface.co/alvdansen/h3-keyframe-animation" rel="nofollow">https://huggingface.co/alvdansen/h3-keyframe-animation
pumanoir · · focus · HN ↗
bobcatsmith · · focus · HN ↗
avereveard · · focus · HN ↗
Lucasoato · · focus · HN ↗
Imagine online procedural MMO with old gen final fantasy / chrono trigger styles :)
fuzzythrowaway · · focus · HN ↗
Trufa · · focus · HN ↗
doctorpangloss · · focus · HN ↗
vlyan · · focus · HN ↗
mdp2021 · · focus · HN ↗
If you do not consider contemplation, for which we have filled our homes with works of art for centuries,
and if you do not consider those job that involve graphics production e.g. in the Madison Avenue "Mad Men" business (advertising - the payments are authentic),
you could consider "authentic" that when we do video production through generative models some start from stills (image generation) and then have other models animate those stills.
matthew-wegner · · focus · HN ↗
amelius · · focus · HN ↗
minimaxir · · focus · HN ↗
pwillia7 · · focus · HN ↗
vunderba · · focus · HN ↗
<a href="https://bfl.ai/pricing" rel="nofollow">https://bfl.ai/pricing
gAI · · focus · HN ↗
mromanuk · · focus · HN ↗
yorwba · · focus · HN ↗
skybrian · · focus · HN ↗
htx619 · · focus · HN ↗
[dead]
swiftcoder · · focus · HN ↗
imgbenchdude · · focus · HN ↗
[dead]
k12sosse · · focus · HN ↗
Doohickey-d · · focus · HN ↗
cyanmagenta · · focus · HN ↗
mkl · · focus · HN ↗
hmstx · · focus · HN ↗
What else?
- Bunch of visible shape inconsistencies: building window sizes and alignment, fence railing, rivets on the bench, coins, accordion features, perspective violations etc...
- Metal clips on the bottom and top edges of the carrying case are on different sides. Said clips look inconsistent.
- One of the case straps originates inside (from under the velvet, even), the further one originates outside. The carrying handle is oddly mis-centered (as with most things).
- I'm not sure that accordion, compressed, fits in that case. She can just about stand in it (diagonally, on one foot).
- It has weird folds that disappear partway through and turn into adjacent folds going the other way. Between this, the keyboards and the dots, I'm sure there are other things wrong with accordion anatomy.
- Awkward left hand: not sure what's going on with the metacarpal joints, and index finger appears to begin forking into two tips pressing into two (misshapen) buttons. Her right hand is getting there too.
- Weird hair midline. I haven't seen a "sub-midline" like this in the wild :)
- The burgundy coat clips through the bench under her right leg. It's like the front slat goes through a hole in the side of the coat.
- We don't usually have that side curling bar on each side of the sitting area. It looks like it's clipping through the flats next to the leather handbag
- Ill-placed buckle on that bag.
- The green metal fence on the right begins behind her for one vertical bar, next one disappears into the plants, no further vertical bars, no top horizontal railing. Actually, it appears to only have fencing on one side, that's not something I'd see often. On the side where it does have a top, its top rail has very irregular width.
- Sand looks like... breadcrumbs? Very repetitive coarse shapes, like it was poorly done with a clone stamp.
- Something weird about the leaves.
- The foremost lamp isn't sure whether it has flat sides or not.
- The other lamp emerges out of nowhere from behind one of the blob-people in the alley.
- A bunch of unfortunate tangents and occlusions which make you wonder whether the diffusion hallucinated detail out of bigger shapes, or if it's been trained on photos where the tangents were done on purpose.
ie. center bottom bar of the bench being exactly aligned with the gray edge tiles, side curly railing-not-present-on-our-real-benches fully formed on her right, almost completely foreshortened out of existence where the bench's perspective forbids it...
Most people are fooled by much cruder fakes though. This is a lot better but keeps plenty of tells for observant people (and most people aren't)
motoxpro · · focus · HN ↗
Or just put in more than 10 words of effort, e.g. a better prompt and bounding boxed prompt edits, and all that stuff would be fixed now.
skybrian · · focus · HN ↗
When I was a kid, there was a game in the funny pages that asked you to find six things wrong in a picture. Someone should run a contest like that for AI-generated images.
It might be an interesting benchmark for vision models, too.
imgbenchdude · · focus · HN ↗
Exact prompt:
Generate a photorealistic image of M81 urban BDU camouflage cargo trousers, shown by themselves. One trouser leg should be posed with the knee lifted 30° from vertical.
Accurate reproduction of the M81 urban camouflage pattern is critical. Match its colors, shapes, scale, distribution, and overall appearance as faithfully as possible.
No person, other clothing, or props.
Ground truth swatch: <a href="https://commons.wikimedia.org/wiki/File:US_City_Camo_(M81_Urban).png" rel="nofollow">https://commons.wikimedia.org/wiki/File:US_City_Camo_(M81_Ur...
Gemini 3 Pro Image (stronger pattern): <a href="https://i.postimg.cc/bZNQYYjx/2026-10-02-google-gemini-3-pro-image-n1.png" rel="nofollow">https://i.postimg.cc/bZNQYYjx/2026-10-02-google-gemini-3-pro... Flux 3 (this run): <a href="https://i.postimg.cc/Xr7wNN0H/2026-10-02-black-forest-labs-flux-3-image-n1.png" rel="nofollow">https://i.postimg.cc/Xr7wNN0H/2026-10-02-black-forest-labs-f...
Flux 3 gets greyscale urban-ish trousers and a lifted knee, but the blotches aren’t real M81 Urban — softer / wrong geometry vs the swatch. Not the worst I’ve seen on this prompt; clearly behind the Gemini 3 Pro Image example above on pattern.
Curious what other models do on the same prompt.
thomasikzelf · · focus · HN ↗
pks016 · · focus · HN ↗
Tried adjusting exposure; Didn't work as well.
spaceman_2020 · · focus · HN ↗
No signal in digital images anymore. If you didn’t see it with your own eyes, it likely never happened
adammarples · · focus · HN ↗
spiderfarmer · · focus · HN ↗
Nothing in life is ever perfect. Doesn't mean imperfect stuff can't have a lot of impact.
spaceman_2020 · · focus · HN ↗
Me, casually scrolling on my phone, don’t
I have no way of knowing that you wanted the image to have a red wetsuit. I would just assume its a real image of a surfer in a red wetsuit
numlock86 · · focus · HN ↗
The other examples range from mostly good to okay'ish at least.
soundworlds · · focus · HN ↗
As I'm finding with the best GenAIs, this allows granular iteration, which is where it becomes useful in an industry-wide manner.
Zufriedenheit · · focus · HN ↗
mdp2021 · · focus · HN ↗
That would be to compare e.g. Qwen Image 3.0 with FLUX 3, with Midjourney etc.
vunderba · · focus · HN ↗
Prompt results are graded based a weighted calculation which includes: adherence to the prompt, image fidelity, and steerability.
My comparison benchmark also tends to favor prompt adherence, which a lot of others don’t. Most of ones that I've seen tend towards rather simplistic prompts (e.g. "neon-lit city facing a robotic uprising, with high-tech battles, in anime style"), whereas the prompts I've created try to test high specificity.
I’ve been running them all the way back to SDXL.
You can compare specific models using the "View All Models" so if you want to see the progression of open-weight models, or model X vs model Y, you can do so.
Just a heads up - I haven't added Flux 3 as I'm waiting until BFL drops the open-weights version.
Generative Comparisons:
<a href="https://genai-showdown.specr.net" rel="nofollow">https://genai-showdown.specr.net
Editing Comparisons:
<a href="https://genai-showdown.specr.net/image-editing" rel="nofollow">https://genai-showdown.specr.net/image-editing
tangotaylor · · focus · HN ↗
grobibi · · focus · HN ↗
psunavy03 · · focus · HN ↗