FLUX 3 Image

(bfl.ai)

211 points by minimaxir1 day ago

19 comments

  • vunderba3 hours ago
    One of the things they seem to be emphasizing here is the UX around being able to place specific elements where you want them in an image. If the positions of the components in the overall composition are very important, this seems to make that a lot easier and kind of reminds me of InvokeAI.<p>Ideogram V4, an open-weight model released back in June can also do this [1], but you have to use a relatively cumbersome JSON structure to describe all the different bounding boxes. So it’s definitely a bit of a hassle.<p>I&#x27;ll probably be waiting until it goes open-weight (hopefully soon) like they did with Flux.2 &#x2F; Klein.<p>[1] - <a href="https:&#x2F;&#x2F;docs.ideogram.ai&#x2F;using-ideogram&#x2F;getting-started&#x2F;prompting-guide&#x2F;4.-json-prompting-ideogram-4.0" rel="nofollow">https:&#x2F;&#x2F;docs.ideogram.ai&#x2F;using-ideogram&#x2F;getting-started&#x2F;prom...</a>
    • kranke1552 hours ago
      You just get an LLM to do the bounding box stuff or use the ComfyUI node that provides a GUI for bounding box generation
      • CuriouslyC2 hours ago
        I always hated Comfy&#x27;s node based UI, but agents make it tolerable. Now I just have them set up a workflow and I go in and tweak it manually if the results aren&#x27;t where I want them. I even have agents cherry doing multiple runs and cherry picking the best outputs, models have gotten good enough that it&#x27;s a real time saver, assuming you have references they can and a rubric to check against.
        • swiftcoder47 minutes ago
          &gt; I always hated Comfy&#x27;s node based UI<p>It&#x27;s one of the most uniquely hostile user experiences I&#x27;ve ever had the (dis)pleasure of working with
        • hdjrudni47 minutes ago
          How do you get agents to set up a workflow? You just get them to modify the JSON directly and then import it, or do you have a tighter integration (e.g. in the UI)?
          • CuriouslyC18 minutes ago
            The agents can interact with Comfy via API pretty well, which afaik ends up being directly with JSON.
      • vunderba2 hours ago
        Yeah, given how much better Ideogram v4 outputs are when you use the proper structured JSON (background, elements, etc.) I think most users probably have stuck some kind of Qwen&#x2F;Gemma-based LLM between their raw prompt and the CLIP encoder.
    • NBJack42 minutes ago
      Adobe has had something like this for a while with Generative Fill. Drag a box, specify contents.
    • popalchemist2 hours ago
      Ideogram 4.5 released 2 days ago with more features along these lines<p><a href="https:&#x2F;&#x2F;www.youtube.com&#x2F;watch?v=2mecWZgbaEg" rel="nofollow">https:&#x2F;&#x2F;www.youtube.com&#x2F;watch?v=2mecWZgbaEg</a>
  • arnaudsm3 hours ago
    The UX looks amazing and very steerable, congrats to the team for focusing on the interface.<p>Chats can be awful user interfaces.
  • vergessenmir3 hours ago
    I think we are all waiting for the open weights or local model releases.
    • pixelesque2 hours ago
      Has that been announced?<p>The website mentions:<p>&gt; FLUX 3 Image is available under a commercial weights license for companies running image generation at scale. Fine-tune and deploy it on your own infrastructure. Reach out to us to learn more.<p>I guess the open ones would be non-commercial?
      • JimDabell1 hour ago
        &gt; Open Weights version of FLUX 3 Image is launching in the coming weeks.<p>— <a href="https:&#x2F;&#x2F;x.com&#x2F;bfl_ai&#x2F;status&#x2F;2105734605621825738" rel="nofollow">https:&#x2F;&#x2F;x.com&#x2F;bfl_ai&#x2F;status&#x2F;2105734605621825738</a>
      • vunderba2 hours ago
        If it is anything like the previous release Flux.2 [dev] - then yeah it&#x27;ll probably be a non-commercial license.<p><a href="https:&#x2F;&#x2F;bfl.ai&#x2F;legal&#x2F;non-commercial-license-terms" rel="nofollow">https:&#x2F;&#x2F;bfl.ai&#x2F;legal&#x2F;non-commercial-license-terms</a>
        • LordDragonfang2 hours ago
          There&#x27;s been an increasing trend of previously-open-weight models going closed-source once they reach a certain size (and size is proportional to capital investment). That the latest release will be open-weight is not necessarily a given just because the previous one was.
          • vunderba41 minutes ago
            It was literally in the announcement from Black Forest Labs:<p>&quot;Open Weights version of FLUX 3 Image is launching in the coming weeks.&quot;<p><a href="https:&#x2F;&#x2F;nitter.cf&#x2F;bfl_ai&#x2F;status&#x2F;2105734605621825738" rel="nofollow">https:&#x2F;&#x2F;nitter.cf&#x2F;bfl_ai&#x2F;status&#x2F;2105734605621825738</a>
          • trentor46 minutes ago
            What trend? BFL was always non commercial. The &quot;only&quot; trend would be qwen image not being apache anymore.
  • skybrian1 hour ago
    This isn&#x27;t much of a test, but I bought $10 in credits on their playground and generated a test image. Not bad, but it didn&#x27;t get the accordion keyboard right. Haven&#x27;t tried editing yet.<p><a href="https:&#x2F;&#x2F;pages.skybrian.com&#x2F;flux3-image-test&#x2F;" rel="nofollow">https:&#x2F;&#x2F;pages.skybrian.com&#x2F;flux3-image-test&#x2F;</a>
    • swiftcoder46 minutes ago
      I wish they would work on fixing the &quot;studio lighting&quot; sheen that all these AI-generated humans have
    • Doohickey-d31 minutes ago
      I also don&#x27;t think there is a park that looks like that, in that location relative ton the Eiffel Tower (although I could be wrong).
    • htx6191 hour ago
      [dead]
  • KazaNLP2 hours ago
    Agreed with other comments about the UX. I&#x27;m more interested in that than the model itself. Would like to start seeing UI like this where you get to choose the model and compare different models. Can&#x27;t jump all over the internet to each model developers sandbox just to test their models. Doing it from one place would be nice.
  • armcat2 hours ago
    Does anyone know if it can be used to generate accurate frame-by-frame sprite sequences? I found that no image model can do this well (with sufficient fidelity) - neither with one shot (full spritesheet), nor single frame conditioning. It would be great if an imagegen model could do this. What I do now (I use my own tool <a href="https:&#x2F;&#x2F;github.com&#x2F;acatovic&#x2F;ai-game-studio" rel="nofollow">https:&#x2F;&#x2F;github.com&#x2F;acatovic&#x2F;ai-game-studio</a>) is basically generate a reference image, then condition on that image to generate a very short video, then extract and prune frames. Then I get indie-level sprite fidelity about 90% of the time.
    • popalchemist2 hours ago
      The task you&#x27;re describing is a video model task, not an image model task. It&#x27;s inherently temporal.<p>Generate a sprite in an image editor, then use a video model to make the loop you want; then turn the resulting video back into individual sprite images.
      • armcat1 hour ago
        Sure and that&#x27;s what I do, but a video can be seen as a causal generation on discreet sequence of images, each image conditioned on the one before it. It can also be seen as a series of image editing tasks. It would be cool to get this working in imagegen because of the amount of control you would get. Right now with video generation you can at best specify start and end frame and hope for the best.
        • popalchemist52 minutes ago
          Image edit models can probably do a grid, but the temporal accuracy &#x2F; coherence will never match what a video model, which is really a world model, can do.
  • gAI2 hours ago
    Love to see an AI lab outside of US&#x2F;China releasing good models.
  • minimaxir2 hours ago
    Of note is the OpenRouter endpoint has a promotional 50% discount, which is rare on image models: <a href="https:&#x2F;&#x2F;openrouter.ai&#x2F;black-forest-labs&#x2F;flux-3-image" rel="nofollow">https:&#x2F;&#x2F;openrouter.ai&#x2F;black-forest-labs&#x2F;flux-3-image</a>
    • vunderba48 minutes ago
      I think that’s less a product of OpenRouter’s generosity and more a result of promotional pricing coming directly from BFL, since other third-party vendors have it as well (Fal.ai, etc.).<p><a href="https:&#x2F;&#x2F;bfl.ai&#x2F;pricing" rel="nofollow">https:&#x2F;&#x2F;bfl.ai&#x2F;pricing</a>
    • pwillia71 hour ago
      open weights or bust
  • neals3 hours ago
    It&#x27;s this a new model or a new ui?
    • swiftcoder44 minutes ago
      Porque no los dos? Presumably you need a model conditioned on the bounding box input to make effective use of the new UI
  • Trufa2 hours ago
    So much negativity as usual and so little talk about the product, this is pretty impressive, well done, it seems to be filling decently a gap that everyone that has worked enough generating images with AI has faced.
    • doctorpangloss1 hour ago
      there basically isn&#x27;t any authentic use for image generation, i would hardly say the negativity is unfounded
  • mromanuk1 hour ago
    For a moment I was confused that this was a release of a new stable diffusion. What happened with Stable Diffusion?
    • yorwba1 hour ago
      Some of the original developers left Stability AI to found Black Forest Labs (so in a sense this is a successor to Stable Diffusion) and Stability AI pivoted to audio and milking their existing models.
  • fuzzythrowaway2 hours ago
    I like the interface; very useful for some use cases that would otherwise be quite frustrating. Dislike that it&#x27;s yet another platform held back by arbitrary moderation. You can&#x27;t make a bicycle for the mind that locks if you try to ride it in the wrong direction.
  • htrp3 hours ago
    <a href="https:&#x2F;&#x2F;news.ycombinator.com&#x2F;item?id=49031796">https:&#x2F;&#x2F;news.ycombinator.com&#x2F;item?id=49031796</a><p>What&#x27;s new from the last post? GA?
    • zamadatix3 hours ago
      The original article mentions:<p>&gt; We will open up an early access phase for FLUX 3 Image in the following weeks.<p>Not sure if there was a separate post for early access or if they just skipped to this.
  • reilly300019 hours ago
    That sort of steering ability that has been possible with the latest Gemini releases has been nice to work with over previous generations. It’s great to see this improve on the platform with declarative controls built into the API and coming soon as an open model.
  • assimpleaspossi2 hours ago
    I shouldn&#x27;t have to scroll all the way down, click to use the thing, then fumble around to figure out what this does. It should be clear at the top of the first page (so I know right away that I don&#x27;t need this).
    • thorum1 hour ago
      Not sure what you mean. The top of the linked page is a video showing the product being used. Immediately below are the words “Compose images from scratch” and more examples, followed by “Text-to-image with strong prompt following and a native understanding of composition.” This is all within the first 100 words of content on the page.
      • bonoboTP48 minutes ago
        I agree that it&#x27;s clear for me and most techies, but maybe for a more general audience they could have added &quot;AI image generation&quot; or some variant of &quot;generative AI&quot; or &quot;image generation model&quot; or something like that. You can also compose images from scratch in Photoshop and Blender.
    • Mashimo56 minutes ago
      &gt; Just write a prompt. Text-to-image with strong prompt following and a native understanding of composition.<p>How is this not clear?
    • richardfulop2 hours ago
      It is a version 3 of a product.If you actually used or needed these, you would have known right away what this is and how it works.
      • nazgulsenpai2 hours ago
        Unless the version 3 announcement is your onboarding point, like it is for me just now. I can go research FLUX from here, but the original poster&#x27;s point still holds.
      • assimpleaspossi1 hour ago
        But I don&#x27;t use it or need it and wasted my time trying to figure out what it did. Clarity should be up front.
  • amelius2 hours ago
    Yes, you can do that with AI now.
  • Grimblewald21 hours ago
    looks cool, eternally greatful these models are marked for open weight releases. Pretty excited
  • imgbenchdude1 hour ago
    Tried Flux 3 on a tiny subjective image benchmark I’m calling One Knee Wonder.<p>Exact prompt:<p>Generate a photorealistic image of M81 urban BDU camouflage cargo trousers, shown by themselves. One trouser leg should be posed with the knee lifted 30° from vertical.<p>Accurate reproduction of the M81 urban camouflage pattern is critical. Match its colors, shapes, scale, distribution, and overall appearance as faithfully as possible.<p>No person, other clothing, or props.<p>Ground truth swatch: <a href="https:&#x2F;&#x2F;commons.wikimedia.org&#x2F;wiki&#x2F;File:US_City_Camo_(M81_Urban).png" rel="nofollow">https:&#x2F;&#x2F;commons.wikimedia.org&#x2F;wiki&#x2F;File:US_City_Camo_(M81_Ur...</a><p>Gemini 3 Pro Image (stronger pattern): <a href="https:&#x2F;&#x2F;i.postimg.cc&#x2F;bZNQYYjx&#x2F;2026-10-02-google-gemini-3-pro-image-n1.png" rel="nofollow">https:&#x2F;&#x2F;i.postimg.cc&#x2F;bZNQYYjx&#x2F;2026-10-02-google-gemini-3-pro...</a> Flux 3 (this run): <a href="https:&#x2F;&#x2F;i.postimg.cc&#x2F;Xr7wNN0H&#x2F;2026-10-02-black-forest-labs-flux-3-image-n1.png" rel="nofollow">https:&#x2F;&#x2F;i.postimg.cc&#x2F;Xr7wNN0H&#x2F;2026-10-02-black-forest-labs-f...</a><p>Flux 3 gets greyscale urban-ish trousers and a lifted knee, but the blotches aren’t real M81 Urban — softer &#x2F; wrong geometry vs the swatch. Not the worst I’ve seen on this prompt; clearly behind the Gemini 3 Pro Image example above on pattern.<p>Curious what other models do on the same prompt.
  • myself2483 hours ago
    Unrelated to the flux images used for floppy disk archiving? Sigh.
    • pizzafeelsright2 hours ago
      Every new piece of software needs a name.<p>Flux is in the top 9000 of the most common words.