13 ms·
FLUX1.1 [pro] – New SotA text-to-image model from Black Forest Labs
- simeon989 2y ago[flagged]
- Mashimo 2y agoOh neat. I wonder if they also improve .schnell and .dev soon. That would be nice :)
- evrim189111 2y agoI think Flux is better than SDXL and Dall e. I tried the models from here https://apps.apple.com/us/app/art-x-a-i-art-generator-aiart/id1644315225 https://apps.apple.com/us/app/art-x-a-i-art-generator-aiart/...
- bobdenver8008 2y ago[flagged]
- davidddef223 2y ago[flagged]
- pieter2222 2y ago[flagged]
- pieter2222 2y ago[flagged]
- pieter2222 2y ago[flagged]
- deleted 2y ago[deleted]
- deleted 2y ago[deleted]
- byteknight 2y agoI won't pay for a model, but that cake image looks dang good.
- mainframed 2y agoAlthough culinarily incorrect :)
- in3d 2y agoBetter link https://blackforestlabs.ai/announcing-flux-1-1-pro-and-the-bfl-api/ https://blackforestlabs.ai/announcing-flux-1-1-pro-and-the-b...
- doctorpangloss 2y agoI'm worried about what happens when more people find out about Ideogram. There are a lot of things that don't appear in ELO scores. For one, they will not reflect that you cannot prompt women's faces in Flux. We can only speculate why.
- giancarlostoro 2y agoHow locked down is it? My problem with a lot of these is I like to make really ridiculous meme type images, but I run into walls for dumb reasons. Like if I want to make something thats "copyrighted" like a mix of certain characters from one franchise or whatever, I cannot sometimes I get told that the model cannot generate copyrighted content, even though courts ruled that AI generated stuff cannot be copyrighted either way... I feel like AI should just be treated as fair use as long as its not 100% blatantly a literal clone of the original work.
- doctorpangloss 2y ago> How locked down is it? ... I get told that the model cannot generate copyrighted... AI should just be treated as fair use Ideogram and Flux both have their own broad set of limitations that are non-technical and unpublished. IMO they are not really motivated by legal concerns, other than the lack of transparency itself. So maybe the issue is that transparency, and that the hazy legal climate means no transparency. You can't go anywhere and see the detailed list of dataset collection and captioning opinions for proprietary models. Open Model Initiative, trying to make a model, did publish their opinions, and they're not getting sued anytime soon. However, their opinions are an endless source of conflict.
- jjordan 2y agoI've been using Venice.ai which offers afaik the most uncensored service currently available, outside of running your own instances. No problem with prompts that include copyrighted terms.
- sdenton4 2y agoIt's perfectly happy to make an imperial storm trooper riding a dragon, for what it's worth
- Jackson__ 2y agoAh, that was one short gravy train even by modern tech company standards. Really wish the space was more competitive and open so it wouldn't just be one company at the top locking their models behind APIs.
- sharkjacobs 2y ago"state of the art" has become such tired marketing jargon. "our most advanced and efficient model yet" "a significant step forward in our mission to empower creators" I get it, you can't sell things if you don't market them, and you can't make a living making things if you don't sell them, but it's exhausting.
- minimaxir 2y agoThe official blog post justifies the marketing copy a bit more with metrics.
- sharkjacobs 2y agoThe point is that the metrics say the thing, this stuff doesn't say actually anything. What does "state of the art" mean? That it's using the latest "cutting edge" model technology? When Apple releases a new iPhone Pro Max, it's "state of the art". When they release a new iPhone SE, there's an argument to be made that it's not because it uses 2 year old chips. But what would it even mean for BFL to release a model which wasn't "state of the art" > our most advanced and efficient model yet Yes, likewise, this is how technology companies work. They release something and then the next thing they release is more advanced. > a significant step forward in our mission to empower creators Going from 12 seconds to 4 seconds is a significant speed boost, but does it move the needle on their mission to empower creators? These are their words, not mine, it's a technical achievement and impressive incremental progress, but are there users out there who are more empowered by this? significantly more empowered!?
- throwaway314155 2y agoHoly shit the level of pedantry. State of the art in this context means it out performs all other models to date on standard evaluations, which is precisely what it does. Did you miss the first flux release? Black forest labs aren't screwing around. The team consists of many of the _actual_ originators of Stable Diffusion's research (which was effectively co-opted by Emad Mostaque who is likely a sociopath).
- jchw 2y agoThe generated images look impressive of course but I can't help but be mildly amused by the fact that the prompt for the second example image insists strongly that the image should say 1.1: > ... photo with the text "FLUX 1.1 [Pro]", ..., must say "1.1", ... ...And of course, it does not.
- thisisnotauser 2y ago[flagged]
- nirav72 2y agoAre there any projects that allow for easy setup and hosting Flux locally? Similar to SD projects like InvokeAI or a1111
- minimaxir 2y agoFlux is more weird than old SD projects since Flux is extremely resource dependant and won't run on most hardware.
- Filligree 2y agoThe GGUF quantisations do run on most recent hardware, albeit at increasingly concerning quality tradeoffs.
- tripplyons 2y agoI haven't noticed any quality degradation with the 8-bit GGUF for Flux Dev, but I'm sure the smaller quantizations perform worse.
- ziddoap 2y agoPeople have Flux running on pretty much everything at this point, assuming you are comfortable waiting 3+ minutes for a 512x512 image. I managed to get it running on an old computer with a 2060 Super, taking ~1.5 minutes per image gen. People are generating on a 1080.
- waffletower 2y agoDoesn't take a lot of effort to get Flux dev/schnell to run on 3090s unquantized, but I agree that 24gb is the consumer GPU memory limit and there are many with less than that. Flux runs great on modern Mac hardware as well, if you have at least 32gb of unified memory.
- stoobs 2y agoI'm running Flux dev fine on a 3080 10GB, unquantised, on windows the nvidia drivers have a function to let it spill over into system ram. It runs a little slower, but it's not a deal-breaker unlike nvidia's pricing and power requirements at the moment
- melvinmelih 2y agoIn case you want to try it out without hassling with the API, I've set up a free tool for it so you can try it out on WhatsApp: https://instatools.ai/products/fluxprovisions https://instatools.ai/products/fluxprovisions
- vessenes 2y agoFlux is so frustrating to me. Really good prompt adherence, strong ability to keep track of multiple parts of a scene, it's technically very impressive. However it seems to have had no training on art-art. I can't get it to generate even something that looks like Degas, for instance. And, I can't even fine tune a painterly art style of any sort into Flux dev. I get that there was working, living artist backlash at SD and I can therefore imagine that the BFL team has decided not to train on art, but, it's a real loss. Both in terms of human knowledge of, say composition, emotion, and so on, but also for style diversity. For goodness sake, the MET in New York has a massive trove of open CC0 type licensed art. Dear BFL, please ease up a bit on this, and add some art-art to your models, they will be better as a result.
- weebull 2y agoI wonder if part of the reason it's good is because it's been trained for a more specific task. I can only imagine that if your concept of a "house" includes range from a stately home to "a pineapple under the sea" you're going to end up with a very generalised concept. It's then takes specific prompting to remove the influences you're not interested in. I suspect the same goes for art styles. There's such huge variety that really they'd be better surveys by separate models.
- throwup238 2y agoI’ve had the same problem with photography styles, even though the photographer I’m going for is Prokudin-Gorskii who used emulsion plates in the 1910s and the entire Library of Congress collection is in the public domain. I’m curious how they even managed to remove them from the training data since the entire LoC is such an easy dataset to access.
- throwaway314155 2y agoi'm fairly confident they did a broad FirstName LastName removal.
- vessenes 2y agoYes, exactly. I think they purposely did not train on stuff like this. I'd bet that you could do a LoRa of Prokudin-Gorskii though; there's a lot of photographic content in flux's training set.
- skybrian 2y agoIt doesn’t get piano keyboards right, but it’s the first image generator I’ve tried that sometimes get “someone playing accordion” mostly right. When I ask for a man playing accordion, it’s usually a somewhat flawed piano accordion, but If I ask for a woman playing accordion, it’s usually a button accordion. I’ve also seen a few that are half-button, half-piano monstrosities. Also, if I ask for “someone playing accordion”, it’s always a woman.
- Der_Einzige 2y agoFar more interesting will be when pony diffusion V7 launches. No one in the image space wants to admit it, but well over half of your user base wants to generate hardcore NSFW with your models and they mostly don’t care about any other capabilities.
- ks2048 2y agoIs there a good site that compares text-to-image models - showing a bunch of examples of text w/ output on each model?
- jeffbee 2y agoI asked for a simple scene and it drew in the exact same AI girl that every text-to-image model wants to draw, same face, same hair, so generic that a Google reverse image search pulls up thousands of the exact same AI girl. No variety of output at all.
- ilaksh 2y agoPretty smart model. Here's one I made: https://replicate.com/p/6ez0x8xqvsrga0cjadg8m7bah0 https://replicate.com/p/6ez0x8xqvsrga0cjadg8m7bah0
- loufe 2y agoThat is astoundingly good adherence to the description. I already liked and was impressed by Flux1 but that is perhaps the most impressive image generation I've ever seen.
- miohtama 2y agoIs it going be able to go head-to-head against Midjourney?
- vunderba 2y agoMJ is by far the worst model for complex prompt ADHERENCE, though it has excellent compositional quality. Comparisons of similar prompt using Midjourney 6.1 https://imgur.com/a/WBnPl7I https://imgur.com/a/WBnPl7I Also, flux (schnell, dev) can be run on your local machine. If you really want to use a paid service, Ideogram is probably the best one out there that balances quality with adherence. DALL-E 3 also has good adherence as well though the quality can sometimes be iffy, and it's very puritanical in terms of censorship.
- drdaeman 2y agoYet, it doesn't seem to know how a Tektronix 4010 actually looks like... ;) I had similar issues trying to paint a "I cast non-magic missile" meme with a fantasy wizard using a missile launcher. No model out there (I've tried SD, SDXL, FLUX.1dev and now this FLUX1.1pro) knows how a missile launcher looks like (neither as a generic term, nor any specific systems) and even has no clue how it's held, so they all draw really weird contraptions.
- morbicer 2y agoIsn't it because the shoulder launched weapon is usually called rocket launcher, rpg or bazooka? Never heard it referred as misille launcher.
- ChrisArchitect 2y agoAnnouncement post: https://blackforestlabs.ai/announcing-flux-1-1-pro-and-the-bfl-api/ https://blackforestlabs.ai/announcing-flux-1-1-pro-and-the-b... (https://news.ycombinator.com/item?id=41730626 https://news.ycombinator.com/item?id=41730626)
- whitehexagon 2y agoI'm running Asahi Linux on a 32GB M1 Pro. Any chance of being able to run text-to-image models locally? I've had some success with LLMs, but only the smaller models. No idea where to start with images, everything seems geared towards msft+nvda.
- collinvandyck76 2y agoDiffusionBee will let you do this quite easily. edit: nevermind, it's a macos app
- lagniappe 2y agoIs DiffusionBee still in development? I had stopped using it because it seemed like the dev interest had stalled.
- collinvandyck76 2y agoIt gets periodic releases, but the source isn't typically updated at the same time.
- LeoPanthera 2y ago"Draw Things" is a native Mac app for text to image. It's a a lot more advanced than DiffusionBee, it will download the models for you, and it's free. It's also available for iOS. (!)
- smcleod 2y agoDraw things is neat but it's so damn slow compared to other tools (e.g. invokeai), I'm not sure why it takes so long to generate images with any model?
- LeoPanthera 2y agoIt's not any slower than invokeai for me. Maybe check the settings, and try using the GPU instead of CoreML.
- ionwake 2y agoSorry to be a noob, but how does this relate to fastflux.ai which seems to work great and creates an image in less than a second? Is this a new model on a slower host?
- jamesc4 2y ago[flagged]
- jamesc4 2y ago[flagged]
- jamesc4 2y ago[flagged]
- nubinetwork 2y agoI tried using schnell, it won't fit in a 16gb GPU, and I couldn't get it to run on CPU.
- TobTobXX 2y agoI've sucessfully run schnell and dev on a 12G GPU. They do take 40s/60s repectively, but it works. I used ComfyUI and didn't have to tweak anything.
- washadjeffmad 2y agoTry an fp8: https://huggingface.co/Kijai/flux-fp8/ https://huggingface.co/Kijai/flux-fp8/
- kindkang2024 2y agoI really enjoy its service. It's promising for UI design. My advocacy website pages' UI design was bootstrapped using it. It is quite good for developers without much design ability. Ironically, I am afraid to type the website out and will keep it unknown here. My account could be suspended because of this. It had already reached -1 karma. It's better to keep my account alive.
- fortran77 2y agoI've been playing with Flux.Dev and such a big step forward from Stable Diffusion and all the other Generative AIs that could run on consumer GPUs. I just tried this Flux1.1 pro page (prompt: "A sad Macintosh user who is upset because his computer can't play games") and was very impressed by the detail and "understanding" this model has.
- basitsoomro123 2y ago[flagged]
- simeon989 2y ago[flagged]