5 ms·
Sounds like it’ll be downloadable. FTA: > In early, unoptimized inference tests on consumer hardware our largest SD3 model with 8B parameters fits into the 24G
by sen 3y ago
Sounds like it’ll be downloadable. FTA:
> In early, unoptimized inference tests on consumer hardware our largest SD3 model with 8B parameters fits into the 24GB VRAM of a RTX 4090 and takes 34 seconds to generate an image of resolution 1024x1024 when using 50 sampling steps. Additionally, there will be multiple variations of Stable Diffusion 3 during the initial release, ranging from 800m to 8B parameter models to further eliminate hardware barriers.
- finnjohnsen2 3y agoThanks for pointing that out. Super promising.
- nuz 3y agoThe 800m model is super exciting
- jncfhnb 3y agoIt will probably suck. These models aren’t quite good enough for most tasks (other than toy fun exploration). They’re close in the sense that you can get there with a lot of work and dice rolling. But I would be pessimistic about a smaller model actually getting you where you want.
- cooper_ganglia 3y agoYeah, SDXL is probably better than SD3 800M if I had to guess. I’m looking forward to the quality advancements with LCM Loras, or an SD3 Turbo!
- tripplyons 3y agoIt's has about as many parameters as SD v1.5, but hopefully with a better architecture, so I think it could end up being better for VRAM-constrained users than SD v1.5.
- bufferoverflow 3y agoNot really. Look at SDXL Turbo. Sure, it's fast, but the images it produces are not very good. I'm not even sure what the use case is.
- nuz 3y agoSDXL turbo is great in my experience! Way better than any alternative at that speed (e.g. sd1.4 or sd2). For img2img it's fantastic. And since the 800m model is using new insights since then and generally a cleaner dataset from the looks of it, I could imagine it's decent for some tasks (or better than sd2 turbo at least, which is enough to be fun and useful in my eyes).
- bufferoverflow 3y agoBut what's the use case? I'd rather wait 30 seconds and get a much higher quality image than some mediocre image in 1 second. Hell, even if it took 5 minutes per image, and produced even better images, I would prefer that.
- declaredapple 3y agoA lot of use cases are cost-limited. Dalle3 makes great images but costs $0.12 per image so it would get extremely expensive at scale (1k generations is already 120$). The cost is by gpu time, the faster you can generate it the cheaper it is. We can get images under 250ms now, which is fast enough to fit in a web request. Some use cases might be generating profile pictures or banners for users or unique profile pictures for bots in online games. Discord, steam, social media, whatnot, you could just type what you want your profile picture to be and make it on the fly. They're small, aren't expected to be extremely high quality, and cheap enough. Testing on https://fastsdxl.ai/ https://fastsdxl.ai/ - "high quality profile picture of a cartoon cat holding a Bouquet of flowers" To be clear it's not perfect, but this is a fairly complex prompt and I find the majority of seeds would be "good enough" for thumbnail profile pictures. I think we're almost there for "cheap good enough" usecases.
- 3y ago