2 ms·
That's a very interesting result. Did you happen to capture the seed for either of those first two images? It would be interesting to try to reproduce.
by cheald 4y ago
That's a very interesting result. Did you happen to capture the seed for either of those first two images? It would be interesting to try to reproduce.
- shagie 4y agoAlas no. And I haven't been able to tickle it again in the right way to get those images out. The invocation of that run is still in my scroll back: (venv) shagie@MacM1 stable-diffusion % python scripts/txt2img.py --prompt "wolf with bling walking down a street" --n_samples 6 --n_iter 1 --plms Global seed set to 42 Loading model from models/ldm/stable-diffusion-v1/model.ckpt Global Step: 470000 LatentDiffusion: Running in eps-prediction mode DiffusionWrapper has 859.52 M params. making attention of type 'vanilla' with 512 in_channels Working with z of shape (1, 4, 32, 32) = 4096 dimensions. making attention of type 'vanilla' with 512 in_channels That's the only spot I see the seed mentioned and then it goes on with lots of other logging but nothing seed related that would indicate a way to reproduce it. --- (late edit) you can fairly accurately (so far 1 image out of 20) get that image out with the prompt "Rick Astley Never Gonna Give You Up"
- cheald 4y agoI'm thus far unable to reproduce it. Given: Rick Astley Never Gonna Give You Up Steps: 20, Sampler: PLMS, CFG scale: 7, Seed: 4231695436, Size: 512x512, Batch size: 2, Batch pos: 0 I ran a couple of batches of 32 (64 images total): https://imgur.com/a/74IbCuD https://imgur.com/a/74IbCuD (The images with the nonsensical but obvious Impact font that was learned from memes are quite funny, though) If you can get a full set of parameters (size, sampler, seed, prompt, cfg scale) then I should hopefully be able to reproduce your results, though.