4 ms·
Important note from the paper - the resolution is limited to 384x384 currently.
by cube2222 2y ago
Important note from the paper - the resolution is limited to 384x384 currently.
- just-ok 2y agoSeems like a massive buried lede in an “outperforms the previous SoTA” paper.
- vunderba 2y agoOuch, that's even smaller than the now-ancient SD 1.5 which is mostly 512x512.
- franktankbank 2y agoGreat for generating favicons!
- jimmyl02 2y agodon't most architectures resolve this via superscaling / some up scaling pipeline after that adds the details? iirc stable diffusion xl uses a "refiner" after initial generation
- dragonwriter 2y agoThe SDXL refiner is not an upscaler, it's a separate model with the same architecture used at the same resolution as the base model that is focussed more on detail and less on large scale generation (you can actually use any SDXL-derived model as a refiner, or none; most community SDXL derivatives use a single model with no refiner and beat the Stability SDXL base/SDXL refiner combination in quality.)
- ilaksh 2y agoThe obvious point of a model that works like this is to see if you can get better prompt understanding. Increasing the resolution in a small model would decrease the capacity for prompt adherence.