4 ms·
Finally someone who did this, I've always thought this was a low hanging fruit. I wonder if you could make interesting and quickly trained diffusion models with
by naillo 3y ago
Finally someone who did this, I've always thought this was a low hanging fruit. I wonder if you could make interesting and quickly trained diffusion models with this trick.
- waldarbeiter 3y agoNote that it is from 2018. As someone here already mentioned there is a paper that applies the same idea to Vision Transformers published this year [1]. [1] https://openaccess.thecvf.com/content/CVPR2023/papers/Park_RGB_No_More_Minimally-Decoded_JPEG_Vision_Transformers_CVPR_2023_paper.pdf https://openaccess.thecvf.com/content/CVPR2023/papers/Park_R...
- liuliu 3y agoSorta of? Latent diffusion models operate in a compressed latent space, which is just a richer / learnable representation than DCT.