4 ms·
Using the free playground link, and it is in fact extremely fast. The "diffusion mode" toggle is also pretty neat as a visualization, although I'm not sure how
by chc4 1y ago
Using the free playground link, and it is in fact extremely fast. The "diffusion mode" toggle is also pretty neat as a visualization, although I'm not sure how accurate it is - it renders as line noise and then refines, while in reality presumably those are tokens from an imprecise vector in some state space that then become more precise until it's only a definite word, right?
- PaulHoule 1y agoIt's insane how fast that thing is!
- maelito 1y agoLink : https://chat.inceptionlabs.ai/ https://chat.inceptionlabs.ai/
- sexy_seedbox 1y agoStill cannot pass the stRawbeRRy or the Sally's 1 sister tests unfortunately...
- icyfox 1y agoSome text diffusion models use continuous latent space but they historically haven't done that well. Most the ones we're seeing now typically are trained to predict actual token output that's fed forward into the next time series. The diffusion property comes from their ability to modify previous timesteps to converge on the final output. I have an explanation about one of these recent architectures that seems similar to what Mercury is doing under the hood here: https://pierce.dev/notes/how-text-diffusion-works/ https://pierce.dev/notes/how-text-diffusion-works/
- chc4 1y agoOh neat, thanks! The OP is surprisingly light on details on how it actually works and is mostly benchmarks, so this is very appreciated :)