3 ms·
It’s not live, but it’s in the realm of outputs I would expect from a GPT trained on video embeddings. Implying they’ve solved single token latency, however, i
by valine 3y ago
It’s not live, but it’s in the realm of outputs I would expect from a GPT trained on video embeddings.
Implying they’ve solved single token latency, however, is very distasteful.
- zozbot234 3y agoOP says that Gemini had still images as input, not video - and the dev blog post shows it was instructed to reply to each input in relevant terms. Needless to say, that's quite different from what's implied in the demo, and at least theoretically is already within GPT's abilities.
- valine 3y agoHow do you think the cup demo works? Lots of still images?
- watusername 3y agoA few hand-picked images (search for "cup shuffling"): https://developers.googleblog.com/2023/12/how-its-made-gemini-multimodal-prompting.html https://developers.googleblog.com/2023/12/how-its-made-gemin...
- valine 3y agoHoly crap that demo is misleading. Thanks for the link.