4 ms·
At what resolution? And also, does the output actually resembles the original image? Examples with background other than uniform? Would be nice if they provided
by dhdhhdd 6y ago
At what resolution? And also, does the output actually resembles the original image? Examples with background other than uniform? Would be nice if they provided more than just screenshots
It's not uncommon to see video calls at 100kbs-150kbps, which is ~10KB/s, and this is for 7fps or so, including audio. So "per frame" that would be 1KB or so (more for key frames, less for I frames).
So they say it can be 0.1KB, so better than that... Exciting, if realistic.
Also, add on top audio, and packet overhead :-) there is at least 0.1KB overhead for sending the packet (bundle it with audio if possible!)
- motoboi 6y agoThis is probably first-order-model[1] using keyframes. You send only one image each 2 seconds and the mesh of face 30 times per seconds. Then use first-order-model to extrapolate 2 seconds of video from the keyframe. Rinse, repeat. Very doable. AMAZING! The original first-order-model could not do 30 frames per second, but maybe this Nvidia model has some improvements. 1 - https://aliaksandrsiarohin.github.io/first-order-model-website/ https://aliaksandrsiarohin.github.io/first-order-model-websi...
- dhdhhdd 6y agoNow we are just guessing. This is a press release, which does not contain enough technical info. So one can't say how groundbreaking this is. It's like with those news "the new battery type has been discovered", with very little actual data, just guesswork.