3 ms·
I wonder how much more a model would learn about subtitles from including audio AND video in training. Sure, the costs would be way bigger (parsing video even d
by btdmaster 3y ago
I wonder how much more a model would learn about subtitles from including audio AND video in training. Sure, the costs would be way bigger (parsing video even deterministically is 1.5 orders of magnitude worse than audio) but it might help with the edge cases where the speech is so unclear even the subtitle scene can't agree.
- innovatorved 3y ago[flagged]
- socks 3y agothis sounds like chatgpt drivel
- innovatorved 3y agonhh, Google Bard