3 ms·
What you can say is if there are N videos and k types of transforms that are used to "hide" content this would be hypothetically perfect with O(n + k) train exa
by highd 10y ago
What you can say is if there are N videos and k types of transforms that are used to "hide" content this would be hypothetically perfect with O(n + k) train examples - naive "reporting only" would be O(nk). This means that you should just need one example of each transform and one known instance of a particular copyright infringing media in order to get high performance.
Mirroring etc. would just be one example of the transforms an approach like this would be robust to.
It's hard to say what the error rates will be, but you should be on the last 5-10% Pd/Pfa pretty much instantly and then if you want more you can just acquire more true labelled data (if a user can find a copyrighted video on your site, I would think a contractor could, too!).
- rifung 10y agoThis well beyond what I understand so excuse me if I'm misunderstanding but doesn't this imply that they'd need to guess what k transforms might be used? This just seems like a cat and mouse type ordeal which I can only assume is something they would like to avoid as the complexity of and costs of running the system only gets worse, while the complexity does not add up on the side of those making pirated videos.