5 ms·
Other commenters have talked about labelling. Maybe labelling of real life data is something they're trying to do; but from my experience with hCaptcha the chal
by bearbin 5y ago
Other commenters have talked about labelling. Maybe labelling of real life data is something they're trying to do; but from my experience with hCaptcha the challenges are _NOT_ real life data. They're AI-generated images which bear a passing resemblance to the targets but if you look closer nothing adds up at all.
Here are a couple of examples:
https://bearbin.net/images/captcha/1.png https://bearbin.net/images/captcha/1.png
https://bearbin.net/images/captcha/2.png https://bearbin.net/images/captcha/2.png
https://bearbin.net/images/captcha/3.png https://bearbin.net/images/captcha/3.png
https://bearbin.net/images/captcha/4.png https://bearbin.net/images/captcha/4.png
- jonnycomputer 5y agoThat is interesting.
- dylan604 5y agoI particularly like the boats on water images where the horizon is just wrong.
- zapt02 5y agoFascinating insight!
- dyeje 5y agoSeems like that is another kind of labeling to me: is our generated image good enough to fool a human?
- monkeybutton 5y agoGANs with human evaluation of the discriminator.
- nicce 5y agoExactly. That is another way to improve accuracy once you have done it ”in a regular” way already. You can look for synthetic image generation and it’s benefits on model accuracy and optimization.
- FanaHOVA 5y agoNot sure they are fooling anyone. It's more like "is our generated image good enough to make a human recognize what it is to get rid of an annoying pop up?". If there were actually consequences to getting it right/wrong people would pay more attention I'm sure.
- rg111 5y agoFor many generative models, this is on the way to become a standard- using Humans as a judge of generated material, and this is not limited to Computer Vision either. I am about to use this technique to judge the sanity of text generated by a Transformer model for a paper that I am writing (with a small group). There are also attempts to properly standardize it, and this is called- HYPE [0]. And there are big names like Fei-Fei Li and Michael Bernstein behind it. [0]: https://arxiv.org/abs/1904.01121 https://arxiv.org/abs/1904.01121
- amirhirsch 5y agoThe generated images are those that provide the neural network with optimal loss reduction when tagged.
- renewiltord 5y agoReally cool observation!
- jspaetzel 5y agoThe broken images to me look like instances where two or more cameras or images were used and then stitched together. Probably also done while the camera and object are moving making it more likely to be wonky.
- gwern 5y agoNo way. The letters/writing look exactly like mirrored GAN output. That's not what would happen with blur or stitching together (there would be no mirroring symmetry or all the '8' letters), or with synthetic 'machine teaching' datapoints either (as far as I've ever seen). Look at the cat StyleGAN sometime if you don't know what I'm talking about. Which leaves me wonder what the point is. If you are generating GAN images per CIFAR or ImageNet class, you know what the label is and don't need to label it. Perhaps they just generate lots of images to fill up the pipeline for the CAPTCHAs, to avoid reuse which could be exploited by spammers, when they have too little paying work?
- 01acheru 5y agoI think that might be something they actually do. Lately it happened to me see pictures of boats on land, bicycles merged with surrounding objects, weird proportions, and usually those strange images are extremely pixelated with those strange reddish or greenish fluo pixels that appear on generative network images. But other times the pictures are 100% real life images.
- austinjp 5y agoVery nice. It looks like some sort of training to defend against bots that can currently pass hCaptcha..? If so, I wonder how long that particular arms race can last. I've not seen anything like this in the wild. And.... well, now I'm curious about how you had these examples to hand. When and why did you start collecting them?