this post was submitted on 27 Sep 2024
787 points (98.4% liked)
Technology
59742 readers
2229 users here now
This is a most excellent place for technology news and articles.
Our Rules
- Follow the lemmy.world rules.
- Only tech related content.
- Be excellent to each another!
- Mod approved content bots can post up to 10 articles per day.
- Threads asking for personal tech support may be deleted.
- Politics threads may be removed.
- No memes allowed as posts, OK to post as comments.
- Only approved bots from the list below, to ask if your bot can be added please contact us.
- Check for duplicates before posting, duplicates may be removed
Approved Bots
founded 2 years ago
MODERATORS
you are viewing a single comment's thread
view the rest of the comments
view the rest of the comments
Aren't these Captchas designed to get training data for AI models anyway?
"System does what it was designed to do" doesn't feel that surprising...
Yes and no, the captchas are just meant to be hard for computers to solve but easier for humans. People saw that, and thought that "if we're making people do this might as well have them do something useful" not meant to be malevolent- and the purpose is still stopping bots, training them is a side-effect.
No, you're wrong, the Traffic Light examples ARE specifically to gather data to train models. Being a good Captcha was just a byproduct of that. If people just wanted a good captcha they wouldn't need hundreds of millions of photos of street lights and bicycles.
No you're wrong, because the sites that embed those captchas on their page are not doing that to help good.
Yes, they are getting something productive out of the human labor that would be done anyways. Trust me as a web developer, and web scraper, some kind of captcha is necessary for many free services to be useful/economically viable. The core of a good captcha is just making it marginally more expensive for the scraper/bot than it is for you.
The sites don't create the captcha, you yourself just said it was embedded there.
They embed for a reason... And the captchas wouldn't exist if they weren't embedded anywhere
Finitebanjo is right. Yes they are used to fight spam and bots but they way they do it us is picked intentionally to train ai.
https://medium.com/@yennhi95zz/how-google-trains-ai-with-your-help-through-captcha-876cb4eb4d01
Also from the Wikipedia article "Google profits from reCAPTCHA users as free workers to improve its AI research." https://en.m.wikipedia.org/wiki/ReCAPTCHA
Yes like I said, the challenges were picked to be useful. But some form of challenge would've been chosen regardless.