fattyfoods@feddit.nl to Open Source@lemmy.ml · 2 days agoThe Open-Source Software Saving the Internet From AI Bot Scraperswww.404media.coexternal-linkmessage-square107fedilinkarrow-up1596arrow-down19cross-posted to: technology@beehaw.orgopensource@programming.dev
arrow-up1587arrow-down1external-linkThe Open-Source Software Saving the Internet From AI Bot Scraperswww.404media.cofattyfoods@feddit.nl to Open Source@lemmy.ml · 2 days agomessage-square107fedilinkcross-posted to: technology@beehaw.orgopensource@programming.dev
minus-squarekcweller@feddit.nllinkfedilinkarrow-up86·2 days agoRobots.txt expects that the client is respecting the rules, for instance, marking that they are a scraper. AI scrapers don’t respect this trust, and thus robots.txt is meaningless.
Robots.txt expects that the client is respecting the rules, for instance, marking that they are a scraper.
AI scrapers don’t respect this trust, and thus robots.txt is meaningless.