Andy Reid@lemmy.world to Technology@lemmy.worldEnglish · 10 months agoAI companies are violating a basic social contract of the web and and ignoring robots.txtwww.theverge.comexternal-linkmessage-square173fedilinkarrow-up11.05Karrow-down114cross-posted to: technology@beehaw.orgwolnyinternet
arrow-up11.04Karrow-down1external-linkAI companies are violating a basic social contract of the web and and ignoring robots.txtwww.theverge.comAndy Reid@lemmy.world to Technology@lemmy.worldEnglish · 10 months agomessage-square173fedilinkcross-posted to: technology@beehaw.orgwolnyinternet
minus-squareEcho Dot@feddit.uklinkfedilinkEnglisharrow-up16arrow-down1·10 months agoLoads of crawlers don’t follow it, i’m not quite sure why AI companies not following it is anything special. Really it’s just to stop Google indexing random internal pages that mess with your SEO. It barely even works for all search providers.
minus-squareGeneral_Effort@lemmy.worldlinkfedilinkEnglisharrow-up3·10 months agoThe Internet Archive does not make a useful villain and it doesn’t have money, anyway. There’s no reason to fight that battle and it’s harder to win.
Loads of crawlers don’t follow it, i’m not quite sure why AI companies not following it is anything special. Really it’s just to stop Google indexing random internal pages that mess with your SEO.
It barely even works for all search providers.
The Internet Archive does not make a useful villain and it doesn’t have money, anyway. There’s no reason to fight that battle and it’s harder to win.