Andy Reid to TechnologyEnglish • 11 months agoAI companies are violating a basic social contract of the web and and ignoring robots.txtwww.theverge.comexternal-linkmessage-square199arrow-up11.09Karrow-down115cross-posted to: [email protected][email protected][email protected]
arrow-up11.08Karrow-down1external-linkAI companies are violating a basic social contract of the web and and ignoring robots.txtwww.theverge.comAndy Reid to TechnologyEnglish • 11 months agomessage-square199cross-posted to: [email protected][email protected][email protected]
minus-squareEcho DotlinkfedilinkEnglish16•11 months agoLoads of crawlers don’t follow it, i’m not quite sure why AI companies not following it is anything special. Really it’s just to stop Google indexing random internal pages that mess with your SEO. It barely even works for all search providers.
minus-square@General_EffortlinkEnglish3•11 months agoThe Internet Archive does not make a useful villain and it doesn’t have money, anyway. There’s no reason to fight that battle and it’s harder to win.
Loads of crawlers don’t follow it, i’m not quite sure why AI companies not following it is anything special. Really it’s just to stop Google indexing random internal pages that mess with your SEO.
It barely even works for all search providers.
The Internet Archive does not make a useful villain and it doesn’t have money, anyway. There’s no reason to fight that battle and it’s harder to win.