Andy Reid@lemmy.world to Technology@lemmy.worldEnglish · 9 months agoAI companies are violating a basic social contract of the web and and ignoring robots.txtwww.theverge.comexternal-linkmessage-square200fedilinkarrow-up11.08Karrow-down115cross-posted to: [email protected][email protected][email protected]
arrow-up11.07Karrow-down1external-linkAI companies are violating a basic social contract of the web and and ignoring robots.txtwww.theverge.comAndy Reid@lemmy.world to Technology@lemmy.worldEnglish · 9 months agomessage-square200fedilinkcross-posted to: [email protected][email protected][email protected]
minus-squarewise_pancake@lemmy.calinkfedilinkEnglisharrow-up57·edit-29 months agorobots.txt is a file available in a standard location on web servers (example.com/robots.txt) which set guidelines for how scrapers should behave. That can range from saying “don’t bother indexing the login page” to “Googlebot go away”. IT’s also in the first paragraph of the article.
robots.txt is a file available in a standard location on web servers (example.com/robots.txt) which set guidelines for how scrapers should behave.
That can range from saying “don’t bother indexing the login page” to “Googlebot go away”.
IT’s also in the first paragraph of the article.