Andy Reid@lemmy.world to Technology@lemmy.worldEnglish · 9 months agoAI companies are violating a basic social contract of the web and and ignoring robots.txtwww.theverge.comexternal-linkmessage-square197fedilinkarrow-up11.09Karrow-down115
arrow-up11.08Karrow-down1external-linkAI companies are violating a basic social contract of the web and and ignoring robots.txtwww.theverge.comAndy Reid@lemmy.world to Technology@lemmy.worldEnglish · 9 months agomessage-square197fedilink
minus-squarewise_pancake@lemmy.calinkfedilinkEnglisharrow-up58·edit-29 months agorobots.txt is a file available in a standard location on web servers (example.com/robots.txt) which set guidelines for how scrapers should behave. That can range from saying “don’t bother indexing the login page” to “Googlebot go away”. IT’s also in the first paragraph of the article.
robots.txt is a file available in a standard location on web servers (example.com/robots.txt) which set guidelines for how scrapers should behave.
That can range from saying “don’t bother indexing the login page” to “Googlebot go away”.
IT’s also in the first paragraph of the article.