• wise_pancake@lemmy.ca
    link
    fedilink
    English
    arrow-up
    58
    ·
    edit-2
    10 months ago

    robots.txt is a file available in a standard location on web servers (example.com/robots.txt) which set guidelines for how scrapers should behave.

    That can range from saying “don’t bother indexing the login page” to “Googlebot go away”.

    IT’s also in the first paragraph of the article.