Internet Archive - Archive-It
GoodBotVerifiedBOTInternet Archive’s Archive-It service preserves publicly accessible web pages for the historical record. Operated by Archive-It.
What is Internet Archive - Archive-It?
Internet Archive - Archive-It is a web crawler operated by Archive-It. Internet Archive’s Archive-It service preserves publicly accessible web pages for the historical record.
Identification
Archive-ItWhy it crawls your site
Internet Archive - Archive-It crawls websites to archive and preserve web content. It is operated by Archive-It as part of their web archival infrastructure. If you see this bot in your server logs, it is a verified crawler and generally safe.
Block or allow Internet Archive - Archive-It
Add a Disallow rule for Archive-It in your robots.txt file. You can also block at the server level using your web server configuration or CDN firewall rules to filter requests matching the user-agent string.
# Block Internet Archive - Archive-It from crawling your entire site User-agent: Archive-It Disallow: / # Allow Internet Archive - Archive-It full access User-agent: Archive-It Allow: /
Ensure your robots.txt allows Archive-It. Verify requests are genuine by checking the user-agent string and referring to Archive-It's documentation.
Traffic & trends
No Cloudflare traffic data available.
Loading trend data…
Data sources
This profile is compiled from the following sources.