Skip to content
Directory/Categories/Archiver/Arquivo Web Crawler
AR

Arquivo Web Crawler

GoodBotVerifiedBOT

Web crawler archives the Portuguese web Operated by Arquivo.

GoodBot
Assessment
Archiver
Category
BOT
Kind
1
Data sources
AI-generated summary, drawn from Cloudflare, operator docs, and community data

What is Arquivo Web Crawler?

Arquivo Web Crawler is a web crawler operated by Arquivo. Web crawler archives the Portuguese web

Helpful — Verified, safe crawler. Respects robots.txt and provides operator documentation.

Identification

Bot name
Arquivo Web Crawler
Operator
Arquivo
Category
Archiver
Kind
BOT
Verification
Verified bot
Cloudflare slug
arquivo
User-agent patterns
Arquivo-web-crawler
User-agent string
Arquivo-web-crawler (compatible; heritrix/3.4.0-20200304 +https://arquivo.pt/faq-crawling)
Main use cases
Web crawlingData collection

Why it crawls your site

Arquivo Web Crawler crawls websites to archive and preserve web content. It is operated by Arquivo as part of their web archival infrastructure. If you see this bot in your server logs, it is a verified crawler and generally safe.

Block or allow Arquivo Web Crawler

Block it

Add a Disallow rule for Arquivo-web-crawler in your robots.txt file. You can also block at the server level using your web server configuration or CDN firewall rules to filter requests matching the user-agent string.

# Block Arquivo Web Crawler from crawling your entire site
User-agent: Arquivo-web-crawler
Disallow: /

# Allow Arquivo Web Crawler full access
User-agent: Arquivo-web-crawler
Allow: /
Allow & verify

Ensure your robots.txt allows Arquivo-web-crawler. Verify requests are genuine by checking the user-agent string and referring to Arquivo's documentation.

Traffic & trends

Cloudflare Radar traffic

No Cloudflare traffic data available.

Google Trends interest

Loading trend data…

Data sources

This profile is compiled from the following sources.

CloudflareRadar
Last updated September 8, 2026

Something wrong with this entry? Suggest a correction