SpiderLing
GoodBotBotwhois.orgA specialized web crawler used to collect text data for building linguistic corpora and conducting lexicographical analysis.
Identification
How we rated this bot
Identifies itself honestly, and we found no credible reports of abuse.
What we read
- Meta Has New Web-Crawling Bots That Sneakily Get Around Robots.txt - Business Insider
- What Is a robots.txt File? (And When To Use One)
- SPIDERLING STUDIOS LIMITED overview - Find and update company information - GOV.UK
- Outlook for Windows | Microsoft 365
- Microsoft Outlook Personal Email and Calendar | Microsoft 365
Block or allow SpiderLing
Add a Disallow rule for Mozilla/5.0 (compatible; SpiderLing; +https://www.sketchengine.eu/crawler/) in your robots.txt file. You can also block at the server level using your web server configuration or CDN firewall rules to filter requests matching the user-agent string.
# Block SpiderLing from crawling your entire site User-agent: Mozilla/5.0 (compatible; SpiderLing; +https://www.sketchengine.eu/crawler/) Disallow: / # Allow SpiderLing full access User-agent: Mozilla/5.0 (compatible; SpiderLing; +https://www.sketchengine.eu/crawler/) Allow: /
Ensure your robots.txt allows Mozilla/5.0 (compatible; SpiderLing; +https://www.sketchengine.eu/crawler/). Verify requests by checking the user-agent string.
Traffic & trends
No Cloudflare traffic data available.
Loading trend data…
Links & references
Data sources
This profile is compiled from the following sources.
Something wrong with this entry? Suggest a correction