What is ArchiveBot?

ArchiveBot is an intelligence gatherer operated by Wikimedia. Agent Analytics can track when it visits your website.

Overview

Operated By Wikimedia
Source Official Website
Expected To Follow Robots.txt Yes
Insights Last Updated August 11, 2026

Do you operate this agent? Contact us to suggest an update.

Category

Intelligence Gatherer
Analyzes web content for brand safety, competitive insights, and ad targeting

Expected Behavior

ArchiveBot usually returns to a small set of pages rather than crawling an entire site. Visits may follow an hourly, daily, or weekly schedule, while websites its clients do not monitor may never encounter it.

ArchiveBot's User Agent

User Agent ArchiveTeam ArchiveBot/20250806.050c783 (wpull 2.0.3) and not Mozilla/5.0 (Windows NT 6.1; WOW64) AppleWebKit/537.36 (KHTML, like Gecko) Chrome/42.0.2311.90 Safari/537.36

How To Block ArchiveBot With Robots.txt

Add this rule to your robots.txt file to block ArchiveBot from accessing your entire website, or use Automatic Robots.txt to block all intelligence gatherers at once. You can customize which pages are blocked by swapping out / for a different path.

User-agent: ArchiveBot # https://knownagents.com/agents/archivebot
Disallow: /

Global Statistics for ArchiveBot

As of August 11, 2026, this data reflects agent visits measured across thousands of websites using Agent Analytics, combined with daily scans of the top 1000 websites and their robots.txt files.

Robots.txt Blocked Percentage

1%
1% of top websites are blocking ArchiveBot
Learn How →

Country of Origin

United States
ArchiveBot normally visits From the United States

Robots.txt Blocking Trend

1% of top websites block ArchiveBot in their robots.txt files.

Overall Intelligence Gatherer Traffic

3.5% of all web traffic came from intelligence gatherers.

Top Visited Website Categories

News
Health
Business and Industrial
Arts and Entertainment
Law and Government

The types of websites most frequently visited by ArchiveBot.

Frequently Asked Questions

Should I Block ArchiveBot?

If you run ads, confirm ArchiveBot's purpose before blocking it. Some intelligence gatherers measure ad placements or evaluate pages for brand safety, and blocking those agents can interfere with advertising systems. If this one serves a different purpose, weigh any research visibility against sharing business or page-level data with third parties. For context, 1% of the top websites we track currently have robots.txt rules for ArchiveBot.


Does ArchiveBot Follow Robots.txt Rules?

Yes. ArchiveBot is expected to follow robots.txt directives, so a disallow rule is the appropriate first step. Automatic Robots.txt can add and maintain that rule, while Agent Analytics lets you verify whether ArchiveBot respects it.


Does ArchiveBot Access Private Content?

No special access. ArchiveBot can read publicly available pages but cannot bypass authentication. Assume anything you publish openly may be collected.


Why Is ArchiveBot Visiting My Website?

A client of Wikimedia may be gathering competitive intelligence, measuring advertisements, or evaluating content for brand safety. The pages requested reflect that client's objective.


How Can I Tell if ArchiveBot Is Visiting My Website?

Agent Analytics tracks ArchiveBot visits in real time alongside every other known AI agent, crawler, and scraper. You can also check your server logs for requests whose user-agent string contains "ArchiveBot". Look for repeated requests to the same small group of commercial or ad-supported pages. Because ArchiveBot does not publish a verification method, any client can claim its identity and a log match is only a clue.