What is Automattic Analytics Crawler?

Automattic Analytics Crawler is a web crawler operated by WordPress (Automattic) that gathers analytics and intelligence data from websites. This bot collects information for WordPress.com's analytics and data gathering purposes. You can set up Agent Analytics to see when Automattic Analytics Crawler visits your website.

Overview

Operated By WordPress
Source Official Website
Expected To Follow Robots.txt Yes
Insights Last Updated September 14, 2026

Do you operate this agent? Contact us to suggest an update.

Category

Intelligence Gatherer
Analyzes web content for brand safety, competitive insights, and ad targeting

Expected Behavior

Automattic Analytics Crawler's traffic pattern depends on the intelligence being collected. It may revisit specific pages on a schedule, crawl broader sections of your site, or appear only when information is requested. Its request scope and frequency reflect the underlying objective.

Automattic Analytics Crawler's User Agent

User Agent Automattic Analytics Crawler/0.2; http://wordpress.com/crawler/

How To Block Automattic Analytics Crawler With Robots.txt

Add this rule to your robots.txt file to block Automattic Analytics Crawler from accessing your website, or use Automatic Robots.txt to block all intelligence gatherers at once. You can customize which pages are blocked by swapping out / for a different path.

User-agent: Automattic Analytics Crawler # https://knownagents.com/agents/automattic-analytics-crawler
Disallow: /

Insights for Automattic Analytics Crawler

As of September 14, 2026, this data reflects agent visits measured across thousands of websites using Agent Analytics, combined with daily scans of the top 1000 websites and their robots.txt files.

Robots.txt Blocked Percentage

6%
6% of top websites are blocking Automattic Analytics Crawler
Learn How →

Country of Origin

United States
Automattic Analytics Crawler normally visits From the United States

Robots.txt Blocking Trend

6% of top websites block Automattic Analytics Crawler in their robots.txts.

Overall Intelligence Gatherer Traffic

3.3% of all web traffic came from intelligence gatherers.

Top Visited Website Categories

Home and Garden
Hobbies and Leisure
Books and Literature
Online Communities
People and Society

The types of websites most frequently visited by Automattic Analytics Crawler.

Frequently Asked Questions

Should I Block Automattic Analytics Crawler?

Confirm Automattic Analytics Crawler's purpose before blocking it. Intelligence gatherers may collect competitive research, measure advertisements, or evaluate pages for brand safety. Weigh any benefit against sharing business or page-level data with its operator or clients. For context, 6% of the top websites we track currently have robots.txt rules for Automattic Analytics Crawler.


Does Automattic Analytics Crawler Follow Robots.txt Rules?

Yes. Automattic Analytics Crawler is expected to follow robots.txt rules, so a disallow rule is the appropriate first step. Automatic Robots.txt can add and maintain that rule, while Agent Analytics lets you verify whether Automattic Analytics Crawler respects it.


Does Automattic Analytics Crawler Access Private Content?

No special access. Automattic Analytics Crawler can read publicly available pages but cannot bypass authentication. Assume anything you publish openly may be collected.


How Can I Tell if Automattic Analytics Crawler Is Visiting My Website?

Agent Analytics tracks Automattic Analytics Crawler visits in real time. You can also check your server logs for requests whose user-agent string contains "Automattic Analytics Crawler". Traffic may range from repeated requests to the same pages to broader crawls or one-off fetches. Because Automattic Analytics Crawler does not publish a verification method, any client can claim its identity and a log match is only a clue.


Why Is Automattic Analytics Crawler Visiting My Website?

WordPress or one of its clients may be gathering competitive intelligence, measuring advertisements, or evaluating content for brand safety. The pages requested reflect the underlying objective.