What is scraping@nytimes.com?
NYTimes.com newsroom scraping bot collects publicly available, non-copyrighted data for journalistic projects including election result tracking, COVID-19 data aggregation, and other news analytics initiatives. You can set up Agent Analytics to see when scraping@nytimes.com visits your website.
Overview
| Operated By | The New York Times |
| Source | Official Website |
| Expected To Follow Robots.txt | Yes |
| Insights Last Updated | September 15, 2026 |
Do you operate this agent? Contact us to suggest an update.
Category
Expected Behavior
scraping@nytimes.com's traffic pattern depends on the intelligence being collected. It may revisit specific pages on a schedule, crawl broader sections of your site, or appear only when information is requested. Its request scope and frequency reflect the underlying objective.
How To Block scraping@nytimes.com With Robots.txt
Add this rule to your robots.txt file to block scraping@nytimes.com from accessing your website, or use Automatic Robots.txt to block all intelligence gatherers at once. You can customize which pages are blocked by swapping out / for a different path.
User-agent: scraping@nytimes.com # https://knownagents.com/agents/scrapingnytimes-com
Disallow: /
Insights for scraping@nytimes.com
As of September 15, 2026, this data reflects agent visits measured across thousands of websites using Agent Analytics, combined with daily scans of the top 1000 websites and their robots.txt files.
Robots.txt Blocked Percentage
Country of Origin
Robots.txt Blocking Trend
7% of top websites block scraping@nytimes.com in their robots.txts.
Overall Intelligence Gatherer Traffic
3.2% of all web traffic came from intelligence gatherers.
Frequently Asked Questions
Should I Block scraping@nytimes.com?
Confirm scraping@nytimes.com's purpose before blocking it. Intelligence gatherers may collect competitive research, measure advertisements, or evaluate pages for brand safety. Weigh any benefit against sharing business or page-level data with its operator or clients. For context, 7% of the top websites we track currently have robots.txt rules for scraping@nytimes.com.
Does scraping@nytimes.com Follow Robots.txt Rules?
Yes. scraping@nytimes.com is expected to follow robots.txt rules, so a disallow rule is the appropriate first step. Automatic Robots.txt can add and maintain that rule, while Agent Analytics lets you verify whether scraping@nytimes.com respects it.
Does scraping@nytimes.com Access Private Content?
No special access. scraping@nytimes.com can read publicly available pages but cannot bypass authentication. Assume anything you publish openly may be collected.
How Can I Tell if scraping@nytimes.com Is Visiting My Website?
Agent Analytics tracks scraping@nytimes.com visits in real time. You can also check your server logs for requests whose user-agent string contains "scraping@nytimes.com". Traffic may range from repeated requests to the same pages to broader crawls or one-off fetches. Because scraping@nytimes.com does not publish a verification method, any client can claim its identity and a log match is only a clue.
Why Is scraping@nytimes.com Visiting My Website?
The New York Times or one of its clients may be gathering competitive intelligence, measuring advertisements, or evaluating content for brand safety. The pages requested reflect the underlying objective.