What is nyt_scraping?
nyt_scraping is a web scraping bot operated by The New York Times. It is used to collect content from websites for The New York Times' purposes. You can set up Agent Analytics to see when nyt_scraping visits your website.
Overview
| Operated By | The New York Times |
| Source | Official Website |
| Expected To Follow Robots.txt | Yes |
| Insights Last Updated | September 14, 2026 |
Do you operate this agent? Contact us to suggest an update.
Category
Expected Behavior
nyt_scraping behaves according to its operator's configuration, and scrapers are among the least predictable bots on the web. It may range from slow, careful extraction to rapid, aggressive fetching, ignore robots.txt, or change its user-agent string when blocked. Its request volume and speed are more informative than its label.
How To Block nyt_scraping With Robots.txt
Add this rule to your robots.txt file to block nyt_scraping from accessing your website, or use Automatic Robots.txt to block all scrapers at once. You can customize which pages are blocked by swapping out / for a different path.
User-agent: nyt_scraping # https://knownagents.com/agents/nyt-scraping
Disallow: /
Insights for nyt_scraping
As of September 14, 2026, this data reflects agent visits measured across thousands of websites using Agent Analytics, combined with daily scans of the top 1000 websites and their robots.txt files.
Robots.txt Blocked Percentage
Country of Origin
Robots.txt Blocking Trend
6% of top websites block nyt_scraping in their robots.txts.
Overall Scraper Traffic
0.1% of all web traffic came from scrapers.
Frequently Asked Questions
Should I Block nyt_scraping?
Often. nyt_scraping collects content at scale, and scrapers in this category may republish or resell it. The site being scraped rarely benefits, while duplicate copies can compete with its pages in search results. For context, 6% of the top websites we track currently have robots.txt rules for nyt_scraping.
Does nyt_scraping Follow Robots.txt Rules?
Yes. nyt_scraping is expected to follow robots.txt rules, so a disallow rule is the appropriate first step. Automatic Robots.txt can add and maintain that rule, while Agent Analytics lets you verify whether nyt_scraping respects it.
Does nyt_scraping Access Private Content?
It may try. Scrapers often ignore robots.txt, and some attempt to collect valuable paywalled or gated content. Authentication is the strongest protection; robots.txt alone does not restrict access.
How Can I Tell if nyt_scraping Is Visiting My Website?
Agent Analytics tracks nyt_scraping visits in real time. You can also check your server logs for requests whose user-agent string contains "nyt_scraping". Look for rapid, sequential requests across many pages. Because nyt_scraping does not publish a verification method, any client can claim its identity and a log match is only a clue.
Why Is nyt_scraping Visiting My Website?
Your site contains data nyt_scraping's operator wants, such as prices, listings, contact details, or articles. Repeated visits generally indicate that your content remains a collection target.