What is scraping@nytimes.com?
NYTimes.com newsroom scraping bot collects publicly available, non-copyrighted data for journalistic projects including election result tracking, COVID-19 data aggregation, and other news analytics initiatives. Agent Analytics can track when it visits your website.
Overview
| Operated By | The New York Times |
| Source | Official Website |
| Expected To Follow Robots.txt | Yes |
| Insights Last Updated | August 11, 2026 |
Do you operate this agent? Contact us to suggest an update.
Category
Expected Behavior
scraping@nytimes.com usually returns to a small set of pages rather than crawling an entire site. Visits may follow an hourly, daily, or weekly schedule, while websites its clients do not monitor may never encounter it.
How To Block scraping@nytimes.com With Robots.txt
Add this rule to your robots.txt file to block scraping@nytimes.com from accessing your entire website, or use Automatic Robots.txt to block all intelligence gatherers at once. You can customize which pages are blocked by swapping out / for a different path.
User-agent: scraping@nytimes.com # https://knownagents.com/agents/scrapingnytimes-com
Disallow: /
Global Statistics for scraping@nytimes.com
As of August 11, 2026, this data reflects agent visits measured across thousands of websites using Agent Analytics, combined with daily scans of the top 1000 websites and their robots.txt files.
Robots.txt Blocked Percentage
Country of Origin
Robots.txt Blocking Trend
0% of top websites block scraping@nytimes.com in their robots.txt files.
Overall Intelligence Gatherer Traffic
3.5% of all web traffic came from intelligence gatherers.
Frequently Asked Questions
Should I Block scraping@nytimes.com?
If you run ads, confirm scraping@nytimes.com's purpose before blocking it. Some intelligence gatherers measure ad placements or evaluate pages for brand safety, and blocking those agents can interfere with advertising systems. If this one serves a different purpose, weigh any research visibility against sharing business or page-level data with third parties. Almost none of the top websites we track currently have robots.txt rules for scraping@nytimes.com.
Does scraping@nytimes.com Follow Robots.txt Rules?
Yes. scraping@nytimes.com is expected to follow robots.txt directives, so a disallow rule is the appropriate first step. Automatic Robots.txt can add and maintain that rule, while Agent Analytics lets you verify whether scraping@nytimes.com respects it.
Does scraping@nytimes.com Access Private Content?
No special access. scraping@nytimes.com can read publicly available pages but cannot bypass authentication. Assume anything you publish openly may be collected.
Why Is scraping@nytimes.com Visiting My Website?
A client of The New York Times may be gathering competitive intelligence, measuring advertisements, or evaluating content for brand safety. The pages requested reflect that client's objective.
How Can I Tell if scraping@nytimes.com Is Visiting My Website?
Agent Analytics tracks scraping@nytimes.com visits in real time alongside every other known AI agent, crawler, and scraper. You can also check your server logs for requests whose user-agent string contains "scraping@nytimes.com". Look for repeated requests to the same small group of commercial or ad-supported pages. Because scraping@nytimes.com does not publish a verification method, any client can claim its identity and a log match is only a clue.