scraping@nytimes.com
NYTimes.com newsroom scraping bot collects publicly available, non-copyrighted data for journalistic projects including election result tracking, COVID-19 data aggregation, and other news analytics initiatives.
At a glance
- Operator: The New York Times
- Type: Research
How Centinel checks it
- User agent: The request calls itself this crawler. Anyone can send the same string.
scraping@nytimes.com publishes nothing Centinel can check a source against, so a match reports the name and leaves the source unconfirmed. The match tokens, verification domains, and address feeds are not published here.
Allowing or blocking it
The crawler object in the /validate response sets access_allowed to true only for a verified source that your tenant allowlists. A policy rule can allow or block this crawler by its category, Research.