What is SemanticScholarBot?

SemanticScholarBot is a search engine crawler operated by Ai2. You can set up Agent Analytics to see when SemanticScholarBot visits your website.

Overview

Operated By Ai2
Source Official Website
Expected To Follow Robots.txt Yes
Insights Last Updated September 14, 2026

Do you operate this agent? Contact us to suggest an update.

Category

Search Engine Crawler
Systematically scans and indexes web pages to include in search results

Expected Behavior

SemanticScholarBot follows an algorithmic schedule shaped by your site's popularity, update cadence, and domain authority. It generally spreads requests across recurring crawl passes, throttles itself to limit server load, and respects crawl rules more reliably than most other bots.

SemanticScholarBot's User Agent

User Agent Mozilla/5.0 (compatible) SemanticScholarBot (+https://www.semanticscholar.org/crawler)

How To Block SemanticScholarBot With Robots.txt

Add this rule to your robots.txt file to block SemanticScholarBot from accessing your website, or use Automatic Robots.txt to block all search engine crawlers at once. You can customize which pages are blocked by swapping out / for a different path.

User-agent: SemanticScholarBot # https://knownagents.com/agents/semanticscholarbot
Disallow: /

Insights for SemanticScholarBot

As of September 14, 2026, this data reflects agent visits measured across thousands of websites using Agent Analytics, combined with daily scans of the top 1000 websites and their robots.txt files.

Robots.txt Blocked Percentage

7%
7% of top websites are blocking SemanticScholarBot
Learn How →

Country of Origin

United States
SemanticScholarBot normally visits From the United States

Robots.txt Blocking Trend

7% of top websites block SemanticScholarBot in their robots.txts.

Overall Search Engine Crawler Traffic

14.8% of all web traffic came from search engine crawlers.

Top Visited Website Categories

Health
Jobs and Education
News
Science
Business and Industrial

The types of websites most frequently visited by SemanticScholarBot.

Frequently Asked Questions

Should I Block SemanticScholarBot?

No, unless you want to remove your pages from that search engine. SemanticScholarBot supplies its index, so blocking it can eliminate your visibility and organic traffic there. Nearly every website should allow this category. For context, 7% of the top websites we track currently have robots.txt rules for SemanticScholarBot.


Does SemanticScholarBot Follow Robots.txt Rules?

Yes. SemanticScholarBot is expected to follow robots.txt rules, so a disallow rule is the appropriate first step. Automatic Robots.txt can add and maintain that rule, while Agent Analytics lets you verify whether SemanticScholarBot respects it.


Does SemanticScholarBot Access Private Content?

No special access. SemanticScholarBot can reach only content available to anonymous visitors and does not sign in. Make sure sensitive content is not public by mistake.


How Can I Tell if SemanticScholarBot Is Visiting My Website?

Agent Analytics tracks SemanticScholarBot visits in real time. You can also check your server logs for requests whose user-agent string contains "SemanticScholarBot". Look for recurring passes that revisit changed pages and discover new ones. Because SemanticScholarBot does not publish a verification method, any client can claim its identity and a log match is only a clue.


Why Is SemanticScholarBot Visiting My Website?

SemanticScholarBot is adding new or changed pages to its search index. It may have discovered your site through links, a sitemap, or direct submission.