The Agentic Web Index

The internet is rapidly evolving from an environment built primarily for humans, into one increasingly used by machines. See how AI agents, crawlers, scrapers, and other bots are reshaping the way information is discovered, accessed, and used across the web.

Overview

Key ecosystem metrics across 5,000+ websites using Agent Analytics and AI Chat Referral Tracking.

Bot vs. Human Traffic

33%
↓ 4%
Compared to the previous 90 days
The amount of visits from known agents vs. humans

Agentrification

28%
↑ 9%
Compared to the previous 90 days
The percentage of bot traffic that's AI-related

AI Chat Referral Volume

0.1%
↓ 4%
Compared to the previous 90 days
The percentage of human website visits that come from AI chat
See Referral Trends ↓

Robots.txt Effectiveness

96.4%
How often bots follow robots.txt rules
See Per Agent ↓

Traffic by Agent Type

AI Agent
AI Agent
Uses an actual web browser to autonomously complete complex tasks on behalf of a human user
AI Assistant
AI Assistant
Fetches website content in response to a user prompt, to include in an AI-generated answer
AI Coding Agent
AI Coding Agent
Fetches documentation and other resources to help build software
AI Data Provider
AI Data Provider
Crawls websites to supply structured content to AI systems as a third-party service
AI Data Scraper
AI Data Scraper
Downloads website content to include in datasets used for training AI models such as LLMs
AI Search Crawler
AI Search Crawler
Indexes website content to possibly include as citations in AI-powered search results
Archiver
Archiver
Captures and stores historical website snapshots for long-term digital preservation
Automated Agent
Automated Agent
Automates browser interactions programmatically without direct human supervision
Developer Helper
Developer Helper
Assists with testing, debugging, and ensuring website functionality
Fetcher
Fetcher
Retrieves web page metadata to power app features like link previews or feeds
Intelligence Gatherer
Intelligence Gatherer
Analyzes web content for brand safety, competitive insights, and ad targeting
Scraper
Scraper
Extracts large amounts of web data, often without explicit website permission
Search Engine Crawler
Search Engine Crawler
Systematically scans and indexes web pages to include in search results
Security Scanner
Security Scanner
Scans websites for security vulnerabilities, threats, and configuration weaknesses
SEO Crawler
SEO Crawler
Analyzes website structure and content to identify SEO improvement opportunities
Uncategorized
Uncategorized
Not yet assigned a type
Undocumented AI Agent
Undocumented AI Agent
Crawls websites without disclosing its purpose, collecting data for an unknown AI use case

Hover over each agent type for more information about what they do

Top Visiting Agents

bingbot
SRCH
Search Engine Crawler
Systematically scans and indexes web pages to include in search results
8.5%
AhrefsBot
SEO
SEO Crawler
Analyzes website structure and content to identify SEO improvement opportunities
6.7%
Googlebot
SRCH
Search Engine Crawler
Systematically scans and indexes web pages to include in search results
6.6%
Known Agent
DEV
Developer Helper
Assists with testing, debugging, and ensuring website functionality
5.9%
PetalBot
SRCH
AI Search Crawler
Indexes website content to possibly include as citations in AI-powered search results
3.3%
ClaudeBot
SCRP
AI Data Scraper
Downloads website content to include in datasets used for training AI models such as LLMs
3.3%
SemrushBot
SEO
SEO Crawler
Analyzes website structure and content to identify SEO improvement opportunities
3.0%
ChatGPT-User
ASST
AI Assistant
Fetches website content in response to a user prompt, to include in an AI-generated answer
2.9%
facebookexternalhit
FTCH
Fetcher
Retrieves web page metadata to power app features like link previews or feeds
2.6%
meta-externalagent
SCRP
AI Data Scraper
Downloads website content to include in datasets used for training AI models such as LLMs
2.4%
Amazonbot
SCRP
AI Data Scraper
Downloads website content to include in datasets used for training AI models such as LLMs
2.3%
Amzn-SearchBot
SRCH
AI Search Crawler
Indexes website content to possibly include as citations in AI-powered search results
2.2%
Baiduspider
SRCH
Search Engine Crawler
Systematically scans and indexes web pages to include in search results
2.2%
MJ12bot
SEO
SEO Crawler
Analyzes website structure and content to identify SEO improvement opportunities
2.2%
DotBot
SEO
SEO Crawler
Analyzes website structure and content to identify SEO improvement opportunities
2.2%

Agents with the most activity

Top Operators

Microsoft
9.2%
Google
8.9%
Ahrefs
8.3%
Meta
7.1%
OpenAI
5.6%
Anthropic
5.3%
Amazon
4.9%
Semrush
3.9%
Huawei
3.5%
Baidu
2.5%
Moz
2.4%
Majestic
2.3%
Automattic
2.1%
Apple
2.0%

Operators with the most activity

AI Scraping Activity

These bots scrape website content to train AI models. Some belong to AI companies, while others belong to third-party services that resell the data. Automatic Robots.txt can block unwanted scraping. Included agent types include AI Data Providers and AI Data Scrapers.

AI Data Provider
AI Data Provider
Crawls websites to supply structured content to AI systems as a third-party service
AI Data Scraper
AI Data Scraper
Downloads website content to include in datasets used for training AI models such as LLMs

AI scraping activity by agent type over time

Top Visited Website Categories

Law and Government
6.3%
Online Communities
6.2%
Home and Garden
5.8%
Games
5.8%
Computers and Electronics
5.7%
Business and Industrial
5.3%
Pets and Animals
5.2%
Jobs and Education
5.2%
Internet and Telecom
5.2%
Autos and Vehicles
5.0%
Shopping
4.9%
Hobbies and Leisure
4.9%
Beauty and Fitness
4.5%

Website categories with most activity

Top Agents

ClaudeBot
SCRP
AI Data Scraper
Downloads website content to include in datasets used for training AI models such as LLMs
26.8%
meta-externalagent
SCRP
AI Data Scraper
Downloads website content to include in datasets used for training AI models such as LLMs
19.3%
Amazonbot
SCRP
AI Data Scraper
Downloads website content to include in datasets used for training AI models such as LLMs
19.0%
GPTBot
SCRP
AI Data Scraper
Downloads website content to include in datasets used for training AI models such as LLMs
8.2%
Bytespider
SCRP
AI Data Scraper
Downloads website content to include in datasets used for training AI models such as LLMs
7.3%
ShapBot
PVDR
AI Data Provider
Crawls websites to supply structured content to AI systems as a third-party service
6.2%
GoogleOther
SCRP
AI Data Scraper
Downloads website content to include in datasets used for training AI models such as LLMs
3.1%
CCBot
SCRP
AI Data Scraper
Downloads website content to include in datasets used for training AI models such as LLMs
2.3%
YouBot
PVDR
AI Data Provider
Crawls websites to supply structured content to AI systems as a third-party service
1.9%
Timpibot
SCRP
AI Data Scraper
Downloads website content to include in datasets used for training AI models such as LLMs
1.7%
Diffbot
PVDR
AI Data Provider
Crawls websites to supply structured content to AI systems as a third-party service
1.0%
DeepSeekBot
SCRP
AI Data Scraper
Downloads website content to include in datasets used for training AI models such as LLMs
1.0%
VelenPublicWebCrawler
SCRP
AI Data Scraper
Downloads website content to include in datasets used for training AI models such as LLMs
0.8%
FacebookBot
SCRP
AI Data Scraper
Downloads website content to include in datasets used for training AI models such as LLMs
0.4%
TerraCotta
PVDR
AI Data Provider
Crawls websites to supply structured content to AI systems as a third-party service
0.3%

Agents doing the most AI scraping

Top Operators

Anthropic
26.8%
Meta
19.7%
Amazon
19.0%
OpenAI
8.2%
ByteDance
7.3%
Parallel
6.2%
Google
3.1%
You.com
1.9%
Timpi
1.7%
Diffbot
1.0%
DeepSeek
1.0%
Velen
0.8%
Ceramic
0.3%
Lyrenth
0.2%

Operators doing the most AI scraping

AI Fetching Activity

These bots fetch website content in real time to power AI assistants, coding agents, and other retrieval-augmented generation (RAG) tasks. Pages inform responses on the spot, such as when an assistant summarizes an article or a coding agent references documentation. Included agent types include AI Assistants and AI Coding Agents.

AI Assistant
AI Assistant
Fetches website content in response to a user prompt, to include in an AI-generated answer
AI Coding Agent
AI Coding Agent
Fetches documentation and other resources to help build software

AI fetching activity by agent type over time

Top Visited Website Categories

Reference
3.8%
Science
3.7%
Finance
3.3%
Law and Government
2.5%
Computers and Electronics
2.4%
Beauty and Fitness
2.1%
Travel and Transportation
1.9%
Internet and Telecom
1.9%
Business and Industrial
1.8%
Health
1.6%
Autos and Vehicles
1.5%
Sports
1.3%
Real Estate
1.3%

Website categories with most activity

Top Agents

ChatGPT-User
ASST
AI Assistant
Fetches website content in response to a user prompt, to include in an AI-generated answer
90.0%
DuckAssistBot
ASST
AI Assistant
Fetches website content in response to a user prompt, to include in an AI-generated answer
3.9%
Claude-User
ASST
AI Assistant
Fetches website content in response to a user prompt, to include in an AI-generated answer
2.2%
Claude-Code
CODE
AI Coding Agent
Fetches documentation and other resources to help build software
1.5%
Shap-User
ASST
AI Assistant
Fetches website content in response to a user prompt, to include in an AI-generated answer
1.2%
Cursor
CODE
AI Coding Agent
Fetches documentation and other resources to help build software
0.4%
Google-NotebookLM
ASST
AI Assistant
Fetches website content in response to a user prompt, to include in an AI-generated answer
0.4%
Perplexity-User
ASST
AI Assistant
Fetches website content in response to a user prompt, to include in an AI-generated answer
0.3%
Gemini-Deep-Research
ASST
AI Assistant
Fetches website content in response to a user prompt, to include in an AI-generated answer
0.0%
MistralAI-User
ASST
AI Assistant
Fetches website content in response to a user prompt, to include in an AI-generated answer
0.0%
GoogleAgent-URLContext
ASST
AI Assistant
Fetches website content in response to a user prompt, to include in an AI-generated answer
0.0%
meta-externalfetcher
ASST
AI Assistant
Fetches website content in response to a user prompt, to include in an AI-generated answer
0.0%
Code
CODE
AI Coding Agent
Fetches documentation and other resources to help build software
0.0%
opencode
CODE
AI Coding Agent
Fetches documentation and other resources to help build software
0.0%
Amzn-User
ASST
AI Assistant
Fetches website content in response to a user prompt, to include in an AI-generated answer
0.0%

Agents doing the most AI fetching

Top Operators

OpenAI
90.0%
DuckDuckGo
3.9%
Anthropic
3.6%
Parallel
1.2%
Google
0.5%
xAI
0.4%
Perplexity
0.3%
Mistral
0.0%
Meta
0.0%
GitHub
0.0%
Anomaly
0.0%
Amazon
0.0%
Qualified
0.0%
Kagi
0.0%
Poggio
0.0%

Operators doing the most AI fetching

AI Search Indexing Activity

These bots crawl website content so it can be surfaced in AI search engines and AI-generated answers. Those answers often include citations or links back to the source pages. Included agent types include AI Search Crawlers.

AI Search Crawler
AI Search Crawler
Indexes website content to possibly include as citations in AI-powered search results

AI search indexing activity by agent type over time

Top Visited Website Categories

Travel and Transportation
7.0%
Reference
6.7%
Health
5.4%
Games
5.2%
Books and Literature
5.0%
Food and Drink
5.0%
Jobs and Education
4.9%
Arts and Entertainment
4.7%
Computers and Electronics
4.7%
People and Society
4.5%
Law and Government
4.4%
Hobbies and Leisure
4.4%
Science
4.3%

Website categories with most activity

Top Agents

PetalBot
SRCH
AI Search Crawler
Indexes website content to possibly include as citations in AI-powered search results
27.0%
Amzn-SearchBot
SRCH
AI Search Crawler
Indexes website content to possibly include as citations in AI-powered search results
18.2%
Applebot
SRCH
AI Search Crawler
Indexes website content to possibly include as citations in AI-powered search results
15.1%
Claude-SearchBot
SRCH
AI Search Crawler
Indexes website content to possibly include as citations in AI-powered search results
11.5%
meta-webindexer
SRCH
AI Search Crawler
Indexes website content to possibly include as citations in AI-powered search results
11.3%
OAI-SearchBot
SRCH
AI Search Crawler
Indexes website content to possibly include as citations in AI-powered search results
10.7%
LinkupBot
SRCH
AI Search Crawler
Indexes website content to possibly include as citations in AI-powered search results
3.1%
PerplexityBot
SRCH
AI Search Crawler
Indexes website content to possibly include as citations in AI-powered search results
2.1%
Google-CloudVertexBot
SRCH
AI Search Crawler
Indexes website content to possibly include as citations in AI-powered search results
0.3%
AzureAI-SearchBot
SRCH
AI Search Crawler
Indexes website content to possibly include as citations in AI-powered search results
0.2%
xAI-SearchBot
SRCH
AI Search Crawler
Indexes website content to possibly include as citations in AI-powered search results
0.2%
ExaSearchBot
SRCH
AI Search Crawler
Indexes website content to possibly include as citations in AI-powered search results
0.1%
AddSearchBot
SRCH
AI Search Crawler
Indexes website content to possibly include as citations in AI-powered search results
0.0%
Anomura
SRCH
AI Search Crawler
Indexes website content to possibly include as citations in AI-powered search results
0.0%
MistralAI-Index
SRCH
AI Search Crawler
Indexes website content to possibly include as citations in AI-powered search results
0.0%

Agents doing the most AI search indexing

Top Operators

Huawei
27.0%
Amazon
18.2%
Apple
15.1%
Anthropic
11.5%
Meta
11.3%
OpenAI
10.7%
Linkup
3.1%
Perplexity
2.1%
Google
0.3%
Microsoft
0.2%
xAI
0.2%
Exa
0.1%
AddSearch
0.0%
Direqt
0.0%
Mistral
0.0%

Operators doing the most AI search indexing

AI Browsing Activity

These bots use browsers to autonomously navigate websites, click through pages, and make decisions to complete tasks for people. Agentic UX best practices and Google PageSpeed Insights help evaluate how well websites support them. Included agent types include AI Agents.

Average Session Duration

1 minute
↑ 29%
Compared to the previous 90 days
The average duration of a session

Average Pages per Session

11.5
↑ 64%
Compared to the previous 90 days
The average number of pages visited per session
AI Agent
AI Agent
Uses an actual web browser to autonomously complete complex tasks on behalf of a human user

AI browsing activity by agent type over time

Top Visited Website Categories

Law and Government
0.2%
Computers and Electronics
0.1%
Shopping
0.1%
Business and Industrial
0.0%
Arts and Entertainment
0.0%
Health
0.0%
Autos and Vehicles
0.0%
Internet and Telecom
0.0%
Travel and Transportation
0.0%
Books and Literature
0.0%
Beauty and Fitness
0.0%
Finance
0.0%
People and Society
0.0%

Website categories with most activity

Top Agents

Manus-User
AGNT
AI Agent
Uses an actual web browser to autonomously complete complex tasks on behalf of a human user
47.6%
Google-Agent
AGNT
AI Agent
Uses an actual web browser to autonomously complete complex tasks on behalf of a human user
29.2%
ChatGPT Agent
AGNT
AI Agent
Uses an actual web browser to autonomously complete complex tasks on behalf of a human user
23.2%
NovaAct
AGNT
AI Agent
Uses an actual web browser to autonomously complete complex tasks on behalf of a human user
0.0%
GoogleAgent-Mariner
AGNT
AI Agent
Uses an actual web browser to autonomously complete complex tasks on behalf of a human user
0.0%
AmazonBuyForMe
AGNT
AI Agent
Uses an actual web browser to autonomously complete complex tasks on behalf of a human user
0.0%
TwinAgent
AGNT
AI Agent
Uses an actual web browser to autonomously complete complex tasks on behalf of a human user
0.0%

Agents doing the most AI browsing

Top Operators

Google
29.2%
OpenAI
23.2%
Amazon
0.0%
Twin
0.0%

Operators doing the most AI browsing

Robots.txt & Compliance

See which robots.txt rules are set across the web and how well agents follow them. An agent's Robots.txt Effectiveness measures the effectiveness of a disallow rule for it by estimating how much the agent reduces its traffic after it's blocked.

Effectiveness (Overall)

96.4%
How often all agents across all agent types follow robots.txt rules

Effectiveness (AI Scrapers & Data Providers)

95.3%
How often AI Data Scrapers and AI Data Providers follow robots.txt rules

Top Rule-Following Agents

Barkrowler
SEO
SEO Crawler
Analyzes website structure and content to identify SEO improvement opportunities
100.0%
serpstatbot
SEO
SEO Crawler
Analyzes website structure and content to identify SEO improvement opportunities
100.0%
bingbot
SRCH
Search Engine Crawler
Systematically scans and indexes web pages to include in search results
100.0%
Googlebot
SRCH
Search Engine Crawler
Systematically scans and indexes web pages to include in search results
100.0%
SiteAuditBot
SEO
SEO Crawler
Analyzes website structure and content to identify SEO improvement opportunities
100.0%
proximic
INT
Intelligence Gatherer
Analyzes web content for brand safety, competitive insights, and ad targeting
100.0%
Known Agent
DEV
Developer Helper
Assists with testing, debugging, and ensuring website functionality
100.0%
YouBot
PVDR
AI Data Provider
Crawls websites to supply structured content to AI systems as a third-party service
100.0%
Sucuri Uptime Monitor
UNC
Uncategorized
Not yet assigned a type
100.0%
AffsignalCrawler
INT
Intelligence Gatherer
Analyzes web content for brand safety, competitive insights, and ad targeting
100.0%
PerplexityBot
SRCH
AI Search Crawler
Indexes website content to possibly include as citations in AI-powered search results
99.9%
SERankingBacklinksBot
SEO
SEO Crawler
Analyzes website structure and content to identify SEO improvement opportunities
99.9%
Dataprovider.com
SCRP
Scraper
Extracts large amounts of web data, often without explicit website permission
99.9%
AhrefsSiteAudit
SEO
SEO Crawler
Analyzes website structure and content to identify SEO improvement opportunities
99.8%
AwarioBot
INT
Intelligence Gatherer
Analyzes web content for brand safety, competitive insights, and ad targeting
99.7%

Agents with the best Robots.txt Effectiveness percentages

Top Rule-Breaking Agents

Baiduspider
SRCH
Search Engine Crawler
Systematically scans and indexes web pages to include in search results
71.2%
ShapBot
PVDR
AI Data Provider
Crawls websites to supply structured content to AI systems as a third-party service
75.9%
facebookexternalhit
FTCH
Fetcher
Retrieves web page metadata to power app features like link previews or feeds
83.0%
YandexBot
SRCH
Search Engine Crawler
Systematically scans and indexes web pages to include in search results
85.5%
Sogou web spider
SRCH
Search Engine Crawler
Systematically scans and indexes web pages to include in search results
87.2%
DeepSeekBot
SCRP
AI Data Scraper
Downloads website content to include in datasets used for training AI models such as LLMs
92.6%
DotBot
SEO
SEO Crawler
Analyzes website structure and content to identify SEO improvement opportunities
94.3%
DataForSeoBot
SEO
SEO Crawler
Analyzes website structure and content to identify SEO improvement opportunities
94.5%
FacebookBot
SCRP
AI Data Scraper
Downloads website content to include in datasets used for training AI models such as LLMs
95.2%
ChatGPT-User
ASST
AI Assistant
Fetches website content in response to a user prompt, to include in an AI-generated answer
96.0%
SemrushBot
SEO
SEO Crawler
Analyzes website structure and content to identify SEO improvement opportunities
96.2%
Applebot
SRCH
AI Search Crawler
Indexes website content to possibly include as citations in AI-powered search results
96.2%
VelenPublicWebCrawler
SCRP
AI Data Scraper
Downloads website content to include in datasets used for training AI models such as LLMs
96.2%
PetalBot
SRCH
AI Search Crawler
Indexes website content to possibly include as citations in AI-powered search results
96.4%
OAI-SearchBot
SRCH
AI Search Crawler
Indexes website content to possibly include as citations in AI-powered search results
96.7%

Agents with the worst Robots.txt Effectiveness percentages

Top Blocked Agents

CCBot
SCRP
AI Data Scraper
Downloads website content to include in datasets used for training AI models such as LLMs
24.6%
Bytespider
SCRP
AI Data Scraper
Downloads website content to include in datasets used for training AI models such as LLMs
22.6%
GPTBot
SCRP
AI Data Scraper
Downloads website content to include in datasets used for training AI models such as LLMs
22.1%
ClaudeBot
SCRP
AI Data Scraper
Downloads website content to include in datasets used for training AI models such as LLMs
21.5%
meta-externalagent
SCRP
AI Data Scraper
Downloads website content to include in datasets used for training AI models such as LLMs
19.7%
omgili
SCRP
AI Data Scraper
Downloads website content to include in datasets used for training AI models such as LLMs
19.2%
Diffbot
PVDR
AI Data Provider
Crawls websites to supply structured content to AI systems as a third-party service
18.7%
anthropic-ai
UND
Undocumented AI Agent
Crawls websites without disclosing its purpose, collecting data for an unknown AI use case
18.6%
cohere-ai
UND
Undocumented AI Agent
Crawls websites without disclosing its purpose, collecting data for an unknown AI use case
18.1%
PerplexityBot
SRCH
AI Search Crawler
Indexes website content to possibly include as citations in AI-powered search results
17.6%
Claude-Web
UND
Undocumented AI Agent
Crawls websites without disclosing its purpose, collecting data for an unknown AI use case
17.4%
Amazonbot
SCRP
AI Data Scraper
Downloads website content to include in datasets used for training AI models such as LLMs
16.8%
Omgilibot
UNC
Uncategorized
Not yet assigned a type
16.4%

Agents blocked by the most top websites

Spoofing & Security

See which agents are most frequently impersonated, and how spoofing activity changes over time. A visit is considered spoofed when it claims a recognized agent identity but fails that agent's supported authentication method, such as verified IP or Web Bot Auth.

Active Threat: AI Bot Spoofing Campaign
We are observing a widespread campaign impersonating AI bots to scan websites for vulnerabilities. The attacker appears to be targeting credential and configuration paths used by AI coding tools. Contact us for more information, or inspect your own traffic.

Top Spoofed Agent Identities

Googlebot
SRCH
Search Engine Crawler
Systematically scans and indexes web pages to include in search results
0.5%
ChatGPT-User
ASST
AI Assistant
Fetches website content in response to a user prompt, to include in an AI-generated answer
0.2%
OAI-SearchBot
SRCH
AI Search Crawler
Indexes website content to possibly include as citations in AI-powered search results
0.1%
PerplexityBot
SRCH
AI Search Crawler
Indexes website content to possibly include as citations in AI-powered search results
0.1%
GPTBot
SCRP
AI Data Scraper
Downloads website content to include in datasets used for training AI models such as LLMs
0.1%
ClaudeBot
SCRP
AI Data Scraper
Downloads website content to include in datasets used for training AI models such as LLMs
0.1%
Applebot
SRCH
AI Search Crawler
Indexes website content to possibly include as citations in AI-powered search results
0.1%
Amzn-SearchBot
SRCH
AI Search Crawler
Indexes website content to possibly include as citations in AI-powered search results
0.0%
Perplexity-User
ASST
AI Assistant
Fetches website content in response to a user prompt, to include in an AI-generated answer
0.0%
MistralAI-User
ASST
AI Assistant
Fetches website content in response to a user prompt, to include in an AI-generated answer
0.0%
bingbot
SRCH
Search Engine Crawler
Systematically scans and indexes web pages to include in search results
0.0%
Claude-User
ASST
AI Assistant
Fetches website content in response to a user prompt, to include in an AI-generated answer
0.0%
GoogleOther
SCRP
AI Data Scraper
Downloads website content to include in datasets used for training AI models such as LLMs
0.0%
Google-Agent
AGNT
AI Agent
Uses an actual web browser to autonomously complete complex tasks on behalf of a human user
0.0%
Amazonbot
SCRP
AI Data Scraper
Downloads website content to include in datasets used for training AI models such as LLMs
0.0%

The most impersonated agent identities

Recent Top Targeted Paths

/.config/anthropic/credentials/default.json
/.claude/settings.json
/.claude.json
/.hermes/.env
/.openclaw/.env
/.codex/config.toml
/.continue/config.json
/.aider.conf.yml
/service-account.json
/serviceaccountkey.json
/service_account.json
/firebase-adminsdk.json
/firebase-service-account.json
/.aws/credentials
/.aws/config
/.s3cfg
/.boto
/.npmrc
/.env.example
/.env.local
/.env.production
/.env.backup
/.env.old
/backend/.env
/api/.env
/admin/.env
/dockerfile
/docker-compose.yaml
/.docker/config.json
/terraform.tfstate
/credentials.json
/secrets.json
/secrets.yml
/key.json
/rclone.conf

Examples of recent top targeted request paths

AI Chat Referrals

See which AI platforms like ChatGPT, Perplexity, and Gemini cite websites and send them human referral traffic. Citations are estimated. Google's guide explains how websites can optimize their content to be more visible in AI chat responses (GEO).

Traffic by AI Chat Platform

ChatGPT
Claude
Copilot
DeepSeek
Gemini
Meta AI
Mistral
Perplexity

Referral activity by AI chat platform over time

Top Mentioned (Cited) Website Categories

Reference
3.8%
Science
3.7%
Finance
3.1%
Law and Government
2.5%
Computers and Electronics
2.3%
Beauty and Fitness
2.1%
Travel and Transportation
1.9%
Internet and Telecom
1.9%
Business and Industrial
1.8%
Health
1.6%
Autos and Vehicles
1.5%
Sports
1.3%
Real Estate
1.3%
Games
1.1%
Jobs and Education
1.0%

Website categories most frequently cited in AI chat responses

Top Clicked (Referred) Website Categories

Travel and Transportation
0.3%
Autos and Vehicles
0.1%
Beauty and Fitness
0.1%
Finance
0.1%
Computers and Electronics
0.1%
Jobs and Education
0.1%
Sports
0.1%
Business and Industrial
0.1%
Home and Garden
0.1%
Health
0.0%
Real Estate
0.0%
Internet and Telecom
0.0%
Food and Drink
0.0%
Shopping
0.0%
Science
0.0%

Website categories receiving the most referrals from AI chat

Methodology

Data Scope

The Index is updated daily with completed days of traffic, security, and referral data from more than 5,000 websites using Agent Analytics and AI Chat Referral Tracking. The current partial day is excluded. Percentage-change tags compare the current period with the preceding period of the same duration. Agent names, operators, and classifications come from the Agent Directory, which is updated as new agents are discovered or existing agents change. Website categories follow the taxonomy used by Google AdSense. Participating websites are not a random sample of the entire web, and the qualifying set can change as websites connect, disconnect, or cross activity thresholds. Results characterize the observed network and broader directional trends; they should not be interpreted as a precise census of global web traffic.

Qualification & Aggregation

Only websites meeting minimum activity and data-quality requirements are included. Internal, test, incomplete, or anomalous data is excluded. Bot traffic percentages use total server traffic as their denominator. AI chat referral percentages use estimated human traffic, calculated by excluding identified bot visits from total server traffic. Rates are calculated for each qualifying website first, then averaged across websites and completed days. This gives each website equal weight regardless of traffic volume and prevents a small number of high-traffic websites from dominating the results. Daily charts are not smoothed, allowing normal seasonality to remain visible.

Measuring Robots.txt Effectiveness

Robots.txt Effectiveness estimates how much an agent's share of website traffic falls on websites that block it sitewide using Disallow: /. It measures the observed outcome of the rule, not the agent's intent. Each day, Known Agents averages the agent's share of traffic across all qualifying websites where it is allowed, whether or not the agent visited them that day. This baseline is scaled to each blocked website's traffic and compared with observed disallowed visits. A day qualifies only with sufficient agent activity across multiple allowed websites and expected traffic across multiple blocked websites. Scores require qualifying observations consistently across multiple recent completed days. For agents supporting authentication, only verified traffic is attributed.

Each qualifying website-day contributes equally. A 0% score means no measurable traffic reduction; 100% means no disallowed visits were observed. Zero observed traffic counts only when the baseline predicts meaningful activity, and agents without enough broadly distributed traffic receive no score. The headline metric gives qualifying agents equal weight. Because this is observational, it cannot distinguish technical enforcement from voluntary behavior or prove causation. Top Blocked Bots is calculated separately using robots.txt scans of Similarweb's top 1,000 websites.

Identifying Spoofed Bots

Spoofing statistics measure traffic from visits that claim the identity of a known agent but fail a supported authentication method, such as published IP verification or HTTP message signatures. Each agent's daily rate is calculated against total server traffic for every qualifying website, then averaged across websites. A failed check indicates that the visit was likely impersonating the named agent; it does not identify the software or operator that actually made the request. Agents without a supported authentication method are not included in these measurements.

AI Chat Citations & Referrals

AI chat referral statistics count directly observed human visits carrying a recognized AI platform in the referring URL or campaign source. Visits without usable referral information cannot be attributed to an AI platform. Citation statistics are estimates based on requests from agents known to retrieve content for AI platforms. Those requests indicate that content may have informed a response, but they do not confirm that a source appeared as a citation to a user. Because AI platforms do not provide a complete public record of their sources, citation results should be interpreted as directional patterns rather than exact citation counts.

Frequently Asked Questions

Can journalists and media organizations use this data?

Absolutely. You may cite The Agentic Web Index with attribution and a link to this page. For interviews, fact-checking, background context, or a more specific breakdown for a story, contact us and include your deadline.


Do you work with researchers?

Absolutely. We welcome thoughtful research into how agents and bots are changing the web. Tell us about your research question, timeframe, and intended use. Depending on the scope and data constraints, we may be able to provide additional context, compare approaches, or explore a joint analysis.


Can I request a specific analysis?

Yes. If you need a breakdown by agent, operator, activity type, website category, or time period that is not shown here, contact us. When the underlying data supports it, we can examine the question and provide a focused analysis.