Tooling Landscape · neutral, comparison-grade reviews
The Agentic-Web Tooling Landscape
The agentic-web tooling landscape spans four jobs — measuring adoption (Cloudflare Radar), analyzing crawlers (Known Agents), tracking AI visibility (Profound, Ahrefs Brand Radar, Semrush), and checking readiness (Cloudflare's scanner, AgentReady.org, llms.txt generators, specification.website). This page reviews each one neutrally — a genuine strength and an honest weakness apiece — and places our own Agent-Readiness Audit among them, not above them.
What is an agentic-web tool?
An agentic-web tool is a product that measures, analyzes, tracks, or audits how a website participates in the agentic web. This landscape reviews eight tools across four jobs — adoption measurement, crawler analytics, AI-visibility tracking and readiness checking — with the same fields for each: what it does, strength, weakness, monetization and when to use it. Every price, plan name and count was checked against the vendor’s own site on 7 September 2026 (URLs in the Sources list); where a vendor publishes no number — Known Agents shows no bot total — this page says so instead of estimating.
A tool is not a standard and not a service: a tool measures or checks something (Cloudflare Radar, Known Agents), a standard specifies a file or protocol that tools read and write (llms.txt, Web Bot Auth), and a service engineers and certifies the result (the Agent-Readiness Audit).
Which four categories organize the agentic-web tooling landscape?
Agentic-web tools fall into four categories — adoption measurement, crawler analytics, GEO / AI-visibility and readiness checking — and each category feeds a different Almanac pillar, as the table shows.
| Category | What it does | Tools | Bridges to pillar |
|---|---|---|---|
| Adoption-measurement | Counts what crawls the agentic web | Cloudflare Radar | State of the Agentic Web |
| Crawler-analytics | Names who is crawling your site | Known Agents (formerly Dark Visitors) | Crawler & Agent Registry |
| GEO / AI-visibility | Tracks whether AI cites you | Profound · Ahrefs Brand Radar · Semrush | GEO guide |
| Readiness-checking | Tests whether your site is agent-ready | Cloudflare Agent Readiness score · AgentReady · llms.txt generators · specification.website | Services / Audit |
What does Cloudflare offer: the Agent Readiness score, Radar data and pay-per-crawl?
Cloudflare is the broadest agentic-web tool: since 17 April 2026 it has published a free Agent Readiness score with a public scanner at isitagentready.com, its Radar service reports network-wide crawl traffic by bot, and its AI Crawl Control product — formerly AI Audit, available on all plans — lets a site see which AI services access its content, control access with granular policies and, in a closed beta, charge AI crawlers (Cloudflare, 2026).
- What it does: scores four dimensions — Discoverability, Content, Bot Access Control and Capabilities — and, through pay-per-crawl, answers an unpaid crawler with HTTP 402 and requires it to sign requests with HTTP Message Signatures, the mechanism behind Web Bot Auth (Cloudflare, July 2025).
- Strength: grounded in a large sample — Cloudflare scanned the 200,000 most visited domains and found a robots.txt on 78 % of them, declared AI-usage preferences on 4 %, markdown content negotiation on 3.9 %, and MCP or API-catalog standards on fewer than 15 sites (Cloudflare, April 2026). Radar’s 2025 Year in Review put GPTBot at about 7.5 % of verified-bot traffic and all AI bots at 4.2 % of HTML requests on average (Cloudflare, December 2025); this Almanac’s State of the Agentic Web dataset records Radar’s May 2026 AI-crawler shares — GPTBot 11.48 %, Bytespider 10.25 %, Applebot 7.01 %, Claude-SearchBot 2.22 % — with source and last-verified date (2026-06-15).
- Weakness: ecosystem-locked — enforcement and pay-per-crawl only work for sites on Cloudflare; Cloudflare-flavored framing; no teaching or reference layer for what the numbers mean.
- Monetization: free scanner and Radar data; AI Crawl Control on all plans; pay-per-crawl in closed beta with Cloudflare as merchant of record (Cloudflare, 2025).
- Our take: the best free first scan and the canonical adoption data source. Cloudflare’s pay-per-crawl prices AI access at the edge, and its verification relies on Web Bot Auth, which verifies which crawler the analytics tools count.
What does Known Agents (formerly Dark Visitors) do?
Known Agents — the crawler-analytics service formerly named Dark Visitors — shows a site every crawler, scraper and AI agent visiting it in real time, measures human referrals from ChatGPT, Claude, Gemini and other AI chat platforms, serves a robots.txt that updates automatically as new bots are discovered, identifies agents from incoming requests through an API, and lists agents in a public directory filtered by 17 categories — the deepest crawler list in this landscape. It publishes no total bot count, and its free plan covers one project (Known Agents, verified 7 September 2026).
- Strength: breadth and detail — a page per agent across 36 directory pages (7 September 2026); the community block-list ai.robots.txt credits Known Agents as a source for its crawler entries (ai.robots.txt README, read 2026-09-07); an official Node.js SDK is open source on GitHub under Apache-2.0 (Known Agents, node-sdk).
- Weakness: crawler-only — no protocols, models or readiness reference; any bot total you read elsewhere is reported, not audited; the Agent Identification API and CSV export start at the paid Startup tier.
- Monetization: freemium — Free ($0: one project, 100K events, automatic robots.txt), Startup ($29 per month: three projects, 1M events), Growth ($299 per month: ten projects, 10M events) and Enterprise (custom), as listed on 7 September 2026.
- Our take: the tool to identify and allow or block specific crawlers; our registry organizes the AI crawlers the analytics tools track in your logs neutrally and agent-first.
How do Profound, Ahrefs Brand Radar and Semrush track GEO visibility?
Profound, Ahrefs Brand Radar and Semrush are GEO / AI-visibility tools that track whether and how often AI answer engines mention or cite your brand — they measure the outcome (visibility) but do not engineer the cause (the readiness signals that earn citations). As of 7 September 2026, entry prices start at $99 per month (Profound Starter, billed yearly; Semrush AI Visibility Toolkit, billed annually), and Ahrefs includes Custom Prompts in every paid plan. The table names what each one tracks, what it costs to start, and who it fits.
| Tool | What it tracks | Entry price | Best fit |
|---|---|---|---|
| Profound | AI-answer visibility across nine engines — ChatGPT, Perplexity, Google AI Overviews, Google AI Mode, Google Gemini, Microsoft Copilot, Grok, DeepSeek and Anthropic Claude — plus Agent Analytics of AI-sourced traffic | Starter $99/month billed yearly (ChatGPT only, 50 prompts); Growth $399/month billed yearly; Enterprise custom | Teams whose primary KPI is answer-engine presence |
| Ahrefs Brand Radar | Brand visibility across AI Overviews, AI Mode, Gemini, Perplexity, ChatGPT, Copilot and Claude, plus YouTube, TikTok and Reddit | Custom Prompts included in every paid Ahrefs plan (5 prompts on Lite); AI Visibility Index add-on from €179/month (EUR price shown 7 September 2026) | Existing Ahrefs users adding AI visibility |
| Semrush | AI Visibility Toolkit: mentions from ChatGPT, Google AI, Gemini and Perplexity, AI-visibility reports, competitor analysis and prompt research | $99/month per domain billed annually (25 custom prompts, one domain) | Marketing teams already on Semrush |
- Strength: quantifies citation presence across engines, turning “are we cited?” into a tracked metric — and the metric moves: in Profound’s study of about 80,000 prompts per platform, 40–60 % of the domains cited for the same prompts had changed one month later (June to July 2025: Google AI Overviews 59.3 %, ChatGPT 54.1 %, Microsoft Copilot 53.4 %, Perplexity 40.5 %; Profound, July 2025).
- Weakness: measurement-only — they tell you whether you are cited, not how to get cited; this is the GEO measurement-vs-engineering split. Prices scale with tracked prompts and engines.
- Monetization: SaaS subscriptions (entry prices in the table).
- Our take: use these to measure AI visibility; use the GEO engineering guide and the Audit to change it. The trackers measure what GEO (Generative Engine Optimization) defines.
What does an llms.txt generator do?
An llms.txt generator writes the /llms.txt file that Jeremy Howard proposed on 3 September 2024: a markdown file at a site’s root with brief background information, guidance and links to detailed markdown files (llmstxt.org, 2024) — and, where the generator supports it, a full-text companion such as the /llms-full.txt this site serves. It is the cheapest first step toward agent readiness — a starting point, not a finished posture.
- Strength: near-zero effort and immediate machine-readability; the specification itself is free and open.
- Weakness: single-signal — it covers only the declaration layer, validates nothing else about the site, and output quality varies by generator. Cloudflare’s Agent Readiness scan does not check llms.txt by default; the scan can be customized to include it (Cloudflare, April 2026).
- Monetization: varies by generator; the standard itself is free.
- Our take: generate the file, then verify the whole readiness posture — llms.txt is only Layer-1 discovery: a generator writes the file the llms.txt standard specifies, then you implement llms.txt after a generator drafts it.
What do AgentReady and specification.website specify?
AgentReady is an open, MIT-licensed, versioned standard for agent readiness that Ora published on 20 March 2026 (v1.0.0 on 24 April 2026), with stable requirement identifiers such as AR-DISC-01 and AR-CAPA-01 across five sections — Discoverability, Content for agents, Capabilities, Identity & Access, Commerce — and a free reference scanner, Deep Scan, that returns a 0–100 score with a letter grade (Ora, 2026). specification.website is Joost de Valk’s “platform-agnostic, full specification of the technical features a good website should have”: MIT-licensed code, CC BY 4.0 content, a source cited on every page, and agent readiness as one of its ten domains (de Valk, 2026).
- What it does: AgentReady publishes the requirements a site must meet so an agent can find, read and act — as a page and as
spec.jsonat agentready.org — and scores sites through Deep Scan (Ora, 2026). - Strength: open, citable by identifier, and free to run.
- Weakness: a standard plus a scanner rather than a reference or glossary; its output is a score and a letter grade.
- Monetization: standard and scanner free; published by a vendor, Ora.
- Our take: cite it as a neutral peer and the closest public overlap with our Audit; its five sections map almost one-to-one onto the six readiness dimensions this site audits.
- What it does: specification.website collects the technical features a good website should have — HTML foundations, SEO, accessibility, security, well-known URIs, agent readiness, performance, privacy, resilience and internationalisation (de Valk, 2026).
- Strength: strong E-E-A-T (authored by Joost de Valk, the Yoast founder), open licences and a source on every page.
- Weakness: it specifies, it does not scan — it tells you what a site should have, not whether yours has it.
- Monetization: open source (MIT code, CC BY 4.0 content).
- Our take: cite it as a neutral peer; neither a specification nor a scanner issues the validated audit and certification badge this site does.
In-depth tool reviews
Each core tool has a dedicated, neutral review:
- Cloudflare AI Crawl Control — see, allow, block and charge AI crawlers at the edge.
- Known Agents (formerly Dark Visitors) — AI-agent analytics and robots.txt generation.
- Profound — how answer engines represent your brand (GEO visibility).
- Otterly.AI — brand-mention and citation monitoring across AI search.
- Firecrawl — turn the live web into LLM-ready markdown for agents.
Which tools does agentswelcome.dev itself expose to agents?
agentswelcome.dev exposes 12 WebMCP tools of its own, listed straight from data/tools.json — the file behind /api/tool-registry and the interfaces.webmcp.tools block of /.well-known/agents.json — so the list cannot drift. 10 of the 12 carry the WebMCP read-only annotation, 2 (sign_guestbook, get_premium_playbook) write or unlock something, and 2 run the x402 payment demo (demo — no real money). Every tool has an HTTP twin, so the same result needs no JavaScript.
| Tool | What it does | HTTP twin | Read-only | Price |
|---|---|---|---|---|
list_techniques | List all 12 agent-first web techniques in this site's catalog. | GET /api/techniques | yes | free |
get_technique | Get the full description of one technique by its id (e.g. 'webmcp', 'llms-txt'). | GET /api/techniques | yes | free |
sign_guestbook | Leave a greeting in the agent guestbook. | POST /api/guestbook | no | free |
get_visitor_info | Ask the server what it knows about the current visitor (echo of /api/whoami). | GET /api/whoami | yes | free |
audit_site | Agent-Readiness Audit: live-check a URL for llms.txt, markdown negotiation, AI-aware robots.txt, JSON-LD and more. | POST /api/audit | yes | 3 per hour per IP free, then $0.005 (demo) |
get_premium_playbook | Unlock the Agent-First Playbook (premium content). | GET /api/premium/playbook | no | $0.10 (demo) |
list_directory | List the certified agent-ready sites in the directory (free). | GET /api/directory | yes | free |
search_almanac | Search The Agentic Web Almanac across all five datasets (crawlers, protocols, models, glossary, state-of-the-agentic-web). | GET /api/search?q=… | yes | free |
get_crawler | Get the registry entry for one AI crawler by id (e.g. 'claudebot', 'gptbot', 'perplexitybot') — operator, purpose, robots token, how to verify. | GET /api/crawlers/{id} | yes | free |
identify_user_agent | Given a User-Agent string, identify which known AI crawler it claims to be and how to verify it for real. | GET /api/verify-crawler?ua=… | yes | free |
define_term | Look up a canonical definition from the Agentic Web Lexicon by id (e.g. 'mcp', 'prompt-injection', 'agent-experience'). | GET /api/glossary/{term} | yes | free |
ask_almanac | Ask a natural-language question about the agentic web. | GET /api/ask?q=… | yes | free |
Which agentic-web tool should you use for which job?
Each tool does one job and none certifies; the matrix puts all eight on the same seven columns, and a validating audit closes the loop.
| Tool | Category | What it does | Monetization | Free tier? | Best job | Where it stops |
|---|---|---|---|---|---|---|
| Cloudflare | Readiness + adoption | Agent Readiness score, Radar data, AI Crawl Control, pay-per-crawl | Free scanner and data; AI Crawl Control on all plans; pay-per-crawl in closed beta | Yes (scanner, Radar, AI Crawl Control) | First scan & adoption data | Cloudflare-flavored; no neutral reference |
| Known Agents | Crawler-analytics | Agent directory (17 categories), automatic robots.txt, identification API | Freemium: Free, Startup $29, Growth $299, Enterprise | Yes (one project, 100K events) | Identify & allow/block crawlers | Crawler-only; no total count published |
| Profound | GEO / AI-visibility | AI-answer visibility monitoring across nine engines | SaaS from $99/month billed yearly | No free plan listed | Track answer-engine presence | Measures, does not engineer |
| Ahrefs Brand Radar | GEO / AI-visibility | AI-mention tracking inside the Ahrefs suite | Included in paid Ahrefs plans; index add-on from €179/month | Free preview check only | AI visibility for Ahrefs users | Measures, does not engineer |
| Semrush | GEO / AI-visibility | AI Visibility Toolkit in a marketing suite | From $99/month per domain billed annually | Free AI Visibility Checker tool only | AI visibility for Semrush users | Measures, does not engineer |
| llms.txt generators | Readiness (declaration) | Writes llms.txt | Varies by generator; the spec is free | Varies | Cheapest first readiness step | Single-signal; validates nothing |
| AgentReady | Readiness (standard + scanner) | MIT-licensed standard plus free Deep Scan (0–100, letter grade) | Standard and scanner free; published by Ora | Yes | Requirements to cite by identifier | Scores and grades; not a reference |
| specification.website | Readiness (spec) | Platform-agnostic website specification with an agent-readiness domain | Open source (MIT / CC BY 4.0) | Yes | A reference spec to build against | Specifies; does not scan |
The terms in this matrix are defined in the Lexicon that defines the agentic-web terms these tools use, and every standard a tool touches in the Protocol Atlas that defines the standards these tools touch.
Where does the Agent-Readiness Audit sit among these tools?
Among these tools, the Agent-Readiness Audit is a validating readiness check that — unlike a one-off scanner, a crawler list, or a visibility tracker — ties its result to a vendor-neutral reference and an “Agents Welcome” certification badge, the cleanest fully-unowned lane in the category. It does not claim to win on any single measurement axis; it adds whole-site validation and a certification none of the eight tools issues.
To confirm and certify what a free scanner flags, the Agent-Readiness Audit checks what the scanners flag — then certifies the result. And this is self-demonstrating: this site scores 100 on the readiness the scanners check — verify it live. The agents these tools track run on models scored in the Frontier Model Matrix that scores the models behind the agents.
Agentic-web tools — frequently asked questions
What is an agentic-web tool?
An agentic-web tool is a product that measures, analyzes, tracks, or audits how a website participates in the agentic web. They fall into four jobs: adoption measurement, crawler analytics, AI-visibility tracking, and readiness checking.
What is the best agent-readiness tool?
There is no single best tool — each one does a different job. Cloudflare's scanner is the best free first scan, Known Agents is the deepest crawler list, Profound, Ahrefs Brand Radar and Semrush measure AI visibility, and llms.txt generators are the cheapest first step. A validating audit then confirms and certifies the result, which none of the measurement or checklist tools does.
How do Profound, Ahrefs Brand Radar and Semrush differ?
All three are GEO / AI-visibility tools that track whether AI answer engines cite your brand. Profound is purpose-built for AI-answer monitoring, Ahrefs Brand Radar bolts AI-mention tracking onto an SEO suite, and Semrush offers an AI-visibility module inside a broad marketing platform. All measure the outcome; none engineers the readiness signals that earn citations.
Is an llms.txt generator enough to make a site agent-ready?
No. An llms.txt generator produces a content-declaration file — a fast, cheap first step — but it is a single signal, not a finished readiness posture. It does not validate that the rest of the site is machine-readable, so the file is a starting point, not a destination.
Where does Agents Welcome's own Audit fit among these tools?
The Agent-Readiness Audit is a validating readiness check that ties its result to a vendor-neutral reference and an 'Agents Welcome' certification badge. It does not beat any single tool on its own axis — Cloudflare is broader on edge data, Known Agents deeper on bots, the GEO trackers stronger on visibility — but it occupies the neutral-reference-plus-certification lane no other tool fully owns.
Sources
- Cloudflare: Introducing the Agent Readiness score, 2026. blog.cloudflare.com
- Cloudflare: AI Crawl Control documentation (formerly AI Audit; available on all plans; last updated 2026-08-14), 2026. developers.cloudflare.com
- Cloudflare: Introducing pay per crawl, 2025. blog.cloudflare.com
- Cloudflare: Cloudflare Radar 2025 Year in Review, 2025. blog.cloudflare.com
- AGENTS WELCOME: State of the Agentic Web dataset (Cloudflare Radar May 2026 crawl shares, last verified 2026-06-15), 2026. agentswelcome.dev
- Known Agents: Product overview (homepage, read 2026-09-07), 2026. knownagents.com
- Known Agents: Pricing (verified 2026-09-07), 2026. knownagents.com
- Known Agents: Agent directory (verified 2026-09-07), 2026. knownagents.com
- Known Agents: Dark Visitors Is Now Known Agents (undated post, read 2026-09-07), 2026. knownagents.com
- Known Agents: GitHub organization — node-sdk, the official Node.js SDK, Apache-2.0 (read 2026-09-07), 2026. github.com
- ai.robots.txt: README crediting Known Agents as a source of its crawler list (read 2026-09-07), 2026. github.com
- Profound: Pricing (verified 2026-09-07), 2026. tryprofound.com
- Profound: AI Search Volatility: why AI search results keep changing, 2025. tryprofound.com
- Ahrefs: Brand Radar (verified 2026-09-07; prices displayed in EUR), 2026. ahrefs.com
- Semrush: AI Visibility Toolkit pricing (verified 2026-09-07), 2026. semrush.com
- Jeremy Howard: The /llms.txt file (proposal of 3 September 2024), 2024. llmstxt.org
- Ora: Introducing AgentReady: the first open standard for agent readiness, 2026. ora.ai
- Joost de Valk: The Website Specification (read 2026-09-07), 2026. specification.website
- Joost de Valk: specification.website repository README — MIT code, CC BY 4.0 content, sources on every page (read 2026-09-07), 2026. github.com
- AGENTS WELCOME: WebMCP tool registry (data/tools.json, served live), 2026. agentswelcome.dev
How do you get from measuring the agentic web to making your site ready for it?
Every tool above tells you where you stand; turning that reading into shipped changes — and a certified result — is the job of the Agent-Readiness Engineering guides that turn each scanner finding into a step and of the Audit.
This catalog feeds every pillar its tools touch: the Agent-Readiness Audit checks what the scanners flag, the adoption data these tools report is compiled over time, the crawler data Known Agents sells is organized agent-first, and the GEO guide engineers the citations these trackers measure. This neutral tooling review is one of the guides the agentic web home indexes for agents.
