anthropics/skills180kwebapp-testing
Toolkit for interacting with and testing local web applications using Playwright. Supports verifying frontend functionality, debugging UI behavior, capturing browser screenshots, and viewing browser logs.
浏览器自动化
Audit whether AI agents and AI crawlers can actually use a site — robots.txt rules per AI crawler token (OpenAI, Anthropic and Perplexity bots, Google-Extended, Applebot-Extended), Product/Offer/Organization/FAQ JSON-LD, whether main content is in the no-JavaScript server HTML, Merchant Center feed completeness incl. native_commerce checkout eligibility and conversational attributes, an optional agentic-commerce feed, and an optional experimental WebMCP check. Triggers on "/digital-marketing-pro:agent-readiness-audit", "can AI agents use our site", "are we blocking GPTBot or ClaudeBot", "is our product feed ready for AI Mode shopping", "run an agent-readiness check". Runs agent-readiness-audit.py offline on exports (network only with --fetch) and never recommends llms.txt for Google.
把这段话发给 Claude Code、Codex 或 Cursor。智能体会先检查安全性,你确认后才安装。
读取 https://funcoding.ai/skills/indranilbanerjee/digital-marketing-pro/agent-readiness-audit/install.md ,按里面的步骤帮我安装这个 Skill。
Answer one question with evidence: can AI agents and AI crawlers use this site? Concretely, the audit checks five things:
Every check is deterministic and runs in scripts/agent-readiness-audit.py. The skill adds interpretation and the brand's policy decisions on top of the script's output; it never replaces that output with judgment.
From Google's AI optimization guide (last updated 2026-07-10):
| Token | Vendor | Role in the audit | What blocking it means (vendor's words, paraphrased) | Source |
|---|---|---|---|---|
OAI-SearchBot | OpenAI | search: must be allowed | Site not surfaced in ChatGPT search features | OpenAI bots |
ChatGPT-User | OpenAI | user fetch: must be allowed | User-initiated visits. OpenAI: robots.txt "may not apply" to these | same |
OAI-AdsBot | OpenAI | ads: must be allowed if the brand runs ChatGPT Ads | Validates the safety of pages submitted as ChatGPT ads | same |
GPTBot | OpenAI | training: policy | Content excluded from foundation-model training | same |
Claude-SearchBot | Anthropic | search: must be allowed | Content not indexed for Claude's search results | Anthropic |
Claude-User | Anthropic | user fetch: must be allowed | Claude can't retrieve the page for a user's question. Anthropic honors robots.txt for all three of its bots | same |
ClaudeBot | Anthropic | training: policy | Future content excluded from training datasets | same |
PerplexityBot | Perplexity | search: must be allowed | Not surfaced or linked in Perplexity results (Perplexity says it is not used for model training) | Perplexity bots |
Perplexity-User | Perplexity | user fetch: must be allowed | Perplexity says this fetcher "generally ignores robots.txt rules" | same |
Google-Extended | training: policy | Controls use for Gemini training and grounding. It "does not impact a site's inclusion in Google Search nor is it used as a ranking signal", and it has no separate user-agent string | Google crawlers | |
Applebot-Extended | Apple | training: policy | Controls use in Apple foundation-model training. It does not crawl, and pages that disallow it "can still be included in search results" | Apple |
"Policy" means that blocking training crawlers is the brand's choice. The audit reports it as info unless the brand sets --training-policy allow|block, in which case a mismatch fails the check.
All checks run offline on files the user exports. The script fetches over the network only with --fetch.
--robots FILE, or --site URL --fetch. Add --path /products/ (repeatable) to test the paths that matter, beyond /.--html FILE (repeatable). Save it as the server sends it, for example with curl -L URL > page.html, not "Save as" from a browser, which saves the post-JavaScript DOM. Add --expect "Product name" (repeatable) for text that must be in that HTML, such as a product name or price.--feed FILE (TSV, CSV, or RSS/Atom XML with g: attributes).--acp-feed FILE. Accepts the JSONL "OpenAI format" or the Google-compatible CSV/TSV (spec).--webmcp.--training-policy either|allow|block. The default, either, reports training-crawler rules without judging them.Load brand context. Read ~/.claude-marketing/brands/_active-brand.json, then ~/.claude-marketing/brands/{slug}/profile.json.
--training-policy).Collect inputs. Ask for exports first. If the user only has a URL, confirm they want a live fetch before using --fetch.
Run the audit.
python "${CLAUDE_PLUGIN_ROOT}/scripts/agent-readiness-audit.py" \
--robots robots.txt --path /products/ \
--html home.html --html product.html --expect "Acme Anvil" \
--feed products.tsv --acp-feed openai-feed.jsonl --webmcp \
--training-policy either --format json > "${CLAUDE_PLUGIN_DATA}/{brand}/seo/agent-readiness/{date}/02-audit.json"
Live variant, used only when the user agreed to network access:
python "${CLAUDE_PLUGIN_ROOT}/scripts/agent-readiness-audit.py" --site https://example.com --fetch \
--page https://example.com/products/anvil --format json
Exit codes:
0: no check failed (warnings allowed)1: at least one check failed2: bad input (nothing to audit, a missing or unparseable file, a bad flag)On 2, fix the input and re-run. Do not interpret a partial run.
Interpret each check. Report the script's status per check: pass, warn, fail, info or skipped.
offers is a warn, because merchant-listing rich results require name, image and offers (Google)--min-text, default 250 characters), when an app-shell root (#root, #__next, #app, ...) holds little text, or when an --expect phrase is missing from the server HTML.<html lang>, images without alt, unlabelled inputs, and unnamed buttons.id, title, description, link, image_link, availability, price (product data spec)availability value is outside in_stock, out_of_stock, preorder, backordernative_commerce(checkout_eligibility) value is not TRUE/FALSEnative_commerce: report how many products opt into checkout on Google. Only listings with native_commerce(checkout_eligibility) = TRUE show the "Buy" button. The feature is for select merchants, with products eligible in the US, Canada and Australia (Merchant Center help; UCP guide). The account-level return policy and customer-support contact can't be checked from a feed, so list them as manual checks.question_and_answer, document_link, related_product, item_group_title, variant_option, popularity_rank) are optional (help). Report their coverage; never fail on them.is_eligible_search / is_eligible_checkout counts.info. WebMCP is a Chrome origin trial (from Chrome 149) and "under active discussion and subject to change" (Chrome docs). Report the declarative toolname / tooldescription forms and the inline registerTool calls found. Recommend nothing beyond "optional experiment".Decide policy questions with the user, not for them. Two are the brand's call: whether to block training crawlers, and whether to opt products into checkout on Google. Present the trade-off and the source.
Prioritize fixes. Use the script's recommendations array as the backbone.
/digital-marketing-pro:tech-seo-audit, markup to /digital-marketing-pro:entity-audit, and visibility follow-up to /digital-marketing-pro:aeo-geo.All outputs go to ${CLAUDE_PLUGIN_DATA}/{brand}/seo/agent-readiness/{YYYY-MM-DD}/:
00-input.md domain, which exports were supplied, fetch yes/no, training policy
01-inputs/ the robots.txt, HTML and feed files actually audited (for reproducibility)
02-audit.json raw script output
03-findings.md per-check status, findings, sources (with the 2026-10-04 check date)
04-policy-decisions.md training-crawler stance, checkout opt-in decision — with rationale
05-fix-plan.md prioritized fixes with owners and the skill each routes to
PLAN.md one-page summary: verdict, top 3 fixes, re-audit date
| Gate | Pass when |
|---|---|
| script_ran_clean | Exit code 0 or 1 (never 2). A 2 means the inputs were wrong, not the site |
| inputs_are_server_html | 00-input.md states the HTML was captured as served (curl or --fetch), not from a browser's saved DOM |
| sources_cited | Every finding in 03-findings.md carries the source URL from the script output or this skill |
| policy_explicit | 04-policy-decisions.md records the training-crawler stance and checkout opt-in as the brand's decisions |
| no_llms_txt_for_google | Nothing in the deliverable recommends llms.txt as a Google visibility lever |
ready (no warnings or failures), needs_work (warnings only), or not_ready (any failure). Experimental and skipped checks do not count.allowed / partially_blocked / blocked per tested path, the deciding robots rule, and the vendor's caveat./digital-marketing-pro:tech-seo-audit for that./digital-marketing-pro:geo-monitor (probes plus Bing Webmaster AI Performance) and /digital-marketing-pro:gsc-ai-performance.--fetch, a 4xx robots.txt means allow-all, and a 5xx or network error means every crawler must assume complete disallow (RFC 9309). The first finding says so; if the cause was your own network rather than the site, re-run.skills/paid-advertising/ads-in-ai-answers.md)/digital-marketing-pro:tech-seo-audit: full technical crawl and rendering/digital-marketing-pro:entity-audit: Organization and entity consistency/digital-marketing-pro:aeo-geo: optimizing for AI answers once agents can read the site/digital-marketing-pro:geo-monitor: whether AI answers actually cite the brand
anthropics/skills180kToolkit for interacting with and testing local web applications using Playwright. Supports verifying frontend functionality, debugging UI behavior, capturing browser screenshots, and viewing browser logs.
浏览器自动化
addyosmani/agent-skills103kTests in real browsers via Chrome DevTools MCP. Use when building or debugging anything that runs in a browser. Use when you need to inspect the DOM, capture console errors, analyze network requests, profile performance, or verify visual output with real runtime data. Requires the chrome-devtools MCP server to be configured.
浏览器自动化
ComposioHQ/awesome-claude-skills77kToolkit for interacting with and testing local web applications using Playwright. Supports verifying frontend functionality, debugging UI behavior, capturing browser screenshots, and viewing browser logs.
浏览器自动化
code-yeongyu/oh-my-openagent70kDrives a real browser through the omowright library from the js eval kernel: sites the user is already signed into, forms and clicks, JS-rendered pages, screenshots, web QA, extension popups, a human handoff for login, CAPTCHA or OTP, and a browser you own for scraping, bot-scored targets, network capture and QA traces. Use for any interactive browser task; not for a plain search or an unblocked static fetch.
浏览器自动化
shanraisshan/claude-code-best-practice67kBrowser automation CLI for AI agents. Use when the user needs to interact with websites, including navigating pages, filling forms, clicking buttons, taking screenshots, extracting data, testing web apps, or automating any browser task. Triggers include requests to "open a website", "fill out a form", "click a button", "take a screenshot", "scrape data from a page", "test this web app", "login to a site", "automate browser actions", or any task requiring programmatic web interaction.
浏览器自动化
CherryHQ/cherry-studio52kRun Cherry Studio critical-path system regression tasks through the repository-owned Playwright E2E workflow. Use for full regression, release acceptance, development-branch system validation, or a named cherry-regression-test task on GitHub-hosted macOS and Windows runners.
浏览器自动化