Each category bundles several individual checks. They all run on every scan – weighted by the share shown above.
- AI citability
- Fact density (numbers, percentages, years), definition and answer patterns, paragraph and passage length, lists and tables, clean heading hierarchy, question-style headings.
- Authority & trust
- Freshness and date signals, sameAs profile links, author/byline, contact details, outbound source links, about-us signals, experience signals, visible ratings, knowsAbout topics.
- robots.txt
- robots.txt reachability, access rules per AI crawler (GPTBot, ClaudeBot, PerplexityBot and more), training/scraper bots, wildcard blocks, sitemap reference, Content-Signal directives, noai meta tags.
- Bot blocker
- Live fetches as GPTBot, ClaudeBot and PerplexityBot – is the page served in full, suspiciously throttled, or blocked?
- Technical & performance
- Server rendering (content visible without JavaScript), security headers, canonical tag, viewport, compression, cache-control, image dimensions (CLS), alt text.
- Homepage
- HTTP status, response time (TTFB), HTTPS, X-Robots-Tag, amount of text, meta description, title tag, exactly one H1.
- Schema & structured data
- JSON-LD present and parses cleanly, recommended @types (Organization, WebSite, LocalBusiness, BreadcrumbList, FAQPage), required fields per type, invalid properties (markup errors), required fields for app rich results, speakable, Open Graph tags, hreflang.
- Sitemap
- sitemap.xml reachability, lastmod freshness signal, number of listed URLs.
- llms.txt
- Presence and formal validity of an llms.txt.