Agent readiness report
penbaytechnologygroup.com
What an AI agent can and cannot do on this site, measured by an unauthenticated crawl and scored against the published rubric.
Scanned · rubric v1.0.1 · scored as a e-commerce site · re-scan (results are cached for seven days)
What to fix first
Ranked by impact first, then by how cheap the fix is — the same ordering the engine uses.
Missing: type-specific markup (Product, Offer). For a ecommerce site, Product or Offer is what lets an agent understand what you actually offer rather than that you exist.
Fix: 1 of 6 inputs have no associated label. A placeholder is not a label — an agent filling your form has no way to know what an unlabelled field wants.
Found API documentation link. A second machine surface — an OpenAPI spec or an MCP server — takes this to full marks.
Publish an llms.txt at the root: an H1, a one-line summary, then described links to your most useful pages. It is a text file and it takes an hour — the cheapest points on this rubric.
Fix: no privacy policy linked; shipping and returns pages not both linked.
Every check, with the evidence
No finding without evidence: each result shows what we actually captured during the crawl.
ADiscovery & access
- A1Pass
robots.txt exists and parses
2/2What we captured
{ "groups": 1, "status": 200, "sitemaps": 1 } - A2Pass
Major AI crawlers allowed
6/6What we captured
{ "of": 6, "allowed": 6, "blocked": [] } - A3Pass
XML sitemap valid and referenced
3/3What we captured
{ "exists": true, "urlCount": 6, "referencedInRobots": true } - A4Fail
llms.txt present
0/4Publish an llms.txt at the root: an H1, a one-line summary, then described links to your most useful pages. It is a text file and it takes an hour — the cheapest points on this rubric.
What we captured
{ "status": 404, "present": false } - A5Partial
Bot user-agent HTTP health
1/3Your robots.txt welcomes these agents, but GPTBot received 429 — the block is happening at your CDN or WAF, before robots.txt is ever read. Cloudflare in particular blocks AI crawlers by default on every plan. One setting change, and it is invisible until someone tests it.
What we captured
{ "probes": [ { "name": "GPTBot", "status": 429, "ttfbMs": 60 }, { "name": "ClaudeBot", "status": 200, "ttfbMs": 1544 }, { "name": "PerplexityBot", "status": 200, "ttfbMs": 1297 } ], "medianTtfbMs": 1297, "edgeBlockingDespiteRobots": true } - A6Pass
No hard interstitial
2/2What we captured
{ "markers": [] } - A7Pass
Server-rendered content parity
5/5What we captured
{ "ratio": 1, "rawTextLength": 2690, "renderedTextLength": 1431 }
BUnderstanding
- B1Pass
JSON-LD present and valid
5/5What we captured
{ "pagesCrawled": 6, "pagesWithValidJsonLd": 6, "homepageHasValidJsonLd": true, "pagesWithInvalidJsonLd": 0 } - B2Partial
Correct schema types for the site type
2/5Missing: type-specific markup (Product, Offer). For a ecommerce site, Product or Offer is what lets an agent understand what you actually offer rather than that you exist.
What we captured
{ "matched": [], "missing": [ "type-specific markup (Product, Offer)" ], "siteType": "ecommerce", "typesFound": [ "BreadcrumbList", "EntryPoint", "ImageObject", "ListItem", "Organization", "PropertyValueSpecification", "ReadAction", "SearchAction", "WebPage", "WebSite" ] } - B3Partial
Semantic HTML structure
2/4Headings and landmarks are how a machine builds an outline of the page. One h1, no skipped levels, and real main/nav/footer elements — all template-level fixes, applied once.
What we captured
{ "problems": [ "https://penbaytechnologygroup.com: 0 h1 elements", "https://penbaytechnologygroup.com: skipped heading level", "https://penbaytechnologygroup.com/contact-us: skipped heading level", "https://penbaytechnologygroup.com/government-contracting: skipped heading level" ], "singleH1": 5, "allLandmarks": 6, "pagesCrawled": 6, "noSkippedLevels": 3 } - B4Partial
Meta and Open Graph
1/3Fix: 6 description(s) outside 50–160 characters; 6 page(s) missing og:title, og:description or og:image.
What we captured
{ "goodTitles": 6, "pagesCrawled": 6, "uniqueTitles": true, "goodDescriptions": 0, "completeOpenGraph": 0 } - B5Partial
Clean text ratio
2/3Visible text is 8.6% of your markup. Deeply nested wrapper elements make a page expensive to parse and dilute the content an agent extracts. Script and style contents are already excluded from this measure, so this is markup weight, not framework overhead.
What we captured
{ "perPage": [ 0.067, 0.123, 0.06, 0.076, 0.073, 0.114 ], "averageRatio": 0.086 } - B6Partial
Alt text coverage
1/219 of 71 content images have no alt text. An agent cannot see the image; the alt attribute is the only description it gets.
What we captured
{ "withAlt": 52, "coverage": 0.73, "contentImages": 71 } - B7Pass
Canonical and duplication hygiene
3/3What we captured
{ "hostVariants": [ { "url": "https://www.penbaytechnologygroup.com", "status": 200 }, { "url": "http://penbaytechnologygroup.com", "status": 301, "location": "https://penbaytechnologygroup.com/" } ], "pagesCrawled": 6, "selfConsistentCanonicals": 6 }
CActionability
- C1Partial
Form usability
4.5/5Fix: 1 of 6 inputs have no associated label. A placeholder is not a label — an agent filling your form has no way to know what an unlabelled field wants.
What we captured
{ "forms": 1, "inputs": 6, "realSubmit": true, "unlabelled": 1, "wrongTypes": 0 } - C2Pass
Navigable link graph
4/4What we captured
{ "anchors": 46, "genericLinks": 1, "genericRatio": 0.02, "hasBreadcrumb": true, "clickableNonAnchors": 0 } - C3Pass
Interactive element semantics
4/4What we captured
{ "realButtons": 1, "positiveTabindex": 0, "clickableNonButtons": 0, "keyboardTrapsTested": false } - C4Partial
Machine endpoints and agent manifests
3/5Found API documentation link. A second machine surface — an OpenAPI spec or an MCP server — takes this to full marks.
What we captured
{ "found": [ "API documentation link" ] } - C5Fail
Commerce actionability
0/4Fix: no product page reachable within the crawl depth; no Offer schema carrying price and availability; cart or checkout path did not return 200 without login; no guest or express checkout signal detected. Agentic commerce depends on an agent being able to read a price and reach a checkout without an account.
What we captured
{ "offerSchema": false, "cartReachable": false, "expressCheckout": false, "productPageFound": false } - C6Partial
On-site search
2/3Search exists but posts rather than using a GET query parameter, so an agent cannot construct a results URL directly. Switching to GET is usually a one-line change, and advertising it via WebSite SearchAction schema makes it discoverable.
What we captured
{ "usesGet": false, "searchForm": false, "searchAction": true }
DTrust & freshness
- D1Partial
Transport security
3/4Fix: no Strict-Transport-Security header.
What we captured
{ "hsts": false, "tlsValid": true, "mixedContentReferences": 0 } - D2Pass
Verifiable identity
4/4What we captured
{ "sameAsCount": 3, "contactDiscoverable": true, "hasOrganizationSchema": true } - D3Fail
Policies discoverable
0/4Fix: no privacy policy linked; shipping and returns pages not both linked.
What we captured
{ "hasTerms": false, "siteType": "ecommerce", "hasPrivacy": false, "hasReturns": false, "hasShipping": false } - D4Partial
Freshness signals
1/4Fix: sitemap has no lastmod values; only 0% of sitemap URLs modified in the last 90 days. Staleness is a ranking signal for models deciding what to cite.
What we captured
{ "sitemapUrls": 6, "withLastmod": 0, "visibleDates": true, "modifiedLast90Days": 0 } - D5Partial
Identity consistency
2/3Fix: no og:image. An unexplained name mismatch reads as a signal the site may not be what it claims.
What we captured
{ "title": "Home - PenBay Technology Group", "hasFavicon": true, "hasOgImage": false, "ogSiteName": "PenBay Technology Group", "schemaName": "Home - PenBay Technology Group" } - D6Fail
security.txt
0/2Publish /.well-known/security.txt with a Contact field and a future-dated Expires field. It is a five-line text file.
What we captured
{ "status": 404, "present": false } - D7Pass
Multi-page stability
4/4What we captured
{ "nonOk": [], "slowPages": 0, "pagesCrawled": 6, "consistentTemplate": true }
Keep this report
Get the permanent link and the ranked fix list by email. One message, no list.
Want this fixed?
We audit in depth and implement the fixes — in your codebase or your CMS, scoped by architecture.
Own this site?
This report and its leaderboard entry come from a public crawl. Prove you own the domain with an email address on it, and you can hide the report or publish it again yourself, in minutes.
To block future scans instead, disallow AgentFriendlyRankBot in your robots.txt — see how our crawler works.
Scan your own site
Same rubric, same evidence standard, about a minute.