Agent readiness report
semandex.net
What an AI agent can and cannot do on this site, measured by an unauthenticated crawl and scored against the published rubric.
Scanned · rubric v1.0.1 · scored as a e-commerce site · re-scan (results are cached for seven days)
What to fix first
Ranked by impact first, then by how cheap the fix is — the same ordering the engine uses.
Missing: type-specific markup (Product, Offer). For a ecommerce site, Product or Offer is what lets an agent understand what you actually offer rather than that you exist.
Only 1 of 5 crawled pages carry structured data. Add it at the template level rather than page by page.
Fix: 10 of 20 inputs have no associated label. A placeholder is not a label — an agent filling your form has no way to know what an unlabelled field wants.
No machine-readable endpoint of any kind. Most sites score zero here, which is exactly why it is the cheapest differentiation on this rubric: an agents.json is a static file, and a documented read-only API is usually a subset of something you already run.
Publish an llms.txt at the root: an H1, a one-line summary, then described links to your most useful pages. It is a text file and it takes an hour — the cheapest points on this rubric.
Every check, with the evidence
No finding without evidence: each result shows what we actually captured during the crawl.
ADiscovery & access
- A1Pass
robots.txt exists and parses
2/2What we captured
{ "groups": 1, "status": 200, "sitemaps": 1 } - A2Pass
Major AI crawlers allowed
6/6What we captured
{ "of": 6, "allowed": 6, "blocked": [] } - A3Pass
XML sitemap valid and referenced
3/3What we captured
{ "exists": true, "urlCount": 8, "referencedInRobots": true } - A4Fail
llms.txt present
0/4Publish an llms.txt at the root: an H1, a one-line summary, then described links to your most useful pages. It is a text file and it takes an hour — the cheapest points on this rubric.
What we captured
{ "status": 404, "present": false } - A5Pass
Bot user-agent HTTP health
3/3What we captured
{ "probes": [ { "name": "GPTBot", "status": 200, "ttfbMs": 319 }, { "name": "ClaudeBot", "status": 200, "ttfbMs": 240 }, { "name": "PerplexityBot", "status": 200, "ttfbMs": 639 } ], "medianTtfbMs": 319, "edgeBlockingDespiteRobots": false } - A6Fail
No hard interstitial
0/2Detected: CAPTCHA. An agent cannot dismiss a consent dialog or solve a challenge — whatever sits behind it is unreachable. Cookieless analytics removes the need for a banner entirely.
What we captured
{ "markers": [ "CAPTCHA" ] } - A7Pass
Server-rendered content parity
5/5What we captured
{ "ratio": 1, "rawTextLength": 9543, "renderedTextLength": 8113 }
BUnderstanding
- B1Partial
JSON-LD present and valid
3/5Only 1 of 5 crawled pages carry structured data. Add it at the template level rather than page by page.
What we captured
{ "pagesCrawled": 5, "pagesWithValidJsonLd": 1, "homepageHasValidJsonLd": true, "pagesWithInvalidJsonLd": 0 } - B2Partial
Correct schema types for the site type
2/5Missing: type-specific markup (Product, Offer). For a ecommerce site, Product or Offer is what lets an agent understand what you actually offer rather than that you exist.
What we captured
{ "matched": [], "missing": [ "type-specific markup (Product, Offer)" ], "siteType": "ecommerce", "typesFound": [ "WebSite" ] } - B3Fail
Semantic HTML structure
0/4Headings and landmarks are how a machine builds an outline of the page. One h1, no skipped levels, and real main/nav/footer elements — all template-level fixes, applied once.
What we captured
{ "problems": [ "https://semandex.net: 0 h1 elements", "https://semandex.net: missing footer", "https://semandex.net/about/: 2 h1 elements", "https://semandex.net/about/: skipped heading level", "https://semandex.net/about/: missing footer", "https://semandex.net/about/collaborations: 2 h1 elements", "https://semandex.net/about/collaborations: skipped heading level", "https://semandex.net/about/collaborations: missing footer", "https://semandex.net/contact/: 2 h1 elements", "https://semandex.net/contact/: skipped heading level" ], "singleH1": 0, "allLandmarks": 0, "pagesCrawled": 5, "noSkippedLevels": 1 } - B4Partial
Meta and Open Graph
2/3Fix: 5 page(s) missing og:title, og:description or og:image.
What we captured
{ "goodTitles": 5, "pagesCrawled": 5, "uniqueTitles": true, "goodDescriptions": 5, "completeOpenGraph": 0 } - B5Partial
Clean text ratio
2/3Visible text is 6.9% of your markup. Deeply nested wrapper elements make a page expensive to parse and dilute the content an agent extracts. Script and style contents are already excluded from this measure, so this is markup weight, not framework overhead.
What we captured
{ "perPage": [ 0.105, 0.06, 0.093, 0.036, 0.051 ], "averageRatio": 0.069 } - B6Partial
Alt text coverage
1/220 of 71 content images have no alt text. An agent cannot see the image; the alt attribute is the only description it gets.
What we captured
{ "withAlt": 51, "coverage": 0.72, "contentImages": 71 } - B7Pass
Canonical and duplication hygiene
3/3What we captured
{ "hostVariants": [ { "url": "https://www.semandex.net", "status": 301, "location": "https://semandex.net/" }, { "url": "http://semandex.net", "status": 301, "location": "https://semandex.net/" } ], "pagesCrawled": 5, "selfConsistentCanonicals": 5 }
CActionability
- C1Partial
Form usability
3.5/5Fix: 10 of 20 inputs have no associated label. A placeholder is not a label — an agent filling your form has no way to know what an unlabelled field wants.
What we captured
{ "forms": 12, "inputs": 20, "realSubmit": true, "unlabelled": 10, "wrongTypes": 0 } - C2Partial
Navigable link graph
3/4Fix: 13% of links use contextless text like "click here".
What we captured
{ "anchors": 88, "genericLinks": 11, "genericRatio": 0.13, "hasBreadcrumb": false, "clickableNonAnchors": 0 } - C3Pass
Interactive element semantics
4/4What we captured
{ "realButtons": 3, "positiveTabindex": 0, "clickableNonButtons": 0, "keyboardTrapsTested": false } - C4Fail
Machine endpoints and agent manifests
0/5No machine-readable endpoint of any kind. Most sites score zero here, which is exactly why it is the cheapest differentiation on this rubric: an agents.json is a static file, and a documented read-only API is usually a subset of something you already run.
What we captured
{ "found": [] } - C5Fail
Commerce actionability
0/4Fix: no product page reachable within the crawl depth; no Offer schema carrying price and availability; cart or checkout path did not return 200 without login; no guest or express checkout signal detected. Agentic commerce depends on an agent being able to read a price and reach a checkout without an account.
What we captured
{ "offerSchema": false, "cartReachable": false, "expressCheckout": false, "productPageFound": false } - C6Pass
On-site search
3/3What we captured
{ "usesGet": true, "searchForm": true, "searchAction": false }
DTrust & freshness
- D1Partial
Transport security
2/4Fix: no Strict-Transport-Security header; 1 insecure subresource reference(s).
What we captured
{ "hsts": false, "tlsValid": true, "mixedContentReferences": 1 } - D2Partial
Verifiable identity
2/4Fix: no Organization schema. sameAs links to a company profile are how a model corroborates that you are who the page says you are.
What we captured
{ "sameAsCount": 0, "contactDiscoverable": true, "hasOrganizationSchema": false } - D3Fail
Policies discoverable
0/4Fix: no privacy policy linked; shipping and returns pages not both linked.
What we captured
{ "hasTerms": false, "siteType": "ecommerce", "hasPrivacy": false, "hasReturns": false, "hasShipping": false } - D4Fail
Freshness signals
0/4Fix: sitemap has no lastmod values; only 0% of sitemap URLs modified in the last 90 days; no visible dates on content. Staleness is a ranking signal for models deciding what to cite.
What we captured
{ "sitemapUrls": 8, "withLastmod": 0, "visibleDates": false, "modifiedLast90Days": 0 } - D5Partial
Identity consistency
2/3Fix: no og:image. An unexplained name mismatch reads as a signal the site may not be what it claims.
What we captured
{ "title": "Semantic Software for Information Management | Semandex", "hasFavicon": true, "hasOgImage": false, "ogSiteName": "Semandex", "schemaName": "Semandex" } - D6Fail
security.txt
0/2Publish /.well-known/security.txt with a Contact field and a future-dated Expires field. It is a five-line text file.
What we captured
{ "status": 404, "present": false } - D7Partial
Multi-page stability
2/4Fix: 1 page(s) did not return 200: https://semandex.net/about-us/blog-posts (404).
What we captured
{ "nonOk": [ { "url": "https://semandex.net/about-us/blog-posts", "status": 404 } ], "slowPages": 0, "pagesCrawled": 6, "consistentTemplate": true }
Keep this report
Get the permanent link and the ranked fix list by email. One message, no list.
Want this fixed?
We audit in depth and implement the fixes — in your codebase or your CMS, scoped by architecture.
Own this site?
This report and its leaderboard entry come from a public crawl. Prove you own the domain with an email address on it, and you can hide the report or publish it again yourself, in minutes.
To block future scans instead, disallow AgentFriendlyRankBot in your robots.txt — see how our crawler works.
Scan your own site
Same rubric, same evidence standard, about a minute.