Agent readiness report
vertexgroup.com.co
What an AI agent can and cannot do on this site, measured by an unauthenticated crawl and scored against the published rubric.
Scanned · rubric v1.0.1 · scored as a general site · re-scan (results are cached for seven days)
What to fix first
Ranked by impact first, then by how cheap the fix is — the same ordering the engine uses.
Missing: type-specific markup (Article, FAQPage, Service, SoftwareApplication). For a other site, Article or FAQPage or Service or SoftwareApplication is what lets an agent understand what you actually offer rather than that you exist.
No machine-readable endpoint of any kind. Most sites score zero here, which is exactly why it is the cheapest differentiation on this rubric: an agents.json is a static file, and a documented read-only API is usually a subset of something you already run.
Publish an llms.txt at the root: an H1, a one-line summary, then described links to your most useful pages. It is a text file and it takes an hour — the cheapest points on this rubric.
Your robots.txt welcomes these agents, but GPTBot received 0, ClaudeBot received 0, PerplexityBot received 0 — the block is happening at your CDN or WAF, before robots.txt is ever read. Cloudflare in particular blocks AI crawlers by default on every plan. One setting change, and it is invisible until someone tests it.
Fix: duplicate titles across pages; 1 description(s) outside 50–160 characters.
Every check, with the evidence
No finding without evidence: each result shows what we actually captured during the crawl.
ADiscovery & access
- A1Pass
robots.txt exists and parses
2/2What we captured
{ "groups": 1, "status": 200, "sitemaps": 1 } - A2Pass
Major AI crawlers allowed
6/6What we captured
{ "of": 6, "allowed": 6, "blocked": [] } - A3Pass
XML sitemap valid and referenced
3/3What we captured
{ "exists": true, "urlCount": 36, "referencedInRobots": true } - A4Fail
llms.txt present
0/4Publish an llms.txt at the root: an H1, a one-line summary, then described links to your most useful pages. It is a text file and it takes an hour — the cheapest points on this rubric.
What we captured
{ "status": 404, "present": false } - A5Partial
Bot user-agent HTTP health
1/3Your robots.txt welcomes these agents, but GPTBot received 0, ClaudeBot received 0, PerplexityBot received 0 — the block is happening at your CDN or WAF, before robots.txt is ever read. Cloudflare in particular blocks AI crawlers by default on every plan. One setting change, and it is invisible until someone tests it.
What we captured
{ "probes": [ { "name": "GPTBot", "status": 0, "ttfbMs": 640 }, { "name": "ClaudeBot", "status": 0, "ttfbMs": 350 }, { "name": "PerplexityBot", "status": 0, "ttfbMs": 328 } ], "medianTtfbMs": 350, "edgeBlockingDespiteRobots": true } - A6Pass
No hard interstitial
2/2What we captured
{ "markers": [] } - A7Pass
Server-rendered content parity
5/5What we captured
{ "ratio": 1, "rawTextLength": 8344, "renderedTextLength": 8305 }
BUnderstanding
- B1Pass
JSON-LD present and valid
5/5What we captured
{ "pagesCrawled": 6, "pagesWithValidJsonLd": 6, "homepageHasValidJsonLd": true, "pagesWithInvalidJsonLd": 0 } - B2Partial
Correct schema types for the site type
2/5Missing: type-specific markup (Article, FAQPage, Service, SoftwareApplication). For a other site, Article or FAQPage or Service or SoftwareApplication is what lets an agent understand what you actually offer rather than that you exist.
What we captured
{ "matched": [], "missing": [ "type-specific markup (Article, FAQPage, Service, SoftwareApplication)" ], "siteType": "other", "typesFound": [ "ContactPoint", "Organization", "PostalAddress" ] } - B3Partial
Semantic HTML structure
3/4Headings and landmarks are how a machine builds an outline of the page. One h1, no skipped levels, and real main/nav/footer elements — all template-level fixes, applied once.
What we captured
{ "problems": [ "https://vertexgroup.com.co/contact.html: skipped heading level", "https://vertexgroup.com.co/quality.html: skipped heading level" ], "singleH1": 6, "allLandmarks": 6, "pagesCrawled": 6, "noSkippedLevels": 4 } - B4Partial
Meta and Open Graph
1/3Fix: duplicate titles across pages; 1 description(s) outside 50–160 characters.
What we captured
{ "goodTitles": 6, "pagesCrawled": 6, "uniqueTitles": false, "goodDescriptions": 5, "completeOpenGraph": 6 } - B5Pass
Clean text ratio
3/3What we captured
{ "perPage": [ 0.282, 0.274, 0.229, 0.282, 0.265, 0.263 ], "averageRatio": 0.266 } - B6Pass
Alt text coverage
2/2What we captured
{ "withAlt": 49, "coverage": 1, "contentImages": 49 } - B7Partial
Canonical and duplication hygiene
1/3Fix: 1 page(s) missing or with a mismatched rel=canonical. Duplicate hosts split whatever authority the page has earned.
What we captured
{ "hostVariants": [ { "url": "https://www.vertexgroup.com.co", "status": 200 }, { "url": "http://vertexgroup.com.co", "status": 301, "location": "https://vertexgroup.com.co/" } ], "pagesCrawled": 6, "selfConsistentCanonicals": 5 }
CActionability
- C1Pass
Form usability
6.11/6.11What we captured
{ "forms": 1, "inputs": 7, "realSubmit": true, "unlabelled": 0, "wrongTypes": 0 } - C2Partial
Navigable link graph
3.67/4.89Fix: no breadcrumbs or clear URL hierarchy.
What we captured
{ "anchors": 75, "genericLinks": 1, "genericRatio": 0.01, "hasBreadcrumb": false, "clickableNonAnchors": 0 } - C3Pass
Interactive element semantics
4.89/4.89What we captured
{ "realButtons": 1, "positiveTabindex": 0, "clickableNonButtons": 0, "keyboardTrapsTested": false } - C4Fail
Machine endpoints and agent manifests
0/6.11No machine-readable endpoint of any kind. Most sites score zero here, which is exactly why it is the cheapest differentiation on this rubric: an agents.json is a static file, and a documented read-only API is usually a subset of something you already run.
What we captured
{ "found": [] } - C5Not applicable
Commerce actionability
0/4Not a commerce site; these points are redistributed across C1–C4.
What we captured
{ "siteType": "other" } - C6Fail
On-site search
0/3No on-site search found. If you have search, expose it as a GET form with a query parameter so an agent can jump straight to results.
What we captured
{ "searchForm": false, "searchAction": false }
DTrust & freshness
- D1Partial
Transport security
3/4Fix: no Strict-Transport-Security header.
What we captured
{ "hsts": false, "tlsValid": true, "mixedContentReferences": 0 } - D2Partial
Verifiable identity
2/4Fix: Organization schema has no sameAs links. sameAs links to a company profile are how a model corroborates that you are who the page says you are.
What we captured
{ "sameAsCount": 0, "contactDiscoverable": true, "hasOrganizationSchema": true } - D3Partial
Policies discoverable
2/4Fix: no terms page linked.
What we captured
{ "hasTerms": false, "siteType": "other", "hasPrivacy": true, "hasReturns": false, "hasShipping": false } - D4Fail
Freshness signals
0/4Fix: sitemap has no lastmod values; only 0% of sitemap URLs modified in the last 90 days; no visible dates on content. Staleness is a ranking signal for models deciding what to cite.
What we captured
{ "sitemapUrls": 36, "withLastmod": 0, "visibleDates": false, "modifiedLast90Days": 0 } - D5Pass
Identity consistency
3/3What we captured
{ "title": "Engineering, Construction & Fit-Out | Vertex Group", "hasFavicon": true, "hasOgImage": true, "ogSiteName": "Vertex Group", "schemaName": "Vertex Group S.A.R.L." } - D6Partial
security.txt
1/2security.txt is present but invalid: no Contact field; no Expires field. An expired Expires field makes the whole file non-conformant, so this fails even though the file was correct when it was written.
What we captured
{ "valid": false, "hasContact": false } - D7Pass
Multi-page stability
4/4What we captured
{ "nonOk": [], "slowPages": 0, "pagesCrawled": 6, "consistentTemplate": true }
Keep this report
Get the permanent link and the ranked fix list by email. One message, no list.
Want this fixed?
We audit in depth and implement the fixes — in your codebase or your CMS, scoped by architecture.
Own this site?
This report and its leaderboard entry come from a public crawl. Prove you own the domain with an email address on it, and you can hide the report or publish it again yourself, in minutes.
To block future scans instead, disallow AgentFriendlyRankBot in your robots.txt — see how our crawler works.
Scan your own site
Same rubric, same evidence standard, about a minute.