Scan
Simon Property Group
Score
33/100
8 of 51 tests pass · 67 points still on the table
Not ready
AI crawlers are turned away before they reach the content, so nothing else on this page has had a chance to matter.
Do these three first
The fastest 12 points on this scan
CDN and WAF do not silently reject AI crawlers
8 of 8 AI user-agents are rejected or challenged at the edge.
Do this: Stop rejecting AI crawlers at the CDN or WAF.
No blanket robots.txt blocks on AI crawler user-agents
1 of 4 answer engines are disallowed from content paths: PerplexityBot. 5 training or opt-out agents are also blocked: GPTBot, ClaudeBot, CCBot, Bytespider, Google-Extended.
Do this: Remove inherited robots.txt blocks on AI crawlers.
Progressive disclosure content ships in the initial HTML
0 of 25 pages ship every disclosure panel populated in the raw HTML. 26 pages load more content on scroll with a paginated fallback.
Do this: Render every panel and hide it with CSS.
Nine dimensions, most to least left on the table
32 pages crawled
- simon.com
- shop.simon.com/?utm_source=SIMON&utm_medium=Header&utm_campaign=Simon_Header_SHOPSIMON
- www.simon.com
- www.simon.com/mall
- shop.simon.com/?utm_source=SIMON&utm_medium=Flyout&utm_campaign=SIMON_SHOPONLINE_BANNER
- shop.simon.com/collections/store-coach-outlet?utm_source=SIMON&utm_medium=Flyout&utm_campaign=Simon_WomensMostPopular_CoachOutlet
- shop.simon.com/collections/store-guess-factory?utm_source=SIMON&utm_medium=Flyout&utm_campaign=Simon_WomensMostPopular_GuessFactory
- shop.simon.com/collections/luxe-store-mcm?utm_source=SIMON&utm_medium=Flyout&utm_campaign=Simon_WomensMostPopular_MCM
- shop.simon.com/collections/store-michael-kors-outlet?page=1&utm_source=SIMON&utm_medium=Flyout&utm_campaign=Simon_WomensMostPopular_MichaelKorsOutlet
- shop.simon.com/collections/luxe-store-tods?page=1&utm_source=SIMON&utm_medium=Flyout&utm_campaign=Simon_WomensMostPopular_Tods
- shop.simon.com/collections/women?page=1&utm_source=SIMON&utm_medium=Flyout&utm_campaign=Simon_WomensMostPopular_BestSellers
- shop.simon.com/collections/store-adidas?utm_source=SIMON&utm_medium=Flyout&utm_campaign=Simon_MensMostPopular_Adidas
- shop.simon.com/collections/store-nautica?utm_source=SIMON&utm_medium=Flyout&utm_campaign=Simon_MensMostPopular_Nautica
- shop.simon.com/collections/store-nike?utm_source=SIMON&utm_medium=Flyout&utm_campaign=Simon_MensMostPopular_Nike
- shop.simon.com/collections/store-hugo-boss?utm_source=SIMON&utm_medium=Flyout&utm_campaign=Simon_MensMostPopular_HugoBoss
- shop.simon.com/collections/store-reebok?&utm_source=SIMON&utm_medium=Flyout&utm_campaign=Simon_MensMostPopular_Reebok
- shop.simon.com/collections/men?utm_source=SIMON&utm_medium=Flyout&utm_campaign=Simon_MensMostPopular_BestSellers
- shop.simon.com/collections/luxe-store-mulberry?utm_source=SIMON&utm_medium=Flyout&utm_campaign=Simon_DesignerFavorites_Mulberry
- shop.simon.com/collections/store-philipp-plein?utm_source=SIMON&utm_medium=Flyout&utm_campaign=Simon_DesignerFavorites_PhilippPlein
- shop.simon.com/collections/luxe-store-salvatore-ferragamo?page=1&utm_source=SIMON&utm_medium=Flyout&utm_campaign=Simon_DesignerFavorites_Ferragamo
- shop.simon.com/collections/luxe-store-burberry?page=1&utm_source=SIMON&utm_medium=Flyout&utm_campaign=Simon_DesignerFavorites_Burberry
- shop.simon.com/collections/store-vince?utm_source=SIMON&utm_medium=Flyout&utm_campaign=Simon_DesignerFavorites_Vince
- shop.simon.com/pages/luxe-designers?utm_source=SIMON&utm_medium=Flyout&utm_campaign=Simon_DesignerFavorites_ALL
- shop.simon.com/collections/curated-just-in?utm_source=SIMON&utm_medium=Flyout&utm_campaign=Simon_Collections_NewArrivals
- shop.simon.com/collections/curated-trending-today?utm_source=SIMON&utm_medium=Flyout&utm_campaign=Simon_Collections_TrendingToday
- shop.simon.com/collections/curated-luxe?utm_source=SIMON&utm_medium=Flyout&utm_campaign=Simon_Collections_Luxe
- shop.simon.com/collections/curated-outlet?utm_source=SIMON&utm_medium=Flyout&utm_campaign=Simon_Collections_Outlet
- shop.simon.com/collections/curated-clearance?utm_source=SIMON&utm_medium=Flyout&utm_campaign=Simon_Collections_Clearout
- shop.simon.com/pages/collections?utm_source=SIMON&utm_medium=Flyout&utm_campaign=Simon_Collections_AllCollections
- shop.simon.com/collections/all?ProductType=women-clothing,men-clothing&utm_source=SIMON&utm_medium=Flyout&utm_campaign=Simon_ShopbyCategory_Clothing
- shop.simon.com/collections/shoes?utm_source=SIMON&utm_medium=Flyout&utm_campaign=Simon_ShopbyCategory_Shoes
- shop.simon.com/collections/women-handbags?utm_source=SIMON&utm_medium=Flyout&utm_campaign=Simon_ShopbyCategory_Handbags
Every test · all 67 tests, in order
67 of 67 showing
AIR-1.1 Can AI crawlers reach your site?1 of 4 answer engines are disallowed from content paths: PerplexityBot. 5 training or opt-out agents are also blocked: GPTBot, ClaudeBot, CCBot, Bytespider, Google-Extended. PROBLEM +4
1 of 4 answer engines are disallowed from content paths: PerplexityBot. 5 training or opt-out agents are also blocked: GPTBot, ClaudeBot, CCBot, Bytespider, Google-Extended.
Your robots.txt file tells crawlers what they may read. A line left behind by a previous site or a copied template can quietly shut out the assistants people now use to find you.
AIR-1.2 Does your CDN let AI crawlers through?8 of 8 AI user-agents are rejected or challenged at the edge. PROBLEM +5
8 of 8 AI user-agents are rejected or challenged at the edge.
Even when robots.txt says yes, a firewall or bot-protection rule can turn assistants away before they reach your pages. We sent a request as each one and recorded what came back.
AIR-1.3 Is your page readable without a security challenge?29 of 32 sampled pages return an interactive challenge (Google reCAPTCHA). PROBLEM +1
29 of 32 sampled pages return an interactive challenge (Google reCAPTCHA).
A CAPTCHA or bot check on an ordinary page stops an assistant at the door. These belong on forms and logins, not on pages anyone can read.
AIR-1.4 Have you stated what AI may do with your content?Content Signals declare search=yes, ai-train=no, ai-input=yes. Contradicts robots rules: search=yes but robots.txt disallows PerplexityBot. NEEDS WORK +1
Content Signals declare search=yes, ai-train=no, ai-input=yes. Contradicts robots rules: search=yes but robots.txt disallows PerplexityBot.
There is now a machine-readable way to say whether AI may search, quote or train on your material. Having a position — any position — is the point.
AIR-1.5 Can a person read your AI policy?robots.txt carries a 1899-character policy comment mentioning ai, training, permission. NEEDS WORK +0
robots.txt carries a 1899-character policy comment mentioning ai, training, permission.
A few plain sentences in robots.txt saying what you permit. Journalists and lawyers read this file, and right now most sites say nothing.
AIR-1.6 Are you allowing assistants to quote you?3 of 32 sampled pages suppress excerpting with noarchive, nosnippet or max-snippet:0. NEEDS WORK +0
3 of 32 sampled pages suppress excerpting with noarchive, nosnippet or max-snippet:0.
Some sites carry old settings that tell search engines not to show excerpts. Those same settings stop an assistant quoting your page in an answer.
AIR-1.7 Does this page declare its real address? PASS 2 pt
91% of sampled pages emit a valid, self-consistent canonical. 3 have problems.
A canonical link tells a machine which URL is the true one. Without it the same content can be treated as several competing pages.
AIR-1.8 Is your page structured so a machine can follow it?3% of sampled pages carry one main element, a wrapping article or section, and a nav. 30 exceed 12 divs per semantic element. PROBLEM +1
3% of sampled pages carry one main element, a wrapping article or section, and a nav. 30 exceed 12 divs per semantic element.
Meaningful sections rather than an undifferentiated wall of containers. It is the difference between a document and a soup of boxes.
AIR-1.9 Do your headings tell a clear story?3% of sampled pages have exactly one non-empty h1 and no skipped heading levels. 30 have no h1; 3 skip a level. PROBLEM +1
3% of sampled pages have exactly one non-empty h1 and no skipped heading levels. 30 have no h1; 3 skip a level.
One main heading, then sub-headings in order. Assistants use headings to work out what a page is about and which part answers a question.
AIR-1.10 Do you publish a guide for AI assistants?/llms.txt lists 175 links across 4 sections, covering 1% of the sitemap. NEEDS WORK +1
/llms.txt lists 175 links across 4 sections, covering 1% of the sitemap.
A file at /llms.txt listing your most important pages. It is early and cheap, and it forces a useful conversation about what actually matters on your site.
AIR-1.11 Can a machine tell what your organization is?0 of 60 entity nodes carry an @id (0%). The WebPage to WebSite to Organization chain is broken at WebPage->WebSite, WebSite->Organization. PROBLEM +2
0 of 60 entity nodes carry an @id (0%). The WebPage to WebSite to Organization chain is broken at WebPage->WebSite, WebSite->Organization.
Structured data that links together — your organization, your site, this page — rather than a scattering of disconnected labels.
AIR-1.12 Can a machine be certain it is you?8 sameAs links: 0 tier-1 authority identifiers, 0 tier-2, 8 social. PROBLEM +1
8 sameAs links: 0 tier-1 authority identifiers, 0 tier-2, 8 social.
Links from your markup to authoritative records like Wikidata or ROR. Without one, an assistant is guessing which organization of your name it has found.
AIR-1.13 Can a machine find out who you are?1 of 7 organizational facts are present in the Organization markup: contactPoint. PROBLEM +1
1 of 7 organizational facts are present in the Organization markup: contactPoint.
Founding date, leadership, address, contact details, legal identifiers. Thin About pages are the most common weakness we see.
AIR-2.1 Primary content is present without JavaScript PASS 6 pt
Raw HTML contains 95% of rendered text on average across 32 URLs. The worst template is /mall: 11% across 1 pages.
AIR-2.2 Progressive disclosure content ships in the initial HTML0 of 25 pages ship every disclosure panel populated in the raw HTML. 26 pages load more content on scroll with a paginated fallback. PROBLEM +3
0 of 25 pages ship every disclosure panel populated in the raw HTML. 26 pages load more content on scroll with a paginated fallback.
AIR-2.4 Stable, human-readable anchor IDs on section headings0 of 40 subheadings carry a slug-like id (0%); 7 have generated or unstable ids. NEEDS WORK +1
0 of 40 subheadings carry a slug-like id (0%); 7 have generated or unstable ids.
AIR-2.5 Alt text on content images, empty alt on decorative698 of 1466 images carry appropriate alt text: 614 have no alt attribute, 0 use the filename. NEEDS WORK +0
698 of 1466 images carry appropriate alt text: 614 have no alt attribute, 0 use the filename.
AIR-2.6 Key facts exist as HTML text, not only in images or PDFs32 of 32 sampled pages are thin relative to their images or lead with a document download. PROBLEM +1
32 of 32 sampled pages are thin relative to their images or lead with a document download.
AIR-3.1 JSON-LD is generated from mapped fields, not hardcoded2 of 2 templates repeat identical schema literals across pages whose visible content differs. PROBLEM +2
2 of 2 templates repeat identical schema literals across pages whose visible content differs.
AIR-3.2 BreadcrumbList on every page below the homepage0% of pages below the homepage carry a BreadcrumbList with at least two steps. PROBLEM +2
0% of pages below the homepage carry a BreadcrumbList with at least two steps.
AIR-3.5 isAccessibleForFree, about, and mentions with entity references0 of 30 substantive pages declare isAccessibleForFree; 0 reference subject entities by identifier. PROBLEM +1
0 of 30 substantive pages declare isAccessibleForFree; 0 reference subject entities by identifier.
AIR-3.7 SearchAction on the WebSite nodeNo SearchAction is declared on the WebSite node. PROBLEM +1
AIR-3.8 Schema validation of the structured data on the page PASS 2 pt
62 nodes across 32 pages parse, carry what their types require, and hold well-formed values.
AIR-4.4 Correct X-Robots-Tag on non-HTML resources0 of 25 non-HTML resources carry an X-Robots-Tag. PROBLEM +1
AIR-4.5 Missing pages return 404 or 410 PASS 1 pt
0 of 4 deliberately invalid URLs returned HTTP 200.
AIR-4.6 Redirects resolve in a single hop PASS 1 pt
1 of 32 sampled URLs redirect.
AIR-4.7 Faceted, calendar, and parameterized URLs are controlled PASS 2 pt
29 of 29 parameterized URLs are held in check by noindex or a canonical back to the unfiltered listing.
AIR-4.8 Sitemap lastmod reflects real content changes22573 of 22573 sitemap URLs carry a lastmod across 1 distinct dates; 100% share 2026-09-19. PROBLEM +1
22573 of 22573 sitemap URLs carry a lastmod across 1 distinct dates; 100% share 2026-09-19.
AIR-4.9 Sitemap index split by type, with media sitemaps PASS 1 pt
An index with 3 child sitemaps covering 22573 URLs, with no image or video sitemap.
AIR-4.10 Paginated series are coherently signaled0 of 5 paginated pages carry rel=next or rel=prev. PROBLEM +1
AIR-5.1 Author entity pages with credentialsNo Person entities appear anywhere in the sample. PROBLEM +2
AIR-5.2 Bylines linked to author entities0 of 1 substantive pages reference an author by @id; 0 carry a bare string; 1 are unsigned. PROBLEM +2
0 of 1 substantive pages reference an author by @id; 0 carry a bare string; 1 are unsigned.
AIR-5.4 datePublished and dateModified are accurate0 of 32 pages carry a publication or modification date across 0 distinct dateModified values. PROBLEM +2
0 of 32 pages carry a publication or modification date across 0 distinct dateModified values.
AIR-6.3 Form fields carry label, name, and autocomplete PASS 2 pt
117 of 146 form fields carry an accessible name, a meaningful name attribute and a valid autocomplete token where one applies.
AIR-6.4 No captchas on browse, search, or filter interactions29 of 29 search or filter surfaces return a challenge. PROBLEM +2
AIR-7.1 Markdown companion for every canonical page0 of 21 sampled pages have a Markdown companion at {path}.md returning a Markdown content type. PROBLEM +2
0 of 21 sampled pages have a Markdown companion at {path}.md returning a Markdown content type.
AIR-7.3 Accept: text/markdown content negotiation0 of 25 URLs serve Markdown for Accept: text/markdown. PROBLEM +2
AIR-7.4 Alternate representations are edge-cached0 of 1 generated representations show evidence of an edge cache hit on a repeat request. PROBLEM +2
0 of 1 generated representations show evidence of an edge cache hit on a repeat request.
LI-2 speakable on summaries and ledesNo speakable specification found on any sampled page. NOT SCORED —
LI-3 IndexNow fires on publish and updateNo IndexNow key file was discoverable. The file is named after the key, which is not published, so this is 'not discoverable' rather than proof of absence. NOT SCORED —
No IndexNow key file was discoverable. The file is named after the key, which is not published, so this is 'not discoverable' rather than proof of absence.
LI-4 Editorial policy page referenced via publishingPrinciples0 of 2 Organization nodes reference publishingPrinciples. NOT SCORED —
LI-5 NLWeb endpoint exposed as an MCP server with an ask methodNo NLWeb endpoint was discoverable. NOT SCORED —
LI-6 Domain MCP server over real content APIsNo domain MCP server or tool descriptor was discoverable. NOT SCORED —
BP-3 GA4 channel group for AI assistant referrersWhether a channel group exists for AI assistant referrers is not visible from outside. Analytics presence is; the grouping is not. NOT SCORED —
Whether a channel group exists for AI assistant referrers is not visible from outside. Analytics presence is; the grouping is not.
BP-4 Server-side tagging captures stripped referrersReconciliation between server-side and client-side numbers has to be attested. NOT SCORED —
Reconciliation between server-side and client-side numbers has to be attested.
BP-5 Access log retention with a queryable storeLog retention and queryability are not observable from outside the site. NOT SCORED —
Log retention and queryability are not observable from outside the site.
BP-7 Fixed prompt panel run monthly across modelsAnswer-share tracking happens outside the site and must be attested. NOT SCORED —
Answer-share tracking happens outside the site and must be attested.
BP-8 Extraction-fidelity baseline capturedNo extraction-fidelity baseline is captured by this version of the scanner, which makes no live model calls. NOT SCORED —
No extraction-fidelity baseline is captured by this version of the scanner, which makes no live model calls.
BP-9 Search Console and Bing Webmaster Tools verifiedGoogle Search Console verification found on the page. Whether sitemaps are submitted and alerts reach a person cannot be seen from outside. NOT SCORED —
Google Search Console verification found on the page. Whether sitemaps are submitted and alerts reach a person cannot be seen from outside.
Not applicable — 23 tests answered with a dash, excluded from the score: AIR-2.3, AIR-3.3, AIR-3.4, AIR-3.6, AIR-4.1, AIR-4.2, AIR-4.3, AIR-5.3, AIR-5.5, AIR-5.6, AIR-6.1, AIR-6.2, AIR-7.2, LI-1, LI-7, BP-1, BP-2, BP-3, BP-4, BP-5, BP-6, BP-7, BP-8.
By dimension
Where Simon Property Group's 33 came from. Open a line to see its tests, worst first.
1 · Primary Indicators 9 of 30 9 30
AIR-1.1 Can AI crawlers reach your site?1 of 4 answer engines are disallowed from content paths: PerplexityBot. 5 training or opt-out agents are also blocked: GPTBot, ClaudeBot, CCBot, Bytespider, Google-Extended. PROBLEM
1 of 4 answer engines are disallowed from content paths: PerplexityBot. 5 training or opt-out agents are also blocked: GPTBot, ClaudeBot, CCBot, Bytespider, Google-Extended.
Your robots.txt file tells crawlers what they may read. A line left behind by a previous site or a copied template can quietly shut out the assistants people now use to find you.
AIR-1.2 Does your CDN let AI crawlers through?8 of 8 AI user-agents are rejected or challenged at the edge. PROBLEM
8 of 8 AI user-agents are rejected or challenged at the edge.
Even when robots.txt says yes, a firewall or bot-protection rule can turn assistants away before they reach your pages. We sent a request as each one and recorded what came back.
AIR-1.3 Is your page readable without a security challenge?29 of 32 sampled pages return an interactive challenge (Google reCAPTCHA). PROBLEM
29 of 32 sampled pages return an interactive challenge (Google reCAPTCHA).
A CAPTCHA or bot check on an ordinary page stops an assistant at the door. These belong on forms and logins, not on pages anyone can read.
AIR-1.8 Is your page structured so a machine can follow it?3% of sampled pages carry one main element, a wrapping article or section, and a nav. 30 exceed 12 divs per semantic element. PROBLEM
3% of sampled pages carry one main element, a wrapping article or section, and a nav. 30 exceed 12 divs per semantic element.
Meaningful sections rather than an undifferentiated wall of containers. It is the difference between a document and a soup of boxes.
AIR-1.9 Do your headings tell a clear story?3% of sampled pages have exactly one non-empty h1 and no skipped heading levels. 30 have no h1; 3 skip a level. PROBLEM
3% of sampled pages have exactly one non-empty h1 and no skipped heading levels. 30 have no h1; 3 skip a level.
One main heading, then sub-headings in order. Assistants use headings to work out what a page is about and which part answers a question.
AIR-1.11 Can a machine tell what your organization is?0 of 60 entity nodes carry an @id (0%). The WebPage to WebSite to Organization chain is broken at WebPage->WebSite, WebSite->Organization. PROBLEM
0 of 60 entity nodes carry an @id (0%). The WebPage to WebSite to Organization chain is broken at WebPage->WebSite, WebSite->Organization.
Structured data that links together — your organization, your site, this page — rather than a scattering of disconnected labels.
AIR-1.12 Can a machine be certain it is you?8 sameAs links: 0 tier-1 authority identifiers, 0 tier-2, 8 social. PROBLEM
8 sameAs links: 0 tier-1 authority identifiers, 0 tier-2, 8 social.
Links from your markup to authoritative records like Wikidata or ROR. Without one, an assistant is guessing which organization of your name it has found.
AIR-1.13 Can a machine find out who you are?1 of 7 organizational facts are present in the Organization markup: contactPoint. PROBLEM
1 of 7 organizational facts are present in the Organization markup: contactPoint.
Founding date, leadership, address, contact details, legal identifiers. Thin About pages are the most common weakness we see.
AIR-1.4 Have you stated what AI may do with your content?Content Signals declare search=yes, ai-train=no, ai-input=yes. Contradicts robots rules: search=yes but robots.txt disallows PerplexityBot. NEEDS WORK
Content Signals declare search=yes, ai-train=no, ai-input=yes. Contradicts robots rules: search=yes but robots.txt disallows PerplexityBot.
There is now a machine-readable way to say whether AI may search, quote or train on your material. Having a position — any position — is the point.
AIR-1.5 Can a person read your AI policy?robots.txt carries a 1899-character policy comment mentioning ai, training, permission. NEEDS WORK
robots.txt carries a 1899-character policy comment mentioning ai, training, permission.
A few plain sentences in robots.txt saying what you permit. Journalists and lawyers read this file, and right now most sites say nothing.
AIR-1.6 Are you allowing assistants to quote you?3 of 32 sampled pages suppress excerpting with noarchive, nosnippet or max-snippet:0. NEEDS WORK
3 of 32 sampled pages suppress excerpting with noarchive, nosnippet or max-snippet:0.
Some sites carry old settings that tell search engines not to show excerpts. Those same settings stop an assistant quoting your page in an answer.
AIR-1.10 Do you publish a guide for AI assistants?/llms.txt lists 175 links across 4 sections, covering 1% of the sitemap. NEEDS WORK
/llms.txt lists 175 links across 4 sections, covering 1% of the sitemap.
A file at /llms.txt listing your most important pages. It is early and cheap, and it forces a useful conversation about what actually matters on your site.
Full marks — 1 pass
AIR-1.7 Does this page declare its real address? PASS
91% of sampled pages emit a valid, self-consistent canonical. 3 have problems.
A canonical link tells a machine which URL is the true one. Without it the same content can be treated as several competing pages.
2 · Rendering and Extraction 8 of 15 8 53
AIR-2.2 Progressive disclosure content ships in the initial HTML0 of 25 pages ship every disclosure panel populated in the raw HTML. 26 pages load more content on scroll with a paginated fallback. PROBLEM
0 of 25 pages ship every disclosure panel populated in the raw HTML. 26 pages load more content on scroll with a paginated fallback.
AIR-2.6 Key facts exist as HTML text, not only in images or PDFs32 of 32 sampled pages are thin relative to their images or lead with a document download. PROBLEM
32 of 32 sampled pages are thin relative to their images or lead with a document download.
AIR-2.4 Stable, human-readable anchor IDs on section headings0 of 40 subheadings carry a slug-like id (0%); 7 have generated or unstable ids. NEEDS WORK
0 of 40 subheadings carry a slug-like id (0%); 7 have generated or unstable ids.
AIR-2.5 Alt text on content images, empty alt on decorative698 of 1466 images carry appropriate alt text: 614 have no alt attribute, 0 use the filename. NEEDS WORK
698 of 1466 images carry appropriate alt text: 614 have no alt attribute, 0 use the filename.
Full marks — 1 pass
AIR-2.1 Primary content is present without JavaScript PASS
Raw HTML contains 95% of rendered text on average across 32 URLs. The worst template is /mall: 11% across 1 pages.
3 · Structured Data and the Entity Graph 2 of 15 2 13
AIR-3.1 JSON-LD is generated from mapped fields, not hardcoded2 of 2 templates repeat identical schema literals across pages whose visible content differs. PROBLEM
2 of 2 templates repeat identical schema literals across pages whose visible content differs.
AIR-3.2 BreadcrumbList on every page below the homepage0% of pages below the homepage carry a BreadcrumbList with at least two steps. PROBLEM
0% of pages below the homepage carry a BreadcrumbList with at least two steps.
AIR-3.5 isAccessibleForFree, about, and mentions with entity references0 of 30 substantive pages declare isAccessibleForFree; 0 reference subject entities by identifier. PROBLEM
0 of 30 substantive pages declare isAccessibleForFree; 0 reference subject entities by identifier.
AIR-3.7 SearchAction on the WebSite nodeNo SearchAction is declared on the WebSite node. PROBLEM
Full marks — 1 pass
AIR-3.8 Schema validation of the structured data on the page PASS
62 nodes across 32 pages parse, carry what their types require, and hold well-formed values.
4 · Crawl, Index and Licensing Hygiene 5 of 12 5 42
AIR-4.4 Correct X-Robots-Tag on non-HTML resources0 of 25 non-HTML resources carry an X-Robots-Tag. PROBLEM
AIR-4.8 Sitemap lastmod reflects real content changes22573 of 22573 sitemap URLs carry a lastmod across 1 distinct dates; 100% share 2026-09-19. PROBLEM
22573 of 22573 sitemap URLs carry a lastmod across 1 distinct dates; 100% share 2026-09-19.
AIR-4.10 Paginated series are coherently signaled0 of 5 paginated pages carry rel=next or rel=prev. PROBLEM
Full marks — 4 pass
AIR-4.5 Missing pages return 404 or 410 PASS
0 of 4 deliberately invalid URLs returned HTTP 200.
AIR-4.6 Redirects resolve in a single hop PASS
1 of 32 sampled URLs redirect.
AIR-4.7 Faceted, calendar, and parameterized URLs are controlled PASS
29 of 29 parameterized URLs are held in check by noindex or a canonical back to the unfiltered listing.
AIR-4.9 Sitemap index split by type, with media sitemaps PASS
An index with 3 child sitemaps covering 22573 URLs, with no image or video sitemap.
5 · Authorship, Provenance, Freshness 0 of 12 0 0
AIR-5.1 Author entity pages with credentialsNo Person entities appear anywhere in the sample. PROBLEM
AIR-5.2 Bylines linked to author entities0 of 1 substantive pages reference an author by @id; 0 carry a bare string; 1 are unsigned. PROBLEM
0 of 1 substantive pages reference an author by @id; 0 carry a bare string; 1 are unsigned.
AIR-5.4 datePublished and dateModified are accurate0 of 32 pages carry a publication or modification date across 0 distinct dateModified values. PROBLEM
0 of 32 pages carry a publication or modification date across 0 distinct dateModified values.
6 · Agent Interfaces 2 of 8 2 25
AIR-6.4 No captchas on browse, search, or filter interactions29 of 29 search or filter surfaces return a challenge. PROBLEM
Full marks — 1 pass
AIR-6.3 Form fields carry label, name, and autocomplete PASS
117 of 146 form fields carry an accessible name, a meaningful name attribute and a valid autocomplete token where one applies.
7 · Alternate Representations 0 of 8 0 0
AIR-7.1 Markdown companion for every canonical page0 of 21 sampled pages have a Markdown companion at {path}.md returning a Markdown content type. PROBLEM
0 of 21 sampled pages have a Markdown companion at {path}.md returning a Markdown content type.
AIR-7.3 Accept: text/markdown content negotiation0 of 25 URLs serve Markdown for Accept: text/markdown. PROBLEM
AIR-7.4 Alternate representations are edge-cached0 of 1 generated representations show evidence of an edge cache hit on a repeat request. PROBLEM
0 of 1 generated representations show evidence of an edge cache hit on a repeat request.
What was read
32 pages, sampled from this site's own sitemap and internal links. The cap is 32 — a ceiling, not a certainty: a site with fewer reachable pages yields fewer.
32 other pages, all 200, by template
The receipt
- Cohort
- sp500
- Source
- List of S&P 500 companies (Wikipedia contributors, 2026-09-08)
- Index
- 11.0, in effect 2026-09-24
- Scanner
- 0.1.0
- Scanned
- 2026-09-20T04:39:47.575557+00:00
- Result hash
- aca02a14149c5e6c69364f92ba45153c54fa45c38e192533d8e117e6aefbd70f
- Basis
- public-interest census: a declared bot with a published policy at https://readinessindex.io/bot, honoring robots.txt addressed to AIRBot, reading public pages only. No per-site permission was sought or given.
The hash is the SHA-256 of the full
result.json, recorded when the scan finished rather than
computed now. The raw evidence behind every line above — every
response, cached — is retained and is a median 1.5 MB per site, so
it is not served from here. Ask
us for it if you want to check a number against the bytes it came
from.
Think this is wrong? It might be. Tell us at hello@readinessindex.io and we will look, correct it if it is, and say so publicly either way.
Is your site ready for AI?
Scan a page, get a number you can verify. Free.
About these numbers. Points are whole numbers everywhere on this page. The arithmetic behind them is not — a test scores its weight times its band divided by four, so a row can earn 1.5 points and print as 2. Scores are computed from the exact figures and rounded once, at the end, never from the rounded rows. That means adding the points column up will not give you the score exactly. The rule is published in the standard.