The Census · 2026-Q4 · sp500

AI crawlers cannot reach this site. AIR-1.2 failed, so an assistant asking about this organization gets nothing from here whatever the page contains. The score beside it says how well the site is built; this says none of it is currently reachable. Both are true, and neither replaces the other — which is why the score is no longer capped to say it.

https://wm.com/ · ranked 482 in its cohort · 32 pages read · 51 tests · 20 September 2026

Scan

Waste Management

Score

30/100

7 of 51 tests pass · 70 points still on the table

You · 30
NOT READY 0–39 WEAK 40–59 ADEQUATE 60–74 STRONG 75–89 EXEMPLARY 90–99

We ran all 51 scored tests across 7 dimensions, against the 32 pages we read.

Not ready

AI crawlers are turned away before they reach the content, so nothing else on this page has had a chance to matter.

Do these three first

The fastest 11 points on this scan

1 +5

AIR-1.2 · medium effort

CDN and WAF do not silently reject AI crawlers

8 of 8 AI user-agents are rejected or challenged at the edge.

Do this: Stop rejecting AI crawlers at the CDN or WAF.

2 +3

AIR-1.4 · small effort

Content Signals declared in robots.txt

No Content-Signal line in robots.txt. Missing: search, ai-input, ai-train.

Do this: Declare a Content Signals position in robots.txt.

3 +3

AIR-2.1 · large effort

Primary content is present without JavaScript

Raw HTML contains 56% of rendered text on average across 32 URLs. The worst template is /us/en/mywm/my-services/*: 0% across 1 pages.

Do this: Server-render or statically render the primary content.

Nine dimensions, most to least left on the table

32 pages crawled

Every test · all 67 tests, in order

67 of 67 showing

Show
AIR-1.1 Can AI crawlers reach your site? PASS 5 pt
Why we test this

All 10 AI agents may fetch content paths.

Your robots.txt file tells crawlers what they may read. A line left behind by a previous site or a copied template can quietly shut out the assistants people now use to find you.

AIR-1.2 Does your CDN let AI crawlers through?8 of 8 AI user-agents are rejected or challenged at the edge. PROBLEM +5
What's wrong

8 of 8 AI user-agents are rejected or challenged at the edge.

Even when robots.txt says yes, a firewall or bot-protection rule can turn assistants away before they reach your pages. We sent a request as each one and recorded what came back.

AIR-1.3 Is your page readable without a security challenge?32 of 32 sampled pages return an interactive challenge (Google reCAPTCHA). PROBLEM +1
What's wrong

32 of 32 sampled pages return an interactive challenge (Google reCAPTCHA).

A CAPTCHA or bot check on an ordinary page stops an assistant at the door. These belong on forms and logins, not on pages anyone can read.

AIR-1.4 Have you stated what AI may do with your content?No Content-Signal line in robots.txt. Missing: search, ai-input, ai-train. PROBLEM +3
What's wrong

No Content-Signal line in robots.txt. Missing: search, ai-input, ai-train.

There is now a machine-readable way to say whether AI may search, quote or train on your material. Having a position — any position — is the point.

AIR-1.5 Can a person read your AI policy?robots.txt carries no human-readable policy statement. PROBLEM +1
What's wrong

robots.txt carries no human-readable policy statement.

A few plain sentences in robots.txt saying what you permit. Journalists and lawyers read this file, and right now most sites say nothing.

AIR-1.6 Are you allowing assistants to quote you? PASS 1 pt
Why we test this

No noarchive, nosnippet or max-snippet:0 directives found.

Some sites carry old settings that tell search engines not to show excerpts. Those same settings stop an assistant quoting your page in an answer.

AIR-1.7 Does this page declare its real address? PASS 2 pt
Why we test this

91% of sampled pages emit a valid, self-consistent canonical. 3 have problems.

A canonical link tells a machine which URL is the true one. Without it the same content can be treated as several competing pages.

AIR-1.8 Is your page structured so a machine can follow it?0% of sampled pages carry one main element, a wrapping article or section, and a nav. 25 exceed 12 divs per semantic element. PROBLEM +2
What's wrong

0% of sampled pages carry one main element, a wrapping article or section, and a nav. 25 exceed 12 divs per semantic element.

Meaningful sections rather than an undifferentiated wall of containers. It is the difference between a document and a soup of boxes.

AIR-1.9 Do your headings tell a clear story?28% of sampled pages have exactly one non-empty h1 and no skipped heading levels. 13 have no h1; 6 have several; 7 skip a level. PROBLEM +1
What's wrong

28% of sampled pages have exactly one non-empty h1 and no skipped heading levels. 13 have no h1; 6 have several; 7 skip a level.

One main heading, then sub-headings in order. Assistants use headings to work out what a page is about and which part answers a question.

AIR-1.10 Do you publish a guide for AI assistants?/llms.txt returns HTTP 404. PROBLEM +2
What's wrong

/llms.txt returns HTTP 404.

A file at /llms.txt listing your most important pages. It is early and cheap, and it forces a useful conversation about what actually matters on your site.

AIR-1.11 Can a machine tell what your organization is?4 of 28 entity nodes carry an @id (14%). The WebPage to WebSite to Organization chain is broken at WebPage->WebSite, WebSite->Organization. 2 duplicate definitions of Organization. PROBLEM +1
What's wrong

4 of 28 entity nodes carry an @id (14%). The WebPage to WebSite to Organization chain is broken at WebPage->WebSite, WebSite->Organization. 2 duplicate definitions of Organization.

Structured data that links together — your organization, your site, this page — rather than a scattering of disconnected labels.

AIR-1.12 Can a machine be certain it is you?1 sameAs links: 0 tier-1 authority identifiers, 1 tier-2, 0 social. NEEDS WORK +1
What's half-done

1 sameAs links: 0 tier-1 authority identifiers, 1 tier-2, 0 social.

Links from your markup to authoritative records like Wikidata or ROR. Without one, an assistant is guessing which organization of your name it has found.

AIR-1.13 Can a machine find out who you are?0 of 7 organizational facts are present in the Organization markup: none. PROBLEM +2
What's wrong

0 of 7 organizational facts are present in the Organization markup: none.

Founding date, leadership, address, contact details, legal identifiers. Thin About pages are the most common weakness we see.

AIR-2.1 Primary content is present without JavaScriptRaw HTML contains 56% of rendered text on average across 32 URLs. The worst template is /us/en/mywm/my-services/*: 0% across 1 pages. NEEDS WORK +3
What's half-done

Raw HTML contains 56% of rendered text on average across 32 URLs. The worst template is /us/en/mywm/my-services/*: 0% across 1 pages.

AIR-2.2 Progressive disclosure content ships in the initial HTML0 of 29 pages ship every disclosure panel populated in the raw HTML. 8 pages load more content on scroll with no paginated fallback. PROBLEM +2
What's wrong

0 of 29 pages ship every disclosure panel populated in the raw HTML. 8 pages load more content on scroll with no paginated fallback.

AIR-2.3 Real data tables with header cells1 tables: 0 complete, 0 without scope on header cells, 0 with no header cells, 1 used for layout. PROBLEM +2
What's wrong

1 tables: 0 complete, 0 without scope on header cells, 0 with no header cells, 1 used for layout.

AIR-2.4 Stable, human-readable anchor IDs on section headings0 of 218 subheadings carry a slug-like id (0%); 0 have generated or unstable ids. PROBLEM +2
What's wrong

0 of 218 subheadings carry a slug-like id (0%); 0 have generated or unstable ids.

AIR-2.5 Alt text on content images, empty alt on decorative100 of 160 images carry appropriate alt text: 0 have no alt attribute, 0 use the filename. NEEDS WORK +0
What's half-done

100 of 160 images carry appropriate alt text: 0 have no alt attribute, 0 use the filename.

AIR-2.6 Key facts exist as HTML text, not only in images or PDFs17 of 32 sampled pages are thin relative to their images or lead with a document download. PROBLEM +1
What's wrong

17 of 32 sampled pages are thin relative to their images or lead with a document download.

AIR-3.1 JSON-LD is generated from mapped fields, not hardcoded PASS 2 pt
Why we test this

Schema values track visible content across 1 templates.

AIR-3.2 BreadcrumbList on every page below the homepage0% of pages below the homepage carry a BreadcrumbList with at least two steps. PROBLEM +2
What's wrong

0% of pages below the homepage carry a BreadcrumbList with at least two steps.

AIR-3.3 Vertical-specific schema types NOT SCORED —
AIR-3.4 FAQPage and QAPage on genuine question-and-answer content0 pages carry FAQ or QA markup; 5 pages show visible question-and-answer content. PROBLEM +2
What's wrong

0 pages carry FAQ or QA markup; 5 pages show visible question-and-answer content.

AIR-3.5 isAccessibleForFree, about, and mentions with entity references0 of 18 substantive pages declare isAccessibleForFree; 0 reference subject entities by identifier. PROBLEM +1
What's wrong

0 of 18 substantive pages declare isAccessibleForFree; 0 reference subject entities by identifier.

AIR-3.6 Product, Offer, and Service with real prices3 Offer nodes: 0 complete and matching the page, 0 drifting from visible values, 0 past priceValidUntil. NEEDS WORK +0
What's half-done

3 Offer nodes: 0 complete and matching the page, 0 drifting from visible values, 0 past priceValidUntil.

AIR-3.7 SearchAction on the WebSite node NOT SCORED —
AIR-3.8 Schema validation of the structured data on the pageEverything parses, but nodes on 3 of 29 pages are missing properties their declared type needs. NEEDS WORK +1
What's half-done

Everything parses, but nodes on 3 of 29 pages are missing properties their declared type needs.

AIR-4.1 RSL licensing document published and referenced NOT SCORED —
AIR-4.2 RSL terms propagated to feeds and schema NOT SCORED —
AIR-4.3 Web Bot Auth verification configured NOT SCORED —
AIR-4.4 Correct X-Robots-Tag on non-HTML resources0 of 17 non-HTML resources carry an X-Robots-Tag. PROBLEM +1
What's wrong

0 of 17 non-HTML resources carry an X-Robots-Tag.

AIR-4.5 Missing pages return 404 or 4101 of 3 deliberately invalid URLs returned HTTP 200; 14 sampled pages are thin enough to be not-found templates. PROBLEM +1
What's wrong

1 of 3 deliberately invalid URLs returned HTTP 200; 14 sampled pages are thin enough to be not-found templates.

AIR-4.6 Redirects resolve in a single hop PASS 1 pt
Why we test this

2 of 32 sampled URLs redirect.

AIR-4.7 Faceted, calendar, and parameterized URLs are controlled NOT SCORED —
AIR-4.8 Sitemap lastmod reflects real content changes PASS 2 pt
Why we test this

9810 of 9812 sitemap URLs carry a lastmod across 560 distinct dates; 14% share 2021-03-21.

AIR-4.9 Sitemap index split by type, with media sitemapsA single flat sitemap covering 9812 URLs, with no image or video sitemap. PROBLEM +1
What's wrong

A single flat sitemap covering 9812 URLs, with no image or video sitemap.

AIR-4.10 Paginated series are coherently signaled NOT SCORED —
AIR-5.1 Author entity pages with credentialsNo Person entities appear anywhere in the sample. PROBLEM +2
What's wrong

No Person entities appear anywhere in the sample.

AIR-5.2 Bylines linked to author entities0 of 17 substantive pages reference an author by @id; 0 carry a bare string; 17 are unsigned. PROBLEM +2
What's wrong

0 of 17 substantive pages reference an author by @id; 0 carry a bare string; 17 are unsigned.

AIR-5.3 reviewedBy on YMYL content NOT SCORED —
AIR-5.4 datePublished and dateModified are accurate0 of 32 pages carry a publication or modification date across 0 distinct dateModified values. PROBLEM +2
What's wrong

0 of 32 pages carry a publication or modification date across 0 distinct dateModified values.

AIR-5.5 citation markup on primary-source references NOT SCORED —
AIR-5.6 C2PA Content Credentials on original assets NOT SCORED —
AIR-6.1 OpenAPI specification for public APIs NOT SCORED —
AIR-6.2 /.well-known/ discovery entries for agent capabilities NOT SCORED —
AIR-6.3 Form fields carry label, name, and autocomplete NOT SCORED —
AIR-6.4 No captchas on browse, search, or filter interactions PASS 2 pt
Why we test this

No search or filter surface in the sample returned a challenge.

AIR-7.1 Markdown companion for every canonical page0 of 23 sampled pages have a Markdown companion at {path}.md returning a Markdown content type. PROBLEM +2
What's wrong

0 of 23 sampled pages have a Markdown companion at {path}.md returning a Markdown content type.

AIR-7.2 link rel=alternate advertises the Markdown version NOT SCORED —
AIR-7.3 Accept: text/markdown content negotiation0 of 25 URLs serve Markdown for Accept: text/markdown. PROBLEM +2
What's wrong

0 of 25 URLs serve Markdown for Accept: text/markdown.

AIR-7.4 Alternate representations are edge-cached NOT SCORED —
LI-1 HowTo on procedural content0 of 1 procedural pages carry HowTo with ordered HowToStep. NOT SCORED —
What's wrong

0 of 1 procedural pages carry HowTo with ordered HowToStep.

LI-2 speakable on summaries and ledesNo speakable specification found on any sampled page. NOT SCORED —
What's wrong

No speakable specification found on any sampled page.

LI-3 IndexNow fires on publish and updateNo IndexNow key file was discoverable. The file is named after the key, which is not published, so this is 'not discoverable' rather than proof of absence. NOT SCORED —
What's wrong

No IndexNow key file was discoverable. The file is named after the key, which is not published, so this is 'not discoverable' rather than proof of absence.

LI-4 Editorial policy page referenced via publishingPrinciples0 of 8 Organization nodes reference publishingPrinciples. NOT SCORED —
What's wrong

0 of 8 Organization nodes reference publishingPrinciples.

LI-5 NLWeb endpoint exposed as an MCP server with an ask methodNo NLWeb endpoint was discoverable. NOT SCORED —
What's wrong

No NLWeb endpoint was discoverable.

LI-6 Domain MCP server over real content APIsNo domain MCP server or tool descriptor was discoverable. NOT SCORED —
What's wrong

No domain MCP server or tool descriptor was discoverable.

LI-7 /llms-full.txt where full-text inclusion is appropriate NOT SCORED —
BP-1 Critical flows complete without JS-only interactions NOT SCORED —
BP-2 Stable selectors on critical-flow elements NOT SCORED —
BP-3 GA4 channel group for AI assistant referrersWhether a channel group exists for AI assistant referrers is not visible from outside. Analytics presence is; the grouping is not. NOT SCORED —
What we're measuring

Whether a channel group exists for AI assistant referrers is not visible from outside. Analytics presence is; the grouping is not.

BP-4 Server-side tagging captures stripped referrersReconciliation between server-side and client-side numbers has to be attested. NOT SCORED —
What we're measuring

Reconciliation between server-side and client-side numbers has to be attested.

BP-5 Access log retention with a queryable storeLog retention and queryability are not observable from outside the site. NOT SCORED —
What we're measuring

Log retention and queryability are not observable from outside the site.

BP-6 Scheduled AI crawler activity report NOT SCORED —
BP-7 Fixed prompt panel run monthly across modelsAnswer-share tracking happens outside the site and must be attested. NOT SCORED —
What we're measuring

Answer-share tracking happens outside the site and must be attested.

BP-8 Extraction-fidelity baseline capturedNo extraction-fidelity baseline is captured by this version of the scanner, which makes no live model calls. NOT SCORED —
What we're measuring

No extraction-fidelity baseline is captured by this version of the scanner, which makes no live model calls.

BP-9 Search Console and Bing Webmaster Tools verifiedNeither Search Console nor Bing Webmaster Tools verification was found. NOT SCORED —
What's wrong

Neither Search Console nor Bing Webmaster Tools verification was found.

Not applicable — 24 tests answered with a dash, excluded from the score: AIR-3.3, AIR-3.7, AIR-4.1, AIR-4.2, AIR-4.3, AIR-4.7, AIR-4.10, AIR-5.3, AIR-5.5, AIR-5.6, AIR-6.1, AIR-6.2, AIR-6.3, AIR-7.2, AIR-7.4, LI-7, BP-1, BP-2, BP-3, BP-4, BP-5, BP-6, BP-7, BP-8.

By dimension

Where Waste Management's 30 came from. Open a line to see its tests, worst first.

DimensionShare earnedEarnedScore
1 · Primary Indicators 10 33
Scored zero — 9 problems
AIR-1.2 Does your CDN let AI crawlers through?8 of 8 AI user-agents are rejected or challenged at the edge. PROBLEM
What's wrong

8 of 8 AI user-agents are rejected or challenged at the edge.

Even when robots.txt says yes, a firewall or bot-protection rule can turn assistants away before they reach your pages. We sent a request as each one and recorded what came back.

AIR-1.3 Is your page readable without a security challenge?32 of 32 sampled pages return an interactive challenge (Google reCAPTCHA). PROBLEM
What's wrong

32 of 32 sampled pages return an interactive challenge (Google reCAPTCHA).

A CAPTCHA or bot check on an ordinary page stops an assistant at the door. These belong on forms and logins, not on pages anyone can read.

AIR-1.4 Have you stated what AI may do with your content?No Content-Signal line in robots.txt. Missing: search, ai-input, ai-train. PROBLEM
What's wrong

No Content-Signal line in robots.txt. Missing: search, ai-input, ai-train.

There is now a machine-readable way to say whether AI may search, quote or train on your material. Having a position — any position — is the point.

AIR-1.5 Can a person read your AI policy?robots.txt carries no human-readable policy statement. PROBLEM
What's wrong

robots.txt carries no human-readable policy statement.

A few plain sentences in robots.txt saying what you permit. Journalists and lawyers read this file, and right now most sites say nothing.

AIR-1.8 Is your page structured so a machine can follow it?0% of sampled pages carry one main element, a wrapping article or section, and a nav. 25 exceed 12 divs per semantic element. PROBLEM
What's wrong

0% of sampled pages carry one main element, a wrapping article or section, and a nav. 25 exceed 12 divs per semantic element.

Meaningful sections rather than an undifferentiated wall of containers. It is the difference between a document and a soup of boxes.

AIR-1.9 Do your headings tell a clear story?28% of sampled pages have exactly one non-empty h1 and no skipped heading levels. 13 have no h1; 6 have several; 7 skip a level. PROBLEM
What's wrong

28% of sampled pages have exactly one non-empty h1 and no skipped heading levels. 13 have no h1; 6 have several; 7 skip a level.

One main heading, then sub-headings in order. Assistants use headings to work out what a page is about and which part answers a question.

AIR-1.10 Do you publish a guide for AI assistants?/llms.txt returns HTTP 404. PROBLEM
What's wrong

/llms.txt returns HTTP 404.

A file at /llms.txt listing your most important pages. It is early and cheap, and it forces a useful conversation about what actually matters on your site.

AIR-1.11 Can a machine tell what your organization is?4 of 28 entity nodes carry an @id (14%). The WebPage to WebSite to Organization chain is broken at WebPage->WebSite, WebSite->Organization. 2 duplicate definitions of Organization. PROBLEM
What's wrong

4 of 28 entity nodes carry an @id (14%). The WebPage to WebSite to Organization chain is broken at WebPage->WebSite, WebSite->Organization. 2 duplicate definitions of Organization.

Structured data that links together — your organization, your site, this page — rather than a scattering of disconnected labels.

AIR-1.13 Can a machine find out who you are?0 of 7 organizational facts are present in the Organization markup: none. PROBLEM
What's wrong

0 of 7 organizational facts are present in the Organization markup: none.

Founding date, leadership, address, contact details, legal identifiers. Thin About pages are the most common weakness we see.

Partly scored — 1 needs work
AIR-1.12 Can a machine be certain it is you?1 sameAs links: 0 tier-1 authority identifiers, 1 tier-2, 0 social. NEEDS WORK
What's half-done

1 sameAs links: 0 tier-1 authority identifiers, 1 tier-2, 0 social.

Links from your markup to authoritative records like Wikidata or ROR. Without one, an assistant is guessing which organization of your name it has found.

Full marks — 3 pass
AIR-1.1 Can AI crawlers reach your site? PASS
Why we test this

All 10 AI agents may fetch content paths.

Your robots.txt file tells crawlers what they may read. A line left behind by a previous site or a copied template can quietly shut out the assistants people now use to find you.

AIR-1.6 Are you allowing assistants to quote you? PASS
Why we test this

No noarchive, nosnippet or max-snippet:0 directives found.

Some sites carry old settings that tell search engines not to show excerpts. Those same settings stop an assistant quoting your page in an answer.

AIR-1.7 Does this page declare its real address? PASS
Why we test this

91% of sampled pages emit a valid, self-consistent canonical. 3 have problems.

A canonical link tells a machine which URL is the true one. Without it the same content can be treated as several competing pages.

2 · Rendering and Extraction 4 27
Scored zero — 4 problems
AIR-2.2 Progressive disclosure content ships in the initial HTML0 of 29 pages ship every disclosure panel populated in the raw HTML. 8 pages load more content on scroll with no paginated fallback. PROBLEM
What's wrong

0 of 29 pages ship every disclosure panel populated in the raw HTML. 8 pages load more content on scroll with no paginated fallback.

AIR-2.3 Real data tables with header cells1 tables: 0 complete, 0 without scope on header cells, 0 with no header cells, 1 used for layout. PROBLEM
What's wrong

1 tables: 0 complete, 0 without scope on header cells, 0 with no header cells, 1 used for layout.

AIR-2.4 Stable, human-readable anchor IDs on section headings0 of 218 subheadings carry a slug-like id (0%); 0 have generated or unstable ids. PROBLEM
What's wrong

0 of 218 subheadings carry a slug-like id (0%); 0 have generated or unstable ids.

AIR-2.6 Key facts exist as HTML text, not only in images or PDFs17 of 32 sampled pages are thin relative to their images or lead with a document download. PROBLEM
What's wrong

17 of 32 sampled pages are thin relative to their images or lead with a document download.

Partly scored — 2 need work
AIR-2.1 Primary content is present without JavaScriptRaw HTML contains 56% of rendered text on average across 32 URLs. The worst template is /us/en/mywm/my-services/*: 0% across 1 pages. NEEDS WORK
What's half-done

Raw HTML contains 56% of rendered text on average across 32 URLs. The worst template is /us/en/mywm/my-services/*: 0% across 1 pages.

AIR-2.5 Alt text on content images, empty alt on decorative100 of 160 images carry appropriate alt text: 0 have no alt attribute, 0 use the filename. NEEDS WORK
What's half-done

100 of 160 images carry appropriate alt text: 0 have no alt attribute, 0 use the filename.

3 · Structured Data and the Entity Graph 3 20
Scored zero — 3 problems
AIR-3.2 BreadcrumbList on every page below the homepage0% of pages below the homepage carry a BreadcrumbList with at least two steps. PROBLEM
What's wrong

0% of pages below the homepage carry a BreadcrumbList with at least two steps.

AIR-3.4 FAQPage and QAPage on genuine question-and-answer content0 pages carry FAQ or QA markup; 5 pages show visible question-and-answer content. PROBLEM
What's wrong

0 pages carry FAQ or QA markup; 5 pages show visible question-and-answer content.

AIR-3.5 isAccessibleForFree, about, and mentions with entity references0 of 18 substantive pages declare isAccessibleForFree; 0 reference subject entities by identifier. PROBLEM
What's wrong

0 of 18 substantive pages declare isAccessibleForFree; 0 reference subject entities by identifier.

Partly scored — 2 need work
AIR-3.6 Product, Offer, and Service with real prices3 Offer nodes: 0 complete and matching the page, 0 drifting from visible values, 0 past priceValidUntil. NEEDS WORK
What's half-done

3 Offer nodes: 0 complete and matching the page, 0 drifting from visible values, 0 past priceValidUntil.

AIR-3.8 Schema validation of the structured data on the pageEverything parses, but nodes on 3 of 29 pages are missing properties their declared type needs. NEEDS WORK
What's half-done

Everything parses, but nodes on 3 of 29 pages are missing properties their declared type needs.

Full marks — 1 pass
AIR-3.1 JSON-LD is generated from mapped fields, not hardcoded PASS
Why we test this

Schema values track visible content across 1 templates.

4 · Crawl, Index and Licensing Hygiene 3 25
Scored zero — 3 problems
AIR-4.4 Correct X-Robots-Tag on non-HTML resources0 of 17 non-HTML resources carry an X-Robots-Tag. PROBLEM
What's wrong

0 of 17 non-HTML resources carry an X-Robots-Tag.

AIR-4.5 Missing pages return 404 or 4101 of 3 deliberately invalid URLs returned HTTP 200; 14 sampled pages are thin enough to be not-found templates. PROBLEM
What's wrong

1 of 3 deliberately invalid URLs returned HTTP 200; 14 sampled pages are thin enough to be not-found templates.

AIR-4.9 Sitemap index split by type, with media sitemapsA single flat sitemap covering 9812 URLs, with no image or video sitemap. PROBLEM
What's wrong

A single flat sitemap covering 9812 URLs, with no image or video sitemap.

Full marks — 2 pass
AIR-4.6 Redirects resolve in a single hop PASS
Why we test this

2 of 32 sampled URLs redirect.

AIR-4.8 Sitemap lastmod reflects real content changes PASS
Why we test this

9810 of 9812 sitemap URLs carry a lastmod across 560 distinct dates; 14% share 2021-03-21.

5 · Authorship, Provenance, Freshness 0 0
Scored zero — 3 problems
AIR-5.1 Author entity pages with credentialsNo Person entities appear anywhere in the sample. PROBLEM
What's wrong

No Person entities appear anywhere in the sample.

AIR-5.2 Bylines linked to author entities0 of 17 substantive pages reference an author by @id; 0 carry a bare string; 17 are unsigned. PROBLEM
What's wrong

0 of 17 substantive pages reference an author by @id; 0 carry a bare string; 17 are unsigned.

AIR-5.4 datePublished and dateModified are accurate0 of 32 pages carry a publication or modification date across 0 distinct dateModified values. PROBLEM
What's wrong

0 of 32 pages carry a publication or modification date across 0 distinct dateModified values.

6 · Agent Interfaces 2 25
Full marks — 1 pass
AIR-6.4 No captchas on browse, search, or filter interactions PASS
Why we test this

No search or filter surface in the sample returned a challenge.

7 · Alternate Representations 0 0
Scored zero — 2 problems
AIR-7.1 Markdown companion for every canonical page0 of 23 sampled pages have a Markdown companion at {path}.md returning a Markdown content type. PROBLEM
What's wrong

0 of 23 sampled pages have a Markdown companion at {path}.md returning a Markdown content type.

AIR-7.3 Accept: text/markdown content negotiation0 of 25 URLs serve Markdown for Accept: text/markdown. PROBLEM
What's wrong

0 of 25 URLs serve Markdown for Accept: text/markdown.

What was read

32 pages, sampled from this site's own sitemap and internal links. The cap is 32 — a ceiling, not a certainty: a site with fewer reachable pages yields fewer.

32 other pages, all 200, by template
2× / https://wm.com/
7× /us/en/* https://www.wm.com/us/en/healthcare
1× /us/en/terms/* https://www.wm.com/us/en/terms/terms-and-conditions
1× /us/* https://www.wm.com/us/en
1× /us/en/mywm/my-services/* https://www.wm.com/us/en/mywm/my-services/view-pickup-eta
5× /us/en/home/* https://www.wm.com/us/en/home/residential-waste-recycling-pickup
3× /us/en/home/*/* https://www.wm.com/us/en/home/residential-waste-recycling-pickup/32-gallon-trash-recycling
1× /us/en/home/bulk-trash-pickup/* https://www.wm.com/us/en/home/bulk-trash-pickup/request
1× /us/en/home/bulk-trash-pickup/request/* https://www.wm.com/us/en/home/bulk-trash-pickup/request/city-billed-confirmation
4× /ca/en/support/faqs/billing/* https://www.wm.com/ca/en/support/faqs/billing/what-are-the-charges-on-my-waste-management-invoice
4× /us/en/support/faqs/billing/* https://www.wm.com/us/en/support/faqs/billing/what-are-the-charges-on-my-waste-management-invoice
1× /ca/fr/support/faqs/billing/* https://www.wm.com/ca/fr/support/faqs/billing/what-are-the-charges-on-my-waste-management-invoice
1× /ca/en/support/faqs/* https://www.wm.com/ca/en/support/faqs//how-do-i-report-a-lost-or-stolen-container

The receipt

Cohort
sp500
Source
List of S&P 500 companies (Wikipedia contributors, 2026-09-08)
Index
11.0, in effect 2026-09-24
Scanner
0.1.0
Scanned
2026-09-20T05:07:50.861805+00:00
Result hash
de982cb4864ef850ee805747a65cd3e2e4cd018e1e52fa84dcc99ff776c78f39
Basis
public-interest census: a declared bot with a published policy at https://readinessindex.io/bot, honoring robots.txt addressed to AIRBot, reading public pages only. No per-site permission was sought or given.

The hash is the SHA-256 of the full result.json, recorded when the scan finished rather than computed now. The raw evidence behind every line above — every response, cached — is retained and is a median 1.5 MB per site, so it is not served from here. Ask us for it if you want to check a number against the bytes it came from.

Think this is wrong? It might be. Tell us at hello@readinessindex.io and we will look, correct it if it is, and say so publicly either way.

Is your site ready for AI?

Scan a page, get a number you can verify. Free.

About these numbers. Points are whole numbers everywhere on this page. The arithmetic behind them is not — a test scores its weight times its band divided by four, so a row can earn 1.5 points and print as 2. Scores are computed from the exact figures and rounded once, at the end, never from the rounded rows. That means adding the points column up will not give you the score exactly. The rule is published in the standard.