Home / Research / Crawlability study

Study

Which platforms can AI crawlers actually read?

Whether your words reach an AI crawler is decided by what renders them, so this samples sites by the platform they are built on: Framer, Webflow, Shopify, WordPress, Bubble, and 5 more. Sites were chosen in advance, not gathered from people who came to us. Each is fetched the way an AI crawler fetches it, with no JavaScript, and again in a real browser, and the two are compared. Every figure carries the number of sites behind it.

Snapshot 2026-08-03. 97 of 392 seeded sites measured so far (25% of the frame). The sample is scanned a few sites a night, so this grows.

80 mean AI Visibility Score across 97 sampled sites
79% of rendered content reaches a crawler on the average sampled page
12% ship almost nothing to crawlers (under 30% prose parity)

Grade distribution

Every measured site in the sample, by the grade its score falls into.

32 A
40 B
10 C
10 D
5 F

By platform

Sorted by score, worst first, so the platforms with the most to fix lead. A row publishes figures only once at least 5 of its sites have been measured.

Platform Measured Mean score Mean parity Near invisible
Webflow 8 / 42 65 59% 25%
Bubble and no-code apps 10 / 42 70 65% 30%
SaaS marketing sites 11 / 39 71 67% 18%
Single-page apps 14 / 45 80 76% 21%
Static site generators and docs 8 / 31 80 82% 13%
Shopify storefronts 8 / 34 83 79% 0%
WordPress 12 / 41 85 86% 8%
Framer 10 / 32 86 92% 0%
Media and local news 6 / 32 92 95% 0%
Wix and Squarespace 10 / 54 94 99% 0%

How this was sampled

The frame was chosen before anything was scanned, which is the part that makes it a study rather than a summary of our traffic. Sites were drawn from public showcases, directories and platform galleries, 10 strata in all, then scanned a few per night in the order they were listed. The sources each stratum came from are listed below.

Splitting by platform rather than by industry is the whole point. What decides whether your words reach a crawler is the thing that renders them, not what your business sells: a Framer marketing site and a Bubble app fail the same way as each other and a completely different way from a WordPress blog. It is also the only split a reader can act on, because you can change how a page is built and you cannot change what industry you are in.

The stratum is still a judgement. Identifying a builder from the outside is inference, and a site can be rebuilt after it was listed, so a row describes the sites we filed under it rather than every site on that platform.

One page per site, its home page, counted once. A site that is re-measured later replaces its earlier result rather than adding a second row. Sites that came to Lantad and scanned themselves are excluded from every figure on this page; they are counted in the live index, which says so.

Sites are only fetched where robots.txt allows our crawler, and a site that declines is skipped and not asked again for a month. The conduct rules are on /bot, and the scoring method is on /methodology.

Where each stratum came from

Counts are sites seeded, not sites measured. Published so the frame can be audited rather than taken on trust.

  • Webflow (42 seeded) from createtoday.io/examples?category=community&platform=webflow, createtoday.io/examples?category=education&platform=webflow, createtoday.io/examples?category=media&platform=webflow, createtoday.io/examples?platform=webflow, flowout.com/portfolio, flowzai.com/blog-post/best-webflow-websites, htmlburger.com/blog/webflow-ecommerce-examples/, joinamply.com/post/best-webflow-websites, todaymade.com/blog/websites-built-with-webflow, wedoflow.com/post/webflow-success-stories-popular-and-impactful-websites-built-with-webflow
  • Bubble and no-code apps (42 seeded) from shno.co/blog/bubble-app-examples, shno.co/blog/carrd-websites, shno.co/blog/glide-app-examples, shno.co/blog/softr-app-examples, softr.io/customer-stories/altitude-marketing, softr.io/customer-stories/eight-digit-media, softr.io/customer-stories/lakeshore-windows, softr.io/customer-stories/no-code-week, softr.io/customer-stories/the-board
  • SaaS marketing sites (39 seeded) from failory.com/startups/developer-tools, failory.com/startups/fintech, failory.com/startups/health-care, failory.com/startups/human-resources, failory.com/startups/logistics, failory.com/startups/marketing
  • Single-page apps (45 seeded) from extruct.ai/ycombinator-companies/w25, producthunt.com/leaderboard/monthly/2025/11, producthunt.com/leaderboard/monthly/2025/6, producthunt.com/leaderboard/monthly/2025/8, producthunt.com/leaderboard/monthly/2026/1, producthunt.com/leaderboard/monthly/2026/3, producthunt.com/leaderboard/monthly/2026/4, producthunt.com/leaderboard/monthly/2026/5
  • Static site generators and docs (31 seeded) from docusaurus.io/showcase, www.11ty.dev/speedlify/
  • Shopify storefronts (34 seeded) from ecommerceparadise.com/50-best-shopify-stores-in-2026, fastbundle.co/blog/best-shopify-stores, shopify.com/blog/niche-stores, sitebuilderreport.com/inspiration/shopify-niche-stores
  • WordPress (41 seeded) from delmain.co/blog/best-dental-websites/, web search, www.allianceinteractive.com/50-best-accounting-website-examples/
  • Framer (32 seeded) from brixtemplates.com/blog/best-framer-agencies, goodspeed.studio/framer-website-examples/e-commerce, landing.gallery/website-builder/framer, sitebuilderreport.com/inspiration/framer-websites
  • Media and local news (32 seeded) from lionpublishers.com/lion-welcomes-new-members-from-18-states/, lionpublishers.com/meet-the-51-finalists-for-the-2025-lion-sustainability-awards/, web search, www.anoffgridlife.com/homestead-blogs/
  • Wix and Squarespace (54 seeded) from createtoday.io/examples?category=beauty-salon&platform=squarespace, sitebuilderreport.com/inspiration/local-business-websites, sitebuilderreport.com/inspiration/squarespace-business-websites, sitebuilderreport.com/inspiration/squarespace-charity-websites, sitebuilderreport.com/wix-examples, tooltester.com/en/blog/wix-website-examples/, web search

About this study

Which sites are in this study?

A seeded list of 392 sites across 10 platform strata, chosen in advance from public showcases, directories and platform galleries, and scanned a few per night. Sites that scanned themselves on Lantad are deliberately NOT counted here: they chose to be measured, which makes them a different population, and mixing the two would describe our visitors rather than the web. The wider live index at /research does count them and says so.

Why do some rows say there is not enough data?

Because there is not. A row publishes a figure only once at least 5 of its seeded sites have been measured, and the background scan adds a few sites a night, so strata fill up over weeks. The alternative is a mean over two sites presented as a fact about a whole platform, which is the sort of number this product exists to argue against.

Why split by platform rather than by industry?

Because the platform is what decides the answer. Whether your words reach an AI crawler is determined by the thing that renders them, so a Framer marketing site and a Bubble app fail the same way as each other and a completely different way from a WordPress blog, however unrelated their businesses are. Industry would group sites by something that has nothing to do with why they score what they score. Platform is also the only split a reader can act on: you can change how a page is built, and you cannot change what industry you are in.

How is a site's platform decided?

From how it was sampled: each stratum was drawn from public showcases, directories and galleries for that platform, and those sources are listed on this page. It remains an inference rather than a measurement, and a site can be rebuilt after it was listed, so a row describes the sites we filed under it rather than every site on that platform.

What exactly is measured?

One page per site, its home page, fetched twice: once as an AI crawler sees it, with no JavaScript, and once in a real browser, then diffed. Prose parity is how much of the rendered content survives that first fetch. Each scan also evaluates robots.txt for 15 published AI crawler tokens. The full method is on /methodology.

Can I cite this?

Yes, with a link to this page. The figures move as more of the frame is measured, so quote the snapshot date shown beside them. Licensed CC BY 4.0.

Free to reuse and republish with attribution and a link to this page. These figures are published under CC BY 4.0. The sampling frame behind them is described above.

See where your own site sits

Run the same check on any page and get the graded report, with the ranked fixes.