When webpages become words

You choose words, images and layout to get your message across. When an AI tool reads a text version of your page, the words can remain while some of their meaning is lost.

  1. Browser view
  2. Extracted text
  3. AI analysis

This is one part of AI SEO: checking that your facts and their relationships stay clear when the visual layout is removed.

Slack: feature inclusion markers lost

Slack’s page shows which features are included in each plan. In the flattened text, the names appear under both plans without that distinction.

Slack pricing
On the webpageIllustration
slack

Free

For getting started

  • −AI daily recaps
  • −AI file summaries
Not included

Business+

For growing teams

  • ✓AI daily recaps
  • ✓AI file summaries
Included
After flattening.txt
# Free
 
AI daily recaps
AI file summaries
 
# Business+
 
AI daily recaps
AI file summaries

The text no longer says which plan includes these features.

Simplified illustration · not a live parse

Make the relationship explicit. Pair every feature, price or condition with the product it belongs to.

What the research found

In the 5 October 2026 check, Slack’s en-GB page excluded these features from Free and included them in Business+. The tested web extraction retained the names under Free without their inclusion markers. The comparison table also lost the markers.

This illustrates an observed extraction issue. No incorrect AI answer was demonstrated. The visuals above are simplified reconstructions; “Clearer version” is suggested wording, not a tested change or a description of current plan entitlements.

Read the research notebook ↗

CommBank: hidden and visible copy combined

CommBank’s page hides an alternative set of benefits. The flattened text includes both versions, without showing which one the customer sees.

CommBank home loans
The page and its hidden textIllustration

Simple home loan

Up to two offset accounts
Visible on the page
Link multiple offsetsHidden alternative in the HTML
display: none
After flattening.txt
# Simple home loan
 
Up to two offset accounts
 
Link multiple offsets

Both variants appear, without their visibility context.

Simplified illustration · not a live parse

Check what the whole page says. Remove stale variants or give alternative offers clear context in the HTML.

What the research found

The 5 October 2026 check found “Link multiple offsets” inside hidden Digi and Simple benefit variants in the rendered DOM. The visible Simple benefits specified up to two offsets. Hidden variants appeared in the tested extraction alongside other features.

The earlier blind answer test still recovered the checked Simple-loan terms correctly. Original HTTP response HTML was not verified, so the complete pipeline cause remains open. This is a dated content example, not current product guidance. The illustration condenses the observation rather than reproducing the page or an exact extraction.

Read the research notebook ↗

Notion: plan checkmarks lost

Notion’s pricing table uses checkmarks to connect workspace PDF export to Business and Enterprise. The text keeps the feature name but loses those connections.

Notion pricing
On the webpageIllustration

Notion plans

Export entire workspace as PDF
FreePlusBusinessEnterprise

The checkmarks carry the answer.

After flattening.txt
Export entire workspace as PDF

The feature is named, but the text doesn’t say which plans include it.

Based on the 5–6 October check · not a live parse

Give checkmarks a text equivalent. Keep each inclusion value connected to its feature and plan.

What the research found

Checked 5–6 October 2026 using GPTSee’s static fetcher and parser, with a browser comparison.

The browser and fetched HTML showed SVG checkmarks for whole-workspace PDF export in the Business and Enterprise columns, with blank cells for Free and Plus. GPTSee’s tested extraction retained the feature name and its tooltip description, but not the checkmarks or their column relationships.

The illustration isolates one row; the text excerpt omits the tooltip and surrounding features. Suggested wording reflects the checked row, not a tested site change or a guarantee of current plan entitlements.

Visit the source page ↗

Micicca: measurements lost from size chart images

Micicca’s charts put sizes and body measurements inside images. The extracted text names the images, but doesn’t contain the measurements a shopper needs.

Micicca size guide
On the webpageIllustration

Micicca size guide

After flattening.txt
# Size Guide
 
Image: SizeGuide Graph INCH CM
 
Image: SizeGuide Graph INCH CM
 
How to Measure

The labels survive. The size-to-measurement relationships are missing.

Based on the 5–6 October check · not a live parse

Publish the data behind the image. A text table can carry the same measurements alongside the illustrated guide.

What the research found

Checked 5–6 October 2026 using GPTSee’s static fetcher and parser, with a browser comparison.

The browser displayed image-based charts in inches and centimetres. GPTSee’s static extraction returned two image references labelled “SizeGuide Graph INCH CM”, followed by measuring instructions. The measurements shown in the charts were absent from the extracted text.

The illustration shows selected sizes and measurements from the inches chart. The suggested text is a subset, not a replacement for the full guide. This check did not test an AI workflow that reads image pixels.

Visit the source page ↗

Le Méridien: room layout lost from floor plans

Le Méridien Melbourne shows how its meeting rooms fit together. Room capacities survive as text, but the floor-plan labels don’t describe the layout.

Le Méridien events
On the webpageIllustration

Meeting floor

Selected spaces · layout simplified, not to scale

After flattening.txt
Floor Plans
 
Meeting Room Floor Plan
 
Meeting Room Floor Plan
 
Meeting Room Floor Plan

The text says a floor plan exists, without describing its spatial information.

Based on the 5–6 October check · not a live parse

Describe the relationships that matter. Support a map or floor plan with text explaining where key spaces are in relation to each other.

What the research found

Checked 5–6 October 2026 using GPTSee’s static fetcher and parser, with a browser comparison.

The expanded Floor Plans section showed room layouts, entrances and lifts. GPTSee’s extraction retained repeated “Meeting Room Floor Plan” labels without that spatial detail. The separate capacity tables retained room names, dimensions and seating capacities.

The diagram simplifies selected spaces on the meeting floor and is not to scale. The suggested description covers the relationships shown here, not the whole plan. Other page copy did describe the event precinct and its rooms, so this is a specific loss of spatial detail rather than a loss of all venue information.

Visit the source page ↗

Archie Brothers: iframe content missing from the text

Archie Brothers’ Melbourne venue card displays an address, opening hours and a phone number inside an iframe. The parent-page extraction keeps the frame’s label, without its contents.

Archie Brothers Melbourne
On the webpageIllustration

Venue information

Melbourne Docklands

Address
440 Docklands Dr
Docklands, VIC 3008
Sunday hours
10:00 AM – 10:00 PM
Phone
+61 3 7003 9772

This card is inside an iframe

The address is also repeated in the page’s own text.

From the static HTML.txt
## Venue Information
 
Embed Widget
 
### Getting to our venue

The hours and phone number are missing. The address survives elsewhere on the page.

Based on the 5–6 October check · not a live parse

Keep essential facts outside the embed, too. An iframe is a separate document. Include key details in the surrounding page so they remain available when the frame isn’t read.

What the research found

Checked 5–6 October 2026 using GPTSee’s static fetcher and parser, with a browser comparison.

On 6 October 2026, the browser’s iframe titled “Embed Widget” displayed the Melbourne Docklands address, opening hours and phone number. GPTSee’s static parser reported one excluded iframe and retained only “Embed Widget” between the Venue Information and Getting to our venue headings. The iframe’s internal document was not traversed.

The address was also present in the parent page’s Other locations section and survived extraction. The venue opening hours and phone number did not. This example therefore shows the loss of the embedded venue card and those specific facts, not the disappearance of every copy of the address.

The illustration and suggested text show the checked address, Sunday hours and phone number only. These are dated examples, not live visitor information. Browser tools may be able to inspect the frame separately.

Visit the source page ↗

Showpo: measurement table missing from static HTML

Showpo’s browser view has both size conversions and body measurements. Its static HTML contains the conversions, but the measurement table is populated in the browser.

Showpo size guide

The next two examples show a related issue: content added by JavaScript can be missing before text extraction begins.

On the webpageIllustration

Body measurements

Selected sizes · centimetres
AUBustWaistHip
683.565.592.5
8866895
109173100

Visible once the browser populates the table.

From the static HTML.txt
### body measurements
 
CM
 
IN
 
##### How to measure

The heading and unit controls survive, but the measurement rows are absent.

Based on the 5–6 October check · not a live parse

Include essential data in the initial HTML. Keep interactive controls, while making the underlying facts available before JavaScript runs.

What the research found

Checked 5–6 October 2026 using GPTSee’s static fetcher and parser, with a browser comparison.

The fetched HTML contained an empty table for body measurements. In the browser, that table had size, bust, waist and hip values. GPTSee’s extraction retained the size-conversion and shoe-size tables elsewhere on the same page, but not these measurement rows.

The excerpts omit surrounding fit advice and measuring instructions; the illustration shows three checked rows. This is content missing from the static response before flattening, not a failure to preserve an existing table. Browser-based retrieval may obtain the populated table.

Visit the source page ↗

HubSpot: no body text in the static response

HubSpot’s Marketing Hub page displays plans, prices and features in a browser. The tested static response contained empty page containers, leaving no body text to extract.

HubSpot pricing
On the webpageIllustration

Marketing Hub

Free

Generate leads and measure results

Starter

Core marketing and sales tools

Professional

Campaigns, automation and reporting

Enterprise

Personalised customer journeys

Plans and their details are visible in the browser.

From the static HTML.txt
No body text

0 lines returned in this check

This fetch returned HTML, but no plan content was available as body text.

Based on the 5–6 October check · not a live parse

Check the response behind the page. Make core product content available to tools that read HTML without running JavaScript.

What the research found

Checked 5–6 October 2026 using GPTSee’s static fetcher and parser, with a browser comparison.

The static fetch returned a successful HTML response and the page title, but GPTSee produced zero body-text lines. After scripts and styles were excluded, the body contained empty navigation and page containers. The browser displayed Free, Starter, Professional and Enterprise plans with their details.

Another web retrieval tool obtained HubSpot’s page content during the research. This finding is specific to the tested static fetch; it does not show that AI tools generally cannot read HubSpot. The illustration omits prices and condenses the browser view.

Visit the source page ↗

Check your own messaging

Start with your common page templates. If a pricing card, table or image loses meaning, the same issue could recur wherever you use it.

Clear source content gives you more control over your message. It does not guarantee an accurate answer, citation or recommendation.

Check a page