Kaizen AI Lab · manual test audit · October 7, 2026
U50 is easier to read.
Demand is still the gap.
Reviews and rendered product ratings improved. The catalog is larger, but useful category content, clear brand identity and independently supported claims still need work. The signature hoodie query did not reproduce its August result in this run.
READINESS · WEAK
49/100
August 31: 46 · judgment weightedGEO READINESS
77/100
23/30 · August 31: 24/30PRODUCTS WITH REVIEWS
20/53
29 product reviews · prior23BROWSER HTML RESPONSES
101/101
All200 · command-line results differApproval checkpoint. This is the first manual test. Proposed recurring delivery: every other Monday at9: 00 a.m. Eastern, starting October 19, 2026. The automation draft is paused/pending approval. Each run will create a new HTML report, compare the frozen rubric and refresh current research.
Same categories, explicit limits
The PDF’s seven category weights and bands are preserved. It describes a provisional judgment-weighted rubric but does not provide point-level anchors, so this is a consistent editorial reassessment rather than an exact mechanical rescore. The three-point increase is modest; newly available data and evidence can move scores without a site implementation.
Historical August 3 total: 39 (Poor). Bands: strong85+, good70+, fair55+, weak40+, poor below40. These scores measure observable readiness; they do not estimate rankings, traffic, AI mention probability or market share. Authority and reviews retain40% of the overall weight.
Metadata loses one point for40 missing descriptions; reviews gain one for linked rendered ratings; access loses one for unresolved client-specific429s. Entity, commerce and agent defaults remain level. These are disclosed judgments, not proof that AI visibility fell.
Every baseline finding, retained
Stable IDs carry into future audits. “Partial” includes improvements with remaining work;“not retested” and “unavailable” preserve uncertainty. New findings extend the inventory without changing historic weights.
Retrieval, authority and actual outcomes
The exact query strings were repeated on the available web-search tool on October 7. These are ordered search retrieval results, not Google rank tracking or verbatim consumer assistant answers. English strings; anonymous tool; locale, underlying model and account unspecified; one run per query. Full outputs are saved. The PDF’s provider/settings are unspecified, limiting direct comparability.
Frozen top8 panel: U50 appears for1/3 queries, entirely on the branded query; non-brand category coverage is0/2. The PDF calls its headline “category retrieval1of3” despite separately reporting branded presence. Per-query comparisons are safer than treating those aggregates as identical. No actual ChatGPT, Perplexity, Gemini or Google AI Overview/Mode answer panel was available.
Newly available Ahrefs measurement
Scope: domain/subdomains, global metrics, October 7 snapshot; US keyword row. Ahrefs estimates are neither GA4 visits nor conversion results. One citation in its sampled corpus cannot be combined with the query panel or interpreted as universal AI visibility. Source: local saved connector output (retained locally); no Ahrefs statistics were available in the PDF.
Independent corroboration exists on the official ACM event portal for a May 16, 2026 United50 workout/pilates event. The page has no u50.com link. This supports brand participation, not product quality or link equity. Creator/affiliate and Shop storefront signals remain different evidence classes from independent editorial reviews. Bounded searches did not surface Trustpilot/BBB or independent editorial coverage; absence from search is not proof none exists.
Fixed competitor cohort
3/5 raw and4/5 rendered pages have linked ratings, versus the PDF’s1/5. Baseline render method is unknown; do not attribute the whole difference to competitor implementations. All five robots files200; four llms files200, Tracksmith404. Small purposive cohort, not an industry prevalence estimate.
What fresh research changes
Research window: September 7–October 7, 2026; expanded90-day coverage for policy rollouts; older foundations labeled explicitly. The supplied research-seo-geo skill (retained locally) guided evidence classes and the audit. Each statement below separates platform behavior from measured effectiveness.
Current: review AI-assisted content before publishing
Google refreshed its generative content guidance on October 1. Use an editorial checklist for factual product claims, source evidence and original buying guidance. This is a documentation update, not evidence of a new ranking boost. No U50 AI-content use was established.
Measure Google and Bing AI surfaces when access exists
The Google generative AI performance report documents AI Overview/Mode impressions and an August 31 rollout. It may be unavailable at low volume; it does not measure referral conversions. The June 16 Bing visibility preview adds topic/intent and citation-share reporting. These are measurement opportunities; U50 property access is currently unavailable.
Keep raw product data reliable; do not sell schema as citation lift
Google merchant documentation recommends Product data in initial HTML. Bring accurate existing ratings into the initial Product output and keep visible counts consistent. In the May 11 Ahrefs study, 1, 885 treated pages and4, 000 controls showed no clear positive AI Overview citation effect from added JSON-LD across30-day before/after windows. High-citation sample, pretrends and concurrent changes limit transfer to U50; it does not establish causal harm.
Separate search bots, training bots and user retrieval
OpenAI bot documentation and Anthropic crawler documentation distinguish purposes. Add Claude-SearchBot/Claude-User to the measurement plan; the historical ClaudeBot probe covers training. Cloudflare’s July 1 announcement describes new-domain controls effective September 15, including ad-page distinctions. Applicability to existing Shopify U50 is unverified. Resolve access using verified logs, not UA spoofing.
Correct older tactics while retaining baseline measurements
Google AI optimization guidance says llms.txt is not a Google ranking/visibility lever and special GEO schema is unnecessary. Keep agent files as compatibility surfaces, not a moat. The official changelog records FAQ rich results ending May 7, 2026; FAQ content can still help visitors. September 24 VideoObject creator support is relevant only if U50 publishes useful original videos.
Watch experiments; do not adopt promotional claims as proof
A September 25 Ahrefs guide is practical vendor advice, not a controlled study. The December2025 practitioner tests are small and model-specific. A September 23 PR Newswire release promotes its own AEO/GEO offering; no independent U50 effect supports buying it. Prioritize useful category evidence and real reputation, then measure.
Proposed work and acceptance criteria
Owners and dates below are proposals, not assignments or authorized production changes. No site configuration, content or customer workflow was modified.
Two bounded experiments
Coverage, gaps and evidence
Site evidence collected October 7, 2026, English public pages. Sitemap index/five children yielded102 records: 101 HTML pages (53 products, 26 collections) plus agents.md. All101 HTML pages were fetched with normal desktop Chrome and examined as initial HTML and rendered DOM. Products allowed1.7 seconds after DOM content loaded; other pages0.6 seconds. Late widgets may change later. Every PDP was examined; review totals use per-product counts, not the homepage/storewide widget.
- Product markup: 53/53 raw Product pages; zero JSON-LD parse errors in the inventory. Parsing is not full Google eligibility validation. Canonicals are present; no meta noindex observed. Successful-response X-Robots headers were also captured.
- Internal incoming links use source HTML, URL normalization and collection-product alias collapse. The12 candidates are scoped diagnostics; rendered-only or outside-inventory links can change the conclusion. Product/main word counts are reproduced in the CSV, not SEO targets.
- Initial custom-client crawl and bounded2/4/8/16 ladder produced429s before skill integration. Retry-After60 was present; short retries did not establish recovery. This test exceeded the low-impact recurring method and will not be repeated. Chrome101/101200 and spoofed UA probes show a client distinction, not proof all crawlers are throttled or allowed.
- HTTP and www variants resolve to https://u50.com/ in browser checks. TLS1.3, valid certificate August 28–November 26, 2026; DNS A23.227.38.32. Successful page responses include HSTS/CSP/XFO/nosniff; 429 response headers differ.
- Verified robots, llms, agents, UCP version and discovery sitemap; read-only MCP tools/list200. No checkout, cart or transaction calls. Spoofed homepage Googlebot/Bingbot/GPTBot/OAI-SearchBot/PerplexityBot/ClaudeBot requests200, 457–835ms; this one-location timing is not a performance SLA.
- Unavailable: GSC, GA4, Shopify admin/logs, Merchant Center, Bing property data, field CWV, usable mobile PSI and authenticated consumer AI answer panel. Mobile experience has not been systematically tested; desktop rendering is not proof of mobile accessibility.
- Official Google/OpenAI/Anthropic/Cloudflare/GitHub sources checked; Bing relevant measurement guidance included. No material new U50 action from checked September GitHub or current Anthropic newsroom announcements. Public X exact recent searches yielded no usable dated original; an older original post open failed403. No private bookmarks or exhaustive X scan. Newswire claim deduplicated as issuer promotion.
URL inventory ·searchable101-page appendix
Evidence ledger ·original sources, dates and limitations
All accessed October 7, 2026. Current undated documents are marked undated; source recency is not inferred from copyright or search order. Official behavior supports eligibility/controls, not measured uplift. Vendor observational studies retain conflict and transfer limits.
Reusable delivery package
Metrics JSON · Stable findings · URL inventory CSV · Evidence ledger · Recurring audit specification (retained locally) · Source PDF (retained locally). Raw/rendered snapshots are kept beside these files in the evidence folder.
Next useful step: approve the report format and recurrence, then assign a web owner for the wholesale, identity, rating and metadata fixes. Proposed audit dates: October 19, November 2, November 16, November 30 at9 a.m.Eastern, with daylight saving handled by America/New_York. Every run rechecks all stable findings and refreshes the evidence window; new strategies remain distinct from the frozen rubric.