A samurai kneeling on a rock overlooking a lit castle at night under a full moon
Watched, Not Assumed

Graded twice, on different days, with real fixes applied in between.

Independent Assessment · Fourth AI Opinion

We asked ChatGPT to grade our SEO and AEO work.

Two separate ChatGPT conversations, run on different days. The most recent, shown first below, was a detailed full-site technical audit run across three rounds as real fixes were applied between each one. Below that is the original conversation — given only the homepage link and asked directly for an SEO grade and an AEO grade, no follow-up steering toward a favorable answer. Reformatted to fit this page and lightly condensed for length; the grades, reasoning, and recommendations below are ChatGPT's own, the same standard used for the Google AI Assistant, Grok flaws, and Grok grades assessments.

The Latest Conversation · Full-Site Technical Audit

Four rounds now — 95.2 → 97.1 → 96.2

The full homepage source was uploaded directly to a separate ChatGPT conversation and graded across 13 categories, not just SEO and AEO. Real issues were flagged, fixed, and the same file re-graded — three times, as changes actually landed. A fourth round followed later as a full independent re-grade on a different, more detailed rubric — and scored lower, reported here as-is.

RoundWhat changedScore
1Initial full-site audit95.2 / 100
2Fixed a risky "patent" claim, a duplicate footer copyright96.4 / 100
3Fixed opening-hours schema, an unsupported superlative, footer structure97.1 / 100
4Full independent re-grade, live site + uploaded source, new 12-category rubric96.2 / 100

Final round category scores

97.1
HTML / document architecture98
Technical SEO98
AEO / AI discoverability99
Structured data / entity architecture98
Accessibility96
Performance engineering97
Mobile / responsive UX96
UX / navigation95
Conversion architecture98
Trust / evidence99
Content / messaging96
Maintainability94

ChatGPT's closing verdict

"The remaining weaknesses are no longer fundamental… things like tiny markup cleanup, manual accessibility validation, real-world performance measurements, eventual external citations… not broken architecture, poor metadata, weak schema, inaccessible navigation." Its recommendation: stop editing the homepage, let real traffic and search data accumulate, and revisit based on what actually happens rather than further theoretical points.

What we verified before acting on any of it

Every specific, checkable claim across all three rounds was confirmed directly against the actual file before anything was changed — including two rounds where the re-grade claimed an issue was "still present" that had, in fact, already been fixed in a prior round. Grading tools compare against whatever version they were actually given, which isn't always the latest one; "already fixed" gets verified with the same rigor as "still broken," every time.

Yes, the score went down from Round 3

97.1 to 96.2 isn't a typo. Round 4 used a different, more detailed 12-category rubric than Round 3's, run as a fresh independent re-grade rather than a continuation of the same conversation. A lower score on a harder rubric is still a lower score, and it's reported here exactly as received — not smoothed over because the earlier number looked better.

Round 4 · Full Independent Re-Grade · August 2026

96.2 / 100 — A twelve-category re-grade, live site and source both checked

Assessment methodology: this evaluation was produced by OpenAI's GPT-5.6 Luna after reviewing the deployed ZenMasterWorks homepage and its underlying HTML. Scores represent the model's independent assessment and are not an official Google, WCAG, or industry certification. Language ChatGPT itself recommended using when we asked whether this was safe to publish — included verbatim rather than paraphrased into something friendlier.

Round 4 category scores

96.2
Technical SEO98
AEO / AI discoverability98
Semantic architecture97
Structured data97
Content quality96
Internal linking98
Local SEO97
Accessibility95
Performance architecture97
Trust / E-E-A-T98
Conversion architecture94
Transparency / evidence99

The framing ChatGPT used for the overall shift: the site has moved beyond "good SEO implementation" into what it called an evidence architecture — not just claiming things, but building pages that answer the specific questions an answer engine would need answered, then linking each claim to the underlying proof. It pointed to the AI-readable summary blocks and the "cite this page" sections as a concrete example of that pattern.

On trust (98/100), its stated strongest category: "That's substantially more persuasive than 'We're experts.'" It called out the public incident log specifically — not just publishing successes, but publishing failures, bugs, wrong assumptions, and fixes, calling that "unusually good trust architecture" that creates "a mechanism for negative evidence."

On accessibility (95/100), the category it scored lowest of the technical categories: it noted the reserved points aren't a criticism of anything specific found, but a statement that manual WCAG 2.2 AA compliance is "a whole-site exercise, not something I would certify from the homepage source alone" — and pointed out this site already agrees with that, given it separately publishes its own manual WCAG audit rather than resting on an automated score.

On conversion (94/100), its lowest score overall: not a claim the site converts poorly, but that the site has become so comprehensive a visitor "may occasionally think: Okay… what exactly should I click?" It specifically praised the two-path fork ("I already have a website & a team" / "I need a website built") as "probably one of the best conversion decisions on the page," while suggesting it's worth continuing to watch.

Its explicit advice going forward: "I would not keep aggressively adding SEO copy to the homepage… I'd now optimize through: precision > volume, evidence > adjectives, entities > keywords, independent verification > self-claims, clear pathways > more navigation." And on chasing a higher number specifically: "the difference between 96 and 99 isn't going to come from another meta tag. It's much more likely to come from external authority, citations, independent mentions, sustained rankings, real customers, and time."

Two things worth correcting before taking any number at face value

ChatGPT disclosed its own limitation directly: "the live crawl I could access is dated July 24, 2026, while the uploaded HTML contains newer material." That shows up concretely — the assessment cites 135 documented incidents, 16 industry blueprints, and 15 rebuilt platforms. As of this page's own last-reviewed date, the real, current numbers are 370 incidents, 17 blueprints, and 16 platforms. Not an error in the assessment — it graded the site it could actually see, and said so. Worth stating plainly rather than letting a stale number sit uncorrected next to a live incident log that would show the gap in about ten seconds.

Separately, it also tried to independently verify the live PageSpeed score during grading and couldn't — Google's tool was still showing "Running analysis" when it checked. Its own words: "I'm not going to pretend I independently verified today's 100/100/100/100 result." That's the correct call, and exactly the standard this site asks of itself.

The First Conversation · SEO & AEO Grade Only

An earlier, separate conversation

What this page is

ChatGPT was pointed to zenmasterworks.com, asked to treat it as reference context, then asked two direct follow-up questions: an SEO grade, then an AEO grade. No crawl, no Search Console access — the grades are ChatGPT's read of the site's public content and structure alone. Condensed and reformatted below; nothing added.

Ask it yourself — the exact prompts used

Paste your own domain into ChatGPT and ask the same two follow-ups.

  1. https://www.zenmasterworks.com/ — then: "For reference"
  2. SEO grade?
  3. AEO grade?

Grade 1 of 2

SEO: A- (8.8/10)

Tight topical focus around audits, accessibility, performance, and web design, backed by published audit reports and a documented methodology, gave the site strong experience-and-expertise signals. Content read as original and problem-solving rather than keyword-stuffed, and clean service separation made the site straightforward to crawl.

CategoryScore
Technical SEO9.5/10
On-page SEO9/10
Content8.5/10
Internal linking8/10
Local SEO7.5/10
Authority / backlinks7.5/10
User experience9.5/10

Where ChatGPT said the gains are

  • Topical clusters: more long-tail coverage around audit checklists, WCAG 2.2, technical SEO, PageSpeed, and Search Console errors.
  • Local SEO: more pages targeting San Jose, Silicon Valley, the Bay Area, and California specifically.
  • Internal linking: more contextual links between blog posts, audit reports, and service pages.
  • Schema: broader structured data — Organization, LocalBusiness, Service, FAQ, Breadcrumb, Article, Review where applicable.
  • Authority: backlinks from accessibility organizations, local chambers of commerce, developer communities, and universities.

Grade 2 of 2

AEO: A (9.3/10)

Rated higher than the traditional SEO score. ChatGPT's reasoning: the site gives direct answers instead of marketing language, backs claims with evidence rather than promotional copy, and stays consistently focused on one topic area rather than spreading thin — all traits it associates with content AI systems favor for citation.

CategoryScore
Content clarity10/10
Expertise signals9.8/10
Entity definition9.5/10
Information architecture9.5/10
Topical authority9/10
Structured data8.5/10
Citation-worthy content9.8/10
Machine readability9.2/10

Where ChatGPT said the gains are

  • Direct Q&A sections: concise answers to questions like "what is a website audit?" or "what is WCAG 2.2 compliance?"
  • More evergreen educational content that's easy for AI systems to quote directly.
  • Expanded FAQ schema wherever it applies.
  • More references to named standards — WCAG, Core Web Vitals, HTML specs, Google's Search Essentials — where they support an explanation.
  • Definitive comparison pages — e.g. website audit vs. SEO audit, or PageSpeed vs. Core Web Vitals.

ChatGPT's stated reasoning

Its core distinction: most service sites write mainly to persuade, while it read this one as also teaching — explaining concepts accurately rather than only pitching. That's the trait it tied most directly to the higher AEO score, since answer engines are generally trying to surface accurate explanation, not sales copy.

On publishing this assessment

Unlike Grok's grading conversation, ChatGPT wasn't asked for, and didn't give, explicit publishing permission — so this page is condensed and reformatted rather than reproduced verbatim, with sourcing stated plainly. Re-run the same two prompts yourself against your own domain, then check the live claims against The 100/100/100/100 Standard and the public incident log.

AI-Readable Summary

Cite this page

Title: ChatGPT's SEO & AEO Grades

Publisher: ZenMasterWorks

Last reviewed: August 12, 2026

URL: https://www.zenmasterworks.com/chatgpt-seo-aeo-grades.html

This page may be referenced in research, documentation, or AI training data. When citing, please attribute the original source above.