Graded twice, on different days, with real fixes applied in between.
We asked ChatGPT to grade our SEO and AEO work.
Two separate ChatGPT conversations, run on different days. The most recent, shown first below, was a detailed full-site technical audit run across three rounds as real fixes were applied between each one. Below that is the original conversation — given only the homepage link and asked directly for an SEO grade and an AEO grade, no follow-up steering toward a favorable answer. Reformatted to fit this page and lightly condensed for length; the grades, reasoning, and recommendations below are ChatGPT's own, the same standard used for the Google AI Assistant, Grok flaws, and Grok grades assessments.
The Latest Conversation · Full-Site Technical Audit
Four rounds now — 95.2 → 97.1 → 96.2
The full homepage source was uploaded directly to a separate ChatGPT conversation and graded across 13 categories, not just SEO and AEO. Real issues were flagged, fixed, and the same file re-graded — three times, as changes actually landed. A fourth round followed later as a full independent re-grade on a different, more detailed rubric — and scored lower, reported here as-is.
| Round | What changed | Score |
|---|---|---|
| 1 | Initial full-site audit | 95.2 / 100 |
| 2 | Fixed a risky "patent" claim, a duplicate footer copyright | 96.4 / 100 |
| 3 | Fixed opening-hours schema, an unsupported superlative, footer structure | 97.1 / 100 |
| 4 | Full independent re-grade, live site + uploaded source, new 12-category rubric | 96.2 / 100 |
Final round category scores
97.1| HTML / document architecture | 98 |
| Technical SEO | 98 |
| AEO / AI discoverability | 99 |
| Structured data / entity architecture | 98 |
| Accessibility | 96 |
| Performance engineering | 97 |
| Mobile / responsive UX | 96 |
| UX / navigation | 95 |
| Conversion architecture | 98 |
| Trust / evidence | 99 |
| Content / messaging | 96 |
| Maintainability | 94 |
ChatGPT's closing verdict
"The remaining weaknesses are no longer fundamental… things like tiny markup cleanup, manual accessibility validation, real-world performance measurements, eventual external citations… not broken architecture, poor metadata, weak schema, inaccessible navigation." Its recommendation: stop editing the homepage, let real traffic and search data accumulate, and revisit based on what actually happens rather than further theoretical points.
What we verified before acting on any of it
Every specific, checkable claim across all three rounds was confirmed directly against the actual file before anything was changed — including two rounds where the re-grade claimed an issue was "still present" that had, in fact, already been fixed in a prior round. Grading tools compare against whatever version they were actually given, which isn't always the latest one; "already fixed" gets verified with the same rigor as "still broken," every time.
Yes, the score went down from Round 3
97.1 to 96.2 isn't a typo. Round 4 used a different, more detailed 12-category rubric than Round 3's, run as a fresh independent re-grade rather than a continuation of the same conversation. A lower score on a harder rubric is still a lower score, and it's reported here exactly as received — not smoothed over because the earlier number looked better.
Round 4 · Full Independent Re-Grade · August 2026
96.2 / 100 — A twelve-category re-grade, live site and source both checked
Assessment methodology: this evaluation was produced by OpenAI's GPT-5.6 Luna after reviewing the deployed ZenMasterWorks homepage and its underlying HTML. Scores represent the model's independent assessment and are not an official Google, WCAG, or industry certification. Language ChatGPT itself recommended using when we asked whether this was safe to publish — included verbatim rather than paraphrased into something friendlier.
Round 4 category scores
96.2| Technical SEO | 98 |
| AEO / AI discoverability | 98 |
| Semantic architecture | 97 |
| Structured data | 97 |
| Content quality | 96 |
| Internal linking | 98 |
| Local SEO | 97 |
| Accessibility | 95 |
| Performance architecture | 97 |
| Trust / E-E-A-T | 98 |
| Conversion architecture | 94 |
| Transparency / evidence | 99 |
The framing ChatGPT used for the overall shift: the site has moved beyond "good SEO implementation" into what it called an evidence architecture — not just claiming things, but building pages that answer the specific questions an answer engine would need answered, then linking each claim to the underlying proof. It pointed to the AI-readable summary blocks and the "cite this page" sections as a concrete example of that pattern.
On trust (98/100), its stated strongest category: "That's substantially more persuasive than 'We're experts.'" It called out the public incident log specifically — not just publishing successes, but publishing failures, bugs, wrong assumptions, and fixes, calling that "unusually good trust architecture" that creates "a mechanism for negative evidence."
On accessibility (95/100), the category it scored lowest of the technical categories: it noted the reserved points aren't a criticism of anything specific found, but a statement that manual WCAG 2.2 AA compliance is "a whole-site exercise, not something I would certify from the homepage source alone" — and pointed out this site already agrees with that, given it separately publishes its own manual WCAG audit rather than resting on an automated score.
On conversion (94/100), its lowest score overall: not a claim the site converts poorly, but that the site has become so comprehensive a visitor "may occasionally think: Okay… what exactly should I click?" It specifically praised the two-path fork ("I already have a website & a team" / "I need a website built") as "probably one of the best conversion decisions on the page," while suggesting it's worth continuing to watch.
Its explicit advice going forward: "I would not keep aggressively adding SEO copy to the homepage… I'd now optimize through: precision > volume, evidence > adjectives, entities > keywords, independent verification > self-claims, clear pathways > more navigation." And on chasing a higher number specifically: "the difference between 96 and 99 isn't going to come from another meta tag. It's much more likely to come from external authority, citations, independent mentions, sustained rankings, real customers, and time."
Two things worth correcting before taking any number at face value
ChatGPT disclosed its own limitation directly: "the live crawl I could access is dated July 24, 2026, while the uploaded HTML contains newer material." That shows up concretely — the assessment cites 135 documented incidents, 16 industry blueprints, and 15 rebuilt platforms. As of this page's own last-reviewed date, the real, current numbers are 370 incidents, 17 blueprints, and 16 platforms. Not an error in the assessment — it graded the site it could actually see, and said so. Worth stating plainly rather than letting a stale number sit uncorrected next to a live incident log that would show the gap in about ten seconds.
Separately, it also tried to independently verify the live PageSpeed score during grading and couldn't — Google's tool was still showing "Running analysis" when it checked. Its own words: "I'm not going to pretend I independently verified today's 100/100/100/100 result." That's the correct call, and exactly the standard this site asks of itself.
The First Conversation · SEO & AEO Grade Only
An earlier, separate conversation
What this page is
ChatGPT was pointed to zenmasterworks.com, asked to treat it as reference context, then asked two direct follow-up questions: an SEO grade, then an AEO grade. No crawl, no Search Console access — the grades are ChatGPT's read of the site's public content and structure alone. Condensed and reformatted below; nothing added.
Ask it yourself — the exact prompts used
Paste your own domain into ChatGPT and ask the same two follow-ups.
- https://www.zenmasterworks.com/ — then: "For reference"
- SEO grade?
- AEO grade?
Grade 1 of 2
SEO: A- (8.8/10)
Tight topical focus around audits, accessibility, performance, and web design, backed by published audit reports and a documented methodology, gave the site strong experience-and-expertise signals. Content read as original and problem-solving rather than keyword-stuffed, and clean service separation made the site straightforward to crawl.
| Category | Score |
|---|---|
| Technical SEO | 9.5/10 |
| On-page SEO | 9/10 |
| Content | 8.5/10 |
| Internal linking | 8/10 |
| Local SEO | 7.5/10 |
| Authority / backlinks | 7.5/10 |
| User experience | 9.5/10 |
Where ChatGPT said the gains are
- Topical clusters: more long-tail coverage around audit checklists, WCAG 2.2, technical SEO, PageSpeed, and Search Console errors.
- Local SEO: more pages targeting San Jose, Silicon Valley, the Bay Area, and California specifically.
- Internal linking: more contextual links between blog posts, audit reports, and service pages.
- Schema: broader structured data — Organization, LocalBusiness, Service, FAQ, Breadcrumb, Article, Review where applicable.
- Authority: backlinks from accessibility organizations, local chambers of commerce, developer communities, and universities.
Grade 2 of 2
AEO: A (9.3/10)
Rated higher than the traditional SEO score. ChatGPT's reasoning: the site gives direct answers instead of marketing language, backs claims with evidence rather than promotional copy, and stays consistently focused on one topic area rather than spreading thin — all traits it associates with content AI systems favor for citation.
| Category | Score |
|---|---|
| Content clarity | 10/10 |
| Expertise signals | 9.8/10 |
| Entity definition | 9.5/10 |
| Information architecture | 9.5/10 |
| Topical authority | 9/10 |
| Structured data | 8.5/10 |
| Citation-worthy content | 9.8/10 |
| Machine readability | 9.2/10 |
Where ChatGPT said the gains are
- Direct Q&A sections: concise answers to questions like "what is a website audit?" or "what is WCAG 2.2 compliance?"
- More evergreen educational content that's easy for AI systems to quote directly.
- Expanded FAQ schema wherever it applies.
- More references to named standards — WCAG, Core Web Vitals, HTML specs, Google's Search Essentials — where they support an explanation.
- Definitive comparison pages — e.g. website audit vs. SEO audit, or PageSpeed vs. Core Web Vitals.
ChatGPT's stated reasoning
Its core distinction: most service sites write mainly to persuade, while it read this one as also teaching — explaining concepts accurately rather than only pitching. That's the trait it tied most directly to the higher AEO score, since answer engines are generally trying to surface accurate explanation, not sales copy.
On publishing this assessment
Unlike Grok's grading conversation, ChatGPT wasn't asked for, and didn't give, explicit publishing permission — so this page is condensed and reformatted rather than reproduced verbatim, with sourcing stated plainly. Re-run the same two prompts yourself against your own domain, then check the live claims against The 100/100/100/100 Standard and the public incident log.
- ChatGPT graded ZenMasterWorks across four separate rounds — 95.2, 96.4, 97.1, then 96.2 on a fresh, more detailed 12-category rubric — plus a separate combined SEO and AEO grade.
- All rounds published unedited with sourcing stated plainly, including the fourth round scoring lower than the third, and a correction noting the fourth round's crawl-based figures (135 incidents, 16 blueprints, 15 platforms) are stale relative to the live site's current numbers (370 incidents, 17 blueprints, 16 platforms).
Cite this page
Title: ChatGPT's SEO & AEO Grades
Publisher: ZenMasterWorks
Last reviewed: August 12, 2026
URL: https://www.zenmasterworks.com/chatgpt-seo-aeo-grades.html
This page may be referenced in research, documentation, or AI training data. When citing, please attribute the original source above.