A weekly GEO report should answer four questions:
- What did we ask? (prompt set)
- What did providers return? (evidence)
- Did mentions or citations change? (deltas)
- What do we explicitly not know? (limitations)
If your readout skips question four, you are marketing—not measuring.
This guide mirrors MentionPop's evidence-first posture. For product context see reports and methodology.
Report sections (conceptual map)
| Section | Purpose |
|---|---|
| Prompt inventory | Shows which questions were sampled |
| Mention summary | Brand string presence in parsed answers |
| Citation summary | Your URLs in source attribution |
| Competitor co-mentions | Context for relative narrative |
| Failures / exclusions | API errors removed from scoring |
| Week-over-week delta | Trend, not single-run drama |
Visual pipeline: /resources/evidence-pipeline-diagram.svg.
Deep method: sampled observation methodology.
Step 1: Validate the sample
Before interpreting numbers:
- Confirm prompt version matches last week unless changes were intentional
- Note provider scope — ChatGPT/Gemini on weekly GEO; Google modules on Studio
- Check run timestamps align with your reporting calendar
- Read exclusion notes for failed calls
Failed calls excluded ≠ brand miss. This policy prevents downtime from looking like a visibility crash.
Step 2: Split mentions and citations
Never merge into a single "visibility score."
| Metric | Ask |
|---|---|
| Mentions | Are we named in category prompts—not only branded ones? |
| Citations | Which URLs earned attribution—homepage vs methodology? |
| Both absent | Access issue, disambiguation issue, or prompt mismatch? |
Guide: citations vs mentions.
Step 3: Read deltas with variability in mind
Week-over-week swings may reflect AI output variability—not your Monday blog post.
Use:
- Four-week rolling averages for executive summaries
- Prompt-level drilldown for operator tactics
- Ship date annotations separate from charts
Week 30 delta notes:
- Prompt set v3 unchanged
- ChatGPT: mentions 12/20 (+1), citations 4/20 (0)
- Gemini: mentions 9/20 (-2), citations 3/20 (+1)
- 2 runs excluded (API timeout) — not scored
- No causation claims for homepage tweak on 7/28Step 4: Classify findings
| Finding type | Example response |
|---|---|
| Measurement insight | "Category prompt X never cites us—competitor Y dominates citations" |
| Technical blocker | "robots.txt disallows GPTBot on /guides/" — see crawler access |
| Entity issue | Brand collision with homonym — disambiguation |
| Noise | Single-week mention drop with stable citations—monitor |
Step 5: Present to leadership honestly
Do say:
- "Under our fixed sample, mentions rose from 55% to 60% of prompts."
- "Citations lag mentions on Google AI Overviews—Studio module."
- "Observations are provider-sampled, not universal consumer UI truth."
Do not say:
- "We rank #1 in ChatGPT."
- "GEO caused pipeline uplift" without attribution design.
- "Perplexity confirms…" unless noting Sonar is internal validation only.
Link evidence-first measurement in appendix slides.
Studio add-on: Google modules
When Studio is enabled, add separate tables for:
- AI Overviews — guide
- Featured snippets
- People Also Ask
Do not average Google modules with ChatGPT mention rates.
Cadence recommendations
| Audience | Cadence | Depth |
|---|---|---|
| Operators | Weekly | Prompt-level |
| Leadership | Monthly | Rolling trends + limitations |
| Board | Quarterly | Strategic prompts only |
Agencies: package readouts with GEO for agencies positioning—measurement separate from deliverable optimization work.
Annotating reports for async teams
Remote teams benefit from inline annotations on weekly exports:
| Annotation type | Example |
|---|---|
| Prompt change | "v4 added comparison prompt 2026-07-28" |
| Ship event | "Methodology updated—expect citation lag 1–2 weeks" |
| Known noise | "Gemini timeout cluster Jul 27–28; exclusions applied" |
| Qualitative | "ChatGPT answer framed us as 'screenshot tool'—disambiguation" |
Annotations preserve context when the GEO lead is out of office.
Escalation thresholds (suggested)
Define internal thresholds to avoid panic or complacency:
| Signal | Suggested response |
|---|---|
| Single-week mention drop ≤2 prompts | Monitor; check exclusions |
| Four-week mention decline ≥20% | Technical + entity audit |
| Citations flat, mentions up | Source quality initiative |
| Both flat after major launch | Prompt relevance review |
| Repeated API exclusions | Engineering ticket to provider errors |
Thresholds are internal guardrails, not product guarantees.
Board-safe summary paragraph (template)
"We sample {N} fixed prompts weekly on ChatGPT and Gemini via MentionPop. Mentions appeared in {X}/{N} prompts; citations in {Y}/{N}. Failed API calls are excluded from scoring. Observations are provider-sampled and may differ from any individual consumer session. We do not report AI market share or guaranteed placement."
Customize {N}, {X}, {Y}—never fabricate lift.
Cross-linking weekly GEO to content ops
When reports flag missing citations on methodology prompts:
- File content ticket with URL + prompt evidence
- Ship update to named URL only (avoid sitewide churn)
- Wait for next two weekly samples before declaring impact
- Link ticket ID in weekly annotation
This loop connects fixes culture to measured evidence.
Weekly reports compound value when teams treat them as structured evidence reviews—not as a gamified score to win each Friday.
Frequently asked questions
What providers appear in a standard weekly GEO report?+
Weekly recurring GEO includes provider-sampled ChatGPT and Gemini runs. Google AI Overviews, featured snippets, and PAA appear on Studio—not the base weekly GEO tier.
Why are some runs missing from scoring?+
Failed API calls are excluded rather than counted as brand misses, so error rates do not inflate or deflate visibility metrics.