Research · AI Visibility Measurement

PSOS: the score for being recommended by AI.

An open, auditable 0–100 KPI that measures how strongly your brand appears inside the answers AI assistants give — across ChatGPT, Gemini, Claude, Perplexity and Grok.

Research: AIVO Standard™ Author: Paul Sheals Published Sept 2025 ≈ 8 min read

Discovery is moving from the search box to the AI assistant. When a customer asks ChatGPT or Gemini for a recommendation, there is no page two, no list of ten blue links, no click-through rate to optimise. Your brand is either named in the answer — or it is invisible in the exact moment a decision is made.

That single shift breaks the tools we have relied on for twenty years. Impressions, rankings, backlinks and domain authority were all built for search engines. They say nothing about whether an AI model recommends you, omits you, or quietly hands the moment to a competitor.

“In AI discovery, visibility is binary. A brand is either in the assistant’s answer, or it does not exist in that decision.”

PSOS — the Prompt-Space Occupancy Score — was created to close that gap: the first open, standardised way to measure how much space your brand occupies in the AI recommendation layer, with the audit trail and statistical rigour a boardroom can trust.

The governance blind spot

Three risks you can’t see without a measure

Vendor dashboards give partial snapshots, but they lack transparency, auditability and statistical confidence. That leaves leadership exposed on three fronts:

Governance blind spots. Boards cannot see or verify how their brand is being represented in AI recommendations.
Strategic misallocation. Teams keep over-investing in traditional SEO while under-investing in AI visibility.
Reputation risk. Misinformation, outdated data, or a competitor displacing you inside the answer goes undetected.

The concept

A single, board-grade number

PSOS distils the messy reality of AI recommendations into one composite score from 0 to 100. Just as PageRank gave the early web a shared measure of authority, and GAAP gave finance a shared standard for reporting, PSOS gives AI visibility a measure that is open, versioned and governed — designed to withstand scrutiny from executives, regulators and investors.

Crucially, it is not another black-box vendor metric. Every prompt set, weighting and formula is documented and versioned, so any score can be reproduced and defended.

How it’s built

Five dimensions of AI visibility

A brand’s presence in AI answers is more than a single mention. PSOS measures five distinct, auditable dimensions and combines them into the composite score.

01

Breadth

The share of relevant prompts where your brand appears — weighted by prominence, so being named first counts for more than being named last.

02

Depth

Whether that visibility persists over time. Measured across 30-, 60- and 90-day windows so a score reflects durable recall, not a campaign spike.

03

Resilience

Consistency across engines — ChatGPT, Gemini, Claude, Perplexity, Grok and more — weighted by market share, so you’re not reliant on a single model.

04

Sentiment

The tone of how you’re mentioned, from NLP polarity analysis, applied as a ±10% overlay. Presence matters; the quality of that presence matters too.

05

Decay

How fast visibility erodes without reinforcement — revealing which brands have naturally durable recall and how much investment is needed to hold position.

How it’s measured

Two modes, one standard

PSOS is measured in two complementary ways, so the same standard fits both a global enterprise and a local business.

Enterprise mode · prompt-validated

Test the models directly

  • Curated clusters of 30–75 real discovery prompts per market, versioned and frozen for auditability.
  • Run across the major engines, with each prompt repeated (3+ replicates) to cancel out AI randomness.
  • Every response parsed for mentions, position, sentiment and refusals — logged with timestamps and provenance.
  • Produces a composite with 95% confidence intervals and the five sub-scores reported separately.
SME / local mode · evidence-based

Measure the signals models read

  • For smaller organisations, uses proxy signals: citation diversity, review volume, structured-data freshness and offline credibility.
  • A Confidence Index qualifies how reliable the result is when direct prompt-testing isn’t feasible.
  • A cost-effective benchmark that still maps to the same 0–100 scale.
Enterprise scoring PSOSₑₙₜ = (Breadth × Depth × Resilience) + Sentiment overlayDecay adjustment
SME / local scoring PSOSₛₘₑ = 0.25·Breadth + 0.20·Depth + 0.20·Resilience + 0.15·Sentiment + 0.10·Freshness + 0.15·Offline credibility

Local scores are read against three plain-English bands:

Below 40
Fragile
40 – 69
Moderate
70 and above
Strong
Why it’s different

Built for governance, not just dashboards

Several tools monitor prompts or track brands. They fall short where it counts for a board: transparency, auditability and defensibility.

DimensionTypical vendor dashboardPSOS™
MethodologyProprietary black boxOpen, versioned, auditable
Audit trailLimited or noneFull logs, frozen prompt sets, provenance
Statistical confidenceNot providedConfidence intervals & error bands
Platform coverageSingle-engine focusAggregated across ChatGPT, Gemini, Claude and more
Governance readinessTactical reportingBoard-ready KPI, attributable to ROI

Governance is designed in, not bolted on: prompt clusters are frozen once published, formulas are versioned (e.g. psos_v1.0.0), engine weightings are provenance-logged, and every affirmative mention is checked against a Citable Reference Unit so hallucinated citations are filtered out rather than counted.

From visibility to revenue

The attribution layer

A score only matters if it moves the business. PSOS includes an attribution layer that links visibility changes to outcomes — leads, qualified opportunities and revenue — using difference-in-differences analysis around specific interventions, such as deploying structured data or launching a review campaign. That turns AI visibility from a marketing abstraction into a KPI a board can allocate budget against.

The evidence

Validated at scale

PSOS was first applied in the AIVO 100™ Global Index — a study measuring AI visibility across more than 100,000 prompts, eight sectors and six leading AI platforms.

100,000+
prompts across 8 sectors & 6 AI platforms in the inaugural AIVO 100™ study
94
the top PSOS score (Apple iPhone) — 89% breadth, under 5% decay over 90 days
r = 0.89
test–retest reliability across 30-day intervals — repeatable, not a single snapshot
15–25%
faster PSOS growth for challenger brands (Oatly, Duolingo, Notion) vs incumbents

The score also tracks commercial reality. Early evidence shows PSOS correlating with aided brand recall (r = 0.67), organic referral traffic (r = 0.52) and marketing-qualified leads (r = 0.43) — positioning it as a leading indicator of performance, not just a visibility gauge. Sector patterns were revealing too: technology brands led on breadth but were volatile, while some healthcare brands carried misinformation risk, with outdated data surfacing in a meaningful share of prompts.

Stated openly

Known limitations

Part of being governance-grade is being honest about boundaries. Version 1.0 states its own:

Boundary conditions

  • English-language dominance can skew global rankings toward US/UK brands; regional weighting is planned.
  • Current scope excludes voice assistants (Siri, Alexa) and regional LLMs (Baidu, Yandex) due to access limits.
  • The ±10% sentiment band may understate impact in reputation-sensitive sectors like healthcare and financial services.
  • Standard 90-day windows can under-represent long-term equity in stable categories such as luxury and industrial B2B.
The full methodology

Read the complete PSOS Methodology

This article is an overview. The full v1.0 methodology — governance framework, both scoring modes, the six-tool architecture, validation evidence and certification pathways — is published open-access on Zenodo with a permanent DOI.

Citation: Sheals, P. (2025). PSOS™ Methodology v1.0 — Prompt-Space Occupancy Score (AI Visibility KPI). AIVO Standard. Zenodo. https://doi.org/10.5281/zenodo.17081529 · Licensed CC-BY-4.0.