Research / Methodology / v1.0

AI recommendation visibility methodology

Every census edition published here is run under a numbered version of this method and says which one. This is version 1.0, in force from 18 September 2026. It exists so that two editions can be compared, and so that a reader can tell a method change from a result change.

The four things measured

These are separate quantities. They fail separately and they mean different things, so each edition reports them separately rather than combining them into one visibility score.

TermDefinition
MentionThe organization appears by name in the answer text of a run.
ReachThe number of runs, out of the edition's total, in which the organization is mentioned.
Question breadthThe number of distinct questions in the pack, out of the pack's total, under which the organization is mentioned at least once.
Position 1The number of runs in which the organization is the first named organization in the answer.

Reach and question breadth answer different questions and can disagree sharply. A firm named often under one question and a firm named occasionally under nine are in different positions even at the same reach.

Being cited as a source is a fifth quantity and is not the same as being mentioned. Where an edition can observe it, it is reported under its own name and never folded into reach.

Collection

Each edition runs a frozen prompt pack. The pack is committed to version control and its hash is recorded on every run record, so the strings can be shown not to have changed during or after collection. From version 1.1 onward the pack is committed and pushed on a different day from collection; version 1.0 does not require this, and editions run under 1.0 say so.

Every question is run the same number of times. Runs are logged out, with no conversation history, and each run sends one prompt string exactly as written with no system prompt.

The location parameter available to the instrument is recorded. Where the instrument has no entry below country level, the city is named in the question text instead and the edition says so, because that is a weaker geographic signal than a real local connection.

Where the endpoint does not report a model identifier, the edition states that it measures what that interface returned on that day rather than the behaviour of any named model. Response hashes are recorded and duplicates are reported, which rules out byte-for-byte caching and nothing more.

Extraction

Organization names are extracted from each stored response by a committed script. Every channel the extractor reads is named in the edition, and so is every channel it does not read.

Extraction is audited against a ground truth read by a reader who has not seen the extractor's output. Recall and precision are scored under a rule written and committed before the scorer runs. Ground truth is never produced by the thing being tested; where any part of it was, the measurement is void rather than weakened.

Recall is reported for the entity class the edition is about and for all named organizations, because the two differ substantially and the wider figure is the one that shows what the extractor does with regulators, portals and systems.

Entity resolution

Names as an engine writes them are not identities. An edition resolves each extracted name to an entity in a reference set, and the reference set is a government licence register wherever one exists for the market.

Branches and spelling variants merge onto one identity. Trade names that differ from public brands are bridged by hand, with a reason recorded per entry, and the bridge is published with the edition. Resolution decisions are made after the results arrive, which the edition states, because they cannot be made before the names are known.

Aggregate figures survive imperfect resolution better than individual ones do. An edition reports both the resolved and unresolved version of any aggregate that resolution moves, and treats a single firm's figures as more sensitive to resolution error than the edition median.

Scoring and uncertainty

Answer-set similarity is measured with the Jaccard index over the set of organizations named in each run, reported per question and pooled. Pairs where both runs name nobody are undefined and excluded, and the scorable pair count is stated.

Interval columns give descriptive 95% binomial intervals, calculated as though each run were an independent trial. Runs are clustered within questions and are not independent, so an edition labels these as binomial intervals, says how they were computed, and states that they are not confidence intervals for the engine's general behaviour.

Every rate is published with its numerator and denominator. Every subtraction a reader is asked to accept is closed from figures stated on the page.

What an edition must publish

An edition is not published until all of these exist.

  • A measurement stamp: collection date and timing, engine and interface, intermediary, question and run counts, location and language parameters, model identifier or a statement that none was reported, reference set with its date and hash, extraction figures with their scoring provenance, and the pack id, hash and commit.
  • A reproducibility bundle at a resolving URL, containing the raw responses as stored, the pack, the extraction and resolution tables, the matcher source, the ground truth, and the scoring rule and script. Its manifest is published outside the archive as plain text so hashes can be checked without downloading anything.
  • A limitations section stating the scope limits plainly, including limits that are also strengths.
  • A corrections log, permanent, carried forward into every later edition, recording what was wrong, why, what replaced it and how the replacement was produced. Entries for errors caught before publication belong there too.
  • A disclosure of the commercial interest behind the work.

Third-party reference data is linked and hashed rather than republished. Where the published file is derived from a source file, both hashes appear with the command that rebuilds one from the other.

Naming organizations

Tables are ordered by a stated quantity for readability, and the ordering is not a ranking. An edition says so, and accepts that no caveat survives being summarized by a third party.

Where a register distinguishes registered from active, the edition defines which it means and does not identify individual firms by status in prose. The published data carries the register's own fields.

Where a client of RealEstateSEO.ph appears in a table, the row is marked. The existence of clients is public, names are publishable, and deal specifics are not.

Version history

VersionDateChange
1.018 September 2026First published version. In force for census edition dubai-01.

Method changes are versioned and disclosed per edition. An edition run under a later version says which figures the change affects and whether it can still be compared with earlier editions.

Powered by GodMode Digital

RealEstateSEO.ph is the industry-specific application of the GodMode framework. We focus on search and AI visibility for property developers, brokerages, and agencies. For deep-tier AI integrations, agentic automation, and Fractional CTO leadership, visit our parent consultancy.

Visit GodMode.ph