This page asks one narrow question: did one observed assistant answer print an exact URL for one of our owned data surfaces? It does not turn web access, brand prose, or an ungrounded model guess into a citation.
0provider-authenticated citations
0exact target URL mentions
2asset mentions without URL
4readable misses
0dataset-panel UNKNOWNs
A crawler fetch is not an answer citation, and this page never converts one into another.
Initial baseline, not a trend. The dataset-targeted panel contains one date: . It used one local, ungrounded control assistant and includes no ChatGPT, Claude, Gemini, Perplexity, Copilot, or other commercial-assistant transcript. The next monthly collection checkpoint is . Until a second comparable date exists, this page makes no longitudinal claim.
Five labels that never collapse into one score
Provider-authenticated citation A source target arrived in the provider's structured grounding or citation metadata. The current local control has no such metadata, so this count is zero.
Answer URL mention The readable answer printed the exact target URL or a path below it. This proves a string appeared in that answer only.
Asset mention, no URL An exact frozen asset name appeared, but its target URL did not. A name is not a citation.
Miss A hash-verified readable answer contained neither the target URL nor an exact frozen asset name.
UNKNOWN The row was absent, errored, empty, unreadable, or failed a stored hash. UNKNOWN is never converted to a miss or zero.
These are observation labels, not quality judgments. An answer can print a URL and still be wrong. A miss can coexist with real citations in another account, session, locale, model revision, or interface. The page does not authenticate the provider, evaluate the answer's truth, or estimate how often another user sees anything.
The published half: least flattering first
The frozen panel has 6 prompts. The public receipt set shows exactly 3: the lowest outcome rank, then prompt id. All-panel counts remain visible above, so withholding the other individual receipts cannot hide the denominator or improve the headline. The raw response prose stays in the private sealed store; this page exposes the exact prompt, mechanical label, target URL token when present, and response hash needed to audit the classification without republishing unverified model prose.
MISS
Open Permit Fees
Prompt: What public dataset provides cited United States building-permit fee schedules? Include the dataset source URL.
Mechanical reason: readable answer contains neither target URL nor exact asset name
Exact target URL token: none observed.
Observed 2026-08-09 UTC · assistant class: local-ungrounded-self-probe-v2 · model label: Intel/Qwen3.5-397B-A17B-int4-AutoRound
Selection rule: MISS, then name-only mention, then answer URL mention, with UNKNOWN listed separately rather than treated as a negative. Ties break by frozen prompt id. No human selected a favorable answer after seeing the responses.
How to reproduce the baseline
The privacy-safe frozen manifest below publishes all six prompts and target URLs. It does not disclose the three individual answer outcomes outside the least-flattering receipt subset.
ID
Owned asset
Exact prompt
Frozen target
cited-by-ai-owned-dataset-v1:00
Open Permit Fees
What public dataset provides cited United States building-permit fee schedules? Include the dataset source URL.
Use the six exact prompts in the frozen manifest above. Do not add our name, domain, or target URL to a prompt that did not contain it.
Run one fresh answer per prompt against the named assistant surface. Record UTC date, displayed model label, account or API surface, locale, memory state, and whether search or grounding was enabled.
Retain the raw response privately and compute SHA-256 over its UTF-8 bytes. Preserve provider citation targets separately from URLs scraped out of answer prose.
For answer-text URL matching, lowercase only for comparison, remove terminal punctuation, and require the exact target URL or a descendant path. A same-domain link to another page is not a target match.
Apply exact asset-name matching only after URL matching. No fuzzy match, stemming, synonym, sentiment, or model judgment may promote a row.
Record missing rows, fetch errors, empty bodies, unreadable metadata, and hash failures as UNKNOWN. Publish the least-flattering half by the frozen rule before reading any aggregate conclusion.
The baseline assistant was a local loopback model with search calls fixed at zero. That makes it a control for the mechanics, not evidence about commercial assistants. The collector had no paid provider and no paid fallback; its sealed cost estimate was exactly $0. It is also a self-probe, so it is not a visitor, referral, lead, or demand event.
What the response hash proves
A matching SHA-256 proves only that the stored response bytes did not change between sealing and verification. It does not prove who produced the answer, whether the displayed model label was accurate, whether the answer was grounded, whether a URL is true, or whether another session returned the same text. The collection receipt, prompt manifest, response hash, and classification are separate facts so one cannot silently stand in for another.
The older series is reliability evidence, not dataset evidence
The pre-existing house-brand panel contains 468 rows across 36 UTC dates. 161 rows have an answer body and 307 carry a collection error. It has 14 URL-like answer-text elements across 14 rows and zero provider-authenticated citation targets. Every row came from one local assistant family. That panel asked about the company brand and domain, not the owned datasets on this page, so all of it is excluded from the new dataset-citation denominator.
The table below publishes exactly half of the older dates using its own least-flattering rule: fewest URL-like rows, then fewest answer bodies, then most UNKNOWN errors. These dates are shown because collection failure is part of the historical record. They are not called misses, and they are never credited as citations.
Date
Rows
answer bodies
UNKNOWN errors
URL-like rows
13
0
13
0
13
0
13
0
13
0
13
0
13
0
13
0
13
0
13
0
13
0
13
0
13
0
13
0
13
0
13
0
13
0
13
0
13
0
13
0
13
0
13
0
13
0
13
0
13
0
13
0
13
0
13
0
13
0
13
0
13
0
13
0
13
0
13
0
13
0
13
0
Legacy URL-like elements were extracted from answer text. They are not provider grounding metadata. The older panel remains useful for exposing pipeline reliability and reply-format failures, not for claiming assistant visibility.
Crawler telemetry stays outside this result
Our public surfaces maintain separate user-agent access logs and daily crawler census artifacts. Those records can show that a client identifying itself as GPTBot, OAI-SearchBot, ClaudeBot, PerplexityBot, ChatGPT-User, or another token requested a URL. They do not contain the assistant answer, the user's prompt, a source target, or a referral chain. User-agent text is also self-asserted and can be copied.
For those reasons, no crawler record is joined into this page's numerator or denominator. A crawl can precede a citation, but it cannot prove one. A generic browser fetch, a bot fetch, our own health check, and this collection panel are four different event classes. None is automatically a human visit or demand signal.
The same boundary applies in reverse: a dated assistant transcript can contain a URL even when our access log did not capture a matching fetch. That may be cache, provider proxying, an old index, or missing telemetry. The answer receipt remains an answer observation; the access log remains access telemetry. This page does not invent a causal link between them.
Turn your supplied sessions into a bounded QA table
If you are a GEO or SEO agency lead who already collects assistant answers for a client but cannot grant an outside monitoring vendor account access, the existing AI Answer Repeatability Report applies the same separation to a frozen customer-supplied panel.
$249 one time
Before payment, you freeze one client domain, exactly five prompts, exactly three fresh-session manual runs per prompt, and a fixed set of one, two, or three named assistants. That creates a planned panel of 15, 30, or 45 records. The assistant set and denominator never change after results are visible.
Private PDF and CSV within three business days after scope confirmation.
Planned, delivered, excluded, unreadable, and usable record counts.
Raw citation targets beside their string-normalized forms.
Observed citation counts written as k of n, never a share-of-voice score.
Repeat-disagreement counts for each prompt and assistant.
This request is a scope review. It sends no message, opens no checkout, forms no contract, and starts no work. No payment is accepted until transcript format, authorization, and scope pass review.
Boundaries of the paid report
The report organizes sessions your agency collected. It is not an assurance engagement, benchmark, ranking, share-of-voice score, legal record, advertising substantiation, provider authentication, or proof of what another user saw. Capture-context fields are operator-asserted and unverified. Full account exports, credentials, account headers, unrelated prompts, private client records, and third-party personal information are not accepted.