{"schema_version":"2.0","record_type":"research","canonical_url":"https://marketingwiki.ai/research/ai-search-visibility-benchmark","id":"research-ai-search-visibility-benchmark","slug":"ai-search-visibility-benchmark","title":"AI Search Visibility Benchmark","description":"Open protocol for measuring mentions, citations, source diversity, support quality, and answer stability across AI answer systems.","dek":"Protocol published before results. Prompt set, system settings, repetitions, raw answers, and scoring stay visible.","type":"Benchmark protocol","status":"Protocol published","topics":["AI search","benchmarking","citations"],"publishedAt":"2026-08-10","updatedAt":"2026-08-10","lastVerifiedAt":"2026-08-10","readingMinutes":2,"author":"Marketing Wiki Editors","reviewer":"Marketing Wiki Editors","sources":[{"title":"Google AI optimization guide","url":"https://developers.google.com/search/docs/fundamentals/ai-optimization-guide"},{"title":"OpenAI publisher and developer FAQ","url":"https://help.openai.com/en/articles/12627856-publishers-and-developers-faq"}],"wordCount":226,"body":"This research asks which sources AI answer systems surface for common marketing AI questions and how stable those results remain across repeated runs.\n\n## Questions\n\n1. Which domains receive mentions?\n2. Which pages receive clickable citations?\n3. Do citations support nearby claims?\n4. How much do answers change between fresh runs?\n5. Does explicit source request change citation quality?\n\n## Cohort\n\nInitial cohort covers question-led prompts about agent architecture, repository workflows, content verification, and AI search measurement. Vendor-ranking prompts remain excluded from first run.\n\n## Protocol\n\n- Freeze prompt text before collection\n- Use fresh conversation for each run\n- Record product, displayed model, browsing mode, region, language, and timestamp\n- Repeat each prompt at least three times per system\n- Store answer text and visible citations\n- Review whether citation supports associated claim\n\n## Metrics\n\n`mention_rate` counts runs naming source or entity. `citation_rate` counts runs with clickable source. `support_rate` counts citations that substantively support nearby answer. `stability` compares results across repetitions.\n\nThese metrics remain separate. No weighted overall score yet.\n\n## Result status\n\nNo result dataset published. Current page describes protocol only. Publishing protocol first reduces incentive to change method after seeing favorable results.\n\n## Planned artifacts\n\n```text\nresearch/prompts/v1.jsonl\nresearch/runs/<system>/<date>.jsonl\nresearch/scoring/v1.md\nresearch/results/v1.csv\n```\n\n## Limitations\n\nInterfaces, model routing, retrieval indexes, and answer policies change. Result is dated observation, not durable guarantee of citation or rank."}