49
AI Access Score
venuematch.isitemediagroup.comPartially blocked
AI crawlers see an almost empty page on venuematch.isitemediagroup.com. Your homepage text only appears after JavaScript runs, and most AI crawlers read the page without running it. Visitors see your site; the crawlers that feed AI answers mostly do not.
A critical finding caps the composite score at 49.
AI Access Score 49 is higher than 3 of the 84 B2B SaaS sites in our September 2026 benchmark (median 91). See the benchmark
Scanned Sep 10, 2026, 15:45 UTC
This is the permanent public report page for venuematch.isitemediagroup.com. It always shows the latest scan of the domain and stays online.
75
Live Answer Visibility
96
Search Index Foundation
51
Training Data Presence
How this score is measured

The overall score weights Live Answer Visibility at 45%, Search Index Foundation at 30% and Training Data Presence at 25%, rounded to the nearest point. Critical findings can cap the score lower; any applied cap is stated next to it. Live Answer Visibility and Training Data Presence reflect crawler access plus on-page structure, with Training Data Presence also counting Common Crawl. Search Index Foundation reflects Googlebot and bingbot access plus indexability, sitemap and response health.

Our crawler tests send each bot's user agent from our own servers, not from the AI companies' verified IP ranges. Some firewalls, including Cloudflare, verify crawlers by network identity rather than user agent. That means a site can challenge our test request while allowing the real, verified crawler. When that happens this scan over-reports blocking; it never under-reports it. The Googlebot result is the most exposed to this, because fake Googlebots are the most common spoof and are policed hardest.

This scan of venuematch.isitemediagroup.com checked 15 AI crawlers across 21 automated checks on 2 pages. None of the 13 scored crawlers are blocked. AI Access Score: 49/100 (Partially blocked).

Critical: 1 finding hiding you from AI

CRITICALContent appears almost entirely JavaScript-rendered

Most AI crawlers read your raw HTML and never run scripts; they are seeing a nearly blank page.

Raw HTML contains only 11 words of readable text alongside empty SPA mount node. Most of this page appears to exist only after JavaScript runs. (Static analysis; a full render diff is on our roadmap.)

Verify it yourselfView source on your page (Ctrl+U) and search for a sentence from your homepage. If you cannot find your own copy in the source, neither can most crawlers.

The bot matrix

robots.txt policy next to what actually happened when each bot knocked. Rows where your policy says yes but the server says no are highlighted; those are the silent failures.

Live-answer botsFetch content in real time to build the answers and citations users see today.
BotOwnerrobots.txt policyReal fetchVerify
OAI-SearchBot
Powers ChatGPT search results
OpenAIPartial
line 4: Disallow: /api/
Reachable (200)
ChatGPT-User
User-triggered fetches from ChatGPT; Cloudflare classes this as Agent traffic, blocked by default on ad pages for new Cloudflare domains
OpenAIPartial
line 4: Disallow: /api/
Reachable (200)
PerplexityBot
Index for Perplexity answers
PerplexityPartial
line 4: Disallow: /api/
Reachable (200)
Perplexity-Userinformational
User-triggered; may ignore robots.txt
PerplexityPartial
line 4: Disallow: /api/
Reachable (200)
Claude-SearchBot
Claude search indexing
AnthropicPartial
line 4: Disallow: /api/
Reachable (200)
Claude-User
User-triggered fetches from Claude; Cloudflare classes this as Agent traffic, blocked by default on ad pages for new Cloudflare domains
AnthropicPartial
line 4: Disallow: /api/
Reachable (200)
DuckAssistBot
DuckAssist answers
DuckDuckGoPartial
line 4: Disallow: /api/
Reachable (200)
Training crawlersCollect content for model training. Blocking them removes you from future models' memory.
BotOwnerrobots.txt policyReal fetchVerify
GPTBot
Training corpus
OpenAIPartial
line 4: Disallow: /api/
Reachable (200)
ClaudeBot
Training corpus
AnthropicPartial
line 4: Disallow: /api/
Reachable (200)
CCBot
Feeds many labs' training data; blocking it is the quietest way to disappear from future models
Common CrawlPartial
line 4: Disallow: /api/
Reachable (200)
Meta-ExternalAgent
Llama training
MetaPartial
line 4: Disallow: /api/
Reachable (200)
Amazonbot
Alexa and Amazon AI features
AmazonPartial
line 4: Disallow: /api/
Reachable (200)
Bytespiderinformational
Known to sometimes ignore robots.txt
ByteDancePartial
line 4: Disallow: /api/
Reachable (200)
Search-index foundationCopilot rides Bing's index and AI Overviews rides Google's. If these starve, the AI layer above them starves.
BotOwnerrobots.txt policyReal fetchVerify
Googlebot
Google Search index; AI Overviews rides this. Cloudflare treats it as mixed Search and Training, so as of September 15, 2026 Cloudflare AI training blocks can hit it
GooglePartial
line 4: Disallow: /api/
Reachable (200)
bingbot
Bing index; Copilot and ChatGPT search lean on it. Cloudflare treats it as mixed Search and Training, so as of September 15, 2026 Cloudflare AI training blocks can hit it
MicrosoftPartial
line 4: Disallow: /api/
Reachable (200)
Training opt-out tokensNot crawlers. Directives controlling whether already-crawled content trains AI.
TokenControlsYour robots.txt
Google-ExtendedGemini training use of Google's crawl
Blocking Google-Extended does NOT affect your Google Search ranking, but it does remove your content from Gemini training use.
Partial
Applebot-ExtendedApple AI training use
Blocking Applebot-Extended does not affect Siri or Spotlight search.
Partial

Your scanned pages

Real defects live on money pages, not just homepages. We scanned 2 pages: your homepage plus pages picked from your sitemap, preferring product, pricing and services pages, because those are the pages AI assistants quote when buyers ask.

PageStatusIndexableCanonicalTitleWords
/200YesOKVenue Matcher™ | AI Sports Venue Media…11
/how-it-works200YesOKVenue Matcher™ | AI Sports Venue Media…11

Warnings (4)

WARNINGOnly 11 words of extractable content on the homepage

AI systems have almost nothing to quote or cite from this page.

Verify it yourselfView source (Ctrl+U) and look for your page copy as plain text.
WARNINGHeading hierarchy issues: no H1

LLM retrieval splits pages into chunks along heading structure; broken hierarchy produces orphan chunks with no context.

Verify it yourselfIn your browser console: [...document.querySelectorAll("h1,h2,h3,h4,h5,h6")].map(h => h.tagName + ": " + h.innerText.trim().slice(0,60)).join("\n")
WARNINGTitle or description needs work: meta description on / is 166 chars (aim for 50 to 160)

These are the raw material that citation snippets are built from. Your meta description is over 160 characters, so AI systems and search snippets truncate it mid-sentence. Character counts are for the decoded text, so an entity like & counts as one character, not five.

Verify it yourself
curl -s "https://venuematch.isitemediagroup.com/" | grep -ioE '<title[^>]*>[^<]*</title>|<meta[^>]*name=.?description[^>]*>'
WARNINGNot present in Common Crawl

Common Crawl feeds training data across the AI industry; absence here means most future models never read you during training.

Verify it yourselfSearch your domain at index.commoncrawl.org

Good to know (6)

INFONo semantic HTML landmarks

main, article and nav elements help parsers separate your content from your chrome.

Verify it yourself
curl -s "https://venuematch.isitemediagroup.com/" | grep -ioE "<(main|article|nav|header)" | sort | uniq -c
INFOOpen Graph incomplete: missing og:image

OG tags control how shared and cited links render on every platform.

Verify it yourself
curl -s "https://venuematch.isitemediagroup.com/" | grep -ioE '<meta[^>]*property=.?og:[^>]*>'
INFOOrganization schema is thin: fewer than 2 sameAs links

sameAs links to LinkedIn, Crunchbase and similar profiles are the entity-corroboration signal that confirms you are you.

Verify it yourselfPaste your URL into validator.schema.org and inspect the Organization block.
INFOBrand naming mismatch: schema says "iSite Media", og:site_name says "Venue Matcher"

Inconsistent entity naming fragments the machine's understanding of who you are.

INFONo llms.txt found

llms.txt is an emerging convention some AI crawlers read for site guidance. Adding one is a five minute quick win.

Verify it yourself
curl -sI "https://venuematch.isitemediagroup.com/llms.txt" | head -1
INFOVerify your Bing index manually

Copilot and ChatGPT search lean on Bing's index; if Bing has not indexed you, two of six AI platforms cannot cite you. Automated Bing verification is unreliable without authenticated APIs, so we will not fake it.

Verify it yourselfSearch "site:venuematch.isitemediagroup.com" on bing.com and confirm your key pages appear.

Recommended robots.txt additions (free, no email needed)

# Recommended AI crawler access, generated by the AI Crawler Access Report by Citant.ai
# Place these lines ABOVE any general User-agent: * block in robots.txt

User-agent: OAI-SearchBot
Allow: /
User-agent: ChatGPT-User
Allow: /
User-agent: PerplexityBot
Allow: /
User-agent: Claude-SearchBot
Allow: /
User-agent: Claude-User
Allow: /
User-agent: DuckAssistBot
Allow: /
User-agent: GPTBot
Allow: /
User-agent: ClaudeBot
Allow: /
User-agent: CCBot
Allow: /
User-agent: Meta-ExternalAgent
Allow: /
User-agent: Amazonbot
Allow: /
User-agent: Googlebot
Allow: /
User-agent: bingbot
Allow: /

# Optional training opt-outs. These are business decisions, not defects:
# blocking Google-Extended removes you from Gemini training but does NOT
# affect Google Search ranking. Leave them out to stay fully open.
# User-agent: Google-Extended
# Disallow: /

Your prioritized fix plan

Criticals first. CMS not detected; generic instructions shown. Open the print-ready report to save it as a PDF.

1. Reduce JavaScript dependency of your content

Fixes: Content appears almost entirely JavaScript-rendered

  1. Most AI crawlers read raw HTML and do not execute JavaScript, so client-rendered content is invisible to them.
  2. Move key copy to server-side rendering or static generation; on React stacks enable SSR/SSG for marketing and product pages.
  3. Verify by viewing source (Ctrl+U) and searching for a sentence from the page; if it is not in the source, crawlers cannot see it.
Verify it yourselfView source on your page (Ctrl+U) and search for a sentence from your homepage. If you cannot find your own copy in the source, neither can most crawlers.

2. Add extractable main content

Fixes: Only 11 words of extractable content on the homepage

  1. The page carries under 150 words of extractable main content; AI systems have almost nothing to quote or cite.
  2. Write the page copy into the HTML itself rather than images or embedded widgets.
Verify it yourselfView source (Ctrl+U) and look for your page copy as plain text.

3. Repair the heading hierarchy

Fixes: Heading hierarchy issues: no H1

  1. Use exactly one H1 per page and do not skip levels; LLM retrieval splits pages into chunks along heading structure, and broken hierarchy produces orphan chunks with no context.
  2. Restructure headings in the page template rather than styling text to look like headings.
Verify it yourselfIn your browser console: [...document.querySelectorAll("h1,h2,h3,h4,h5,h6")].map(h => h.tagName + ": " + h.innerText.trim().slice(0,60)).join("\n")

4. Fix title and meta description (/)

Fixes: Title or description needs work: meta description on / is 166 chars (aim for 50 to 160)

  1. Write a title of 15 to 70 characters containing your brand or the page topic, and a meta description of 50 to 160 characters.
Verify it yourself
curl -s "https://venuematch.isitemediagroup.com/" | grep -ioE '<title[^>]*>[^<]*</title>|<meta[^>]*name=.?description[^>]*>'

5. Get into Common Crawl

Fixes: Not present in Common Crawl

  1. Make sure CCBot is allowed in robots.txt and not firewall-blocked, then wait for the next crawl cycle; presence follows access plus inbound links.
Verify it yourselfSearch your domain at index.commoncrawl.org

6. Complete Open Graph tags

Fixes: Open Graph incomplete: missing og:image

  1. Add og:title, og:description and og:image meta tags to the page head.
Verify it yourself
curl -s "https://venuematch.isitemediagroup.com/" | grep -ioE '<meta[^>]*property=.?og:[^>]*>'

7. Add Organization schema with sameAs links

Fixes: Organization schema is thin: fewer than 2 sameAs links

  1. Add the JSON-LD below to your homepage head, filling in your real details.
  2. The sameAs links are the entity-corroboration signal that connects your site to your LinkedIn and other profiles.
  3. Validate at validator.schema.org.
Starter Organization schema (edit the details)
<script type="application/ld+json">
{
  "@context": "https://schema.org",
  "@type": "Organization",
  "name": "Venuematch",
  "url": "https://venuematch.isitemediagroup.com",
  "logo": "https://venuematch.isitemediagroup.com/logo.png",
  "description": "What Venuematch does, in one sentence.",
  "sameAs": [
    "https://www.linkedin.com/company/venuematch",
    "https://x.com/venuematch"
  ]
}
</script>
Verify it yourselfPaste your URL into validator.schema.org and inspect the Organization block.

8. Align your brand naming

Fixes: Brand naming mismatch: schema says "iSite Media", og:site_name says "Venue Matcher"

  1. Your Organization schema name, title tag brand, and og:site_name disagree; inconsistent naming fragments the machine's understanding of who you are.
  2. Pick one canonical brand string and use it in all three places.

9. Add an llms.txt file

Fixes: No llms.txt found

  1. Save the generated draft below as /llms.txt at your site root.
  2. It is a plain-markdown guide that tells AI systems what your most important pages are; adoption is emerging, and it costs nothing.
Your llms.txt draft, generated from your scanned pages
# Venuematch

> One-sentence description of what venuematch.isitemediagroup.com does, written for AI systems. Edit this line.

## Key pages

- [Venue Matcher™ | AI Sports Venue Media Planning | iSite Media](https://venuematch.isitemediagroup.com/): Venue Matcher™ is an AI-powered media planning tool that identifies the stadiums and arenas best aligned with your
- [Venue Matcher™ | AI Sports Venue Media Planning | iSite Media](https://venuematch.isitemediagroup.com/how-it-works): Venue Matcher™ is an AI-powered media planning tool that identifies the stadiums and arenas best aligned with your

## Contact

- [Contact](https://venuematch.isitemediagroup.com/contact): How to reach Venuematch
Verify it yourself
curl -sI "https://venuematch.isitemediagroup.com/llms.txt" | head -1

This scan measures whether AI can READ you. It cannot measure whether AI RECOMMENDS you, or who it recommends instead. That requires prompt-testing against your named competitors across 6 platforms.

Fix what's broken, then measure your visibility

The AI Search Audit by Citant.ai diagnoses everything this scan found, hands you the exact fix for every critical, and measures where you stand against your named competitors across ChatGPT, Claude, Perplexity, Gemini, Copilot and Google AI Overviews. It includes a 30 minute walkthrough call.

Book the $440 AI Search Audit

The $440 is credited in full against the GEO Pilot if you sign within 30 days.