ugccopilot.aiscanned Aug 11, 2026 · 11:030.19s
Public AI visibility report

ugccopilot.aiAI visibilityNeeds Work

This site has a useful foundation, but important gaps still limit AI readability.

Key strengths include AI guidance file and sitemap, while homepage access and crawler policy need attention.

Recommended next step

remove AI crawler Disallow: / rules or replace them with narrower path-level restrictions for private content only.

Overall score

40/100
Needs Work

What happens next

Monitor this score as your site changes

Weekly checks catch robots, sitemap, schema, and content regressions before AI systems stop reusing your pages.

Monitor changes
Download PDF
Go to fixesOverall position1,297 out of 1,724Leaderboard

// score breakdown

Points by check

8 checks

Crawlability0/20
Robots.txt0/15
llms.txt15/15
Sitemap10/10
Markdown support0/15
Semantic HTML0/10
Structured data10/10
Content signals5/5
4pass0warn4fail

Public link

llmscan.dev/scan/LeAh7XCJDumzHMT6F-iQD

Signals checked

8 AI visibility signals

Fix bundle

4 copy-ready files

Share badge

Needs Work · 40/100

Add a polished proof badge

A compact badge for footer, press, or trust sections that links visitors to this public report.

Embed codellmscan.dev/scan/LeAh7XCJDumzHMT6F-iQD
<a href="https://www.llmscan.dev/scan/LeAh7XCJDumzHMT6F-iQD"
  target="_blank"
  rel="noopener"
>
  <img
    src="https://www.llmscan.dev/scan/LeAh7XCJDumzHMT6F-iQD/badge.png"
    alt="LLM Scan AI visibility score badge"
    width="460"
    height="120"
    style="width: 260px; max-width: 100%; height: auto;"
  />
</a>
Open badge
L
LLM Scan
Needs Work
Score
40/100

Share your score

Post the public report with: “We scored 40/100 for AI-readability.”

Download fixes

Grab generated files and implementation notes for the highest-impact gaps.

Rescan weekly

Save this domain to catch regressions after content, sitemap, or robots changes.

Monitor weekly

// signal breakdown

8 signals AI systems depend on

The homepage is reachable, but robots.txt blocks Claude-Web from crawling the site.

Signal weight

0/20
Fail

Evidence

url
https://ugccopilot.ai/
finalUrl
https://ugccopilot.ai/
status
200

Recommendation

Next step: Remove AI crawler Disallow: / rules or replace them with narrower path-level restrictions for private content only.

robots.txt explicitly blocks Claude-Web from the whole site.

Signal weight

0/15
Fail

Evidence

robotsTxtUrl
https://ugccopilot.ai/robots.txt
exists
true
rawRobotsTxt
# UGC Copilot — robots.txt # Strategy: allow indexing, AI citation, AND selective training opt-in. # The major three training crawlers (GPTBot, Google-Extended, Applebot-Extended) # are opted IN on all paths except /compare-* — comparison pages stay out of # training data to protect competitive positioning. Long-tail training scrapers # remain blocked. Policy decision logged 2026-05-01; # see docs/distribution-strategy-priorities.md for rationale. # ----------------------------------------------------------------------------- # Default rules — apply to all crawlers including search engines (Googlebot, # Bingbot, DuckDuckBot, etc.) and AI crawlers without an explicit block below. # ----------------------------------------------------------------------------- User-agent: * Allow: / Disallow: /dashboard Disallow: /dashboard/ Disallow: /admin Disallow: /admin/ Disallow: /auth Disallow: /auth/ Disallow: /__/ # ----------------------------------------------------------------------------- # AI training crawlers — major three opted IN (with /compare protection). # These power ChatGPT (GPTBot), Gemini and AI Overviews (Google-Extended), and # Apple Intelligence (Applebot-Extended) — the highest-leverage surfaces for # durable brand recall in future model snapshots. /compare-* is disallowed to # keep competitive comparison pages out of training data. # ----------------------------------------------------------------------------- User-agent: GPTBot Allow: / Disallow: /compare Disallow: /compare/ Disallow: /dashboard Disallow: /dashboard/ Disallow: /admin Disallow: /admin/ Disallow: /auth Disallow: /auth/ Disallow: /__/ User-agent: Google-Extended Allow: / Disallow: /compare Disallow: /compare/ Disallow: /dashboard Disallow: /dashboard/ Disallow: /admin Disallow: /admin/ Disallow: /auth Disallow: /auth/ Disallow: /__/ User-agent: Applebot-Extended Allow: / Disallow: /compare Disallow: /compare/ Disallow: /dashboard Disallow: /dashboard/ Disallow: /admin Disallow: /admin/ Disallow: /auth Disallow: /auth/ Disallow: /__/ # ----------------------------------------------------------------------------- # AI training crawlers — long tail still blocked. Lower brand-recall value, # less transparent training-set practices, and several are content scrapers # (Common Crawl resale, Diffbot data products) more than first-party model # trainers. Revisit individually as practices evolve. # ----------------------------------------------------------------------------- User-agent: anthropic-ai Disallow: / User-agent: CCBot Disallow: / User-agent: Bytespider Disallow: / User-agent: FacebookBot Disallow: / User-agent: meta-externalagent Disallow: / User-agent: Amazonbot Disallow: / User-agent: cohere-ai Disallow: / User-agent: cohere-training-data-crawler Disallow: / User-agent: Diffbot Disallow: / User-agent: PanguBot Disallow: / User-agent: Timpibot Disallow: / User-agent: omgili Disallow: / User-agent: ImagesiftBot Disallow: / User-agent: AI2Bot Disallow: / User-agent: Claude-Web Disallow: / # ----------------------------------------------------------------------------- # AI search / citation crawlers — explicitly ALLOWED so we appear in # answer engines (ChatGPT search, Perplexity, Claude search, Bing Copilot). # ----------------------------------------------------------------------------- User-agent: OAI-SearchBot Allow: / Disallow: /dashboard Disallow: /dashboard/ Disallow: /admin Disallow: /admin/ Disallow: /auth Disallow: /auth/ Disallow: /__/ User-agent: ChatGPT-User Allow: / Disallow: /dashboard Disallow: /dashboard/ Disallow: /admin Disallow: /admin/ Disallow: /auth Disallow: /auth/ Disallow: /__/ User-agent: PerplexityBot Allow: / Disallow: /dashboard Disallow: /dashboard/ Disallow: /admin Disallow: /admin/ Disallow: /auth Disallow: /auth/ Disallow: /__/ User-agent: Perplexity-User Allow: / Disallow: /dashboard Disallow: /dashboard/ Disallow: /admin Disallow: /admin/ Disallow: /auth Disallow: /auth/ Disallow: /__/ User-agent: ClaudeBot Allow: / Disallow: /dashboard Disallow: /dashboard/ Disallow: /admin Disallow: /admin/ Disallow: /auth Disallow: /auth/ Disallow: /__/ User-agent: Claude-SearchBot Allow: / Disallow: /dashboard Disallow: /dashboard/ Disallow: /admin Disallow: /admin/ Disallow: /auth Disallow: /auth/ Disallow: /__/ User-agent: Applebot Allow: / Disallow: /dashboard Disallow: /dashboard/ Disallow: /admin Disallow: /admin/ Disallow: /auth Disallow: /auth/ Disallow: /__/ # ----------------------------------------------------------------------------- # Forward-looking signal (IETF draft draft-romm-aipref-contentsignals). # Not yet honored by major crawlers — explicit User-agent rules above are the # enforcement mechanism — but harmless to include for future compatibility. # ----------------------------------------------------------------------------- Content-Signal: ai-train=yes, search=yes, ai-input=yes # ----------------------------------------------------------------------------- # Sitemap and AI documentation pointers # ----------------------------------------------------------------------------- Sitemap: https://ugccopilot.ai/sitemap.xml # llms.txt — structured product summary for AI crawlers # llms-full.txt — comprehensive product documentation # Tracking parameter URLs (utm_*, ref=*) intentionally NOT disallowed — # blocking them triggers "Indexed, though blocked by robots.txt" warnings in # Search Console; canonical tags handle dedup instead.

Recommendation

Next step: Remove AI crawler Disallow: / rules or add narrower Allow/Disallow rules if AI crawlers should be able to discover public content.

The llms.txt file was found and includes the expected text, length, heading, and URL signals.

Signal weight

15/15
Pass

Evidence

llmsTxtUrl
https://ugccopilot.ai/llms.txt
present
true
accessible
true

The sitemap.xml file is valid and contains URL entries.

Signal weight

10/10
Pass

Evidence

sitemapUrl
https://ugccopilot.ai/sitemap.xml
sitemapUrls
[https://ugccopilot.ai/sitemap.xml]
robotsSitemapUrls
[]

The homepage returned HTML when requested with Accept: text/markdown, so the server appears to ignore markdown content negotiation.

Signal weight

0/15
Fail

Evidence

url
https://ugccopilot.ai/
acceptHeader
text/markdown
status
200

Recommendation

Next step: Add content negotiation for Accept: text/markdown on the homepage and return a markdown representation with Content-Type: text/markdown. Keep the HTML response for regular browser requests.

The homepage HTML is missing important title, metadata, structure, content, or link text signals.

Signal weight

0/10
Fail

Evidence

url
https://ugccopilot.ai/
quality
poor
score
43

Recommendation

Next step: Shorten the title tag to 70 characters or fewer. Shorten the meta description to 160 characters or fewer. Add exactly one h1 element that describes the page topic. Add missing semantic elements: main, article, nav, footer.

Valid JSON-LD structured data was found with core Organization or WebSite schema.org types.

Signal weight

10/10
Pass

Evidence

url
https://ugccopilot.ai/
quality
good
hasStructuredData
true

AI content usage signals detected via Content-Signal robots.txt directives.

Signal weight

5/5
Pass

Evidence

url
https://ugccopilot.ai/
hasContentSignals
true
hasContentSignalHeader
false

Recommendation

Next step: Consider adding Content-Signal HTTP header, AI-specific head meta tags, robots noai/noimageai directive so AI systems can consistently discover content usage preferences across robots.txt, HTTP headers, and HTML metadata.

What happens next

Keep watching this score

Scores move when sites and models change. This one was measured once, on Aug 11, 2026.

Watch weekly

// generated fixes

Downloadable fix files

Preview the generated files below. Enter your email to reveal the full fixes, download the bundle, or copy the agent-ready implementation prompt.

Done-for-you

Agency package

Not sure how to ship the technical fixes? Book a call and we can help turn this report into implemented updates.

Fix planning from your scan

Implementation guidance

AI visibility monitoring

llms.txtMarkdown
# AI UGC Ads Generator — Trend Analysis + Video in 5 Minutes | UGC Copilot > AI UGC ads generator — turn your product URL into AI UGC video ads in minutes. Trend analysis, viral scripts, AI Twins. Renders with Sora 2, Veo 3.1, Kling 3.0 & Seedance 2.0. From $29/mo. This llms.txt file summarizes the public, canonical resources that AI assistants and crawlers should use to understand this site. ## Site Overview - Canonical URL: https://ugccopilot.ai/- Site type: web site
robots.txtTXT
# robots.txt additions# Copy these blocks into the existing robots.txt file. Keep current rules unless a note calls out a conflicting Disallow. # AI crawler access# Add explicit Allow rules for blocked AI crawlers; remove or narrow conflicting Disallow rules if your crawler target requires precedence.User-agent: GPTBotAllow: / User-agent: ChatGPT-UserAllow: /
schema.jsonJSON
{  "@context": "https://schema.org",  "@graph": [    {      "@type": "Organization",      "@id": "https://ugccopilot.ai/#organization",      "name": "AI UGC Ads Generator — Trend Analysis + Video in 5 Minutes | UGC Copilot",      "description": "AI UGC ads generator — turn your product URL into AI UGC video ads in minutes. Trend analysis, viral scripts, AI Twins. Renders with Sora 2, Veo 3.1, Kling 3.0 & Seedance 2.0. From $29/mo.",      "url": "https://ugccopilot.ai/",
head metaHTML
# Content-Signal recommendations Use these directives to make AI-use preferences explicit for compliant crawlers and AI systems. They are advisory signals, so keep them aligned with robots.txt, terms, and access controls. ## Recommended values - ai-train=no: AI model training, fine-tuning, and dataset creation.- search=yes: AI search indexing, snippets, and discovery.- ai-input=yes: AI answer grounding, retrieval, and generated-response context.

What happens next

Monitor this score as your site changes

Weekly checks catch robots, sitemap, schema, and content regressions before AI systems stop reusing your pages.

Watch weekly

Related scans

Similar AI visibility reports

Browse leaderboard