skipcut.comscanned Jul 1, 2026 · 11:010.45s
Public AI visibility report

skipcut.comAI visibilityNeeds Work

This site has a useful foundation, but important gaps still limit AI readability.

Key strengths include sitemap and structured data, while plain-text page access and homepage access need attention.

Recommended next step

add content negotiation for Accept: text/markdown on the homepage and return a markdown representation with Content-Type: text/markdown.

Keep the HTML response for regular browser requests.

Turn this scan into weekly monitoring.

Send yourself a magic link that opens a workspace already watching this domain for weekly AI visibility changes.

Overall score

57/100
Needs Work
Download PDF
Go to fixesOverall position460 out of 1,169Leaderboard

// score breakdown

Points by check

8 checks

Crawlability10/20
Robots.txt7.5/15
llms.txt7.5/15
Sitemap10/10
Markdown support0/15
Semantic HTML7.1/10
Structured data10/10
Content signals5/5
3pass4warn1fail

Public link

llmscan.dev/scan/e_2O3VvSROeOAbXv2xwGq

Signals checked

8 AI visibility signals

Fix bundle

4 copy-ready files

Share badge

Needs Work · 57/100

Add a polished proof badge

A compact badge for footer, press, or trust sections that links visitors to this public report.

Embed codellmscan.dev/scan/e_2O3VvSROeOAbXv2xwGq
<a href="https://www.llmscan.dev/scan/e_2O3VvSROeOAbXv2xwGq"
  target="_blank"
  rel="noopener"
>
  <img
    src="https://www.llmscan.dev/scan/e_2O3VvSROeOAbXv2xwGq/badge.png"
    alt="LLM Scan AI visibility score badge"
    width="460"
    height="120"
    style="width: 260px; max-width: 100%; height: auto;"
  />
</a>
Open badge
L
LLM Scan
Needs Work
Score
57/100

Share your score

Post the public report with: “We scored 57/100 for AI-readability.”

Download fixes

Grab generated files and implementation notes for the highest-impact gaps.

Rescan weekly

Save this domain to catch regressions after content, sitemap, or robots changes.

Monitor weekly

// signal breakdown

8 signals AI systems depend on

The homepage is reachable, but robots.txt contains AI crawler restrictions for ClaudeBot, PerplexityBot, and Google-Extended.

Signal weight

10/20
Warn

Evidence

url
https://skipcut.com/
finalUrl
https://skipcut.com/
status
200

Recommendation

Next step: Review AI crawler Disallow rules and keep only the paths that should be excluded from AI crawler access; serve a non-empty HTML homepage with a canonical link tag.

The robots.txt file was found, but it contains formatting issues.

Signal weight

8/15
Warn

Evidence

robotsTxtUrl
https://skipcut.com/robots.txt
exists
true
rawRobotsTxt
User-agent: * Allow: / Content-Signal: ai-train=yes, search=yes, ai-input=yes # Allow all search engines to crawl the site Allow: /blogs/ Allow: /v/ Allow: /ft/ Allow: /sitemap.xml Allow: /robots.txt Allow: /config/manifest.json Allow: /site.webmanifest Allow: /watch/ Allow: /playlist/ Allow: /live/ Allow: /trending.html Allow: /live-tv.html Allow: /404.html # Disallow certain directories that don't need indexing Disallow: /ad/ Disallow: /functions/ Disallow: /clean/ Disallow: /*.js$ Disallow: /*.css$ Disallow: /*.png$ Disallow: /*.jpg$ Disallow: /*.jpeg$ Disallow: /*.gif$ Disallow: /*.svg$ Disallow: /*.ico$ Disallow: /*.webp$ Disallow: /*.woff$ Disallow: /*.woff2$ Disallow: /*.ttf$ Disallow: /*.eot$ # Crawl delay for respectful crawling Crawl-delay: 1 # Sitemap locations Sitemap: https://skipcut.com/sitemap.xml Sitemap: https://skipcut.com/news-sitemap.xml # Additional directives for better SEO # Allow Googlebot to access all content User-agent: Googlebot Allow: / # Allow Bingbot to access all content User-agent: Bingbot Allow: / # Allow all major search engines User-agent: Slurp Allow: / User-agent: DuckDuckBot Allow: / User-agent: Baiduspider Allow: / User-agent: YandexBot Allow: / User-agent: facebookexternalhit Allow: / User-agent: Twitterbot Allow: / User-agent: LinkedInBot Allow: / User-agent: WhatsApp Allow: / User-agent: TelegramBot Allow: / # Block AI training bots (optional - remove if you want AI training) User-agent: GPTBot Allow: / User-agent: ChatGPT-User Allow / User-agent: CCBot Allow: / User-agent: anthropic-ai Allow: / User-agent: Claude-Web Allow: / User-agent: Omgilibot Allow: / User-agent: Applebot Allow: / # Block archive.org (optional - remove if you want archiving) User-agent: ia_archiver Disallow: / # Block common scrapers User-agent: AhrefsBot Disallow: / User-agent: SemrushBot Disallow: / User-agent: MJ12bot Disallow: / User-agent: dotbot Disallow: / User-agent: rogerbot Disallow: / User-agent: Screaming Frog SEO Spider Disallow: / # Allow legitimate SEO tools User-agent: Googlebot-Image Allow: / User-agent: Googlebot-Mobile Allow: / User-agent: Mediapartners-Google Allow: / User-agent: AdsBot-Google Allow: / User-agent: Googlebot-News Allow: / User-agent: Googlebot-Video Allow: /

Recommendation

Next step: Fix robots.txt syntax issues so each rule uses Field: value format, directives appear under a User-agent, and Sitemap entries use absolute URLs.

The llms.txt file exists but is missing one or more expected quality signals.

Signal weight

8/15
Warn

Evidence

llmsTxtUrl
https://skipcut.com/llms.txt
present
true
accessible
true

Recommendation

Next step: Publish /llms.txt as text or markdown with more than 200 characters, markdown headings, and at least one absolute URL.

The sitemap.xml file is valid and contains URL entries.

Signal weight

10/10
Pass

Evidence

sitemapUrl
https://skipcut.com/sitemap.xml
sitemapUrls
[https://skipcut.com/sitemap.xml]
robotsSitemapUrls
[]

The homepage returned HTML when requested with Accept: text/markdown, so the server appears to ignore markdown content negotiation.

Signal weight

0/15
Fail

Evidence

url
https://skipcut.com/
acceptHeader
text/markdown
status
200

Recommendation

Next step: Add content negotiation for Accept: text/markdown on the homepage and return a markdown representation with Content-Type: text/markdown. Keep the HTML response for regular browser requests.

The homepage has some semantic HTML signals, but one or more title, metadata, heading, landmark, content, or link text checks need improvement.

Signal weight

7/10
Warn

Evidence

url
https://skipcut.com/
quality
partial
score
71

Recommendation

Next step: Avoid skipped heading levels so sections progress from h1 to h2 to h3 without gaps. Add missing semantic elements: main, article.

Valid JSON-LD structured data was found with core Organization or WebSite schema.org types.

Signal weight

10/10
Pass

Evidence

url
https://skipcut.com/
quality
good
hasStructuredData
true

AI content usage signals detected via Content-Signal robots.txt directives.

Signal weight

5/5
Pass

Evidence

url
https://skipcut.com/
hasContentSignals
true
hasContentSignalHeader
false

Recommendation

Next step: Consider adding Content-Signal HTTP header, AI-specific head meta tags, robots noai/noimageai directive so AI systems can consistently discover content usage preferences across robots.txt, HTTP headers, and HTML metadata.

// generated fixes

Downloadable fix files

Preview the generated files below. Enter your email to reveal the full fixes, download the bundle, or copy the agent-ready implementation prompt.

Done-for-you

Agency package

Not sure how to ship the technical fixes? Book a call and we can help turn this report into implemented updates.

Fix planning from your scan

Implementation guidance

AI visibility monitoring

llms.txtMarkdown
# YouTube Ad Blocker | Watch YouTube Without Ads | SkipCut > Watch YouTube without ads using SkipCut YouTube Ad Blocker. Background play, sponsorblock, and local history. No install, no APK, no extension. Fast and safe. This llms.txt file summarizes the public, canonical resources that AI assistants and crawlers should use to understand this site. ## Site Overview - Canonical URL: https://skipcut.com/- Site type: software application
robots.txtTXT
# robots.txt additions# Copy these blocks into the existing robots.txt file. Keep current rules unless a note calls out a conflicting Disallow. # AI crawler access# Add explicit Allow rules for blocked AI crawlers; remove or narrow conflicting Disallow rules if your crawler target requires precedence.
schema.jsonJSON
{  "@context": "https://schema.org",  "@graph": [    {      "@type": "Organization",      "@id": "https://skipcut.com/#organization",      "name": "YouTube Ad Blocker | Watch YouTube Without Ads | SkipCut",      "description": "Watch YouTube without ads using SkipCut YouTube Ad Blocker. Background play, sponsorblock, and local history. No install, no APK, no extension. Fast and safe.",      "url": "https://skipcut.com/",      "logo": "https://skipcut.com/img/skipcut-favicon.png",
head metaHTML
# Content-Signal recommendations Use these directives to make AI-use preferences explicit for compliant crawlers and AI systems. They are advisory signals, so keep them aligned with robots.txt, terms, and access controls. ## Recommended values - ai-train=no: AI model training, fine-tuning, and dataset creation.- search=yes: AI search indexing, snippets, and discovery.- ai-input=yes: AI answer grounding, retrieval, and generated-response context.

Related scans

Similar AI visibility reports

Browse leaderboard