Now availableVisibilityKit CrawlerDownload for Mac

robots.txt

A text file at the root of a site that tells crawlers which paths they may and may not fetch.

CrawlerCrawling & Technical SEO

Definition

robots.txt lists user agents and the paths they are allowed or disallowed from crawling. It can also point to your XML sitemap.

Well-behaved crawlers, including search engines and most AI crawlers, read it before crawling. It is a request, not access control.

Why It Matters

A single wrong line can block search engines or AI crawlers from your whole site. It is also where you decide which AI crawlers may read your content.

How VisibilityKit Helps

VisibilityKit Crawler reports URLs blocked by robots.txt during a crawl, and its AI readiness check reports which AI crawlers your robots.txt allows.

Frequently Asked Questions

Does robots.txt stop a page from being indexed?

No. It stops crawling. Use noindex to keep a page out of the index.

Where does robots.txt live?

At the root of the host, for example example.com/robots.txt. Each subdomain needs its own.

Browse the glossary

301 vs 302 Redirect404 Error
Action CenterAEO AuditAEO vs GEOAEO vs SEOAI Answer EngineAI Brand MonitoringAI CitationAI Competitive AnalysisAI Content GapAI Content OptimizationAI Crawler AccessAI CrawlingAI DiscoverabilityAI Engine Optimization (AEO)AI GroundingAI HallucinationAI IndexingAI Marketing ROIAI MentionAI RecommendationAI Reputation ManagementAI Search RankingAI Search TrendsAI Search vs GoogleAI SEO StrategyAI SERPAI Share of VoiceAI Traffic AttributionAI VisibilityAI Visibility Action PlanAI Visibility SnapshotAI Visibility vs SEO RankingAI-First Content StrategyAI-First SEOAI-Powered SearchAI-Ready ContentAnswer EngineAnswer EvidenceAnswers Matrix
Brand FidelityBrand GraphBrand Visibility ScoreBranded PromptBroken LinkBuyer IntentBuyer Question
Canonical TagChatGPT SEOCitation RateCitation TrackingClaude OptimizationClient Viewer RoleClient WorkspacesContent CitabilityConversational SearchCrawl BudgetCrawl ComparisonCrawl DepthCrawl SnapshotCustom Extraction
DeepSeek SearchDuplicate Content
Entity Optimization
Fix VerificationFound but Not Cited
Gatekeeper SourceGemini SearchGenerative Engine Optimization (GEO)Google AI OverviewsGoogle Analytics 4 (GA4)Google Search ConsoleGrok Search
Heading Structure (H1 to H6)HreflangHTTP Status Codes
Image Alt TextIndexabilityInlinksInternal Link SuggestionsInternal LinkingInternal PageRank
JavaScript Rendering
Knowledge Graph
Large Language Model (LLM)llms.txtLocal-Only Crawl DataLooker Studio Export
MCP ServerMention vs Recommendation vs CitationMeta AI SearchMeta DescriptionMicrosoft Copilot SearchModel Context Protocol (MCP)Multi-Brand Portfolio
Natural Language Processing (NLP)Noindex (Meta Robots)
Orphan PageOutlinks
Perplexity RankingPrompt OptimizationPrompt Set
Redirect ChainRedirect LoopRedirect MapRendered vs Raw HTMLResponse Timerobots.txt
Scheduled CrawlSchema MarkupSearch Console Winners and LosersSemantic SearchSEO CrawlerSoft 404Structured Data for AI
Technical SEO AuditThin ContentTitle TagTopical Authority
URL List Mode
VisibilityKit vs Alternatives
Web CrawlerWhite-Label Report
XML Sitemap
Zero-Click Search

Put these terms to work

VisibilityKit Crawler

A desktop SEO crawler for macOS, Windows and Linux. Find what is broken, assign the work, then verify every fix. Crawl data stays on your machine.

Get Crawler free