Publish a robots.txt file
SkillWeb & browsingrobots-txt is a skill for AI agents that checks a site's robots.txt file. It is used when auditing technical SEO foundations or diagnosing crawl coverage gaps on any public website, showing which pages search engines are allowed to crawl.
Use Publish a robots.txt file in Claude, ChatGPT or Ahel Desktop
Free. Sign in, add Publish a robots.txt file and connect your AI. About a minute.
Also: Claude Code · Cursor · Codex
Then ask your AI: use the Publish a robots.txt file skill
Details
Instructions available. Your AI can read the instructions. Execution depends on the setup they require.
Account requirements not reviewed. Check the skill instructions before use; ahel provides instructions and does not run this skill.
No other account needed.
Have the URL of a public website you want to examine.
What your AI can do with it
- Check a site's robots.txt file
- Audit technical SEO foundations
- Diagnose crawl coverage gaps
- Show which pages search engines are allowed to crawl
- Apply to any public website
Getting started
- Have the URL of a public website you want to examine.
- Add the robots-txt skill to your agent's available skills.
- Ask the agent to check the site's robots.txt or audit its crawl coverage.
What this skill tells your AI
The instructions your AI receives, as published by thedaviddias/front-end-checklist in skills/robots-txt/SKILL.md and read by ahel’s review.
robots.txt is the first file crawlers fetch; misconfigured directives can silently block search engines from crawling your entire site, killing organic visibility.
Quick Reference
- Serve a valid
robots.txtat/robots.txton the production domain, returning HTTP 200 - Include a
Sitemap:directive pointing to your XML sitemap - Never disallow crawling of CSS/JS assets that render your pages
- Avoid blocking all crawlers with
Disallow: /on a live site
Check
Fetch /robots.txt on the live domain and verify it returns HTTP 200, uses correct User-agent / Disallow / Allow syntax, and includes a Sitemap: directive pointing to the XML sitemap. Check for accidental Disallow: / directives.
Fix
Create or update robots.txt at the web root with valid directives. Add a live Sitemap: line for the production sitemap URL. Remove any Disallow: / rules that block the whole site or resources needed for rendering.
Explain
Explain how robots.txt controls crawler access, why an accidental Disallow: / can delist a site, and why CSS/JS must remain accessible for rendering-based indexing.
Code Review
Review metadata generation, rendered HTML, structured data, and response headers related to Publish a robots.txt file. Flag exact routes or templates where search-facing output violates the rule, and describe how to verify the final page output.
For full implementation details, code examples, and framework-specific guidance,
see references/rule.md.
Rule page: https://frontendchecklist.io/en/rules/seo/robots-txt
Signals
- GitHub stars
- 74k
- Forks
- 7k
- Last commit
- Oct 2026
Questions
- When should this skill be used?
- Use it when auditing technical SEO foundations or diagnosing crawl coverage gaps, and when it applies to any public website.
- What does the skill actually check?
- It checks a site's robots.txt file so the agent can see which pages search engines are allowed to crawl.
- Does it work on any site?
- It applies to any public website.
Advanced
- Item type
- skill
- Key
robots-txt-thedaviddias- Source
- github.com/thedaviddias/front-end-checklist
github.com/thedaviddias/front-end-checklist
Related picks
Skill · aaron-he-zhu
The pick for SEOseo-audit
Skill · aeonfun
The pick for SEObrowser-use
Skill · browser-use
More in Web & browsingwebapp-testing
Skill · anthropics
More in Web & browsingplaywright-cli
Skill · microsoft
More in Web & browsingbenchmark
Skill · affaan-m
More in Web & browsing