Publish a robots.txt file

SkillWeb & browsing

robots-txt is a skill for AI agents that checks a site's robots.txt file. It is used when auditing technical SEO foundations or diagnosing crawl coverage gaps on any public website, showing which pages search engines are allowed to crawl.

Use Publish a robots.txt file in Claude, ChatGPT or Ahel Desktop

Free. Sign in, add Publish a robots.txt file and connect your AI. About a minute.

Also: Claude Code · Cursor · Codex

Then ask your AI: use the Publish a robots.txt file skill

Details

Instructions available. Your AI can read the instructions. Execution depends on the setup they require.

Have the URL of a public website you want to examine.

Publish a robots.txt fileStart free

What your AI can do with it

  • Check a site's robots.txt file
  • Audit technical SEO foundations
  • Diagnose crawl coverage gaps
  • Show which pages search engines are allowed to crawl
  • Apply to any public website

Getting started

  1. Have the URL of a public website you want to examine.
  2. Add the robots-txt skill to your agent's available skills.
  3. Ask the agent to check the site's robots.txt or audit its crawl coverage.

What this skill tells your AI

The instructions your AI receives, as published by thedaviddias/front-end-checklist in skills/robots-txt/SKILL.md and read by ahel’s review.

robots.txt is the first file crawlers fetch; misconfigured directives can silently block search engines from crawling your entire site, killing organic visibility.

Quick Reference

  • Serve a valid robots.txt at /robots.txt on the production domain, returning HTTP 200
  • Include a Sitemap: directive pointing to your XML sitemap
  • Never disallow crawling of CSS/JS assets that render your pages
  • Avoid blocking all crawlers with Disallow: / on a live site

Check

Fetch /robots.txt on the live domain and verify it returns HTTP 200, uses correct User-agent / Disallow / Allow syntax, and includes a Sitemap: directive pointing to the XML sitemap. Check for accidental Disallow: / directives.

Fix

Create or update robots.txt at the web root with valid directives. Add a live Sitemap: line for the production sitemap URL. Remove any Disallow: / rules that block the whole site or resources needed for rendering.

Explain

Explain how robots.txt controls crawler access, why an accidental Disallow: / can delist a site, and why CSS/JS must remain accessible for rendering-based indexing.

Code Review

Review metadata generation, rendered HTML, structured data, and response headers related to Publish a robots.txt file. Flag exact routes or templates where search-facing output violates the rule, and describe how to verify the final page output.


For full implementation details, code examples, and framework-specific guidance, see references/rule.md.

Rule page: https://frontendchecklist.io/en/rules/seo/robots-txt

Signals

GitHub stars
74k
Forks
7k
Last commit
Oct 2026

Questions

When should this skill be used?
Use it when auditing technical SEO foundations or diagnosing crawl coverage gaps, and when it applies to any public website.
What does the skill actually check?
It checks a site's robots.txt file so the agent can see which pages search engines are allowed to crawl.
Does it work on any site?
It applies to any public website.
Advanced
Item type
skill
Key
robots-txt-thedaviddias
Source
github.com/thedaviddias/front-end-checklist