Choose crawl controls that match the actual goal
Avoid using crawl rules to solve an indexing or privacy problem.
Use this for: Each new block or existing rule whose purpose is unclear.
For each path you plan to block, decide whether the goal is less crawling, exclusion from search, or private access. Use robots.txt for crawl control. For a public HTML page that must stay out of search, use a crawlable noindex directive. For private material, arrange authentication or access control. Neither robots.txt nor noindex keeps a URL secret.
Illustrative choice: /search/?q=... creates unwanted crawl paths; /thank-you/ is a public page intended to stay out of search; an unpublished customer contract needs authenticated access. These require different controls.
- Name the intended outcome for each affected path.
- Choose crawl rules, a readable noindex, or access control for that outcome.
- Check that the proposed mechanism can actually achieve the stated goal.
More help and optional notes
If it fails: Move the requirement to the correct page, response-header or access-control setting before editing robots.txt.
Retest: Check that the proposed mechanism can actually achieve the stated goal.
Not applicable: No robots.txt restriction is used or proposed, and there is no indexing/privacy requirement to classify in this review.
Optional: sample URL, expected result, observation date or a reminder. Keep confidential data out of shared exports.
Changing this result updates your checkmark and progress immediately. Notes are optional.

