Triage the Coverage Report: Which Crawl Errors Actually Block Rankings
Objective: Given a synthetic 12-URL GSC Coverage export mixing 404s, robots.txt blocks, redirect loops, and soft 404s, correctly sort each by whether it blocks indexing and assign a fix and an owner.
You're the marketing coordinator at Bansal Wire Industries, the Ghaziabad-based stainless steel wire manufacturer that exports to 60+ countries. Buyers in each export market find your product-category pages through search, and this month's Coverage report just landed with 12 flagged URLs.
Sort each flagged URL into 'blocks indexing, fix now' vs 'does not block indexing, low priority', then assign a fix and an owner (you vs a developer) for each.
Before you start
What you'll need
Free path (everything below is enough to finish)
It's the direct source of truth for how Google sees each page, and it's free with no verification cost
Free, no install, and easy to hand to a developer as a punch list
No access? Any plain spreadsheet or shared doc works equally well
The process
1 step
Step 01 of 01
The lesson's Step 1 groups crawl errors into four types: 404s (page gone), blocked-by-robots.txt (you blocked the crawler yourself), redirect loops (A to B to A), and soft 404s (a 'not found' page returning a 200 OK status). Search engines cannot rank a page they cannot read, so these get fixed before any content or keyword work.
Of these 12 flagged URLs, which need a fix today because they block indexing, and which can wait?
Procedure
- Export the 12 flagged rows (URL, last crawled, status) from the Coverage report
- Tag each row with its error type: 404, blocked-by-robots.txt, redirect loop, or soft 404
- Mark 404s and soft 404s on live product-category pages as 'blocks indexing, fix now'
- Mark blocked-by-robots.txt rows as 'blocks indexing, fix now' only if the URL should be public
- Mark redirect loops as 'blocks indexing, fix now' always, they trap both users and crawlers
- Leave already-retired or intentionally-blocked URLs as 'low priority, no action'
- Assign an owner: robots.txt and redirect-loop fixes go to a developer, content/URL fixes go to you
COVERAGE TRIAGE, Bansal Wire Industries, 12 flagged URLs BLOCKS INDEXING, FIX NOW (7) 1. /products/stainless-steel-wire-rod, soft 404 (thin spec page returns 200) -> owner: you, add spec table and pricing tier copy 2. /products/bright-wire-de, redirect loop (/de <-> /de-de) -> owner: developer, collapse to single canonical URL 3. /export/germany, blocked by robots.txt (accidental Disallow: /export/) -> owner: developer, remove the blanket rule 4. /products/spring-wire, 404 (page deleted, still linked from nav) -> owner: developer, 301 redirect to /products/spring-wire-coils 5. /export/vietnam, redirect loop -> owner: developer 6. /products/binding-wire, soft 404 -> owner: you 7. /export/uae, blocked by robots.txt -> owner: developer LOW PRIORITY, NO ACTION (5) 8. /old-catalog-2019.pdf, 404 (intentionally retired, no inbound links) -> no action 9. /internal/pricing-draft, blocked by robots.txt (correctly private) -> no action 10. /cart, blocked by robots.txt (correctly private) -> no action 11. /products/discontinued-galvanized-wire, 404 (product line ended, no replacement) -> no action 12. /test-page-staging, blocked by robots.txt (correctly private) -> no action
Healthy
7 of 12 rows are correctly sorted into 'fix now' because they sit on live, linked pages that buyers or search engines still expect to reach.
Unhealthy
Treating every flagged row as equally urgent, or leaving an accidental robots.txt block on a live export-market page unfixed because 'it's just a warning'.
What this means
A Coverage flag is not automatically a problem, it's a prompt to check whether the URL is supposed to be reachable.
So what do I do about it?
| Symptom | Action | Effort |
|---|---|---|
| An export-market landing page is blocked by robots.txt and buyers in that country can't find you in search | Remove the Disallow rule for that path and request re-indexing via URL Inspection | 5 min |
| A retired product page still gets crawled and flagged as 404 | Leave it alone if nothing links to it; only fix pages with real inbound links or traffic history | 5 min |
Final deliverable
A triaged crawl-error log covering all 12 flagged URLs, sorted into 'blocks indexing, fix now' vs 'low priority, no action', each with an assigned fix and owner.
See a reference example
COVERAGE TRIAGE, Concord Biotech, 8 flagged URLs (excerpt) BLOCKS INDEXING, FIX NOW (5) 1. /api-portfolio/fermentation-based, soft 404 -> owner: you, expand thin spec page 2. /regulatory/us-fda, blocked by robots.txt (accidental) -> owner: developer ... LOW PRIORITY, NO ACTION (3) 6. /old-investor-deck-2019.pdf, 404, no inbound links -> no action ...
Success criteria
You're done when you can:
- All 12 rows classified as either 'blocks indexing, fix now' or 'low priority, no action', with a stated reason
- Every 'fix now' row has both a specific fix and an owner (you or developer)
- No live, linked page is left in 'no action' just because the report showed a warning label