
Search engines discover your pages by crawling them. When a crawler arrives at a URL and something goes wrong — the page is missing, blocked, or returns the wrong response — that page can drop out of the index, or never make it in at all. Nothing dramatic happens overnight, which is exactly why crawl errors are so easy to ignore. You carry on publishing, your traffic slowly sags, and nobody can quite explain why.
The good news is that most crawl errors are boring, fixable housekeeping. A stray link here, a forgotten blocking rule there. Once you know where to look, a small business site can usually be tidied up in an afternoon.
Start with the crawl and coverage reports in your search console, then cross-check against your XML sitemap. You are looking for a handful of recurring problem types:
Group the list by cause rather than fixing URLs one at a time. Ten 404s from a single deleted category are one job, not ten.
When you find a broken internal link, fix it at the source. Edit the page that contains the link so it points directly at the live URL. It is tempting to add another redirect and move on, but redirects pile up, slow pages down, and eventually break.
For external links pointing at pages you have removed, set up a single 301 redirect to the closest relevant page — never to the homepage by default, because that frustrates visitors and muddies your signals. Then check for chains: a redirect that points at another redirect that points at a third. Flatten each chain so there is only one hop.
Accidental blocking is the most common cause of "my page has vanished" panic. Three culprits account for most of it: a robots.txt rule, a noindex tag, and a server-level restriction.
The subtlety worth remembering is that these do different things. A robots.txt block stops a crawler from fetching the page, which means it will never see a noindex tag on that page. So if you disallow a URL and also add noindex, the noindex is invisible and the URL may still appear in results from external links. Remove the block, let the page be crawled, and the noindex will do its job.
Also check for staging leftovers. It is very common for a site to go live while a "discourage indexing" setting from the development version is still switched on. Look at your live pages as a crawler sees them and confirm the tags are the ones you actually want.
Duplicates dilute your efforts because crawlers have to choose between several versions of the same content, and they do not always choose the one you prefer. Typical causes on small business sites include www and non-www addresses, http and https versions, trailing and non-trailing slashes, uppercase URLs, tracking parameters, and paginated category pages.
Pick one preferred version of every address and enforce it with a redirect, not a hope. Then add a canonical tag on each page pointing to that preferred version. Keep your internal links consistent — if half your menu points to one version and half to another, you are creating the duplication yourself.
Crawl errors are not a one-off project. New ones appear whenever you publish, delete, restructure or change hosting. A short monthly routine keeps them from stacking up:
Do this consistently and you will spend far less time chasing mystery traffic drops. Clean crawling is not glamorous, but it is the foundation everything else in your search visibility is built on — and it is entirely within your control.
Plan topics by month, assign clear owners and leave room for timely updates so your blog remains active and relevant.
Comments
ALina Kelian
19th May 2018 ReplyAddaurl Simple, genuine and looked after with care — the kind of place worth returning to. ullamco laboris nisi ut aliquip ex ea commodo consequat.
Rlex Kelian
19th May 2018 ReplyAddaurl Simple, genuine and looked after with care — the kind of place worth returning to. ullamco laboris nisi ut aliquip commodo.
Roboto Alex
21th May 2018 ReplyAddaurl No rushing, no fuss — just thoughtful notes and practical help, written by people who care. ullamco laboris nisi ut aliquip ex ea commodo consequat.