You published new pages, waited, and checked Search Console — and Google still has not crawled them. No crawl date, no index, no traffic. It is one of the more anxious problems Malaysian business owners bring to ZenWeb, because a page Google never visits cannot rank for anything.
The reassuring part: Google not crawling your site is nearly always a setup or signal problem, not a dead end. Our SEO team clears it on client sites most weeks, and the causes are a short, repeatable list. If crawling stalled alongside a wider slump, our guide on why rankings drop suddenly is worth a read too.
The video below explains how Googlebot discovers and crawls pages; the guide after it walks through why your site is not being crawled and how to fix each cause in order.
Source video: Google Search Central on YouTube.
Google not crawling your site rarely means anything is broken beyond repair. It usually means Googlebot cannot reach your pages, has been told to stay away, or has not found a reason to visit yet. Each of those is a separate, fixable problem.
This guide keeps it practical: what crawling is, why yours has stalled, and a clear order of fixes — so you know what to check first and roughly how long a re-crawl takes.
Quick Answer: Crawling is when Googlebot visits a page and reads its content. Indexing is when Google stores that page to show in search. Google not crawling your site means the first step never happens — so the page cannot be indexed or ranked. Crawling comes first; everything else depends on it.
Before you fix anything, separate three stages that often get lumped together:
This matters because the fix differs at each stage. If Google is crawling but not storing pages, that is an indexing problem — our guide on pages not getting indexed covers it. If Googlebot never even visits, that is a crawl problem, and that is what this guide fixes.
Quick Answer: Google usually fails to crawl a site for five reasons: robots.txt or a noindex tag blocking access, server errors that time out the crawler, orphan pages with no internal links, thin or duplicate content Google skips, or a new site with no authority yet. Almost every case is one of these, not a mystery.
Across the sites we audit, the same causes come up again and again:
These are access and signal gaps, not penalties. A blocking rule or a broken robots.txt is the most common culprit — our guide on robots.txt blocking Google shows how to check yours.
Not sure which of these is blocking your crawl?
We trace it to the exact cause and clear it fast. See our SEO services →
Quick Answer: Across ZenWeb audits of Malaysian SME sites, blocking rules in robots.txt or a noindex tag are the single biggest cause of Google not crawling a site, followed by thin pages Google skips and orphan pages with no internal links. Fixing access and internal links clears most cases.
When we audit why a site is not being crawled, the causes cluster in a predictable order:
| Root cause | Share of cases |
|---|---|
| Blocked by robots.txt or noindex | 28% |
| Thin or duplicate pages Google skips | 22% |
| Orphan pages with no internal links | 18% |
| Server errors or timeouts (5xx) | 16% |
| New site with no authority yet | 16% |
Source: ZenWeb client audits, Malaysian SME sites, 2024–2026. Directional, not guarantees.
Quick Answer: Confirm crawling in Google Search Console. Use the URL Inspection tool to see a page’s last crawl date, and the Crawl Stats report (Settings) to see whether Googlebot is visiting your site at all. A site:yourdomain.com search shows what is already indexed. Together these tell you if crawling stopped, slowed, or never started.
You do not have to guess. Three checks in Search Console give a clear picture:
site: search. Type site:yourdomain.com into Google to see roughly how many pages are already in the index.If Crawl Stats shows requests collapsing to near zero, or errors climbing, the problem is access or server health. Our guide on crawl errors in Search Console explains how to read each error type.
site: search tell you fast whether crawling stopped, slowed, or never began — so you fix the real problem, not a guess.Quick Answer: Google decides what to crawl using internal links, sitemap inclusion, site authority, server health, and content freshness. Internal links carry the most weight — they are how Googlebot finds and prioritises pages. Strengthen the heavy signals first and crawling improves fastest.
Google has never published a formula, but its guidance and our testing point to a clear order of influence:
| Crawl signal | Relative influence (0-100) |
|---|---|
| Internal links pointing to the page | 100 |
| Inclusion in the XML sitemap | 85 |
| Site authority and backlinks | 78 |
| Fast, error-free server response | 68 |
| Fresh, regularly updated content | 60 |
Source: ZenWeb testing aligned to Google’s stated crawl signals, 2024–2026. Directional weighting, not a Google-published formula.
Internal links and the sitemap do most of the discovery work, which is why a missing sitemap or a broken one so often sits behind a crawl problem.
Quick Answer: To get Google crawling again, clear any robots.txt or noindex block, submit a clean XML sitemap, add internal links to the stranded pages, then request indexing with the URL Inspection tool. Fix the block first — requesting a crawl on a page Google is told to avoid changes nothing.
Six steps, in order, that take a page from ignored to crawled and indexed.
Disallow and the page’s HTML for a noindex tag, then remove whatever is shutting Googlebot out.Work top to bottom. Requesting indexing is the last step, not the first — if a page is still blocked, our guide on pages not getting indexed covers what to check next.
Quick Answer: The hard stops are a Disallow in robots.txt, a noindex meta tag, HTTP 5xx server errors, and content hidden behind scripts Googlebot cannot run. Each one silently prevents crawling until removed. Fix these before touching content or links, because nothing else works while a hard block is in place.
Some blockers stop crawling outright. Clear these first:
Disallow: / blocks your entire site. Check it at yourdomain.com/robots.txt and remove any line that should not be there.noindex from a staging site tells Google to drop the page — our guide on the noindex tag blocking pages shows how to find it.These are also common after a redesign or migration, when old rules get carried over. If Google reaches your pages but changes how they appear — like when it rewrites your meta description — that is on-page, not a crawl block. Media has its own quirks too: if pictures stay missing from search, see why images are not ranking.
Stuck on a technical block you can’t find?
We audit robots.txt, tags, and server logs to pinpoint what’s stopping the crawl. Get a technical SEO audit →
Quick Answer: Each crawl blocker shows a telltale sign in Search Console. “Blocked by robots.txt” points to a Disallow rule; “Excluded by noindex tag” points to a meta tag; “Discovered — currently not indexed” usually means orphan or thin pages. Matching the symptom to the fix is the fastest route back to being crawled.
Use this table to match what Search Console shows to the fix that clears it:
| Blocker | Search Console symptom | Fix |
|---|---|---|
| robots.txt Disallow | Blocked by robots.txt | Remove or narrow the Disallow line |
| noindex meta tag | Excluded by noindex tag | Delete the noindex from the page head |
| Server error | Server error (5xx) | Fix hosting or timeout, retest live URL |
| Orphan page | Discovered — currently not indexed | Add internal links from strong pages |
| Thin or duplicate content | Crawled — currently not indexed | Improve depth and make the page unique |
Source: ZenWeb client audits, Malaysian SME sites, 2024–2026. Directional.
Quick Answer: A requested crawl can happen within hours to a few days; a sitemap resubmission or a robots.txt fix usually clears in one to two weeks. Building authority for a brand-new site is slowest, at four to eight weeks. Google must re-crawl before anything changes, so plan in days and weeks, not minutes.
Set expectations by the fix. Google has to re-crawl and re-assess before results move:
| Action | Typical time | Effort |
|---|---|---|
| Request indexing via URL Inspection | Hours–a few days | Low |
| Submit or resubmit the XML sitemap | A few days–2 weeks | Low |
| Fix robots.txt or remove noindex | 1–2 weeks | Low |
| Add internal links to orphan pages | 1–3 weeks | Medium |
| Build authority for a new site | 4–8 weeks | High |
Source: ZenWeb operational data, 500+ Malaysian SME campaigns, 2024–2026. Typical ranges, not guarantees.
Quick Answer: Clearing a single robots.txt line or requesting indexing on a few pages is a safe DIY job. Bring in help when a whole site will not crawl, when errors return after every fix, or when the block sits in server config or a JavaScript build you cannot safely edit.
Decide by scale and how deep the blocker sits:
At ZenWeb, we fix crawling at the root — access rules, server health, internal links, and sitemaps — and confirm each page is crawled and indexed. If it is part of a bigger problem, such as a site that is new and not ranking yet, our SEO team can sort the whole picture.
Google not crawling your site looks alarming but almost never is. Googlebot either cannot reach your pages, has been told to stay away, or has not found a reason to visit — and each of those is a clear, fixable cause.
Confirm the block in Search Console, clear it, strengthen internal links and your sitemap, then request a crawl and give Google a few weeks. If the site still will not crawl, or the block sits somewhere you cannot safely reach, ZenWeb’s SEO team can trace it and confirm your pages are found, crawled, and indexed.
New sites often get crawled slowly because Google has no history or backlinks to trust yet. Make sure the site is not blocked in robots.txt or by a leftover noindex tag, submit an XML sitemap in Search Console, and request indexing for key pages. Earn a few links so Google sees a reason to visit, and crawling usually picks up over the first few weeks.
Open Google Search Console, paste the URL into the URL Inspection tool, and look at the last crawl date. If it shows a recent date, Google crawled it; if it says the page is not on Google or was never crawled, something is blocking it. The Crawl Stats report under Settings shows whether Googlebot is visiting your site overall.
After you request indexing through URL Inspection, a crawl can happen within a few hours to a few days, though it is not guaranteed. Sitemap resubmissions and robots.txt fixes usually clear within one to two weeks. Requesting a crawl does not promise indexing — Google still decides whether the page is worth keeping.
Yes. A single line like Disallow: / in your robots.txt file blocks Googlebot from your entire site. This often happens by accident when a staging-site rule gets pushed live. Check the file at yourdomain.com/robots.txt, remove any rule that should not be there, then request a fresh crawl to recover.
Google spends limited crawl effort where it sees the most value. Pages with strong internal links, a place in the sitemap, and unique content get crawled first. Orphan pages, thin or duplicate pages, and slow-loading URLs get crawled last or skipped. Adding internal links and improving the page usually gets it crawled.
Still can’t get Google to crawl your site?
Book a free 30-minute strategy session. We’ll check your robots.txt, tags, server health, and internal links, then give you a clear plan to get your pages crawled and indexed.
Complete the form and our team will contact you to discuss your goals. Let’s grow your business.

Online