← Back to list

Discovered vs Crawled (But Not Indexed): Understanding Google’s Indexing Behavior

If you’ve ever spent time inside Google Search Console (GSC), you’ve likely seen the frustrating statuses: “Discovered — currently not…

Qaushik · 2025-08-17 14:14 · 0 claps · 2.0 min read
#seo #seo-tips #seo-tips-and-tricks #seo-tips-for-beginners #seo-tips-for-website
Open on Medium ↗
Wiki topics: SEO · SEO & SEM

What is Crawl budget in SEO?

What is Crawl budget in SEO?

Discovered vs Crawled (But Not Indexed): Understanding Google’s Indexing Behavior

If you’ve ever spent time inside Google Search Console (GSC), you’ve likely seen the frustrating statuses: “Discovered — currently not indexed” and “Crawled — currently not indexed.” At first glance, they look similar, but in reality, they tell very different stories about how Google views your site.

Let’s break them down.

1. Discovered — Currently Not Indexed

What it means: Google knows your page exists (through your sitemap, backlinks, or internal links), but it hasn’t actually crawled it yet.

Why it happens:

  • Crawl budget limitations — Google prioritizes high-value pages, and if your site has too many URLs or low authority, it may push crawling to “later.”
  • Low trust in your site — New websites or those with poor authority may not get deep crawls.
  • Duplicate or thin content — If your site has lots of near-identical pages, Google won’t rush to crawl every one.

Who’s responsible? Mostly on the site owner/SEO side, because crawl signals like site speed, URL structure, sitemap hygiene, and authority influence whether Google decides to crawl.

2. Crawled — Currently Not Indexed

What it means: Google visited the page, read its content, but decided not to include it in the index.

Why it happens:

  • Thin or low-quality content — Pages with little unique information.
  • Duplicate content — Google may prefer another version of the same content.
  • Lack of value — If Google sees no reason for the page to rank, it won’t index it.

Who’s responsible? Almost always the site owner/content creator, because this is a quality problem. Google crawled the page but rejected it.

Crawl Budget: The Hidden Player

Both scenarios tie back to something called crawl budget: the number of pages Googlebot is willing and able to crawl on your site within a set timeframe. For small sites, it’s usually not an issue, but for large e-commerce or news sites, wasting crawl budget on duplicate or junk URLs can leave valuable pages stuck in “Discovered — not indexed.”

How to Fix These Issues

✅ For Discovered — not indexed:

  • Improve site speed and server response.
  • Submit clean XML sitemaps.
  • Strengthen internal linking to important pages.
  • Block unnecessary or duplicate URLs.

✅ For Crawled — not indexed:

  • Improve content depth and uniqueness.
  • Eliminate thin/duplicate pages.
  • Add value with multimedia, FAQs, and better structure.
  • Earn backlinks to signal importance.

Key Takeaway

  • Discovered — not indexed = Google hasn’t crawled it yet (crawl/authority issue).
  • Crawled — not indexed = Google crawled but rejected (quality issue).

As SEOs and marketers, our job is to make crawling efficient and content valuable. Only then will Google decide our pages deserve a place in the index — and in front of users.

Have you noticed these statuses on your site? What’s your go-to strategy to move them into “Indexed” territory?

Follow Qaushik Labs for more such information.


메타데이터
post_id
06bcb2f375e9
slug
discovered-vs-crawled-but-not-indexed-understanding-googles-indexing-behavior-06bcb2f375e9
url
https://medium.com/@info_19469/discovered-vs-crawled-but-not-indexed-understanding-googles-indexing-behavior-06bcb2f375e9
canonical_url
https://medium.com/@info_19469/discovered-vs-crawled-but-not-indexed-understanding-googles-indexing-behavior-06bcb2f375e9
author_url
https://medium.com/@info_19469
status
ok
fetched_at
2026-07-22 21:40:45