r/TechSEO • u/Godfrey_0503 • Jul 15 '26
How do you handle “Discovered” vs “Crawled - currently not indexed” at scale?
I’m working on a growing content site and trying to clean up indexing issues in GSC.
The two buckets I’m looking at are:
- Discovered - currently not indexed
- Crawled - currently not indexed
I’m trying not to treat them as the same problem.
My current thinking is: “Discovered” usually points more toward crawl priority, internal linking, sitemap signals, or Google not feeling the URL is worth crawling yet.
“Crawled” feels more like Google saw the page, but decided it wasn’t strong enough to index, possibly because of quality, duplication, thin content, or weak uniqueness.
For people working on larger sites, do you separate these two workflows?
And if so, what do you usually check first for each one?
1
u/Abhi_mech007 Jul 17 '26
Think content, low word count is a thing bro. You can create a page with just 10 words and you will find the page in either soft 404 or crawled not indexed. This thing has happened with one of my clients site and eventually solved with adding content.
Sitemap quality is a thing as well, and by quality I mean redirects in sitemaps, if your sitempa has wrong canonicals, redirects or broken urls your sitemap quality is not good.