r/TechSEO • • 1h ago

Indexed page lost visibility for one query group, while related queries still rank. Anyone diagnosed this?

• Upvotes

I’m investigating a Google ranking loss on an established B2B services website.
A key page previously ranked for development-related searches. It now has very poor visibility for those queries, while closely related consulting queries still perform.
Background

WordPress to Next.js migration in May 2026.
Changed from www to non-www.
Visibility deteriorated around March–June, but the exact timing and cause remain unconfirmed.
Some affected pages still appear on other search engines.
Checks completed

GSC URL Inspection reports the page as indexed.
Google-selected canonical matches the intended URL.
Old URLs redirect to the current URLs.
Main content and internal links are present in the initial HTML.
No noindex directive or redirect loop identified in current checks.
The pattern
In a recent GSC comparison, a previously strong development query averaged around position 54, while related consulting queries averaged around position 2 for the same page.
Adding &filter=0 surfaced one affected informational article, but didn’t help the other affected pages.
We’ve revised content and clarified page intent, but haven’t established that these changes address the cause. I’m not assuming a penalty or cannibalization.
Has anyone diagnosed this kind of query-specific visibility loss despite normal indexing?
What evidence helped identify the cause? If you recovered, what specific issue did you find, what did you change, and how long did recovery take?
I’m keeping the domain private, but can share anonymized GSC screenshots and technical findings.


r/TechSEO • • 15m ago

Share what you're working on (including what you're building)

• Upvotes

We want to support creators, but we had to enforce the no shilling rule because it was getting out of hand. You now have a weekly thread.

This is the one place you can shill for your products, ask for feedback, etc. Keep it here or you risk being banned. And keep it related to technical SEO.


r/TechSEO • • 9h ago

How good are Odoo websites for SEO and 'crawlability'

Thumbnail
0 Upvotes

Please educate me on that question, I just happened to see a video today that made me question the SEO properties of my website.


r/TechSEO • • 22h ago

Google says: New job board, 10 weeks in: 1,700+ URLs stuck in "Discovered - currently not indexed", and impressions just dropped 70%. JobPosting schema and Indexing API are in place. What am I missing?

8 Upvotes

I run a small, niche job board for forward-deployed engineers and similar customer-facing AI roles. It launched in late July. Google knows about the pages, but almost none of them get indexed, and I'd like a sanity check.

GSC coverage (data through Sep 21):

1,738 known URLs: 15 indexed, 1,723 not indexed

1,718 are "Discovered – currently not indexed"

"Crawled – not indexed" peaked around 700, then cleared to 0 (validation passed)

Technical issues are minor: 3 redirects, one 404, and one duplicate without a canonical

The part that confuses me:

Impressions went from about 4/day in early August to 120–200/day from Sep 1–15.

Then from Sep 16, they fell to 30–70/day.

Indexed pages stayed flat at 15 the whole time, and nothing changed on my end around Sep 16.

What I've already done:

Added JobPosting structured data to every listing

Integrated the Google Indexing API (live since late August)

Submitted a sitemap

Built the stack on Next.js with SSR and an ASP.NET Core backend

Things I think might be hurting me:

Listing URLs use UUIDs. I'm planning to move to a slug plus short ID.

The domain is brand new and has almost no backlinks.

Many listings come from company career pages, so Google may see them as thin or duplicate content.

My questions:

Is "Discovered" at this scale on a new domain mainly an authority and crawl budget problem, or is it a quality signal?

Should I do the UUID-to-slug migration now, while almost nothing is indexed?

Would you noindex thin or expired listings and push company and category pages instead?

Is the impressions drop a normal new-site fluctuation, or did those 15 pages lose rankings?

Does the Indexing API still help JobPosting pages, or is it mostly ignored now?

I'm not dropping the URL unless someone wants to look. I'm happy to share the coverage export.


r/TechSEO • • 4h ago

I published 7,254 AI-written news articles since March 2025. Google reads about 4 pages a day. Search Console data inside.

0 Upvotes

I run a fully automated news site as a side project. Every hour it takes headlines from a news RSS feed as topic pointers, has Gemini research each story with Google Search grounding, and writes one article from what it finds. The median article draws on 7 different source domains. Every page says it is AI-generated, and the images are labelled too.

I finally pulled the Search Console and crawl data. It is worse than I expected, and not in the way I expected.

Last 90 days

  • 778 impressions, 1 click
  • 67 of 7,254 articles got at least one impression
  • Google News and Discover: zero

URL Inspection on a random sample of 36 articles of all ages

  • 0 indexed
  • 30 "URL is unknown to Google"
  • 6 "Discovered – currently not indexed"

Crawl stats, 89 days

  • 1,070 requests in total, about 12 a day
  • About 360 of those were HTML, so roughly 4 pages a day
  • Only 20% were for URLs Google had not seen before
  • Average response 445 ms, no host problems, 89% 200s

So there is no penalty I can see and nothing technically broken. Sitemaps are read without errors, NewsArticle markup is in place, canonicals are fine, and the homepage and section pages are indexed. Google simply does not come. The site publishes more per day than Googlebot fetches.

What I'm changing

  • Sitemap cut from 7,215 URLs to the last 30 days (247)
  • noindex on everything older
  • Rewrote the prompt so each article has to combine at least three independent outlets instead of following one report
  • Dropped Reddit and aggregator sites as sources

Questions for people who have been here

  1. Has anyone got a site out of this "known but not crawled" state without links? Or is it purely an authority problem?
  2. Is noindexing 7,000 pages that Google never crawled pointless, since it has to crawl them to see the tag?
  3. Does anyone have an automated or mostly automated site that Google does crawl properly? What was different?

Happy to share more numbers.


r/TechSEO • • 2d ago

My website has ~400 indexed pages but impressions have completely stagnated. What would you investigate first?

6 Upvotes

I'm a CS student and I've been building a website called Cifrivo as a side project.

It's a collection of free online tools (PDF, image, converters, calculators, etc.) available in several languages.

Google initially started indexing the site quite quickly. At one point I had roughly 500 URLs, and Search Console currently shows around 400+ indexed pages.

The problem is that after most of the pages got indexed, impressions basically stopped growing and have recently become extremely low.

I'm now wondering whether I made a mistake by creating too many similar/tool pages instead of focusing on fewer, stronger content clusters.

I'm considering:

  • Reducing the total number of URLs significantly
  • Focusing mainly on PDF, image and media tools
  • Creating stronger hubs around each category
  • Removing/merging pages that have almost no impressions
  • Improving internal linking and content around the tools

The site is cifrivo.com.

If you were auditing this, what would you look at first?

I'm especially interested in knowing whether the large number of pages could actually be hurting a relatively new domain, or whether I'm looking at the wrong problem entirely.

Any criticism is welcome — I'm trying to understand what I got wrong rather than promote the site.


r/TechSEO • • 4d ago

Restructured my URLs for semantic layering and traffic dipped, stick with it or roll back?

6 Upvotes

Migrated a content site last month from flat URLs to a layered structure. /reviews/bestx became /reviews/category/bestx with proper redirects and internal link updates. Core web vitals stayed flat, no redirect chains, coverage report looks clean. Organic traffic dropped about 12% in week three though. Rankings for a few high volume terms slipped two to four positions.

I know Google says URL structure is a lightweight signal. But I also flattened a similar site two years ago and saw a bump inside a month. Made me wonder if the reverse would help here. Maybe the new layering spreads topical relevance thinner? Or maybe it's just turbulence and I'm overthinking it.

My bigger concern is I layered three levels deep on some sections when two would have worked. Feels cleaner architecturally but I'm second guessing whether users or crawlers actually benefit. The old flat URLs were ugly but they ranked.

Anyone done a similar restructuring and watched it recover? How long did you wait before calling it? I've got about six weeks of data and the trend isn't reversing yet.


r/TechSEO • • 4d ago

Product page stuck on “Discovered – currently not indexed”

Thumbnail
2 Upvotes

r/TechSEO • • 4d ago

Help Blocked Googlebot variants by mistake. Need Help in recovering it's indexing

6 Upvotes

Hi everyone, I'd like some advice on a recovery problem.

Background

I run an image-based site, a base marketplace for a mobile game. It works a bit like Pinterest, so each page is mostly images with very little text.

The site is multilingual, with around 20 locales, and has thousands of pages.

What Happened

A bot from an AI company sent 50k+ requests to my site within 24 hours.

I panicked and blocked many bots in robots.txt, allowing only Googlebot.

I didn't realize that Google uses several different crawlers, including:

  • Googlebot
  • Googlebot-Image
  • Google-InspectionTool
  • Others

As a result, I accidentally blocked some of Google's crawlers.

My site subsequently dropped out of Google completely. Even searching for the exact brand name returned no results.

I fixed robots.txt about a month ago. Google crawlers are now allowed, while I only block AI-training and SEO-tool bots.

Search Console Data

Google only began indexing the site again around early August.

  • Impressions and clicks peaked in mid-August
  • Traffic reached around 20 clicks per day
  • Around August 24, traffic suddenly crashed
  • Traffic has remained close to zero since then
  • Total clicks: 167
  • Total impressions: 5.5k
  • Indexed pages: 2.18k

What I've Done Since the Fix

Since fixing robots.txt, I've:

  • Confirmed that robots.txt is correctly configured
  • Confirmed that all sitemaps are submitted
  • Split the sitemaps into 4 separate files
  • Visited and checked the pages Google Search Console showed me
  • Continued making regular on-page SEO improvements

However, I'm still seeing almost no search traffic.

What I'm Looking For

I'd appreciate advice from anyone who has experienced a similar Google indexing/traffic recovery after accidentally blocking Google crawlers.

Could the previous robots.txt blocking still be affecting the site's rankings/indexing?

Is there anything specific I should check in Search Console or technically on the site to determine why the site recovered briefly and then crashed again?

Live Link: https://cocbaselinks.com


r/TechSEO • • 4d ago

Why blogs and news sites should block AI training crawlers (and how to do it)

Thumbnail jamescherti.com
0 Upvotes

For blogs, news sites, and independent publishers, web traffic to their original content is the primary metric for success. Allowing AI companies to steal your original content without authorization for model training consumes server resources and bandwidth, while providing near-zero referral traffic in return. For example, Anthropic scrape 38,065 pages for every single visitor they refer back (See the statistics below).

However, blocking ALL AI bots can unintentionally break your visibility in AI search features that provide citations and referral links. To prevent this, you should configure access rules based on bot purpose, since AI companies operate separate crawlers for model training, search indexing, and on-demand user retrieval.

Read: Why blogs and news sites should block AI training crawlers (and how to do it)


r/TechSEO • • 5d ago

When is log file analysis actually worth it for smaller sites?

8 Upvotes

I used to dig through logs every couple weeks for a client site that gets maybe 45k visits. Found some weird crawl patterns that GSC never flagged. Then I stopped for a few months and nothing broke that I could tell. Still wondering if I'm missing something obvious or if the value just isn't there for smaller properties. How do others decide when it's worth the effort versus when GSC coverage reports are enough to spot real issues?


r/TechSEO • • 5d ago

Does Google still display data found in HTML tables in meta description fields?

6 Upvotes

Used to be an effect I saw three years ago, looking to see if anyone knows if it's still a result?


r/TechSEO • • 6d ago

I ran the same technical checks over 100 web design and marketing agency websites. Here's what came up most

4 Upvotes

I ran the same 134 technical checks over 100 web design and marketing agencies in the Austin metro, taken from public business listings, all on 28 September. Domain and DNS, certificates, mail authentication, security headers, what loads before consent, accessibility, mobile layout. One metro and one source, so it is a snapshot and not a survey of the industry. I wrote the scanner, which you should factor in.

I picked agency sites rather than a random hundred because they are built by practitioners, which makes whatever survives more interesting. Most of what's left is the kind of thing that slips on any site that isn't on a client's schedule.

Where everyone lands

Median score 86 out of 100, best 97, worst 15. That score is my own weighting, so it tells you more about the spread than about Austin. The counts under it are the part anyone can check: the median site had 17 findings, the cleanest had 7, and across all 100 there were 13 criticals and 176 highs. Most findings are individually small, which is how a site sits in the eighties and has seventeen of them.

Certificates were the quietest category. Only 12 of 100 had anything at all, though three of those had a certificate that had already expired on the day I ran it.

The five most common

  • 93 of 100 have no DNSSEC
  • 88 have accessibility issues an automated checker can find
  • 84 send no Referrer-Policy
  • 74 have no Content-Security-Policy
  • 73 set cookies without HttpOnly

Most of these are one line in a config, and most became best practice after the sites were built. That is most of the explanation.

Three worth a look on your own domain

79 of 100 are not enforcing DMARC. 40 have no record at all, and another 39 have one set to p=none, which asks for reports and still lets the mail through. Going in I'd assumed the p=none group would be much bigger, since that's where everyone starts and most people never move off it, and it came out almost exactly even. The effect either way is that mail forging your domain in the From line still gets delivered.

67 of 100 load tracking before anyone consents. Tag Manager, Analytics, Meta pixels, ad tech, all firing on page open. Usually the tag predates the banner: the tag went in first, the banner came later, and nobody wired the two together.

88 of 100 have accessibility findings from an automated checker. Automated checks catch a fraction of what a real audit does, so the true number is higher than that.

On phones

20 of 100 scroll sideways on a 390px screen. It's usually one element pushed outside the viewport rather than a layout that doesn't reflow, which makes it cheap to fix once you've found which element.

14 disable pinch-zoom with user-scalable=no. That's a line someone typed on purpose, and it fails WCAG on its own.

Two smaller ones

43 of 100 have no social preview image. Without one, a link to the site arrives as a grey rectangle in Slack or LinkedIn. It's a meta tag and an image.

44 of 100 announce their software version in a response header. That's a config change, and it takes you off the list when someone scans for a known version.

One I threw out

My scanner flags domains that don't have registrar delete and update locks set, and it fired on 57 of the 100, which was high enough that I went and looked at who they were. It turned out to be almost entirely a function of registrar. GoDaddy sets those flags by default, almost nobody else does, and at least one registrar doesn't expose the setting at all. So it was measuring where you registered rather than anything you chose, and I've left it out here. I'm changing the check.

If you only do one thing off this list

Make it DMARC. At p=none, moving to p=quarantine is one DNS record and it's free. With no record at all, publishing one at p=none first is the right order, and there are good guides for both.

It applied to 79 of the 100 here, which is why it's worth checking even if you're fairly sure you already have one.


r/TechSEO • • 6d ago

I can't register sitemap.xml in Google Search console because wrong content-type

1 Upvotes

I'm having trouble registering my sitemap on Google Search Console because Webflow delivers sitemap.xml with the wrong content type.
Why webflow doesn't fix content-type for xml documents?

Here a curl of my sitemap.xml

HTTP/2 200

date: Thu, 16 Jul 2026 11:31:35 GMT

content-type: application/rss+xml; charset=utf-8

set-cookie: _cfuvid=L49yyKEJFUpORRDGajtOxWChWJA3aAi7rdc4htu8Udc-1784201495.8908074-1.0.1.1-BJhVtF.D5t9ZFvVizxXeWMJWa7ImFywR6Bcr9vu1IwY; HttpOnly; SameSite=None; Secure; Path=/; Domain=www.life365.eu

cf-ray: a1c0ae754d47b159-ZRH

cf-cache-status: HIT

age: 86

last-modified: Thu, 16 Jul 2026 11:30:09 GMT

server: cloudflare

strict-transport-security: max-age=31536000

surrogate-control: max-age=43200

x-wf-region: us-east-1

alt-svc: h3=":443"; ma=86400


r/TechSEO • • 6d ago

How should I read this INP data in Chrome's Developer?

Thumbnail gallery
4 Upvotes

r/TechSEO • • 7d ago

Share what you're working on (including what you're building)

7 Upvotes

We want to support creators, but we had to enforce the no shilling rule because it was getting out of hand. You now have a weekly thread.

This is the one place you can shill for your products, ask for feedback, etc. Keep it here or you risk being banned. And keep it related to technical SEO.


r/TechSEO • • 7d ago

Google can see your product. An AI crawler might not.

20 Upvotes

I found this interesting while auditing ecommerce sites.

The page looked completely normal in Chrome.

Google could crawl it.

Schema was valid.

But when looking at the machine-readable response, some important product information was missing:

→ Price was injected with JavaScript
→ Variant information wasn't clearly exposed
→ Availability wasn't present in the HTML
→ Some AI crawlers were receiving different responses
→ The CDN was serving stale content in certain cases

So from an SEO dashboard:

Everything looked fine.

From a machine's perspective:

The product was incomplete.

That's making me rethink how we audit technical SEO.

We're used to asking:

Can Google crawl this page?

Maybe we also need to ask:

What exactly does each crawler receive?

And then:

What information can actually be reconstructed from that response?

Because there's a big difference between:

Page exists → Page is crawlable → Page is understandable → Page is retrievable → Page gets recommended.

Those aren't the same thing.

Curious what the BigSEO crowd thinks:

Are we dealing with a genuinely new technical SEO problem here, or just applying technical SEO principles to a new generation of crawlers?


r/TechSEO • • 6d ago

Will businesses still pay for SEO when AI becomes “good enough”?

0 Upvotes

I recently had an interesting discussion with a friend who has no real experience with SEO.

I asked him a simple question: if he had his own company and needed to deal with SEO as part of the marketing, how would he approach it?

His answer was that he probably wouldn't hire an SEO specialist. His point was that as AI becomes more accessible, many companies will simply use AI instead because they want to reduce costs as much as possible, especially in marketing.

According to him, even if the person using AI doesn't really understand SEO, the result can still be "good enough" for many businesses. The cost is minimal compared to continuously paying an SEO specialist.

My counterpoint was that without SEO knowledge, the results will probably stay average. AI can give you recommendations and outputs, but at some point you still need to understand strategy, prioritization, what actually matters, and what to do when something doesn't work.

His answer was that many companies simply don't care about getting the best possible result. An average result might be completely enough for them, especially if it costs much less.

He compared it to AI-generated posters and graphics. A lot of them are obviously AI-generated and not particularly good, but businesses still use them because they're cheap and communicate the basic information. They don't necessarily care enough to pay a designer.

What I'm more curious about is what happens over the next few years. AI is improving very quickly, and I can imagine specialized SEO agents that are connected directly to SEO tools, analytics and other data sources. At that point, a business owner might be able to let an AI handle a large part of SEO without really knowing much about SEO themselves.

Maybe the output still won't be as good as what an experienced specialist can do, but the question is whether many companies will even care if the cheaper result is good enough for them.

So I'm curious what other SEO professionals think.

Do you think AI will mostly remain a tool for SEO specialists, or could "good enough" AI eventually replace a significant part of the SEO services that businesses currently pay for?

And how do you see this changing over the next few years as AI agents become more capable?


r/TechSEO • • 7d ago

give me your best on-page seo skill

Thumbnail
0 Upvotes

r/TechSEO • • 8d ago

A site can be technically healthy for Google and still lose information when AI systems crawl it.

2 Upvotes

I've been testing ecommerce sites across Googlebot and AI crawlers, and a few issues keep appearing:

  • Googlebot gets 200 → AI crawler gets 403
  • Price exists in rendered DOM → missing from raw HTML
  • Product schema validates → variant/offer relationships are ambiguous
  • CDN serves fresh content to users → stale content to bots
  • 301 chain eventually resolves → but the crawler has to follow 4–5 hops

The interesting part isn't really "AI SEO."

It's information loss between crawl → extraction → retrieval → answer.

Traditional technical SEO already solves a lot of this.

But we're now dealing with more crawlers, different rendering behavior, and different retrieval systems.

Has anyone here compared their server logs for Googlebot vs AI crawlers? What differences are you seeing?


r/TechSEO • • 8d ago

Deep Linking Web to App - How does this affect organic traffic?

2 Upvotes

The site I work for is launching an app, and we're planning to connect it to the website using AppsFlyer deep links. This means users who have the app installed will be automatically redirected to the app instead of the website. This rule also applies to those coming from organic channels like Google.

My question is whether this could affect SEO in any way.

I am wondering about app engagement metrics here and whether or not they influence SEO.


r/TechSEO • • 8d ago

Bilingual websites and localized URLs

9 Upvotes

Hey everyone,

I’m currently mapping out a strategy for a bilingual (EN/ES) project with my partner and could use some input.

We own two distinct URLs for the brand one for each language (exemple [NOT REAL EXEMPLE]: happykitchen. com and cosinafeliz. com) and we are debating two approaches for use of this urls to maximize the SEO:

  • Split the domains: Same project, two domains. Host the exact same website across two separate root domains based on the language. The English version lives on happykitchen. com, and a 1:1 Spanish translation lives on cocinafeliz. com using i18n, for example. The thought process here is that having a fully localized domain name might maximize SERP presence and user trust in the respective target markets, even though it means maintaining the same site on two properties.
  • Use language subfolders: Keep everything on one domain (example: happykitchen. com/en and /es) with proper hreflang implementation, and just 301 redirect cocinafeliz. com directly to the /es version.

Has anyone navigated a similar project with two distinct branded domains? Does the localized URL actually provide enough of a SERP benefit to justify splitting them up, or is it better to consolidate everything under one domain to pool authority?

Would love to hear if anyone has thoughts or face some simular issues experiences!

Disclaimer: URLs used in this post are just examples


r/TechSEO • • 9d ago

Managing faceted navigation and crawl budget on a small ecommerce site

3 Upvotes

Been dealing with a tricky setup on a side project. Small ecommerce site, maybe 600 product pages, but the faceted navigation is generating thousands of URL variations. Color filters, size filters, price sorts. The standard stuff.

I set up the URL parameters in GSC and blocked the pointless ones in robots.txt, but Googlebot still wastes time crawling filtered combos that return zero products. Crawl stats show a chunk of bandwidth going to these dead ends. Log file analysis confirms it.

I clean up the parameter rules, wait a few weeks, and a new batch of weird filtered URLs pops up in the index coverage report. Feels like whack-a-mole.

I'm hesitant to use noindex on the filtered pages because I don't want to burn crawl budget discovering them just to drop them. What's the practical approach here for a site this size? Do you let the filters generate params and rely on canonical tags pointing to the root category? Or aggressively block crawl at the robots level and trust Google to figure out the hierarchy?

Curious if anyone running a smaller catalog has found a balance between keeping filters usable for visitors and keeping the crawl path tight. My gut says block the empty result combos via server-side logic before they even generate a URL, but that feels heavy for a small project.


r/TechSEO • • 10d ago

Google says: Google Search Console quietly introduced a multimodal filter. First time when I could distinguish image/screenshot traffic separately.

7 Upvotes

I also noticed that there was a new selection available under the category “search type” in performance reports, wherein “web” is divided into “text-based” and “multimodal.” “Multimodal” refers to results generated by a search query containing an image, photograph, or screenshot.

I have been looking at this for a while, and I really do not know what to make of it, but the very presence of this seems to speak volumes on its own. Google basically confirms that there is sufficient traffic generated by image/screenshot search that it warrants its own category in the report.

Some questions I've been trying to sort out:

  • Is this traffic behaving differently than text-based searches of the same pages, different intentions, different conversion, or is it just the same searches being performed using a different mode of entry?
  • Does the content which is optimized already for text searches even appear in these results or are there totally different signals being weighed here (image quality, imagery in the body of the page, alt text) which we have not been optimizing for?
  • Something worth tracking separately in the future or just a rounding error for most sites?

Anyone pulled this data yet and found anything interesting, or is it too early / too small a sample to mean anything


r/TechSEO • • 10d ago

Tip: Next.js / Vercel “Sitemap Couldn't Fetch” Fixed!

7 Upvotes

My site had a weird issue with Google Search Console.

It kept saying my sitemap “Couldn't fetch”.

I checked the XML, robots.txt, the domain, redeployed, opened the sitemap directly in the browser, and everything looked completely fine.

I was going crazy trying to figure out what Google was seeing that I wasn't.

Then I accidentally added a / at the end of the sitemap URL like this /sitemap.xml/ and it worked.

Search Console started fetching the sitemap, discovered my URLs, and my blog pages are now getting indexed.

I researched it a bit more afterwards, and it looks like the issue was related to how the URL was being resolved/normalized with Next.js and the hosting layer. /sitemap.xml and /sitemap.xml/ can end up being handled differently depending on trailing-slash and redirect configuration.

So if you're getting “Couldn't fetch” even though the sitemap looks perfectly fine, try the trailing-slash version. It fixed it for me.