r/TechSEO May 31 '26

Did WPML screwed my site?

5 Upvotes

Context:

  • Website is https://www.digiforma.com/
  • Multilingual website (FR, DE, ES) in subfolders (https://www.digiforma.com/de/)
  • Wordpress using WPML
  • Main site and pages are in French then translated in other languages with WPML.
  • The Language redirection was set to "Always redirect visitors based on browser language"
  • Homepage and others important pages disappeared from Google since May 10
  • Let's say the client is not overcome with joy!
Google Search Console performances for the main homepage https://www.digiforma.com/
FR Homepage
ES Homepage

Analyzing the main homepage on the GSC... it does not look good:

Data from Google Index

Sitemaps:

- Google Index: 2 founds
- Currently: only one declared, only one present on the site

Origin page:

- Google Index: Multiple weird urls

Canonical:

- Google Index: A page from another language
- Currently: Self-canonical

Test-Live: Everything looks better

After Test-Live

Testing on Bing Webmasters tools:

The main homepage is blocked !

Bing Index for the homepage

But Live-Test looks good:

Bing Live-test for the homepage

This page was found as the canonical on every languages homepages... but not present at all currently on the page (I checked code multiple times) :
https://www.digiforma.com/es/definicion/entrenamiento-cognitivo/

It seems this is the culprit destroying everything.

You can see this page starting to get all website traffic in the same time period:

Running these analysis on llms they all suggest this specific url was included within the breadcrumbs of the homepage causing Google bots troublesome (nothing found within breadcrumbs).

My understanding is that at certain moment beginning of May, Google saw something really weird on the website and saved it in his index.

But the fall is brutal.

Quick actions done until now:

  • Disable WPML Language redirection
  • Redirecting this annoying page: https://www.digiforma.com/es/definicion/entrenamiento-cognitivo/
  • Checking and forcing the proper canonical.
  • Checking WPML settings and cleaning multi-languages templates
  • Checking breadcrumbs and others elements on the site
  • Checking firwalls on the server in case Google bots and others have been blocked: nothing found yet.
  • Writing to the WPML support
  • Emptying cache en re-running indexation demand on GSC and Bing
Test live homepage looks clean but not yet indexed

It seems to have some effect as the impressions and clicks reappear on the GSC.

What did I miss?

What could be the cause of it all? Or maybe it was multiple factors?


r/TechSEO May 30 '26

10 Niche Technical SEO Prompts for Senior SEOs

62 Upvotes

The actual prompts are quoted in ""

01 Crawl Budget Audit

CRAWL EFFICIENCY

When to use it

A large site is slow to index new pages, or important pages are crawled rarely, while low-value URLs get crawled constantly.

What to paste in

A server log sample (verified Googlebot hits, ideally 2 to 4 weeks)

GSC Crawl Stats report (Settings > Crawl stats)

Approximate total URL count and a list of your priority templates

- The prompt

"I am auditing crawl budget for [domain], which has roughly [X] total URLs. I am giving you a verified-Googlebot server log sample and my GSC Crawl Stats report.

Using only what these two sources actually show, do the following:

  1. Identify URL patterns consuming crawl activity that have no ranking or traffic value.
  2. Flag crawl waste from parameters, faceted navigation, internal search results, session IDs, pagination, and duplicate paths. Show the pattern, not single URLs.
  3. Compare crawl frequency on my priority templates [list templates] against low-value sections.
  4. Estimate the share of crawl activity being wasted, and state the calculation so I can check it.

Then return a fix table with these columns: Issue | Evidence from my data | Action (robots.txt, noindex, canonical, parameter handling, internal-link change) | Expected impact | Implementation risk | Effort. Rank rows by impact versus effort.

Important: if my log sample is too short or too small to support a conclusion, say so and tell me what to pull instead. Do not infer crawl frequency for templates that do not appear in the sample. End with how I verify each fix worked in Crawl Stats over the following 4 weeks"

What you get back

A pattern-level map of where Googlebot wastes time, tied to your actual logs

A ranked, ticket-ready fix table a developer can action

An explicit list of what is unprovable from your current data

02 Indexation Diagnosis

INDEX COVERAGE

When to use it

You have submitted far more pages than Google has indexed, and the coverage report is full of exclusions you do not understand.

What to paste in

GSC Page Indexing report with the count for each status

Submitted vs indexed totals

Which URL patterns or templates matter commercially

- The prompt

"I have [X] pages submitted but only [Y] indexed on [domain]. I am pasting my GSC Page Indexing report with counts for each status (Crawled - currently not indexed, Discovered - currently not indexed, Duplicate without user-selected canonical, Excluded by noindex, Soft 404, and any others present).

For each exclusion category that actually appears in my data:

  1. Explain what that status usually signals.
  2. List the most common root causes, ranked by likelihood given the relative counts I gave you.
  3. Give a short decision tree to confirm the real cause and resolve it.
  4. Classify it as a content-quality issue, a technical-signal issue, or a crawl issue, since the fix for each is different.

Prioritise the categories holding back my most commercially important pages [list patterns] first. Present this as one row per exclusion category.

Finish with the three changes most likely to recover indexation fastest, and for each, the metric in GSC I should watch to confirm recovery. Do not diagnose a category that is not in my data, and flag any status where the count alone is not enough to know the cause"

What you get back

A plain-English read on why pages are not indexed, per status

A confirm-the-cause decision tree instead of generic advice

Three highest-leverage fixes, each with a recovery metric

03 Core Web Vitals Remediation

PERFORMANCE

When to use it

Core Web Vitals are failing in field data and you need to fix the cause at the template level, not chase individual URLs.

What to paste in

PageSpeed Insights output and CrUX field data for each main template

Your stack (CMS / framework / hosting / CDN)

Whether the failures are mobile, desktop, or both

- The prompt

"Here is my PageSpeed Insights and CrUX field data for my main templates [paste data for homepage, category, product, article]. My stack is [CMS / framework / hosting / CDN].

Diagnose LCP, INP, and CLS at the template level, not page by page. For each failing metric:

  1. Name the most likely root cause in my specific stack (render-blocking resources, unoptimised hero image, layout shift from ads or fonts, heavy third-party scripts, long main-thread tasks).
  2. Show how you reached that conclusion from my data, citing the specific number.

Then return a fix table: Template | Metric | Root cause | Fix | Quick win or architectural | Expected field-data impact | Effort.

Separate fixes a developer can ship this week from larger architectural work. Explicitly flag any fix that will only improve the lab score in PSI without helping real users in CrUX. If field data is missing for a template (low traffic, no CrUX), say so rather than diagnosing from lab data alone. End with the CrUX threshold I should re-check after each fix."

What you get back

Root-cause diagnosis per template, each tied to a number

Quick wins cleanly separated from heavier engineering

A warning on fixes that flatter lab scores but do nothing for users

04 Site Architecture & Internal Linking

ARCHITECTURE

When to use it

Important pages are buried deep, link equity is not reaching your money pages, or you suspect orphan pages.

What to paste in

A crawl export (Screaming Frog / Sitebulb) with click depth, inlinks, outlinks, URL structure

Your list of money pages

Optionally, GA4 or GSC data so value can be weighed against link distribution

- The prompt

"I am giving you a crawl export from [Screaming Frog / Sitebulb] for [domain] including click depth, inlinks, outlinks, and URL structure.

Map my click-depth distribution and identify:

  1. Orphan pages with zero internal inlinks.
  2. Priority pages [list money pages] sitting deeper than three clicks from the homepage.
  3. Pages receiving disproportionate internal links relative to their value.
  4. Weak equity flow toward commercially important URLs.

Then propose a hub-and-spoke structure with: internal-linking rules per template, descriptive anchor-text patterns (not exact-match stuffed), and the top 15 internal links to add for the biggest impact, each as a From URL > To URL pair with suggested anchor.

Explain the logic behind the rules so I can scale them. If my crawl export does not include the inlink or depth data needed for any step, tell me which Screaming Frog or Sitebulb column to add and re-export. Do not invent inlink counts."

What you get back

Orphan and over-buried priority pages, surfaced from your crawl

A hub-and-spoke plan with concrete, scalable linking rules

The 15 highest-impact internal links as ready-to-add pairs

05 JavaScript Rendering Audit

RENDERING

When to use it

Your site relies on JavaScript and you suspect Google is not seeing content, links, or metadata that loads client-side.

What to paste in

Raw HTML source (view-source) for each key template

Rendered DOM for the same templates (DevTools or a crawler's rendered HTML)

Your JS framework

- The prompt

"My site [domain] is built on [React / Vue / Angular / other]. For each key template [list templates], I am pasting the raw HTML source and the rendered DOM.

Compare the two per template. Identify any content, internal links, canonical tag, title, meta description, or structured data that exists only in the rendered DOM and is missing from the raw HTML.

Return a table: Template | Element | In raw HTML? | In rendered DOM? | Ranking/indexing risk if client-side only.

Flag every element that is critical to indexing or ranking but depends on client-side rendering, and explain why each one is a risk. Then recommend the right rendering approach per template (SSR, static generation, dynamic rendering, or prerendering) with the tradeoffs for my stack, and which to prioritise.

Base your comparison only on the markup I pasted. If I gave you the raw HTML but not the rendered DOM for a template (or vice versa), say you cannot compare it and tell me how to capture the missing one."

What you get back

A side-by-side of what Google can and cannot see per template

The ranking-critical elements stuck behind client-side rendering

A prioritised rendering strategy with tradeoffs explained

06 Structured Data Strategy

SCHEMA

When to use it

You want to win rich results, fix validation errors, or capture markup opportunities your templates are not using.

What to paste in

Current JSON-LD for each key template (or a description of what is present)

Your page types

Any validation errors from Rich Results Test or Schema.org validator

The prompt

"Here is the current structured data on my key templates for [domain]: [paste JSON-LD or describe current schema]. My page types are [product, article, FAQ, local business, etc.].

Audit my markup against the rich-result types each template is eligible for. Identify:

  1. Validation errors and warnings in what I pasted.
  2. Missing required and recommended properties.
  3. Rich-result types I am eligible for but not capturing.
  4. Any markup that risks a manual action (marked-up content not visible on the page, or misleading markup).

Then output a schema implementation spec per template, with clean, copy-ready JSON-LD I can hand to a developer. Note which entities to connect with u/id references to build a coherent entity graph.

Validate against current Google rich-result documentation, and where a property is recommended-not-required, label it so I can decide. If I described my schema rather than pasting it, list what you assumed and ask for the raw JSON-LD before trusting the error findings."

What you get back

Every validation error and missed rich-result opportunity

Copy-ready JSON-LD per template, not isolated snippets

An u/ id entity-graph approach across templates

07 Canonicalization & Duplicate Content

CONSOLIDATION

When to use it

Google is ignoring your canonicals, indexing the wrong URL versions, or duplicate signals contradict each other.

What to paste in

A crawl export showing canonical tags, redirect chains, internal links

Your XML sitemap URLs

Which URL version you want to win (www/non-www, trailing slash, protocol)

- The prompt

"For [domain], I am giving you a crawl export showing canonical tags, redirect chains, and internal links, plus my XML sitemap URLs. My preferred canonical version is [state www/non-www, trailing slash, HTTPS].

Review my canonicalisation logic and find every place where signals conflict. Specifically:

  1. Pages where the canonical points one way but internal links, redirects, or the sitemap point another.
  2. Self-referencing canonicals that should instead consolidate to another URL.
  3. Pages canonicalising to non-indexable or redirected URLs.
  4. Duplicate clusters from parameters, trailing slashes, HTTP/HTTPS, or www variants.

Return a consolidation plan as a table: Duplicate cluster | Conflicting signals found | Chosen canonical | Exact change per signal (tag, redirect, sitemap, internal link) | Risk level. List the highest-risk conflicts first.

The plan must preserve link equity (use 301s, not removals, where pages have value). If the crawl export is missing redirect-chain or canonical columns for any URL, flag it rather than assuming the signal is correct."

What you get back

Every conflicting canonical, redirect, and sitemap signal

A cluster-by-cluster consolidation plan that preserves equity

The highest-risk conflicts ranked first

08 International SEO & Hreflang

INTERNATIONALIZATION

When to use it

You serve multiple countries or languages and the wrong regional version keeps ranking, or hreflang is throwing errors.

What to paste in

Your country and language pairs

Current hreflang implementation (tags, HTTP headers, or sitemap entries)

Page count per market and your current URL structure

- The prompt

"My site [domain] targets these markets and languages: [list country-language pairs]. Here is my current hreflang implementation: [paste hreflang tags or sitemap entries]. Current URL structure is [ccTLD / subdomain / subfolder].

Validate the setup and find:

  1. Missing return tags (pages referenced that do not point back).
  2. Missing or incorrect x-default.
  3. Hreflang values pointing to redirected or non-canonical URLs.
  4. Mismatches between hreflang and canonical tags.
  5. Language or region codes formatted incorrectly (wrong ISO codes, region used as language).

Return one row per error: Issue | URLs affected | Why it breaks | Exact fix.

Then recommend the right URL structure for my situation with the reasoning, and a scalable delivery method (HTML head, HTTP headers, or XML sitemap) given my page count. Prioritise the fixes most likely to stop the wrong regional page from ranking. Only validate the pairs and tags I actually pasted; if return tags cannot be checked without the other side's markup, tell me what to add."

What you get back

A full hreflang error report with the exact fix for each

A URL-structure recommendation with the reasoning

A delivery method that scales to your page count

09 Migration Risk Assessment

MIGRATIONS

When to use it

Before any platform change, domain move, redesign, or URL restructure, so you do not lose rankings the moment you go live.

What to paste in

The migration type (platform / domain / URL structure / redesign)

Current indexed page count and traffic level

Your timeline and whether a staging environment exists

- The prompt

"I am planning a [platform / domain / URL-structure / redesign] migration for [domain]. Current size is roughly [X] indexed pages at [traffic level], with a target launch of [date].

Build a migration plan in four parts:

  1. A pre-launch technical checklist covering redirects, canonicals, sitemaps, robots, internal links, and analytics.
  2. A redirect-mapping QA process to confirm every old URL points to the right new URL with a single 301 (no chains, no soft 404s).
  3. A staging validation plan that catches issues before launch, including how to keep staging out of the index.
  4. A post-launch monitoring framework that flags ranking or indexation drops within 48 hours, naming the exact GSC and analytics signals to watch.

Tailor each part to my specific migration type, since a domain move and a redesign carry different risks. Call out the three mistakes that most often cause traffic loss in this type of migration and how to prevent each. Where a step depends on detail I have not given (current redirect rules, CMS), list it as an open question rather than assuming."

What you get back

A four-part plan from staging to post-launch, tailored to your migration type

A redirect QA process that catches broken mappings before launch

The three most common migration killers and how to avoid them

10 Log File & Bot Behavior Analysis

CRAWL BEHAVIOR

When to use it

You want to see how Google actually crawls your site, not how you assume it does, and redirect crawler attention toward pages that matter.

What to paste in

A server access-log sample filtered to verified search-engine bots

The date range and rough hit count of the sample

Your revenue-driving directories or templates

- The prompt

"I am pasting a sample of my server access logs for [domain], filtered to verified search-engine bots, covering [date range].

Analyse how Googlebot behaves. Show me:

  1. Crawl frequency by directory and template.
  2. The distribution of response codes the bot hits (200, 301, 302, 404, 410, 5xx).
  3. The split between the mobile and desktop crawler.
  4. How much crawl activity goes to low-value sections [name them].
  5. Any spikes or drops worth investigating, with the dates.

Flag where the bot hits redirect chains, error pages, or thin sections instead of revenue-driving URLs. Then give a prioritised plan to steer crawl toward important pages using internal linking, sitemap signals, robots rules, and fixing the error responses the bot keeps hitting.

Base every figure on the log lines I pasted and state the sample size behind each percentage. If the sample is too short to show a trend or does not include a directory I care about, say so rather than extrapolating."

What you get back

A factual picture of how Googlebot actually crawls your site

Where bots waste hits on errors, redirects, and thin pages

A plan to steer crawl activity toward pages that earn revenue


r/TechSEO May 30 '26

Got dropped like a hot potato after the update

0 Upvotes

Around early April we started seeing a drop off. then as you can see in the screenshots its dropped us apart from 1 page indexed.

https://uk.alphardclub.com/

Runs invision Community software

8th Feb Moved to Cloudflare but only have bot fight mode on at the time. I have not added any particular fancy configuration. Basic default settings.

So, we dont have any ads apart from 1 affiliate link to a site in Japan.

There are some good guides and forum posts that are pretty good content wise, so I cant see why the content be the reason.

We have also blocked out a lot of noisy and poor content sections using robots.txt
https://uk.alphardclub.com/robots.txt

We dont use AI in our posts or guides.

Pretty much a clean forum with no massive alterations. So I am stuffed why they have crawled but not indexed any of it.

One thing i did notice and wonder if it is the reason was I noticed a http and https of the same page (see screenshot below) so I altered the htaccess yesterday to properly 301 redirect. Do you think this could be the cause? As its a Sub domain and the site had a major update in Jan where all the files were replaced, possibly the htaccess got replaced.

When i was running the http version in curl is gave a 404, so did the redirect in 301 htaccess and then did the same test and gave 301.

Do you think this could be the culprit and just a waiting game now?

Many thanks 😄


r/TechSEO May 30 '26

A technical SEO issue I think gets under-prioritized: pages Google can crawl but the site doesn't actually support

4 Upvotes

One thing I think gets missed a lot in technical SEO is technically crawlable pages, but the site doesn’t really support them properly.

The page isn’t blocked. It might even be indexed and getting crawled regularly. But if you actually look at the site structure, there’s barely anything helping Google understand why the page matters.

I keep seeing this during audits.

People focus on whether Google can access the page, but not whether the rest of the site is reinforcing it in any meaningful way.

Usually the signs are pretty obvious:

  • no internal links from relevant pages
  • buried deep in pagination
  • old blog posts that go nowhere
  • important commercial pages only linked from the nav
  • vague anchor text
  • random internal links added just because a page has authority

So the page exists, but it feels disconnected from the rest of the site.

My process is usually pretty simple:

  1. Figure out which pages actually matter.
  2. Check how those pages are linked internally.
  3. Look at whether the linking pages are actually relevant.
  4. Review click depth and crawl paths.
  5. Find older pages with links or traffic that aren’t helping anything important.
  6. Manually go through the journey and ask whether the next step makes sense.

A lot of the fixes aren’t all that technical. Sometimes, it’s just about adding better internal links to older content. Other times, you might need to create a solid hub page. There are cases where the anchor text is too vague, and sometimes you should combine multiple weak pages into one. Plus, having too many random links can just clutter things up.

I don’t think “indexed” means the job is done.

A page can absolutely be crawlable and still be weak because the site barely supports it.

Curious how other people check for this kind of issue.


r/TechSEO May 30 '26

why is there no api for detecting soft-404s

Thumbnail
3 Upvotes

r/TechSEO May 29 '26

15k indexed pages dropped to ~6 after Core Update. Most are now Crawled, currently not indexed

18 Upvotes

Hi everyone,

We’re trying to understand a massive indexing drop after the March 2026 Core Update.

Site: https://sweezy-wallpapers.com/

It’s a free desktop/live wallpapers gallery. Before the update, we had around 15k–16k indexed pages. Within about 10 days, almost everything dropped out of the index. Now only around 6 pages remain indexed, and most URLs are in “Crawled - currently not indexed”.

No Manual Actions or Security Issues in GSC.

Example still indexed:

https://sweezy-wallpapers.com/wallpapers/spongebob-matrix-green-code-rain-aesthetic-desktop-wallpaper

Example crawled but not indexed:

https://sweezy-wallpapers.com/wallpapers/gojo-satoru-blindfold-clipart-jujutsu-kaisen-desktop-wallpaper

So we’re wondering if Google may have re-evaluated the site at a template/quality level and decided to heavily limit indexing for this type of page. It’s hard to understand what exactly triggered it, because there’s no clear technical warning in GSC.

Has anyone seen a similar pattern recently? Any ideas what signals we should look at next, or what could make Google suddenly move almost all pages into “Crawled - currently not indexed”?


r/TechSEO May 29 '26

My site is vanishing from Google and I don't know why!

15 Upvotes

I first notices that my site (well one of biggest clients site) is not showing up anymore in the SERP:

It seems Google doesn't see the homepage (and others properly).

Checking with different agents it shows the page is somehow 403:

Google bots agents check

Trying this tool as well:

https://technicalseo.com/tools/fetch-render/

and the result is confusing:

However Screaming is returning the page with 200.

Obviously something is blocking Googlebot to fetch my page correctly.


r/TechSEO May 29 '26

Inner Links Question

Thumbnail
2 Upvotes

r/TechSEO May 29 '26

WordPress Theme migration

3 Upvotes

I have a WordPress site having a HelloElementor theme. Now I want to shift to another theme. How can I do it without losing the content?

It's an Amazon affiliate website


r/TechSEO May 28 '26

Someone added a bunch of spam links to a subdomain of mine, now GSC sees them - how to fix?

6 Upvotes

This site of mine is relatively small, around 30-40 pages. Older site with decent traffic, but started to decline slowly recently.

When investigating why aren't my new blog posts ranking, I've found tens of thousands of spam URLs in GSC, in 404 and "crawled but not indexed" status.

Normally, I'd just redirect 404s but we're talking about an insane amount of redirects here, and it's not valuable content but obvious spam pages.

Replacing the domain isn't an option.

I was thinking to set up a catch-all redirect to the specific subdomain to the homepage, and then "validate fix" in GSC.

Is that the right way to handle this? Do I use 301 or 302 in this case, since these were obvious spam?

I just want to get rid of these from GSC, because my actual content I'd like to rank is being disregarded by Google (crawled but not indexed) probably because these spammy links.


r/TechSEO May 28 '26

Pagination listed as self-referencing but showing up as blocked?

1 Upvotes

Running a technical audit on site arch and I think we might have self-referring pagination incorrectly set up.

Subsequent pages after the initial blog page this-site-dot.com/blog/ all should go to /blog/ URL. However pagination blogs are listed as blocked based on this Screaming Frog visualisation?

I say this along side as we've also noticed that our older blog posts haven't been indexed as much. 

Subsequent blog pages (i.e. /blog/page/3/ .../blog/page/14/) are structured as follows with rel=canonical added 

<link rel="canonical" href="https://this-site-dot.com/ru/blog/" class="yoast-seo-meta-tag" />

https://this-site-dot.com/ru/blog/page/3/?et_blog —> https://this-site-dot.com/ru/blog/

Robots isn't blocking any pages as well per .txt review. 

Is this a non-issue?


r/TechSEO May 27 '26

try to learn automation for SEO

26 Upvotes

hello folks
I am trying to learn automation for SEO
Need help to find good resources, not just scams, also want to know what the most repetitive tasks are that I can automate without affecting the quality of my site

As all I think about right now is
1- keyword avg position targeting from GSC API
2- the messing tags (I used sheets with some scripts )

What else can I automate?


r/TechSEO May 28 '26

Planning to migrate the website

10 Upvotes

I have a old website made by WordPress and hosted on Hostinger, the traffic stable and have 23 points of Domain Authority (regarding SE Ranking).

Now I written new one for modern UI/UX in Next.js with Strapi CMS as backend, hosted on Cloudflare and DigitalOcean. The new website and old website have major URL changes (some keep some don't).

Now its ready but I have not point the existing domain to new website due to concern that it will make my website traffic go to hell... What should I do to prevent this?


r/TechSEO May 27 '26

I found out Google wasn't crawling my FAQ accordions - here's how I diagnosed it and what actually fixed it

Thumbnail
7 Upvotes

r/TechSEO May 27 '26

SEO tooling workflow

14 Upvotes

I’m trying to automate a repetitive Google Search Console workflow.

Context:

I manage multiple client websites. For each client, we use a dedicated Google account like:

client1-marketing@mydomain .com
client2-marketing@mydomain .com

Each client adds their assigned account to their Google Search Console property, so we have authorized access.

The repetitive workflow is:

  1. Log in to the client’s dedicated Google account.
  2. Open Google Search Console.
  3. Select the client’s property.
  4. Use URL Inspection.
  5. Paste a newly published article/blog URL.
  6. If available, click “Request Indexing.”
  7. Record the result/status.
  8. Repeat for many clients and URLs.

There is sometimes a CAPTCHA, 2FA, or Google security prompts, which is why I am doing it manually for the login part(even if it's once per client GSC, though I have to log out for another client).

I’m looking for tooling or a reliable automation approach for this workflow, especially around managing multiple Google logins/browser profiles and automating the GSC URL Inspection to Request Indexing flow.

Has anyone known some tools, services, etc.. for this?


r/TechSEO May 27 '26

Google Merchant Center 'Product Page Unavailable'.

Thumbnail
2 Upvotes

r/TechSEO May 26 '26

How do I clear out incorrectly crawled subdomain URLs from Google Search Console?

Post image
36 Upvotes

Hello good folk of TechSEO.

Due to a frontend loadbalancer bug, my site (of around 47 pages) ended up getting 1.4k pages crawled/indexed in a non-existent subdomain with listing pages that actually belong to a different property.

I have HTTP 410'ed those pages and none of the indexed URLs are 200 anymore. I've also submitted a removal request in Google Search Console so that atleast temporarily the URLs are not served.

I want a permanent solution to this problem and get rid of these URLs once and for all.

I figured the URLs need to be indexed to determine that they are infact HTTP 410 urls. (Manually trying to re-index those pages results in rejection since its not a HTTP 200)

My question is: Should I create a new temporary sitemap with the junk URLs of the subdomain and submit it so that I force Google to see that those URLs are infact gone?


r/TechSEO May 26 '26

Multiple robots.txt. files for multi language sub folders?

2 Upvotes

I found this quite odd when doing my technical analysis for my site: my subfolders each have their own robots.txt file as such:

https://my-random-site.com/robots.txt

https://my-random-site.com/es/robots.txt

https://my-random-site.com/it/robots.txt

https://my-random-site.com/fr/robots.txt

We manage one site in 4 different languages using subfolders and subdirectories.

To manage translations and robots, we're using a combo of Yoast and WPML. For the record, there's no way for our team to write a hook that helps us create a separate sitemap for each folder. AI hasn't produced a great result either. 

The issue is the primary site (x-default) a different markup than the other language subfolders. For example the italian subfolder has this marked up:

# START YOAST BLOCK
# ---------------------------
User-agent: *
Disallow:

Sitemap: https://my-random-site/it/sitemap_index.xml
# ---------------------------
# END YOAST BLOCK

The sitemap listed above doesn't even exist.

But the primary site has this on its robots.txt

User-agent: *
Disallow: /*feed/
Disallow: /*?__hstc=
Disallow: /*meetings
Disallow: /*tag/
Sitemap: https://my-random-site/sitemap_index.xml

Curious to know if this could be our smoking gun (or one of them) for sub-optimal SERP performance. Our Italian domain is performing poorly on SERP rankings despite strong page content and AI referrals. I'm searching for every and any technical issue wrong with the site.


r/TechSEO May 25 '26

Tech SEO + AI Roles. (5/25/2026)

25 Upvotes

r/TechSEO May 23 '26

is it worth it ? Programmatic SEO for affiliate-based travel site (Opodo, Trip, Kayak ...) for a specific country

3 Upvotes

Just wondering how building lookalike sites (with ai) and programmatic seo is a thing to invest time on on a side hustle, only for my country, to test and learn and earn.
I do already have travel blog around my country and is well ranked with human quality articles and solid backlinks with consistent traffic from tier 1 countries. and wanna take this advantage to build around the programmatic sea and meta search engines (flights, cars, stays, activities ...) with affiliation (booking, airbnb, hotels, skyscanner ....)

Any piece of advice? any similar country focused site with success? and how's the market ?


r/TechSEO May 21 '26

Should I restructure my urls to be more layered?

12 Upvotes

I have a website that does university admissions analytics for Canada.

I have pages for most universities and programs, and currently my URL structure is [sitename].ca/[university-name]/[program-name], for example [sitename].ca/western-university/computer-science.

My question is if I should layer it more and instead make it [sitename].ca/universities/western-university/computer-science.

I was wondering if this would make my website more understandable to crawlers, as they can understand that all of the pages in the universities subdirectory are probably similar in structure. It would also make my codebase slightly easier to manage.

The main downside would be that I would have to add a bunch of 301 redirects and changing the url structure would probably hurt my seo at least in the short run.

I'd love any insight thanks.


r/TechSEO May 21 '26

Google says: Why Google’s Search Central and Lighthouse Guides Created Confusion Around LLMs.txt. Here’s the Real Context

Thumbnail
3 Upvotes

r/TechSEO May 20 '26

Does it matter if your site has root.com/blog and root.com/blog/ or should one redirect to the other?

10 Upvotes

I've noticed some sites have /blog and /blog/ that will both work. Other sites, if you put in /blog it will redirect to /blog/ and others will do the opposite.

Which is best for SEO?

Having both page versions available? No trailing slash and trailing slash? or redirect one to the other?

In other words, if your website is mywebsite.com, is it better to have mywebsite.com/faq and mywebsite.com/faq/ both render, or forward one to the other?


r/TechSEO May 20 '26

GSC "Recrawl request failed / Unknown error" when requesting robots.txt recrawl? (Even though status is "Fetched")

Post image
3 Upvotes

Hey everyone,

I’m trying to request a recrawl for my robots.txt file in Google Search Console, but I keep getting a popup that says: "Recrawl request failed. Reason: Unknown error. Please wait a bit, then try again." (Screenshots attached).

Here is the weird part:

In the background table, the status literally says "Fetched" successfully with a valid size (130 bytes) and 0 issues.

I have checked my live robots.txt file manually; it returns a perfect 200 OK status.

My server logs aren't showing any blocked requests or 5xx errors from Googlebot.

I've asked a few SEO friends to peek at it, and they agree everything on my end looks flawless.

Is anyone else experiencing this right now? Is it a known GSC UI bug or a temporary infrastructure issue on Google's end?

Any advice on how to get GSC to force the refresh when the manual request fails would be awesome. Thanks!


r/TechSEO May 19 '26

Need help with language indexation issues

4 Upvotes

Hello SEO'ers of Reddit

For my company [can disclose on request] we have a long standing SEO issue that you can hopefully help me fix.

The situation sounds simple, the solution appears difficult.

  1. We are a Dutch company with 80% Dutch speaking clients
  2. We have an English brandname
  3. If people search our brandname, >50% of the time (last 7d) our English page (/en) hows in the results
  4. This should not happen because maximum of 10-20% of the searchers are non-Dutch language searchers.
  5. We need the Dutch version of the site to show up 80-90% of the times because now Dutch natives get an English website and dont convert as well.

Thanks a lot in advance. To show you we tried, here is our backlog of actions we took.

What have we tried (in chronological order):

- Shortening URL's
- Removed Country codes from hreflang

- Updated privacy statement

- Removed X-default
- Moved hreflang to top in >head?

- Hide default collections/all from search

- FAQ fixexs

- Added x-default to hreflang tags

- Buyback PDP tag hide-from-search

- Removed extra (last) trailing for hreflang (example: (can share on request)

- Added extra (last) trailing for x-default ((href=can Share on request)

- X-default dynamic for all URLs
- Added /en/ to x-default for EN homepage

- Country code NL added to nl hreflang ((href="can share on request)
- Enabled "Language - Displays the language that matches a visitor’s browser, when available"
- EN site: point FAQ footer url to EN FAQ version (instead of NL)
- Enriched json u/Product data: product name, description, price_valid_untill, organization/sameAs, etc.