Checking domain history
A domain's history reveals what legacy you are taking on: previous content, operators and uses shape how users, search engines and, in doubt, rights holders see the domain. Since AI answers such as Google's AI Overviews now make up a growing share of search, this history also increasingly decides whether a domain becomes visible to generative AI systems at all. Check the past before you plan the future.
Why domain history has become even more important in 2026
Google is increasingly personalising search through so-called “Preferred Sources”: users can specify which brands and websites they want their information and AI answers to come from. In Google's “AI Mode”, a very large share of searches now end as pure “zero-click searches” without a single click to an external page. If a domain is not on a target audience's favourites list, or is not recognised by AI as a trustworthy source, it stays practically invisible to generative search – regardless of how good its classic ranking is.
That makes a domain's history an even bigger lever than before: a domain with a clean past can bring existing trust with it, while a domain with a spam or penalty history can stay permanently invisible, no matter how good the new content is.
How to proceed
- Web archive: Look at archived snapshots over several years in the Wayback Machine. Watch for gaps – they tell you something too.
- Previous operators: Who was listed in the legal notice? A company, an association, a private individual? Do these operators still exist?
- Previous content: Guide, shop, company page, blog? What expectations do returning visitors have?
- Topic changes: Frequent, radical topic changes or phases with thin content point strongly to purely manipulative interim SEO use.
- Trademark use: Was the domain term previously used as a registered trademark or protected company name? Then trademark conflicts loom – check the DPMA, the EUIPO (TMview) or WIPO – see trademark check.
- Historical contact details: Old physical addresses, names and phone numbers may still be associated with the domain on the web.
- Previous shops: Did customers order here before? Open warranty issues or trust concerns can fall back on you if there is a risk of confusion with the new project.
- Malware and spam: Search specifically for warnings in the Google Transparency Report (Safe Browsing), check email blacklists and look for archived phases where the site was hacked or infected with malware.
- Redirects: Where did the domain point last? Unhealthy redirect histories and chains massively affect which backlink signals and how much trust actually arrive.
- Legal risks: Check whether the domain was previously involved in illegal schemes, gambling or dubious pharma phases. Such historical baggage drastically increases the risk of permanent devaluation.
A high PageRank or a strong Domain Rating in tools like Ahrefs or Semrush is worthless if the domain is dragging a toxic history along in the background. Metrics only show a snippet – the actual check only happens through the points above.
A warning from practice: when Google's history stays invisible
A real example shows how drastic this can be: a caught domain looked like a solid foundation at first glance. Only after verification in Google Search Console (GSC) did a significant problem become visible under the “Security & Manual Actions” tab – a finding that no external SEO tool would have revealed beforehand.
🔴 1 issue detected: Significant spam issues. Description: Pages on this site appear to use aggressive spam techniques such as scaled content abuse, cloaking, or content copied from other websites, and/or repeated or severe violations of Google's spam policies for web search have been identified. Affects: All pages.
For Google, the domain was effectively burned: it had apparently been used for aggressive black-hat techniques such as mass AI spam (scaled content abuse) and cloaking (the bot sees different content than the human). Cleaning up such baggage algorithmically requires months of work, rigorous content audits and a humble reconsideration request to Google – often with no guarantee of success. This is exactly why the GSC check (see checklist above) belongs strictly before any purchase or backorder award, not after.
Entity authority: why some domains get accepted as a “Preferred Source” – and others don't
A second, less obvious factor increasingly matters: whether Google recognises a domain as a standalone entity in the Knowledge Graph at all – independent of trust or backlinks. Our own observations on deinrecatch.de show an at first paradoxical pattern for which domains users can add as a “Preferred Source”:
Adding possible
- domains with an active online shop (e.g. with a cart and checkout)
- domains with a clear local connection (e.g. a regional institution)
- simple landing pages
- pure parking pages
Adding not possible
- active websites with real organic traffic and backlinks – rejected nonetheless
At first glance this looks contradictory: why is a pure parking page accepted while an active site with real traffic is rejected? The answer doesn't lie in traffic or backlinks, but in entity structure:
- Why active sites can still fail: Traffic and backlinks are classic metrics. Without a verified Google Business Profile, without a Google Publisher Center listing and without solid Organization schema, a website initially remains a “collection of texts about a keyword” rather than a recognisable entity for Google.
- Why shops and local sites work: Verified company data in Google Merchant Center or a verified, physical Google Business Profile quickly classify a domain as a real, trustworthy entity.
- Why parking pages sometimes work anyway: Two effects combine here. First, the “ghost effect” of expired domains: if the domain was previously a real association, an authority or an established business, this entity status can persist in the Knowledge Graph for years, even after the domain is re-registered with a simple template. Second, modern parking and sales templates often use clean Schema.org markup (such as
Offer,ProductorOrganization) – technically one of the “languages” AI search systems can evaluate well.
For buyers this means: a look at a domain's old history (see checklist above) not only reveals legal and spam risks, but can also hint at whether an “entity legacy” still exists in the background that could be useful for your own visibility in AI search.
Networks of expired domains: opportunity and risk at once
Because AI systems look for consensus – if several, apparently independent expert portals share the same assessment, that counts as a strong signal – some market participants deliberately buy several topic-relevant expired domains to rebuild them as standalone guide or comparison portals. Historical trust and real backlinks of the domain are meant to help them be perceived faster than a brand-new domain.
Google now consistently penalises blunt deception – invented test results or artificially ranking your own product first – including through algorithmic measures such as scaled content abuse (see the case study above). The decisive difference between a network that works and a network that gets penalised is therefore not just content quality, but above all disclosure. If the economic links between the portals and the promoted product are made transparent – as on our own transparency page – you are operating within the scope of legitimate content partnerships. If several portals are deliberately presented as “independent” even though they belong to the same operator, that is an undisclosed connection – which violates both Google's spam policies and general principles for labelling advertising and affiliate relationships.
Anyone who creates objective comparisons based on verifiable data, clearly labels where the comparison comes from, and works technically cleanly (short load times, clear answers in the first sentences for AI crawlers) is operating within a legitimate framework. A network of covertly aligned sites meant to fake a false consensus is exactly the kind of risk shown in the case study above – only self-inflicted rather than inherited.
Your action plan: entity instead of just a website
Regardless of whether you choose a single project or a transparently labelled network: keyword traffic alone is a metric of the past. To become visible in AI answers and “Preferred Sources” at all, it helps to turn a website into a recognisable entity:
- Google Publisher Center: Register editorial sites there – a direct path to verification as a publisher entity.
- Use Schema.org consistently: Implement solid Organization, Product or LocalBusiness markup and link it via the
sameAsproperty to real, active profiles (e.g. LinkedIn, YouTube, Instagram). - Build third-party mentions transparently: High-quality, editorially independent mentions and backlinks are valuable – make sure sponsored or paid placements are labelled accordingly (
rel="sponsored").
Old content is not your content
Old texts and images must not simply be reused – the rights belong to the previous authors. The previous design must not be copied in an identity-establishing way either, and users must not be misled about the new operator's identity. A fresh start needs its own content and a transparent legal notice.
Continuing the check process
Checking backlinks · Trademark check · Local domain check · Economic relationship and disclosure