
Free Keyword Research Tool: The Free Stack That Works for Service Businesses
Which free keyword research tool should a Singapore clinic, firm, contractor or tutor use? Combine five free tools to find your first 30-50 keywords. See how.
From F&B to fintech, clinics to law firms, startups to enterprise. If your customers search on Google, we make sure they find you first, not your competitors.
One specialist team, focused only on the organic rankings that put you in front of ready-to-buy Singapore customers.
A clear, sequenced path from audit to rankings. You always know what we’re doing and why it matters for your leads.

Quick answer: Identify fix duplicate content Singapore websites suffer from by crawling every URL, grouping pages that share the same content, choosing one canonical version, and pointing every signal at it. The goal is not deletion but consolidation, so links, rankings and crawl attention land on a single page.
Almost every Singapore SME site we crawl has duplicate content, and almost none of the owners know it. The reason is that duplication is rarely created deliberately. Nobody sits down and writes the same service page twice. It appears quietly, as a by-product of how content management systems generate URLs, how e-commerce platforms handle filters, and how well-meaning staff copy an existing page to make a new one. By the time anyone notices, a site with forty real pages is serving four hundred addresses, most of which say the same thing. The single most damaging myth in this area is that Google issues a duplicate content penalty. It does not, in the vast majority of cases. What actually happens is quieter and more expensive: your ranking signals get split across near identical URLs, Google picks a version you did not intend, and your crawl attention gets spent on pages that will never earn a customer. This guide covers how to find duplication, how to decide which version should survive, and which repair tool to use in each situation. If you want the broader context around crawling and indexation, our technical SEO work sits around this problem.
Duplicate content means two or more URLs, whether on your site or across the web, that serve substantially the same body copy. “Substantially” is the operative word. Search engines do not require a byte for byte match. Two product pages that differ only in a colour swatch and a SKU code are duplicates in every way that matters. So are a service page and its printer friendly version, a blog post and its AMP variant, and your homepage served at four different addresses.
The term has been badly mangled by a decade of low quality SEO advice, so let us be precise about what duplicate content is not. It is not a penalty. Google has stated repeatedly that duplication without deceptive intent is handled algorithmically by choosing one version to show, not by demoting your domain. It is not plagiarism detection. It is not a word count threshold. And it is not triggered by having boilerplate text such as a footer address or a delivery policy repeated across your site, which is completely normal and expected.
What it genuinely costs you falls into three buckets. First, signal dilution. If three URLs carry the same content and external sites have linked to all three over the years, the authority those links carry is spread across three addresses instead of concentrated on one. Second, wrong version selection. Google picks a canonical for you, and its pick is frequently the URL with the ugliest address, the thinnest internal linking, or a stray tracking parameter attached. Third, crawl waste, which we will come back to, because it only becomes a serious problem above a certain site size.
There is a fourth cost that rarely gets mentioned and that we see repeatedly on Singapore sites: cannibalised reporting. When two URLs both rank in the top twenty for the same query, your analytics and rank tracking become unreadable. You cannot tell whether a change helped, because impressions keep shuffling between the two.
The sources of duplication are boringly predictable once you know where to look, and the local hosting and platform landscape produces a few that are specific to this market.
Protocol and hostname variants. A site that resolves at http and https, with and without www, is serving four versions of every page. Singapore hosts have been good about forcing HTTPS since free certificates became standard, but www consolidation is still missed constantly, particularly on sites migrated from an older local host.
Trailing slash inconsistency. /services and /services/ are different URLs to a crawler. Most platforms normalise this automatically. Some custom builds do not.
Domain duplication across .sg and .com. This one is genuinely local. A large share of Singapore businesses own both a .com.sg or .sg domain and a .com, often for brand protection, and then serve the full site on both rather than redirecting one to the other. We have seen businesses run more than one live copy of the same site this way for years, with each copy accumulating only part of the links the brand earned.
E-commerce faceted navigation. Filter by brand, then by price, then by size, and your platform generates a unique URL for each combination, all of which return the same products in a different order. On a catalogue of any size the number of generated URLs can exceed the number of real products by an order of magnitude. Our e-commerce SEO audits almost always start here.
Pagination and sort orders. ?sort=price_asc and ?sort=newest return the same items in different sequences. Every one is a separate URL.
Session IDs and tracking parameters. Campaign tags such as utm_source create infinite variants of any page they are appended to. If those tagged URLs get linked or shared publicly, they become crawlable addresses in their own right.
Location page templates. A business with outlets in Jurong, Tampines and Orchard writes one page and swaps the district name. If ninety percent of the body copy is identical, those pages are near duplicates competing with one another rather than three distinct assets. This is one of the most common failure patterns in local SEO work.
Staging and development environments left indexable. A staging. or dev. subdomain with no password and no noindex directive is a full duplicate of your production site. We find one of these roughly once a month.
You do not need expensive software for the first three passes. You need patience and a spreadsheet.
Pass one, the site operator sweep. Search Google for site:yourdomain.com.sg and look at the reported number of results against how many pages you believe you have. If you think you have sixty pages and Google reports six hundred, you have a duplication or parameter problem and you now know its rough scale. Then narrow it: site:yourdomain.com.sg "an exact sentence from a page". If that sentence returns more than one URL, you have found a duplicate cluster by hand.
Pass two, Search Console coverage. Open the Pages report in Google Search Console and read the “Not indexed” reasons carefully. Three of them are duplication signals in plain language: “Duplicate without user selected canonical”, “Duplicate, Google chose different canonical than user”, and “Alternate page with proper canonical tag”. The second of those is the important one. It means you told Google which version you wanted and Google disagreed with you, which usually indicates your internal linking or your redirects contradict your canonical tags.
Pass three, a crawl. Run any desktop crawler over your domain and sort the output by page title, then by meta description, then by word count. Identical titles cluster instantly. Most crawlers also compute a content similarity score, and anything above roughly ninety percent similarity deserves a look. Export that list. It becomes your working document.
Catalogue driven stores are the extreme case here, and our e-commerce case studies show the effect: after 14 faceted navigation parameter combinations were blocked and a clean sitemap of canonical URLs submitted, product page indexation rose from 34% to 79% within two months.
Pass four, the external check. Take two or three distinctive sentences from your most valuable pages and search them in quotation marks. If a supplier, a directory, a franchise partner or a scraper has republished your copy verbatim, you will find it here. Manufacturer supplied product descriptions are the classic case, and they are the reason so many Singapore retailers selling the same imported brands struggle to differentiate.
Work through those four passes and you will have a list of clusters. A cluster is a group of URLs that say the same thing. The next decision is which member of each cluster survives.
This is the step teams skip, and skipping it is why so many duplication clean-ups make rankings worse rather than better. Before you touch a redirect, decide deliberately which URL is the keeper. Four criteria, in this order of weight.
Existing external links. The URL with genuine inbound links from other websites has value you cannot recreate. Keep it unless there is a compelling reason not to.
Current organic performance. Check Search Console for which version already receives impressions and clicks. Google has effectively voted. Overruling that vote is possible but costs time.
URL quality. Shorter, cleaner, keyword relevant and parameter free beats long and messy, all else being equal.
Strategic fit. Sometimes the right answer is a URL that does not yet exist, because two half pages should become one strong page. Agent profile pages are a recurring example, and our property agent results work started with the agent’s single landing page being rebuilt as a proper website with an agent profile and CEA credentials.
In our experience, the most common mistake here is choosing the newest page simply because it is the one someone just finished writing. The newest page usually has no links, no history and no data, and redirecting an established URL into it throws away years of accumulated signal.
There are four instruments. Using the wrong one is worse than doing nothing, because it can deindex pages you wanted to keep.
| Tool | What it does | Use it when | Do not use it when |
|---|---|---|---|
| 301 redirect | Permanently sends users and crawlers to another URL | The duplicate has no independent reason to exist | Visitors still need to see both versions |
| Canonical tag | Tells search engines which version to index, both stay live | Both URLs must remain accessible, such as filtered views | The pages are genuinely different content |
| Noindex | Keeps the page live but out of the index | Thin, internal or utility pages with no search value | The page has external links worth preserving |
| Consolidation rewrite | Merges two pages into one better page | Two thin pages compete for the same query | One page clearly outperforms the other |
A few rules that prevent the usual damage. Never combine noindex with a canonical tag pointing elsewhere, because you are giving two contradictory instructions and the outcome is unpredictable. Never canonicalise to a URL that then redirects somewhere else, which creates a chain search engines may simply ignore. Never use robots.txt to fix duplication. Blocking a URL in robots.txt prevents crawling, which means the canonical tag on that page can never be read, which means the duplication is frozen in place rather than resolved. That last one is the single most widely repeated piece of bad advice in this area, and we still see it in agency audit documents.
For faceted navigation specifically, the correct pattern is usually canonical tags on filter combinations pointing back to the unfiltered category page, combined with parameter handling at the server or CDN level. Blanket blocking filters in robots.txt will leave you with thousands of URLs Google knows about and cannot evaluate.
Because so many Singapore businesses hold multiple domains, this deserves its own treatment. If you run the same site on brand.com and brand.com.sg, you have a strategic decision rather than a technical one.
If you serve only Singapore, pick one domain and 301 redirect the other to it in full, page for page. Do not redirect everything to the homepage, which discards the value of every deep link. Page level mapping takes an afternoon and preserves the signal.
If you serve Singapore plus other markets from separate domains with genuinely different content, pricing or currency, keep both and implement hreflang annotations so search engines understand they are regional variants rather than duplicates. Hreflang is a tag that declares the language and region a page is intended for, and it must be reciprocal, meaning each page points at the other. Half implemented hreflang is common and effectively does nothing.
If the second domain exists purely for brand protection and has no traffic, redirect it and stop thinking about it. We have seen businesses spend months maintaining content on a defensive domain that received under a hundred visits a year. Property and agency sites are particularly prone to this, which is why our real estate SEO engagements usually begin with a domain inventory before anything else.
Duplication fixes are slow to show results, and that delay causes a lot of teams to panic and reverse them prematurely. Set expectations before you start.
In the first two weeks, expect crawl activity on the removed URLs to continue. Google revisits redirected addresses for a long time before it stops. This is normal and is not a sign the redirect failed.
Between weeks three and eight, watch the Search Console Pages report. The “Duplicate” categories should shrink and your indexed page count should fall, which sounds alarming but is exactly what you want. You are trading a large number of weak URLs for a smaller number of strong ones.
From week six onwards, watch impressions and average position for the surviving URL. The pattern to look for is consolidation: the keeper page should pick up the queries the removed pages were ranking for, at a better average position than any of them held individually. If impressions across the cluster fall and do not recover after ten weeks, you probably chose the wrong keeper, and the fix is to reverse that specific redirect rather than the whole project.
Set up a simple before and after sheet with the cluster, the keeper URL, combined pre fix impressions and post fix impressions. We recommend reviewing it monthly for one quarter, then filing it. For smaller sites this whole exercise is usually a two or three day project, which is why it sits inside our small business SEO scope rather than being sold as a separate engagement.
The ecommerce case study above shows what that sequence looks like on a real catalogue. The WooCommerce home and lifestyle store had only 68 of its 200 product pages indexed, with faceted navigation generating duplicate content. Once the parameter combinations were blocked and a clean sitemap of canonical URLs submitted, the work moved on to rewriting 15 category pages. Keywords on page 1 climbed from 12 to 38 by Month 5, two category pages reached #1 by Month 9, and organic revenue reached $28,600 a month.
Field notes: Duplication of some kind turns up in most of the crawls we run, whether through an unredirected www variant, a forgotten second domain, an indexable staging subdomain, or simply repeated metadata. In our medical case study, a GP clinic in Toa Payoh, the first phase resolved duplicate meta descriptions across the service pages, implemented canonical tags sitewide, fixed all 47 crawl errors and submitted an updated XML sitemap, with Core Web Vitals passing by the end of Month 2. With that foundation in place, the content and local work that followed took monthly organic visitors from 180 to 563 by Month 6. We also often find that the URL Google has chosen as canonical is not the one the business considers its main page, so check Search Console before you decide which page to keep.
Duplicate content is not a penalty problem. It is a concentration problem. Every duplicate URL takes a share of the links, the relevance and the crawl attention that should be pointing at one strong page, and the result is a site that looks busy and ranks thinly. The work of fixing it is unglamorous: crawl everything, group what matches, choose a keeper on the basis of links and existing performance rather than recency, then apply the narrowest tool that solves the case. Redirect what should not exist, canonicalise what must stay live, noindex what has no search purpose, and merge what is competing with itself. Resist the temptation to block anything in robots.txt, because a URL that cannot be crawled cannot be consolidated. Give the change eight to twelve weeks before judging it, and measure the cluster rather than the individual page. If you would like a second pair of eyes on which URLs to keep, a consultation is the fastest way to get a defensible answer.
Our clients who consolidate on links and existing performance rather than on which page looks newest tend to recover the combined ranking strength faster than expected, often within the first month after the redirects go live.
Not in the way the phrase implies. There is no automatic penalty for having the same content at two addresses. Google selects one version to index and shows that one, filtering out the rest. Penalties only enter the picture where duplication is deliberately deceptive, such as scraping other sites wholesale or spinning near identical pages purely to occupy more search results. For a normal Singapore business site, the cost is diluted signal and wasted crawl attention rather than a manual action.
Boilerplate repeated across your site, such as a footer, an address block, a delivery policy or a standard disclaimer, is expected and causes no issue. The threshold that matters is whether the unique, substantive part of a page is genuinely different from other pages. If you strip out the header, footer and sidebar and two pages read almost identically, they are duplicates in practice, regardless of what a percentage similarity tool reports.
Usually not. Deleting a page that has external links or search history destroys value. A 301 redirect to the version you are keeping preserves most of it while removing the duplicate from the index. Delete outright only when a page has no links, no traffic and no strategic purpose, and even then a redirect to the closest relevant page is a kinder outcome for anyone who follows an old link.
It is a single line of code in the head of a web page that says, in effect, “if you find several versions of this content, treat this address as the official one”. Search engines treat it as a strong hint rather than an absolute instruction. It is the right tool when both versions need to stay live and clickable, which is why it suits filtered and sorted views on e-commerce sites.
Yes, and it is common. When two pages target the same query, search engines have to pick one, and the pick can change from week to week. The visible symptoms are unstable rankings, impressions moving between URLs, and neither page ever reaching the position a single consolidated page would have earned. Merging them into one stronger page usually resolves it within a month or two.
They count as duplication across domains, yes. If fifty retailers in the region use the identical manufacturer text, none of those pages offers anything unique, and search engines default to the most authoritative domain. The practical fix is not to rewrite every product at once but to rewrite the descriptions for the products that actually drive revenue, adding local detail such as Singapore availability, warranty terms in SGD, and real usage notes.
It should, and that is the intended outcome. You are converting a large number of weak URLs into a smaller number of stronger ones. Watch impressions and clicks at the cluster level rather than the raw indexed count. A site that goes from 3,000 indexed URLs to 400 while holding or growing impressions has had a successful consolidation, not a loss.
Crawling of the old URLs continues for weeks after the change. Search Console coverage categories typically start shifting between three and eight weeks in, and ranking consolidation on the surviving page tends to appear from around week six. Give any consolidation project a full quarter before judging it, and only reverse individual redirects rather than the whole project if something is clearly wrong.
It is if both serve the same content and neither redirects. Pick the domain you want to build, redirect the other page for page, and keep the second domain registered purely defensively. The only case for running both is genuinely different content for different markets, and that requires hreflang annotations to be implemented reciprocally on both sides.
No, and this is the most common mistake we correct. Blocking a URL prevents it from being crawled, which means any canonical tag or noindex directive on that page can never be read. The URL stays in the index in a degraded state, with no way to resolve it. Allow the crawl, then use a canonical tag, a redirect or a noindex to give a clear instruction.
If your indexed page count looks nothing like the number of pages you actually wrote, that gap is worth understanding before you invest in more content. We run a free initial review that maps your live URLs against your intended site structure and shows you exactly where signal is being split. There is no obligation and no sales script attached to it. Send us the domain through our contact page and we will tell you what we find, including whether the problem is worth paying anyone to fix. Our pricing is published openly if you would rather see the numbers before you get in touch.
Natalie leads SEO strategy at Singapore SEO Agency, helping local and regional businesses build organic search programmes that drive qualified leads. She specialises in technical SEO and content-led authority building for Singapore SMEs.
Get a free SEO audit for your Singapore website — we'll show you exactly where you stand, what's holding you back, and what it would take to rank on page 1.
Get Your Free SEO Audit →
Which free keyword research tool should a Singapore clinic, firm, contractor or tutor use? Combine five free tools to find your first 30-50 keywords. See how.

SEO vs SEM is usually the wrong question. Learn what the terms really mean and how to use ads and Search Console data to decide which searches to earn or buy.

Google Search Console login problems usually start with verification and ownership. Learn how to get in, fix access errors and offboard agencies safely.

The click through rate formula is clicks divided by impressions. Learn what each platform counts, the averaging trap and how to set it up in Google Sheets.

A click through rate means nothing on its own. Learn to compare CTR by position, query type and SERP features, and see what low CTR is really telling you.

What a google analytics certification proves, what it misses, how to verify one and the practical questions to ask before you hire a marketer or freelancer.
Fast, no obligation. We reply within 24 hrs.
Singapore’s specialist SEO agency for SMEs. We rank your business on Google — and only Google. No distractions, just results.
© 2026 Singapore SEO Agency. All rights reserved.