What constitutes duplicate content on a page?
-
I am working on SEO for a Shopify store. Their products are very similar, hence the pages are so similar that Moz shows them as duplicate content. The only difference in the product pages is the title and model number. I am going to "go for the gold" and try re-writing all the product descriptions. It's incredibly difficult due to the products being nearly identical with just a minor variation. I know I could go down the road of just creating variants --- but the customer is not down for that.
Here's my question: what constitutes duplicate content? 80% of the content, 90%????
If I can going to re-write the descriptions, what should I aim for?
Thank you!
-
If you're not trying to rank, then you may not need to prioritize fixing this.
Duplicate content isn't a penalty; it's a risk. The risk is that more than one page on your site will seem appropriate in response to a particular search query and Google might 1) rank the wrong one, or 2) "decide" it's not clear and rank neither. If these aren't pages you'd expect or want to show up in search results anyway, then you can feel free simply to put together a page that'll provide the best experience for the user.
-
I am not trying to rank both pages - I am trying to create unique pages for each product. The complicated part is that they are tile. So a tile is made of the same material, same process for making them, used in the same application, and have the same size. So there are 1000 tiles that are very similar. The only slight variance is their color, part number and potentially if they have a pattern on them - such as a flower. Inside the set, there may be 100 tiles that have flowers. so it gets a bit difficult to write a description when so many things are the same about each tile.
-
Hi Steve! I'm not positive what Google considers duplicate, but I can tell you that our tools flag pages as duplicate when ≥90% of the source code (including content) matches. It's likely that we're more sensitive to it than Google is, which is intentional.
Out of curiosity, are you trying to rank both of these pages?
-
Hi-
I wouldn't being to guess the % needed, but I can tell you one way that we try to get around the issue with similar products. We added a "short description" section with bullet points to highlight all the questions people might ask about the product and in there we are very specific about color, shape, flavor etc. When we have similar products (say different flavored gummi bears) we list both basic facts about the product that are all the same (i.e. size, how many per bag) as well as listing the other attributes that are unique to that product (i.e. banana flavored, yellow colored) and that seems to be enough to keep us from a duplicate content penalty. It's also nice because it cuts down customer questions.
Just a thought
Ken
-
Hey Steve,
First question: How similar are these products? Are they the same but with color/trim/size differences? Have you had the conversation about canonicalization or are they not that similar?
Regarding duplication: I wouldn't look at it from a percentage standpoint, and if I did, I'd aim for 0-20% duplication with the assumption that 20% dupe was due to sentence beginnings and common intro phrases such as "If you're looking for....", which even as I look at that, I'd want to fix (because they're ubiquitous). Focus on the five Ws (who, what, where, when, why and how) and answer each one of those the best that you can, with the product's uses and why each model is different in mind.
Is there a way you can discuss the different use cases for each model number? Highlight benefits and applications? When people look for your product, what else are they searching for or concerned about?
Beau
Got a burning SEO question?
Subscribe to Moz Pro to gain full access to Q&A, answer questions, and ask your own.
Browse Questions
Explore more categories
-
Moz Tools
Chat with the community about the Moz tools.
-
SEO Tactics
Discuss the SEO process with fellow marketers
-
Community
Discuss industry events, jobs, and news!
-
Digital Marketing
Chat about tactics outside of SEO
-
Research & Trends
Dive into research and trends in the search industry.
-
Support
Connect on product support and feature requests.
Related Questions
-
Google ranking content for phrases that don't exist on-page
I am experiencing an issue with negative keywords, but the “negative” keyword in question isn’t truly negative and is required within the content – the problem is that Google is ranking pages for inaccurate phrases that don’t exist on the page. To explain, this product page (as one of many examples) - https://www.scamblermusic.com/albums/royalty-free-rock-music/ - is optimised for “Royalty free rock music” and it gets a Moz grade of 100. “Royalty free” is the most accurate description of the music (I optimised for “royalty free” instead of “royalty-free” (including a hyphen) because of improved search volume), and there is just one reference to the term “copyrighted” towards the foot of the page – this term is relevant because I need to make the point that the music is licensed, not sold, and the licensee pays for the right to use the music but does not own it (as it remains copyrighted). It turns out however that I appear to need to treat “copyrighted” almost as a negative term because Google isn’t accurately ranking the content. Despite excellent optimisation for “Royalty free rock music” and only one single reference of “copyrighted” within the copy, I am seeing this page (and other album genres) wrongly rank for the following search terms: “free rock music”
On-Page Optimization | | JCN-SBWD
“Copyright free rock music"
“Uncopyrighted rock music”
“Non copyrighted rock music” I understand that pages might rank for “free rock music” because it is part of the “Royalty free rock music” optimisation, what I can’t get my head around is why the page (and similar product pages) are ranking for “Copyright free”, “Uncopyrighted music” and “Non copyrighted music”. “Uncopyrighted” and “Non copyrighted” don’t exist anywhere within the copy or source code – why would Google consider it helpful to rank a page for a search term that doesn’t exist as a complete phrase within the content? By the same logic the page should also wrongly rank for “Skylark rock music” or “Pretzel rock music” as the words “Skylark” and “Pretzel” also feature just once within the content and therefore should generate completely inaccurate results too. To me this demonstrates just how poor Google is when it comes to understanding relevant content and optimization - it's taking part of an optimized term and combining it with just one other single-use word and then inappropriately ranking the page for that completely made up phrase. It’s one thing to misinterpret one reference of the term “copyrighted” and something else entirely to rank a page for completely made up terms such as “Uncopyrighted” and “Non copyrighted”. It almost makes me think that I’ve got a better chance of accurately ranking content if I buy a goat, shove a cigar up its backside, and sacrifice it in the name of the great god Google! Any advice (about wrongly attributed negative keywords, not goat sacrifice ) would be most welcome.0 -
ECommerce Duplicate content on product pages (eg delivery info, contact details etc)
Hi, Running a Magento site and wanted to check about duplicate page content. We have 1000+ product pages and it has been suggested to remove some of the "duplicated content" which displays on every product page and replace this with an image of the same text content. By this I am talking about content which is for promo/customer purposes and is displayed on every page. eg: "If you find our products cheaper elsewhere then please click below to get your price match...... etc", and a chunk of text for the "Delivery Tab Information" and "Contact Tab Information" on each and every product page. A SEO company has suggested to turn this content into images. Does anyone have thoughts on this please?
On-Page Optimization | | Ampweb0 -
Duplicate content - "Same" profile-information
Hi, I own a casting website with lots of profiles. Some of these profiles only typed in their firstname, email and age, when they registered on the site, and they haven't added more information ever since. From Crawl Diagnostics, I can see that there is "lots" of these profiles, which looks exactly the same (only showing age and firstname), allthought they are not the same. I could add which day the profile were created on the site, to maybe avoid these "duplications". The email will always be hidden. Or, how big an issue is this? Crawl Diagnostics tells me, that there is around 200 of these, and they are "marked" as High Priority. Any ideas on what to do? /Kasper
On-Page Optimization | | KasperGJ0 -
Long list of companies spread out over several pages - duplicate content?
Hi all, I am currently working with a company formation agent. They have a list of every limited company spread over hundreds of pages. What do you guys think? Is there a need for Canonicals? The website is ranking pretty well but I want to make sure there aren't any problems in the future. Here are two pages as examples: http://www.formationsdirect.com/companysearchlist.aspx?start=MULLAGHBOY+CONSTRUCTION+LIMITED&next=1# http://www.formationsdirect.com/companysearchlist.aspx?start=%40a+company+limited&next=1# Also what about the actual company pages? See an example below http://www.formationsdirect.com/companysearchlist.aspx?name=AMNA+CONSTRUCTION+LTD&number=06630333#.U8PW6_ldX1s Thanks in advance Aaron
On-Page Optimization | | AaronGro0 -
Duplicate Content - Deleting Pages
The Penguin update in April 2012 caused my website to lose about 70% of its traffic overnight and as a consequence, the same in volume of sales. Almost a year later I am stil trying to figure out what the problem is with my site. As with many ecommerce sites a large number of the product pages are quite similar. My first crawl with SEOMOZ identified a large number of pages that are very similar - the majority of these are in a category that doesn't sell well anyway and so to help with the problem I am thinking of removing one of my categories (about 1000 products). My question is - would removing all these links boost the overall SEO of the site since I am removing a large chunk of near-duplicate links? Also - if I do remove all these links would I have to put in place a 301 redirect for every single page and if so, what's the quickest way of doing this. My site is www.modern-canvas-art.com Robin
On-Page Optimization | | robbowebbo0 -
Duplicate Content - Potential Issue.
Hello, here we go again, If I write an article somewhere, lets say Squidoo for instance, then post it to my blog on my website will google see this as duplicate content and probably credit Squidoo for it or is there soemthing I can do to prevent this, maybe a linkk back to Squidoo from my website or a dontfollow on my website? Im not sure so any help here would be great, Also If I use other peoples material in my blog and link back to them, obviously I dont want the credit for the original material I am simply collating some of this on my blog for others to have a specific library if you like. Is this going to damage my websites reputation? Thanks again peeps. Craig Fenton IT
On-Page Optimization | | craigyboy0 -
Why Does SEOMOZ Crawl show that i have 5,769 pages with Duplicate Content
Hello... I'm trying to do some analysis on my site (http://goo.gl/JgK1e) and SEOMOZ Crawl Diagnostics is telling me that I have 5,769 pages with duplicate content. Can someone, anyone, please help me understand: how does SEOMOZ determine if i have duplicate content Is it correct ? Are there really that many pages of duplicate content How do i fix this, if true <---- ** Most important ** Thanks in advance for any help!!
On-Page Optimization | | Prime850 -
How do you see a list of URLs with duplicate page titles?
When looking at the Duplicate Page Title report, the Other URLs column has various numbers that presumably indicate the number of pages that share the same title. When I click on one of these numbers, say a URL that shows 4 in that column, the next page reports "No sample duplicate URLs to report". Why isn't it showing me the other 3 URLs with the same page title?
On-Page Optimization | | jkenyon0