Smaller Index
-
Hi guys,
We are a price comparison website with thousands of webpages. Most of them are product webpages with not so good quality content. Only price information and product image, no product details nor costumers reviews.
We are planing to focus on less product categories by adding reviews, details, better images etc... and I would like to know if I should maintain the other "not-so-good" products in other categories or if I should remove it from index to leverage domain average content quality.
Our index size is 200k pages and we are planning to focus on 10k pages max.
Thanks for your help.
-
Hello Pedro,
I think you are making a very wise decision. If you have already been throttled by Panda this could be what you need to bring the site out of it. If not, this could be what you need to save you from a future update. In fact, Matt Cutts recently answered a question about this sort of thing:
http://www.youtube.com/watch?v=adocBLGQoYENote: The question is about "no results" pages but he discusses similar scenarios as well.
These sort of "stub pages" have been a thorn in Google's side for many years, and rest assured they will continue to find ways of keeping them out of the index - including punishing the good content on sites that use them.
As Infant Raj mentioned below, be sure the URLs return a 404 status code in the http header, which will ensure more prompt removal from the index than if they were to redirect or show a 200 status code. I'd ignore the first paragraph in his answer though.
-
If those less significant pages arent entry pages for organic or referral traffic you can remove it. Else its not a good idea to remove those pages just to reduce the number of indexed pages.
If those pages are removed, make sure you add a custom 404 page to handle the 404 errors
Got a burning SEO question?
Subscribe to Moz Pro to gain full access to Q&A, answer questions, and ask your own.
Browse Questions
Explore more categories
-
Moz Tools
Chat with the community about the Moz tools.
-
SEO Tactics
Discuss the SEO process with fellow marketers
-
Community
Discuss industry events, jobs, and news!
-
Digital Marketing
Chat about tactics outside of SEO
-
Research & Trends
Dive into research and trends in the search industry.
-
Support
Connect on product support and feature requests.
Related Questions
-
Indexing but Not Ranking
Hello, I have noticed this with some articles. Here is an example:
Content Development | | SneakerFiles
http://www.sneakerfiles.com/air-jordan-11-low-closing-ceremony-release-date/ For the term "Air Jordan 11 Low Closing Ceremony" I do not show up on the first 3 pages of Google even though the website has a back link to this particular page (maybe more). It's almost like some pages just disappear. Also I will rank very high (page 1) for a sneaker that is going to release for months then all of a sudden, about a week from the release, I disappear entirely. Not sure why this is (other in my field also rank on page 1 but stay). The site has been around for just about 10 years now with a large amount of backlinks (I don't go out and try to obtain backlinks unless it's something exclusive to share with others). Any help would be greatly appreciated.0 -
Sitemap - 200 out of 2100 pages indexed
I submitted the .xml sitemap in Google Webmaster Tools and only 200 out of 2100 pages were indexed.
Content Development | | Madlena
Why is that and what can I do ?0 -
One story stands out for not getting indexed?
We have all our stories published today ( 20-Jun-2013 ) got indexed by google except this ( http://coed.com/2013/06/20/heres-a-video-of-kate-upton-topless-on-a-horse/ ). Do anyone out there have any clue about that? Thanks in advance
Content Development | | COEDMediaGroup0 -
In my website all the pages are not indexed by google..what to do for the same
In my website http://www.dubins.ae, all the pages are not indexed by google. How to make sure that all the pages are indexed by google?
Content Development | | Muna0 -
In Index but not in Serps
Hi, I have a situation with a client site which is quite frustrating. Basically, most "recent" (by that I mean for the last couple of months) blog posts are failing to reach the SERPS (actually, one has and a couple have from the early days but it's taken months for them to arrive). Previously the blog posts were indexed very quickly - often instantly. Now, I've checked WMT etc and I've submitted each post manually but still nothing. The Sitemap is valid etc. However, pages (not blog posts) seem to be getting into the serps very quickly. Another complication is that if I search: site:www.domainname.com and set the date filter to a month I can see some of the earlier blog posts in that result set. However, if I scrape a bit of unique content from one of those posts and search - nothing in the SERPS. And my Moz report tells me that the page is not to be found in the top 50 either (so I'm confident these pages are not in the SERPS). Any ideas why this would happen to just blog posts? Is it something to do with the parent blog landing perhaps being too strong in the rankings? Any ideas appreciated. Thanks.
Content Development | | KMUK0 -
How to make new content Indexed faster by google
I would like to know what can I do. Normally it takes google around 3 days to index my content. I got a site map, swiched the crawling rate to the fastest in my webmaster tools. I also tried crawling my homepage as google bot and sending it to the index with all linked pages but even if I do so my content takes around 3 days if not more to get indexed. I publish around 20 posts a week. My SEOmoz page authority is 48. Some sites of my competition seem to be getting their content indexed in the same day. What else can be done?
Content Development | | sebastiankoch0 -
Blog Posts Not Getting Indexed by SERP's
Our blog posts are typically indexed rather quickly, sometimes within 10 minutes or so. Lately it seems like it is taking much longer, and a post from yesterday afternoon isn't appearing in any of the search engines. Can anyone give me an idea why this may be happening? http://www.360dwellings.com/2011-denver-parade-of-homes-luxury-home-tour/
Content Development | | 360ryan0 -
Please help me stop google indexing https pages on my wordpress site
I added SSL to my wordpress blog because that was the only way to get a dedicated IP address for my site at my host. Now I am noticing Google has started indexing posts both as http and https. Can some one please help how to force google not to index https as I am sure its like having duplicate content. All help is appreciated. So far I have added this to top of htaccess file: RewriteEngine on Options +FollowSymlinks RewriteCond %{SERVER_PORT} ^443$ RewriteRule ^robots.txt$ robots_ssl.txt And added robots_ssl.txt with following: User-agent: Googlebot Disallow: / User-agent: * Disallow: / But https pages are still being indexed. Please help.
Content Development | | rookie1230