Stuffing keywords into URLs
-
The following site ranks #1 in Google for almost every key phrase in their URL path for almost every page on their site. Example: themarketinganalysts.com/en/pages/medical-translation-interpretation-pharmaceutical-equipment-specifications-medical-literature-hippa/ The last folder in this URL uses 9 keywords and I've seen as many as 18 on the same site. Curious: every page is a "default.html" under one of these kinds of folders (so much architecture?).
Question: How much does stuffing keywords into URL paths affect ranking? If it has an effect, will Google eventually ferret it out and penalize it?
-
This was a good answer and deserves to be labeled as such. I decided not to pursue this since I have been lucky to take the top spot for important key phrases. Thank you for such a well crafted answer.
-
Hi Paul, no problem at all. As Ryan says, we all like a mystery.
As for the canonicals they can have a big effect if all variations of the domain are present. i.e.
etc
Not only are these duplicate pages they will most likely split up any inbound link juice as you can see from the PA of the pages you mention. Go to the http:// version and the http://www and you'll see the problem!
Using <link rel="canonical" href="<a href="http://www.vibralogix.com/">http://www.mydomain.com/" /> would probably be sufficient, and should be included, but I think it's best to have the canonicals redirected properly in the htaccess.</link rel="canonical" href="<a>
Very best wishes
Trevor
-
Thanks for the kind words Paul.
If you are looking for outstanding SEOs to follow, I would recommend EGOL and Alan Bleiweiss. I merely ride in the wake of their excellence.
Your response jumped around a bit but a few replies I would offer:
-
You are right. The value of most directories has dropped significantly. There are very few that offer any real value nowadays.
-
MVC is the current best practice for web design, but friendly URLs is a separate item. You can achieve them with or without MVC.
-
Most people who complain about their site's ranking drop actually have issues on their site if you look closely. I can't begin to share how many people I have encountered who were insistent their site was outstanding when their site had numerous issues.
-
Likewise, I have worked with clients who were quite upset about other sites that ranked well who referred to them as "junk" sites when those ranking were earned. Yes, there are exceptions and Google still has work to do, but they are doing a reasonable job. The truly bad sites usually disappear in 4-8 weeks.
-
I know nothing about "The Marketing Analysts" but they could have an offline presence or have undergone a name change which may explain the "Since 1989" claim. Let's remember Al Gore didn't invent the internet until about 1996 and there has been tremendous changes since then.
-
-
Hi Egol,
Thank you for your reply. The long folder names are probably from using WordPress as you pointed. I found a blog on their subdomain using WordPress.
I have to say that I've enjoyed reading your responses throughout the QA forum because your responses are short and to the point, pithy and no-BS. So, I'm curious about your response to my question. Above you responded "I doubt it" to the question about Google ferreting out keyword stuffed URL paths. Instead of trying to read between the lines, let me ask you, how good of a job is Google doing? How are they falling short?
Kindest regards,
Paul
-
Hi Trevor!
Thank you for your response! I'm VERY new to the concept of canonical issues. If you not in my other response, I'm just getting back into the game. How much do you think the canonical issue really plays?
Kind regards,
Paul
-
Hi Ryan!!
Man, I'm thrilled to see you responded, and that you responded so thoroughly. I've been reading threads in this QA forum for a few days, and I've come to think of you as a bit of an SEO celebrity! I have to figure out how to filter for questions you've answered! : )
Okay...the site. I've been away from SEO for about eight years and a lot has changed. In the past, I've enjoyed top positions in the SERPs under highly competitive key phrases, even recently (probably because good legacy websites seem to carry weight). Back then, I placed my primary site in directories thinking that people who visited the directories would see my listing and click on it and visit me (as opposed to getting a link to get "juice"). This is probably what has been giving my site good rankings for a while, and the fact that I've never used web-chicanery to outrank others. Over the years, I've seen spammy and trickster sites appear and disappear. I used to rip those sites (the only way to get a global vision of what's going on), and I studied what they did. I've got a curious little archive of living black hat tricks, all of which failed as Google caught on to them.
Now I turn my reflectors back on to what's going on in SEO and what companies and individuals are doing to position themselves in SERPs. I'm saddened to report this, but for all the overhauls, tweaking and tinkering that Google has done since 2001 when I started, spammy sites and sites with poor content, usability, usefulness, and design are still outranking truly useful, high-quality, high-integrity sites.
Very recently, I read complaints by people who felt like their sites had been unfairly affected by the Panda update (http://www.google.com/support/forum/p/Webmasters/thread?tid=76830633df82fd8e&hl=en). I followed the links to "offending" sites (sites people felt ranked higher than theirs for no good reasons), and I went through the code in the complainants' sites as well. Holy cow...many of the complainants have good reason to complain. Shallow, spammy, zero-effort sites are blowing away robust sites with truly useful content. I've NEVER had a sinking feeling in my gut in 10 years that ranking well was a crapshoot - but I got that feeling after studying those complaints.
Years ago I worked in Macromedia Dreamweaver (remember how cool "Macromedia" was?) with regular HTML and nowadays I work in Visual Studio, just recently creating my first MVC3 site. MVC allows you to manipulate every tiny aspect of your site, including the URL path. There is absolutely no relation between the path you see in your browser and the actual path to the files on the server. And you can change the path and name of any page instantly and almost effortlessly. It's GREAT for SEO. So, I've been paying special attention to directory names and page names out there on the Internet. That's when I came across "themarketinganalysts" site and their unusually high rankings for so many important key phrases. After combing through that site, studying the source code, checking their rankings across many key phrases - I have to say, regardless of PA of 53 and keyword variances, the code reminds me of some of the code from spammy trickster sites from the early 90s.
If you hand code html, you get a certain vision for what the page will look like as you type along, from the mind’s eye of a visitor. When you go to a site and the code is packed with keywords, weird use of elements (like themarketinganalystemarketinganalysts' textless use of the H1tag to render the logo through CSS – an old trick to put the
next to the tag), you get the feeling that whoever wrote that code is telling search engines one thing, and visitors something different. It's duplicitous. Oddly enough, I'm not fazed by a company that outranks me (there is enough work for ALL of us), but I want to see healthy optimization, not one story in the code and another on the rendered page.
I'm going to do a more in-depth review of the code, page by page, look for trends and track down the sources that provide PA coefficients (or try to!). I’ll use the Wayback Machine to study the evolution of the site. Off the bat:
Mar 21, 2009 "This website coming soon"
Mar 31, 2009 "PREDICTIVE WEB ANALYTICS" - nothing about translation
May 25, 2009 Starts taking current formOdd. This is claimed on the current site: "Since 1989, The MARKETING ANALYSTS has built its Language Translation Services business..." That claim in not supported by what Wayback Machine shows. Geesh... Did I stumble across enterprise-wide shadiness? Hope not!
I'll come back to you and share my SEO findings.
-
Yep those PAs are strong even without canonicalization. Let's hope for Paul's sake that the site doesn't get an seo audit anytime soon!
-
Really great catch on the canonical issue Trevor! The entire time I just knew I was missing something, and that's it.
The www version of the URL has a PA of 53 which put's it as even stronger then the wiki page. The links mostly use "medical translation" as the anchor text with some "medical translator" and "medical translation service" variances thrown in. The link profile is varied enough to satisfy me the page has earned it's ranking.
-
Hi Ryan I noticed that the site has a canonical issue with both an http and www version too. Nice and thorough analysis, really interesting regarding the flag. Now I'm back home I might just have to take a look....although really should think about getting some shut eye here in blighty
-
I love a great SEO mystery and, for me at least, you have found one. I think this is a case for the famous SEO forensic analyst Alan "Sherlock" Bleiweiss.
I can confirm your overall findings and cannot explain the results. Specifically, on Google.com I searched for "medical translation". The results are listed below.
Result #19: http://en.wikipedia.org/wiki/Medical_translation
PA: 52, DA 98
Title: Medical translation - Wikipedia, the free encyclopedia
H1: Medical translation
First words of content: Medical translation is the translation of technical, regulatory....
Internal links (2): Anchor text on both links is "medical translation". Lowest PA of a linked page is 61. About 1000 links per page.
Title: Medical Translation Services: Pharmaceutical, Equipment, Specifications, Medical Literature, HIPPA, [99 chars in title so display is cut-off]
PA: 12, DA 60
H1: none. H2: Medical Translation: Medical Translation Services: Pharmaceutical, Equipment, Specifications, Medical Literature, HIPPA
First words of content: When it comes to the medical translation, you can trust THE MARKETING ANALYSTS.
Internal links (3): Anchor text on all three "Medical translation". The highest PA from a page is 15. One of the links is from the home page which has 220 links total.
As I try to reach for some other factor that would allow this site to rank so well compared to the wiki page I notice the following:
-
the site has "medical translation" in it's site's navigation bar
-
the site has a link in the left sidebar on the home page directly to the page. The sitebar is a tad spammy with 43 links.
The above two items are factors, but not enough to do it for me.
I still couldn't explain the ranking so I searched the page for the term "medical". It only appeared twice so I performed a "find" which indicated the term was being used many more times on the page but was not visible. After searching the HTML and CSS I determined there was extra hidden content. I could not find anything suspicious in the CSS and was puzzled on how this content was being hidden then I realized the "trick" involved.
Please notice the US/UK flag in the upper-right area of the page. Press it. Viola! The home page contains extra content directly related to Medical Transcription that no one will ever see. The content includes "Medical Transcription" as a H3 tag, a link to the target page, and a nice paragraph.
This technique is squarely black hat. The purpose of a language button is to offer a translation. There is only one button for the language the page is already being presented in, so no one will ever press it. The content is additional text and links which has nothing to do with a translation.
Even so, I find it interesting this content is enough to yield the #1 ranking in SERP. Either there is another factor remaining that I could not locate (I really don't think that is the case but would love to hear from others) or Google is putting more weight to content on the home page. I have always felt home page content was very strong, but this page just is not strong enough to blow the Wiki page away like this at all, unless Google is weighing this home page content quite strongly.
I like the Yahoo results MUCH better for this search. Wiki is #2 and this page is #13. Bing shows Wiki as #5 with this page as #13. I am ok with those ranking as well.
-
-
WordPress produces similar long URLs that match the post title.
-
How much will it help? Very little, except where competition is very lo.
Will Google ferret it out? I doubt it.
-
Hi Paul
Wow! To me that just looks so spammy and over-optimised. I would think that the SE's would think the same too but as you say the urls rank #1.
What are the other metrics like for the site, perhaps they may show the reasons for high rankings?
Update: Just taken a quick look and it does seem the domain is quite strong with a DA 60. Having said that they have a canonical issue which,, if they sorted may make them even stronger.....so keep that quiet!
Got a burning SEO question?
Subscribe to Moz Pro to gain full access to Q&A, answer questions, and ask your own.
Browse Questions
Explore more categories
-
Moz Tools
Chat with the community about the Moz tools.
-
SEO Tactics
Discuss the SEO process with fellow marketers
-
Community
Discuss industry events, jobs, and news!
-
Digital Marketing
Chat about tactics outside of SEO
-
Research & Trends
Dive into research and trends in the search industry.
-
Support
Connect on product support and feature requests.
Related Questions
-
Do you get penalized for Keyword Stuffing in different page URLs?
If i have a website that provides law services in varying towns and we have pages for each town with unique content on each page, can the page URLS look like the following: mysite.com/miami-family-law-attorney mysite.com/tampa-family-law-attorney mysite.com/orlando-family-law-attorney Does this get penalized when being indexed?
White Hat / Black Hat SEO | | Armen-SEO0 -
Forcing Google to Crawl a Backlink URL
I was surprised that I couldn't find much info on this topic, considering that Googlebot must crawl a backlink url in order to process a disavow request (ie Penguin recovery and reconsideration requests). My trouble is that we recently received a great backlink from a buried page on a .gov domain and the page has yet to be crawled after 4 months. What is the best way to nudge Googlebot into crawling the url and discovering our link?
White Hat / Black Hat SEO | | Choice0 -
Does this URL need rewriting?
Hello, Does this URL need to be rewritten? http://www.nlpca.com/DCweb/modelingwithnlparticleandreas.html Bob
White Hat / Black Hat SEO | | BobGW0 -
Removing/ Redirecting bad URL's from main domain
Our users create content for which we host on a seperate URL for a web version. Originally this was hosted on our main domain. This was causing problems because Google was seeing all these different types of content on our main domain. The page content was all over the place and (we think) may have harmed our main domain reputation. About a month ago, we added a robots.txt to block those URL's in that particular folder, so that Google doesn't crawl those pages and ignores it in the SERP. We now went a step further and are now redirecting (301 redirect) all those user created URL's to a totally brand new domain (not affiliated with our brand or main domain). This should have been done from the beginning, but it wasn't. Any suggestions on how can we remove all those original URL's and make Google see them as not affiliated with main domain?? or should we just give it the good ol' time recipe for it to fix itself??
White Hat / Black Hat SEO | | redcappi0 -
Single-words high keyword density. How many is too many.
Dear All SeoMoz users, I'm a web designer for some time now. Doing some basic SEO from time to time. I just started up with brand new website. The website is not ranking very well for 2nd line keyword (keyword density < 2%), but the problem is not ranking at all for for my main keyword. I think the problem is the keyword density. For phrases that are 3-words long my keyword density is less than 4%. I suspect the problem is that keyword density for single-word phrases is between 8-12%. Please note that the 3 words with highest keyword density make my main 3-words long keyword. Is this the case? Should I be avoiding keyword density larger than 4% for single-word phrases as well? What is you experiences is this matter? Could my single-word phrases be treated as keyword stuffing by Google?
White Hat / Black Hat SEO | | pseefeld0 -
Why is the rankings for certain targeted keywords dropping sharply in the past few weeks?
Excuse me here I can't reveal the website url here due to the confidentiality. Background: This website has about 100+pages, and is a wordpress site. Out of 15 targeted keywords, there are few ranked first on Google.co.uk. others are ranked between 3rd and 50+. In the past couple of weeks, we have been submitting to various websites/directories. and 4 main keywords' rankings are improving steadily from 50+ to 20+. We don't buy links, or pay for link exhcanges. We only exchange links on rare occasions. Symptoms: 1. 2 out of 4 main keywords dropped its rankings from 20+ to 80+ 4 weeks ago over night. 2. We kept on acquiring links from directories and websites, and 2 weeks after the drop, the rankings of 2 affected keywords had gradually made its way back to 20-30. Just we thought the glitch was over, these 2 keywords has dropped rankings to 50+ once again. 3. The change of rankings looks too suspicious to us, so we went to use the Moz tool, and discovered the domain authority has also dropped 25% of its value in a month time. We never experienced such violent changes in terms of rankings in such a short period of time. My question is: what factors there are to have caused such violent shifts? Why the value of the domain authority is dropped by 25%-30%? What elements affect the domain authority the most?
White Hat / Black Hat SEO | | robotseo0 -
How much pain can I expect if I change the URL structure of the site again?
About 3 months ago I implemented a massive URL structure change by 'upgrading' some of the features of our CMS Prior to this URL's for catergorys and products looked something like this http://www.thefurnituremarket.co.uk/proddetail.asp?prod=OX09 I made a few changes but din't implement it fully as I felt it would be better to do it instages as the site was getting indexed more thouroughly. HOWEVER... We have just hit the first page for some key SERP's and I am wary to rock the boat again by changing the URL structures again and all the sitemaps. How much pain do you think we could feel if i went ahead and optimised the URL's fully? and What would you do? 🙂
White Hat / Black Hat SEO | | robertrRSwalters0