Repeated mysterious 404's from ancient site structure killing my rankings
-
Several years ago I changed my site structure to go from a flash based site to a blog based wordpress site. After doing so I went from page 1 to page 30 for my relevant search terms. I have employed people to help me track down the problem and I believe that they have narroed it to the existance of 404's being created from some unknown internal source. I have been for years getting links like this...
<colgroup><col width="792"></colgroup>
||
......regularly showing in webmaster tools, (this is from a top pages report from MOZ where there are hundreds also shown).
When I do a moz crawl of the site, none of these links show up. Therefore I have no way of finding the source of these links (they also do not show me the source in WMT as they should).
We have completely cleared the site and rebuilt it and although it is still only a couple of weeks in it still does not appear to have stopped them.
Does anyone have any way of helping me find the source of these mysterious 404's?
-
Why bother trying to clean anything up? If somewhere out there there are links to your domain, and they're 404'ing, just 301 them to new pages on your site! Capture that link juice, don't let it run out
-
Thanks for your reply EEE3
The ancient link says it is linked from another non existent ancient page that no longer exists and it is always first crawled and last detected on the day that it arrives.
eg. last crawled 4/23/14, first detected 4/23/14
http://www.dfphotographer.com.au/brisbaneweddingphotographer/2011/03/st-kilda-wedding......
linked from
http://dfphotographer.com.au/brisbaneweddingphotographer/index.php/2011/03/st-kilda-wedding.....
and
http://dfphotographer.com.au/brisbaneweddingphotographer/2011/03/st-kilda-wedding....
-
Thanks for your response Keri,
Being staff can you please tell me where does the top pages data come from? Is it from crawling my site (like a google spider) or is it sourced from google or somewhere else. How often is that data refreshed?
In answer to your response, I have tried both screaming frog and xenu and my nice clean site structure is all it picks up. None of the ancient messy site structure appears.
Have been through the list of domains looking for an old sitemap or something similar that may have been scraped off my site but after a long and arduous task could not locate any reference to any of these links that show up in top pages and webmaster tools (which says they are linked from other ancient pages - which I will expand on below)
We have looked at all the usual suspects - old sitemaps, plugins and rebuilt the site just in case we missed anything that was lingering around. I have had really good people looking at it who continue to do so it just never seems to go away.
-
In Webmaster Tools, when you click on the 404 and the popup window appears, what is showing in the Linked from tab?
-
I edited the post so the URLs didn't run together. Still not perfect, but a little easier to read.
I'm not exactly sure where those links are coming from. You might run a tool like Xenu Link Sleuth or Screaming Frog on your site to see if there is an internal linking widget gone awry. The other thought I have is to look at Open Site Explorer to see what sites are linking to you and if they're linking to any of those pages.
Got a burning SEO question?
Subscribe to Moz Pro to gain full access to Q&A, answer questions, and ask your own.
Browse Questions
Explore more categories
-
Moz Tools
Chat with the community about the Moz tools.
-
SEO Tactics
Discuss the SEO process with fellow marketers
-
Community
Discuss industry events, jobs, and news!
-
Digital Marketing
Chat about tactics outside of SEO
-
Research & Trends
Dive into research and trends in the search industry.
-
Support
Connect on product support and feature requests.
Related Questions
-
Does the Moz Pro site crawl, crawl password protected sites?
So i asked Moz Pro site crawl to crawl my page, and a lot of issues came up - but for password protected sites. Does the Moz Pro site crawl do this? A lot of the issues, are not relevant for a site that is password protected.
Link Explorer | | Minlaering.dk0 -
Why aren't my "page social metrics" increasing?
I post a lot to Facebook & twitter, & my "page social metrics" haven't budged in 4 weeks. Â I even stopped using bit.ly & stated the full URL as a test. Â Still no change. Â The fb account is /getgoodgifts & twitter is /giftsing. Â Thoughts on why social metrics aren't increasing?
Link Explorer | | giftsing0 -
GTLDs in Open Site Explorer
Curious to know if anyone has run into a similar problem - Â I entered a site containing a gTLD (.dog extensions) in OSE and it gave me an error (screenshot attached). I checked another tool (ahrefs.com) and was able to extract some data. Does anyone know if these type of domains are not included in OSE and/or why there was an error? Appreciate any insight! RSA04bYsL
Link Explorer | | ATShock0 -
Campaign shows website links from Https. My site is not https but http:// HELP
When looking at my web campaign, I have inbound links pointing from every page located on my menu from https:// Â I do not have an https. website in any way. The weird part is that when I click on the link, it is a replica of my homepage, although in text only format. I have attached a picture for reference. Why and how could my top links be coming from this? I do not have any inbound links non https:// from my website Thank you! BNHHc5u
Link Explorer | | Morg56850 -
When the site explorer will recognise the new TLD?
Hey, I own a website, called tipster.website I noticed that the site explorer doesnt recognise the domain extension. When they will be added too? I dont know if the other extensions work , but this one definitely doesnt work. ps: majestic.com also don't recognise the extension.
Link Explorer | | nyanainc0 -
Open Site Explorer is finding old html Files that havn't been on my site in two years... even after a 301 Redirect. HELP!
Hello!
Link Explorer | | morganlindsaycole
My problem started when I became aware that when I checked my backlinks for the past two years, it states that no backlinks have been found.  When I ran a site analysis on SEMrush - No backlinks are found on the URL, or Domain. There are 7 Backlinks on the Root Domain and those were configured in 2012. I made a second domain www.columbusweddingphotographersreviews.comwhere I linked to my domain at www.morganlindsayphotography.com so I could test that google had crawled both websites and after, still no backlink was found. I have also been published on a dozen or so wedding websites that has linked to my website where they are follow links and still nothing. (http://www.brendasweddingblog.com/blogs/2015/2/23/an-elegant-fall-wedding-in-ohio-with-morgan-lindsay-photography) **Website Background-  **
In 2012 I had two separate websites - One for Seniors that was an HTML website I build in Dreamweaver at www.morganlindsayphotography/seniors  and another for Wedding Clients found at www.morganlindsayphotography.com/Wedding - (wordpress) I had a Splash page wish was found atwww.morganlindsayphotography.com. Two years ago when I became aware splash pages were frowned upon in Google, I combined the two websites and stayed with the Wordpress which was www.morganlindsayphotography.com/WeddingÂ
Because I did not want users to have to go to www.morganlindsayphotography.com/Wedding to view my url, Godaddy moved my wordpress site from thewww.morganlindsyphotography.com/Wedding directory towww.morganlindsayphotography.com When I ran the Open Site Explorer with Moz I found after runningwww.morganlindsayphotography.com the TOP pages on this domain according to Page Authority are old HTML files from my senior website, as well as old Posts from when my wordpress site was found atwww.moragnlindsayphotography.com/WeddingsÂ
No current pots or pages are showing up besideswww.morganlindsyphotography.com I do run a cache management system to speed up my system and recently cleaned out my .htcacess folder and still had no luck. This is difficulty something **Last night I made a 301 Redirect in my htaccess for all the old links pointing to the new links as best as I could.  My htacess folder looks like this.. BEGIN WordPress <ifmodule mod_rewrite.c="">RewriteEngine On
RewriteBase /
RewriteRule ^index.php$ - [L]
RewriteCond %{REQUEST_FILENAME} !-f
RewriteCond %{REQUEST_FILENAME} !-d
RewriteRule . /index.php [L]</ifmodule> END WordPress Permanent URL redirect - generated by www.rapidtables.com Redirect 301 /Wedding http://www.morganlindsayphotography.com/ Permanent URL redirect - generated by www.rapidtables.com Redirect 301 /Wedding/ http://www.morganlindsayphotography.com/ Permanent URL redirect - generated by www.rapidtables.com Redirect 301 /about.html http://www.morganlindsayphotography.com/about-morgan-lindsay/ Permanent URL redirect - generated by www.rapidtables.com Redirect 301 /app.html http://www.morganlindsayphotography.com/blog/ Permanent URL redirect - generated by www.rapidtables.com Redirect 301 /experience.html http://www.morganlindsayphotography.com/senior-sessions/ Permanent URL redirect - generated by www.rapidtables.com Redirect 301 /index.html http://www.morganlindsayphotography.com/ Permanent URL redirect - generated by www.rapidtables.com Redirect 301 /senior.html http://www.morganlindsayphotography.com/ohio-senior-photographer/ Permanent URL redirect - generated by www.rapidtables.com Redirect 301 /seniorsconstruction.html http://www.morganlindsayphotography.com/ohio-senior-photographer/ Permanent URL redirect - generated by www.rapidtables.com Redirect 301 /Wedding/2012/06/22/brittany-reis-jason-mcclaflin-tiffin-ohio-wedding/ http://www.morganlindsayphotography.com/holy-family-church-columbus-wedding/ After I ran the open site moz explorer and the www.morganlindsayphotography/Wedding was still there..0 -
Hi guys. My site, www.x-mini.com attained more links and got better alexa ranking. However, my DA and PA dropped. How can I explain this?
My site, www.x-mini.com attained more links and got better alexa ranking. However, my DA and PA dropped. How can I explain this? Siz5lkp.png
Link Explorer | | Dineshr840 -
Competitor analysis, why they rank so much better [ecommerce/Magento]
For a Dutch e-commerce website, we're having some issuse being found & ranked on important keywords. I have done a detailed view into the content & technical part (html) of our website and a competitor's site. Our site is www.hond.nl  (hond is the Dutcn noun dog, the best domain we could get)
Link Explorer | | Canome79
Competitor: www.obobo.nl (no meaning at all) For searches for instance by dogfood (Dutch: "hondenvoer") we rank really bad while our competitor ranks really well. I've gone trouch the moz-tools intensively but can't figure out why. We got more content, more self-written texts, more incoming root-domain links etc. Any ideas were we could get a solution? Seems that our Splash pages are doing "ok" but especially Category pages are listing badly.0