Moz Q&A is closed.
After more than 13 years, and tens of thousands of questions, Moz Q&A closed on 12th December 2024. Whilst we’re not completely removing the content - many posts will still be possible to view - we have locked both new posts and new replies. More details here.
Crawlers crawl weird long urls
- 
					
					
					
					
 I did a crawl start for the first time and i get many errors, but the weird fact is that the crawler tracks duplicate long, not existing urls. For example (to be clear): there is a page: www.website.com/dogs/dog.html but then it is continuing crawling: 
 www.website.com/dogs/dog.html
 www.website.com/dogs/dogs/dog.html
 www.website.com/dogs/dogs/dogs/dog.html
 www.website.com/dogs/dogs/dogs/dogs/dog.html
 www.website.com/dogs/dogs/dogs/dogs/dogs/dog.htmlwhat can I do about this? Screaming Frog gave me the same issue, so I know it's something with my website 
- 
					
					
					
					
 Answer from Screaming Frog! The reason the SEO spider is crawling these URLs, is due to incorrect relative linking on the site from the login URL. 
 It's actually when the spider crawls the login page, http://www.website.com/login?returnurl=%2F which then leads to this URL http://www.website.com/Home/ctl/SendPassword?returnurl=http:/www.website.com/ and then this /home/ sub directory URL http://www.website.com/Home/ctl/page/dogs.aspx which links to http://www.website.com/Home/ctl/page/page/dogs.aspx and so on and so forth. This is the path to the incorrect relative linking (attached for you).To stop this, you can correct the incorrect relative linking, or easier, simply exclude the login page. 
- 
					
					
					
					
 Wow, Big mistakes are made one Home maybe because of the .aspx. extension? alle pages have seo-friendly urls Thanks Wesley and Paddy Displays 
- 
					
					
					
					
 I see a link to http://www.odin-groep.nl/Home/ctl/OverOdin/OverOdin/HeutinkICT.aspx from http://www.odin-groep.nl/Home/ctl/OverOdin/ReindersICT.aspx. It's the bottom left block which causes this link. This way you will get a big nesting effect. 
- 
					
					
					
					
 OK found one problem on this page http://www.odin-groep.nl/Home/ctl/OverOdin/ReindersICT.aspx you have a link to http://www.odin-groep.nl/Home/ctl/OverOdin/OverOdin/LesscherIT.aspx which i think should be 
- 
					
					
					
					
 ok I did a quick screaming fog and I think I have an idea, you just have to follow the breadcrumbs You said in you example "In Links 9", you need to find out what those pages are and follow it back to the point of origin As I think its just one bad link that cause this nested link effect. eg http://www.odin-groep.nl/Home/ctl/OverOdin/OverOdin/OverOdin/OverOdin/HeutinkICT.aspx is being linked from http://www.odin-groep.nl/Home/ctl/OverOdin/OverOdin/OverOdin/StationtoStation.aspx (as well as others) You just have to follow that trail till you find the source of the problem 
- 
					
					
					
					
 every link, except the hompage itself 
- 
					
					
					
					
 I can't see any source: The pages are like: | URL | www.website.com/page/ | 
 | Status Code | 200 |
 | Status | OK |
 | Type | text/html; charset=utf-8 |
 | Size | 55811 |
 | Title | |
 | Level | 10 |
 | In Links | 9 |
 | Out Links | 38 |
- 
					
					
					
					
 Which URL(s) is/are causing problems? 
- 
					
					
					
					
 please be free to check: http://tinyurl.com/lox7le9 
- 
					
					
					
					
 You don't necessarily have to remove the link. As long as you can verify that it directs to the right page. But curious to see what caused the problem  
- 
					
					
					
					
 I think Screaming Frog will tell you the page it found the weird url, then you can check the source, and find out whats producing that link. 
- 
					
					
					
					
 That is a good one! It's true that I have the same linking to the page itself. I will remove all that kind of links first and crawl again. I'll keep you in touch! 
- 
					
					
					
					
 Are you somehow linking to www.website.com/dogs/dog.html from the page itself? There could be something wrong with that link. 
 I made a small mistake not so long ago with a redirection plugin. I told it to go to domain.com. This plugin was looking at the base + what i told it to. So it went to: domain.com/domain.com. Perhaps you made a similar mistake.Maybe you can send me the URL and i can take a look at it? 
Got a burning SEO question?
Subscribe to Moz Pro to gain full access to Q&A, answer questions, and ask your own.
Browse Questions
Explore more categories
- 
		
		Moz ToolsChat with the community about the Moz tools. 
- 
		
		SEO TacticsDiscuss the SEO process with fellow marketers 
- 
		
		CommunityDiscuss industry events, jobs, and news! 
- 
		
		Digital MarketingChat about tactics outside of SEO 
- 
		
		Research & TrendsDive into research and trends in the search industry. 
- 
		
		SupportConnect on product support and feature requests. 
Related Questions
- 
		
		
		
		
		
		WEbsite cannot be crawled
 I have received the following message from MOZ on a few of our websites now Our crawler was not able to access the robots.txt file on your site. This often occurs because of a server error from the robots.txt. Although this may have been caused by a temporary outage, we recommend making sure your robots.txt file is accessible and that your network and server are working correctly. Typically errors like this should be investigated and fixed by the site webmaster. I have spoken with our webmaster and they have advised the below: The Robots.txt file is definitely there on all pages and Google is able to crawl for these files. Moz however is having some difficulty with finding the files when there is a particular redirect in place. For example, the page currently redirects from threecounties.co.uk/ to https://www.threecounties.co.uk/ and when this happens, the Moz crawler cannot find the robots.txt on the first URL and this generates the reports you have been receiving. From what I understand, this is a flaw with the Moz software and not something that we could fix form our end. _Going forward, something we could do is remove these rewrite rules to www., but these are useful redirects and removing them would likely have SEO implications. _ Has anyone else had this issue and is there anything we can do to rectify, or should we leave as is? Moz Pro | | threecounties0
- 
		
		
		
		
		
		Pages with URL Too Long
 Hello Mozzers! MOZ keeps kindly telling me the URLs are too long. However, this is largely due to the structure of E-commerce site, which has to include 'brand' 'range' and 'products' keyword. For example - Moz Pro | | tigersohelll
 https://www.choicefurnituresuperstore.co.uk/Devonshire-Rustic-Oak-Bedside-Cabinet-1-Drawer-p40668.html MOZ recommends no more than 75 characters. This means we have 25-30 characters for both the brand name and product name. Questions:
 If it is an issue, how to fix it on my site?
 If it's not an issue, how can we turn off this alert from MOZ?
 Anyone know how big an issue URLs are as a ranking factor? I thought pretty low.0
- 
		
		
		
		
		
		How long do changes in title tags take to affect SEO?
 This is kind of a loaded question. I'm completely new to SEO. I think my boss signed up for Moz Pro sometime in February and started adding data to our Ecommerce site to help with rankings. Sometime before this, I changed some of the title tags on the site (trying to help with organic search and CTR). I did not do a site wide change.... just changed maybe 10-20 (just a guess). I did it with keywords in mind but did not make note of when I did it. I didn't really know better at the time, and I did not have access to Google Analytics or Moz Pro. I was looking through the ranking data/graph for February and March. It won't let me look before February 29th (so that's why I think my boss started the Mos Pro subscription around at that time). On that day it said we ranked 12 keywords in the 1-3 spot, and then the following week (march 7) it went down to 6. I don't think or know if any major site changes were implemented, so I'm not sure why that happened and if it has anything to do with my title tag changes I did maybe a week or two before (again I am not sure when I did this unfortunately). Since then the keyword ranking numbers stayed about the same with organic traffic slowly going down (it could be because we are getting out of season for our industry though). The second week of March the site was upgraded and since then the menu has been completely changed around. Last week I did a site wide title tag change. So the minor changes I made in February are no longer in effect anyway. I added more keywords to Moz earlier this week and the number for 1-3 spot keywords went up from 6 to 20. It also says my ranking moved up 4 keywords and down 13 keywords. Anyway, I am wondering how seriously I should take these changes and if I'm damaging the site. I am new to Moz Pro also so all the data you can access is kind of confusing/overwhelming. Moz Pro | | AliMac260
- 
		
		
		
		
		
		Block Moz (or any other robot) from crawling pages with specific URLs
 Hello! Moz reports that my site has around 380 duplicate page content. Most of them come from dynamic generated URLs that have some specific parameters. I have sorted this out for Google in webmaster tools (the new Google Search Console) by blocking the pages with these parameters. However, Moz is still reporting the same amount of duplicate content pages and, to stop it, I know I must use robots.txt. The trick is that, I don't want to block every page, but just the pages with specific parameters. I want to do this because among these 380 pages there are some other pages with no parameters (or different parameters) that I need to take care of. Basically, I need to clean this list to be able to use the feature properly in the future. I have read through Moz forums and found a few topics related to this, but there is no clear answer on how to block only pages with specific URLs. Therefore, I have done my research and come up with these lines for robots.txt: User-agent: dotbot Moz Pro | | Blacktie
 Disallow: /*numberOfStars=0 User-agent: rogerbot
 Disallow: /*numberOfStars=0 My questions: 1. Are the above lines correct and would block Moz (dotbot and rogerbot) from crawling only pages that have numberOfStars=0 parameter in their URLs, leaving other pages intact? 2. Do I need to have an empty line between the two groups? (I mean between "Disallow: /*numberOfStars=0" and "User-agent: rogerbot")? (or does it even matter?) I think this would help many people as there is no clear answer on how to block crawling only pages with specific URLs. Moreover, this should be valid for any robot out there. Thank you for your help!0
- 
		
		
		
		
		
		How to force SeoMoz to re-crawl my website?
 Hi, I have done a lot of changes on my website to comply with SeoMoz advices, now I would like to see if I have better feedback from the tool, how can I force it to re-crawl a specific campaign? (waiting another week is too long :-)) Moz Pro | | oumma0
- 
		
		
		
		
		
		Site Explorer - No Data Available for this URL
 Hi All I have just joined on the trial offer, im not sure if i can afford the monthly payments, but im hoping SEOmoz will show me that i also cannot afford to be without it! In my proses of learning this site and flicking through each section to see what things do. However when i enter my URL into Site Explorer i get the following message "No Data Available for this URL" My site should be crawl-able, so how do i get to see data for my site/s. I wont post my URL here, as the site has a slightly adult theme. Moz Pro | | jonny512379
 If anyone could confirm if i can post "slightly adult" sites. Best Regards
 Jon0
- 
		
		
		
		
		
		Duplicate page titles are the same URL listed twice
 The system says I have two duplicate page titles. The page titles are exactly the same because the two URLs are exactly the same. These same two identical URLs show up in the Duplicate Page Content also - because they are the same. We also have a blog and there are two tag pags showing identical content - I have blocked the blog in robots.txt now, because the blog is only for writers. I suppose I could have just blocked the tags pages. Moz Pro | | loopyal0
- 
		
		
		
		
		
		How long does a crawl take?
 A crawl of my site started on the 8th July & is still going on - is there something wrong??? Moz Pro | | Brian_Worger1
 
			
		 
			
		 
					
				 
					
				 
					
				 
					
				 
					
				 
					
				 
					
				