Moz Q&A is closed.
After more than 13 years, and tens of thousands of questions, Moz Q&A closed on 12th December 2024. Whilst we’re not completely removing the content - many posts will still be possible to view - we have locked both new posts and new replies. More details here.
Setting A Custom User Agent in Screaming Frog
- 
					
					
					
					
 Hi all, Probably a dumb question, but I wanted to make sure I get this right. How do we set a custom user agent in Screaming Frog? I know its in the configuration settings, but what do I have to do to create a custom user agent specifically for a website? Thanks much! - Malika
 
- 
					
					
					
					
 Setting a custom user agent determines things like HTTP/2 so there can be a big difference if you change it to something that might not take advantage of something like HTTP/2 Apparently, it is coming to Pingdom very soon just like it is to Googlebot http://royal.pingdom.com/2015/06/11/http2-new-protocol/ This Is an excellent example of a user agent's ability to modify the way your site is crawled as well as how efficient it is. https://www.keycdn.com/blog/https-performance-overhead/ It is important to note that we didn’t use Pingdom in any of our tests because they use Chrome 39, which doesn’t support the new HTTP/2 protocol. HTTP/2 in Chrome isn’t supported until Chrome 43. You can tell this by looking at the User-Agentin the request headers of your test results. Pingdom user-agent Note: WebPageTest uses Chrome 47 which does support HTTP/2. Hope that clears things up, Tom 
- 
					
					
					
					
 Hi Malika, Think about screaming frog and what it has to detect in order to do that correctly it needs the correct user agent syntax for it will not be able to make a crawl that would satisfy people. Using a proper syntax for a user agent is essential and I have tried to be non-technical in this explanation I hope it works. the reason screaming frog needs the user agent because the user-agent was added to HTTP to help web application developers deliver a better user experience. By respecting the syntax and semantics of the header, we make it easier and faster for header parsers to extract useful information from the headers that we can then act on. Browser vendors are motivated to make web sites work no matter what specification violations are made. When the developers building web applications don’t care about following the rules, the browser vendors work to accommodate that. It is only by us application developers developing a healthy respect When the developers building web applications don’t care about following the rules, the browser vendors work to accommodate that. It is only by us application developers developing a healthy respect It is only by us application developers developing a healthy respect for the standards of the web, that the browser vendors will be able to start tightening up their codebase knowing that they don’t need to account for non-conformances. For client libraries that do not enforce the syntax rules, you run the risk of using invalid characters that many server side frameworks will not detect. It is possible that only certain users, in particular, environments would identify the syntax violation. This can lead to difficult to track down bugs. I hope this is a good explanation I've tried to keep it very to the point. Respectfully, Thomas 
- 
					
					
					
					
 Hi Thomas, would you have a simpler tutorial for me to understand? I am struggling a bit. Thanks heaps in advance  
- 
					
					
					
					
 I think I want something that is dumbed down to my level for me to understand. The above tutorials are great but not being a full time coder, I get lost while reading those. 
- 
					
					
					
					
 Hi Matt, I havent had a luck with this one yet.  
- 
					
					
					
					
 Hi Malika! How'd it go? Did everything work out?  
- 
					
					
					
					
 happy I could be of help let me know if there's any issue and I will try to be of help with it. All the best 
- 
					
					
					
					
 Hi Thomas, That's a lot of useful information there. I will have a go on it and let you know how it went.  Thanks heaps! 
- 
					
					
					
					
 please let me know if I did not answer the question or you have any other questions 
- 
					
					
					
					
 this gives you a very clear breakdown of user agents and their set of syntax rules. The following is valid example of user-agent that is full of special characters, read this please http://www.bizcoder.com/the-much-maligned-user-agent-header user-agent: foo&bar-product!/1.0a$*+ (a;comment,full=of/delimitersreferences but you want to pay attention to the first URL https://developer.mozilla.org/en-US/docs/Web/HTTP/Gecko_user_agent_string_reference | Mozilla/5.0 (X11; Linux i686; rv:10.0) Gecko/20100101 Firefox/10.0 | http://stackoverflow.com/questions/15069533/http-request-header-useragent-variable 
- 
					
					
					
					
 if you formatted it correctly see below User-Agent = product *( RWS ( product / comment ) )and it was received by your headers yes you could fill in the blanks and test it. https://mobiforge.com/research-analysis/webviews-and-user-agent-strings http://mobiforge.com/news-comment/standards-and-browser-compatibility 
- 
					
					
					
					
 No, you Cannot just put anything in there. The site has to recognize it and ask why you are doing this? I have listed how to build and already built in addition to what your browser will create by using useragentstring.com Must be formatted correctly and have it work with a header it is not as easy as it sometimes seems but not that hard either. You can make & use this to make your own from your Mac or PC http://www.useragentstring.com/ Mozilla/5.0 (Macintosh; Intel Mac OS X 10_11_5) AppleWebKit/537.36 (KHTML, like Gecko) Chrome/53.0.2747.0 Safari/537.36 how to build a user agent - https://developer.mozilla.org/en-US/docs/Web/HTTP/Gecko_user_agent_string_reference
- https://developer.mozilla.org/en-US/docs/Setting_HTTP_request_headers
- https://msdn.microsoft.com/en-us/library/ms537503(VS.85).aspx
 Lists of user agents https://support.google.com/webmasters/answer/1061943?hl=en https://msdn.microsoft.com/en-us/library/ms537503(v=vs.85).aspx 
- 
					
					
					
					
 Hi Thomas, Thanks for responding, much appreciated! Does that mean, if I type in something like - HTTP request user agent - Crawler access V2 & Robots user agent Crawler access V2 This will work too? 
- 
					
					
					
					
 To crawl using a different user agent, select ‘User Agent’ in the ‘Configuration’ menu, then select a search bot from the drop-down or type in your desired user agent strings. http://i.imgur.com/qPbmxnk.png & Video http://cl.ly/gH7p/Screen Recording 2016-05-25 at 08.27 PM.mov Or Also see http://www.seerinteractive.com/blog/screaming-frog-guide/ https://www.screamingfrog.co.uk/seo-spider/user-guide/general/#user-agent https://www.screamingfrog.co.uk/seo-spider/user-guide/ https://www.screamingfrog.co.uk/seo-spider/faq/ 
Got a burning SEO question?
Subscribe to Moz Pro to gain full access to Q&A, answer questions, and ask your own.
Browse Questions
Explore more categories
- 
		
		Moz ToolsChat with the community about the Moz tools. 
- 
		
		SEO TacticsDiscuss the SEO process with fellow marketers 
- 
		
		CommunityDiscuss industry events, jobs, and news! 
- 
		
		Digital MarketingChat about tactics outside of SEO 
- 
		
		Research & TrendsDive into research and trends in the search industry. 
- 
		
		SupportConnect on product support and feature requests. 
Related Questions
- 
		
		
		
		
		
		Duplicate without user-selected canonical excluded
 We have pdf files uploaded in the media of wordpress and used in our website. As these pdfs are duplicate content of the original publishers, we have marked links to these pdf urls as nofollow. These pages are also disallowed in robots.txt Now, Google Search Console has shown these pages Excluded as "Duplicate without user-selected canonical" As it comes out we cannot use canonical tag with pdf pages so as to point to the original pdf source If we embed a pdf viewer in our website and fetch the pdfs by passing the urls of the original publisher, would the pdfs be still read as text by google and again create duplicate content issue? Another thing, when the pdf expires and is removed, it would lead to 404 error. If we direct our users to the third party website, then it would add up to our bounce rate. What should be the appropriate way to handle duplicate pdfs? Thanks Intermediate & Advanced SEO | | dailynaukri1
- 
		
		
		
		
		
		Could I set a Cruise as an Event in Schema mark up?
 Hi there, We are now in the process of implementing a JSON-LD mark-up solution and are building cruises as an event. Will this work and can we get away with this without penalty? Previously they have been marking their cruises as events using the data highlighter and this has displayed correctly in the SERP. The ideal schema would be Trip but this is not supported by Google Rich Results yet, hopefully they will support this in the future. Another alternative would be product but this does not display rich-results as we would like. Event has the best result in terms of how the information is displayed. For example someone might search "Cruises to Spain" and the landing page would display the next 3 cruises that go to Spain, with dates & prices. The event location would be the cruise terminal, the offer would be the starting price and the start & end date would be the cruise duration, these are fixed dates. I am interested to hear the communities opinion and experience with this problem. Intermediate & Advanced SEO | | NoWayAsh1
- 
		
		
		
		
		
		Best Permalinks for SEO - Custom structure vs Postname
 Good Morning Moz peeps, I am new to this but intending on starting off right! I have heard a wealth of advice that the "post name" permalink structure is the best one to go with however... i am wondering about a "custom structure" combing the "post name" following the below example structure: Www.professionalwarrior.com/bodybuilding/%postname/ Where "professional" and "bodybuilding" is my focus/theme/keywords of my blog that i want ranked. Thanks a mill, RO Intermediate & Advanced SEO | | RawkingOut0
- 
		
		
		
		
		
		How does Infinite Scrolling work with unique URLS as users scroll down? And is this SEO friendly?
 I was on a site today and as i scrolled down and viewed the other posts that were below the top one i read, i noticed that each post below the top one had its own unique URL. I have not seen this and was curious if this method of infinite scrolling is SEO friendly. Will Google's spiders scroll down and index these posts below the top one and index them? The URLs of these lower posts by the way were the same URLs that would be seen if i clicked on each of these posts. Looking at Google's preferred method for Infinite scrolling they recommend something different - https://webmasters.googleblog.com/2014/02/infinite-scroll-search-friendly.html . Welcome all insight. Thanks! Christian Intermediate & Advanced SEO | | Sundance_Kidd0
- 
		
		
		
		
		
		Screaming frog Advice
 Hi I am trying to crawl my site and it keeps crashing. My sys admins keeps upgrading the virtual box it sits on and it now currently has 8GB of memory, but still crashes. It gets to around 200k pages crawl and dies. Any tips on how I can crawl my whole site, can u use screaming frog to crawl part of a site. Thanks in advance for any tips. Andy Intermediate & Advanced SEO | | Andy-Halliday0
- 
		
		
		
		
		
		Should I set up no index no follow on low quality pages?
 I know it is a good idea for duplicate pages, blog tags, etc. but I remember somewhere that you can help the overall link juice of a website by adding no index no follow or no index follow low quality content pages of your website. Is it still a good idea to do this or was it never a good idea to begin with? Michael Intermediate & Advanced SEO | | Michael_Rock0
- 
		
		
		
		
		
		Setting up 301 Redirects after acquisition?
 Hello! The company that I work for has recently acquired two other companies. I was wondering what the best strategy would be as it relates to redirects / authority. Please help! Thanks Intermediate & Advanced SEO | | Colin.Accela0
- 
		
		
		
		
		
		How to set cannonical link rel to CS CART
 I whant to specify a link rel cannonical for each category page, how to do that without changing the code (just from admin section), because filters and sorting search are making the site dublicate content with their parameters; If there is a way please specify the method, i whant to avoid hours of working in a script like this. Thank's. Intermediate & Advanced SEO | | oneticsoft0
 
			
		 
			
		 
			
		 
					
				 
					
				 
					
				 
					
				 
					
				 
					
				 
					
				