<?xml version="1.0" encoding="UTF-8"?><rss version="2.0"
	xmlns:content="http://purl.org/rss/1.0/modules/content/"
	xmlns:wfw="http://wellformedweb.org/CommentAPI/"
	xmlns:dc="http://purl.org/dc/elements/1.1/"
	xmlns:atom="http://www.w3.org/2005/Atom"
	xmlns:sy="http://purl.org/rss/1.0/modules/syndication/"
	xmlns:slash="http://purl.org/rss/1.0/modules/slash/"
	>

<channel>
	<title>Emma Labrador, Author at Relevance</title>
	<atom:link href="https://www.relevance.com/author/emma-labrador/feed/" rel="self" type="application/rss+xml" />
	<link>https://www.relevance.com/author/emma-labrador/</link>
	<description>Award-winning Marketing Agency</description>
	<lastBuildDate>Thu, 04 Dec 2025 20:56:40 +0000</lastBuildDate>
	<language>en-US</language>
	<sy:updatePeriod>
	hourly	</sy:updatePeriod>
	<sy:updateFrequency>
	1	</sy:updateFrequency>
	<generator>https://wordpress.org/?v=7.1</generator>

<image>
	<url>https://www.relevance.com/wp-content/uploads/2025/12/cropped-relevance_favicon_transparent_new-32x32.png</url>
	<title>Emma Labrador, Author at Relevance</title>
	<link>https://www.relevance.com/author/emma-labrador/</link>
	<width>32</width>
	<height>32</height>
</image> 
	<item>
		<title>A Guide to Log File Analysis – How to Open the Google Blackbox</title>
		<link>https://www.relevance.com/a-guide-to-log-file-analysis-how-to-open-the-google-blackbox/</link>
					<comments>https://www.relevance.com/a-guide-to-log-file-analysis-how-to-open-the-google-blackbox/#respond</comments>
		
		<dc:creator><![CDATA[Emma Labrador]]></dc:creator>
		<pubDate>Tue, 05 Jul 2016 10:30:08 +0000</pubDate>
				<category><![CDATA[SEO]]></category>
		<guid isPermaLink="false">https://www.relevance.com/?p=43214</guid>

					<description><![CDATA[<p>Log file analysis is something you should take into consideration as it can improve your overall digital marketing strategy. Log analysis has an impact on visibility, traffic, conversions, sales and helps reveal new points of SEO improvements. Log files are the only data that are 100% accurate to really get how bots are crawling your&#8230;</p>
<p>The post <a href="https://www.relevance.com/a-guide-to-log-file-analysis-how-to-open-the-google-blackbox/">A Guide to Log File Analysis – How to Open the Google Blackbox</a> appeared first on <a href="https://www.relevance.com">Relevance</a>.</p>
]]></description>
										<content:encoded><![CDATA[<p>Log file analysis is something you should take into consideration as it can improve your overall digital marketing strategy. Log analysis has an impact on visibility, traffic, conversions, sales and helps reveal new points of SEO improvements. Log files are the only data that are 100% accurate to really get how bots are crawling your website.</p>
<p>A log file is actually a file output made from a web server containing ‘hits’ or records of all requests that the server has received. Data are stored and deliver details about the time and date in which the request was made, the URL requested, the user agent, the request ID address, and other interesting details. To explain it quickly, log file analysis allows you to get information about SEO visits and to see what the Googlebot is actually doing on your website. You can thus cross it with your crawl data and see further.</p>
<p>Let’s see the advantages of what you can get with log file analysis.</p>
<h2>The advantages of log file analysis</h2>
<h3>Why is it useful?</h3>
<p>Log file analysis is useful for many reasons:</p>
<ul>
<li>For your audits :
<ul>
<li>You can diagnosis useful and useless pages</li>
<li>You can detect zones that Google crawls</li>
<li>You can know which pages Google does not know</li>
</ul>
</li>
<li>For your monitoring :
<ul>
<li>You can get alerts and avoid waiting for a Google Webmaster Tool message</li>
<li>You can monitor optimizations or deployment more easily</li>
<li>You can anticipate attacks</li>
</ul>
</li>
</ul>
<h3>Why should you use log file analysis?</h3>
<p>You can exactly know what Google does on your website:</p>
<ul>
<li>Which pages are crawled by the Googlebot? Log file analysis helps you monitor the crawl behavior and crawl frequency. You can know what Google is actually crawling and see how page popularity, page depth, load time or any other important metrics can influence Google’s crawl. It can also help you determine if specific new content has increased Google visits on your website. The more interesting your site is, the more often Google will come.</li>
</ul>
<p>&nbsp;</p>
<figure id="attachment_43215" aria-describedby="caption-attachment-43215" style="width: 805px" class="wp-caption aligncenter"><a href="https://www.relevance.com/wp-content/uploads/2016/06/1-crawl-frequency.png"><img fetchpriority="high" decoding="async" class=" wp-image-43215" src="https://www.relevance.com/wp-content/uploads/2016/06/1-crawl-frequency.png" alt="crawl frequency" width="805" height="384" /></a><figcaption id="caption-attachment-43215" class="wp-caption-text">Image source: OnCrawl log analyzer</figcaption></figure>
<p>&nbsp;</p>
<ul>
<li>What are my active pages? Which pages receive the most SEO visits? Are these my most valuable pages? You can know which pages are actually generating SEO traffic, value and, thus, conversions. Logs analysis can also help to determine the most popular pages to Google’s eyes and see which ones are less crawled. For instance, if a user would like to rank a specific post for a targeted query but it is located in a directory that Google only visits once a time every three months, he will miss chances to receive organic traffic from this publication for at least three months. With log analysis, he can know that it could be necessary, for example, to redefine his internal linking to increase the impact of his “most valuable pages”.</li>
<li>Does Google meet errors? Do you have too many 4xx errors on your website that lower Google experience on your website? Log data analysis can also help track errors in status codes like 4xx or 5xx that compromise SEO. Analyzing a website’s status codes also helps to measure their impact on bot hits and their frequency. Too many 404 errors will limit the crawler visit.</li>
<li>In every case, you need to save Google crawl budget to help spend it on the right pages. It helps you improve your money pages’ performance and be sure that Google is actually seeing them! With log file analysis you can, for example, detect if Google spends too much crawl budget crawling resources like images or .css files. This budget is linked to the authority of your domain, the sanity of your website, and is proportional to the flow of link equity through your website. You don’t want that budget to be spent on useless pages.</li>
</ul>
<h2>Imagine if you could do log analysis for free.</h2>
<p>Actually, log file analysis can be seen as expensive for companies and many of them stick to crawl analysis. But this technology is getting more accessible.</p>
<p>Open source solutions also exist. One of them is the free <a href="https://github.com/cogniteev/oncrawl-elk/">OnCrawl open source log analyzer</a>. It is quite easy to install even for non-tech profiles.</p>
<h3>1- Install Docker</h3>
<p>Install <a href="https://www.docker.com/get-docker">Docker Tool Box</a>.</p>
<p>Choose Docker Quickstart terminal to start.</p>
<p>Copy/paste the IP address 192.168.99.100</p>
<p>&nbsp;</p>
<p><a href="https://www.relevance.com/wp-content/uploads/2016/06/2-Docker-console.png"><img decoding="async" class="aligncenter wp-image-43216" src="https://www.relevance.com/wp-content/uploads/2016/06/2-Docker-console.png" alt="Docker console" width="805" height="466" /></a></p>
<p>Then, download oncrawl-elk release: <a href="https://github.com/cogniteev/oncrawl-elk/archive/1.1.zip">https://github.com/cogniteev/oncrawl-elk/archive/1.1.zip</a></p>
<p>Add these lines in the terminal to create a directory and unzip the file:</p>
<ul>
<li>MacBook-Air:~ cogniteev$ mkdir oncrawl-elk</li>
<li>MacBook-Air:~ cogniteev$ cd oncrawl-elk/</li>
<li>MacBook-Air:oncrawl-elk cogniteev$ unzip ~/Downloads/oncrawl-elk-1.1.zip</li>
</ul>
<p>And then, add:</p>
<ul>
<li>MacBook-Air:oncrawl-elk cogniteev$ cd oncrawl-elk-1.1/</li>
<li>MacBook-Air:oncrawl-elk-1.1 cogniteev$ docker-compose -f docker-compose.yml up -d</li>
</ul>
<p>(<strong>Note</strong>: those lines are working for Mac and Linux. If you are under Windows, the process is a little bit more complicated for non techs).</p>
<p>Docker-compose will download all necessary images from docker hub, this may take a few minutes. Once the docker container has started, you can enter the following address in your browser: <span style="text-decoration: underline;"><span style="color: #0000ff; text-decoration: underline;">http://DOCKER-IP:9000</span></span>. *Make sure to replace DOCKER-IP with the IP you copied earlier.*</p>
<p>You should see the OnCrawl-ELK dashboard, but there is no data yet. Let&#8217;s get some data to analyze!</p>
<p>&nbsp;</p>
<figure id="attachment_43218" aria-describedby="caption-attachment-43218" style="width: 625px" class="wp-caption aligncenter"><a href="https://www.relevance.com/wp-content/uploads/2016/06/4-oncrawl-open-source-log-analysis.png"><img decoding="async" class=" wp-image-43218" src="https://www.relevance.com/wp-content/uploads/2016/06/4-oncrawl-open-source-log-analysis.png" alt="oncrawl open source log analysis" width="625" height="704" /></a><figcaption id="caption-attachment-43218" class="wp-caption-text">Image Source: OnCrawl open source log analyzer without data</figcaption></figure>
<p>&nbsp;</p>
<h3>2-Import log files</h3>
<p>Importing data is as easy as copying log access files to the right folder. Logstash starts indexing any file found at logs/apache/*.log, logs/nginx/*.log, automatically.</p>
<p>If your web server is powered by Apache or NGinx, make sure the format is combined for log format. They should look like this:</p>
<p>127.0.0.1 &#8211; &#8211; [28/Aug/2015:06:45:41 +0200] &#8220;GET /apache_pb.gif HTTP/1.0&#8221; 200 2326 &#8220;http://www.example.com/start.html&#8221; &#8220;Mozilla/5.0 (compatible; Googlebot/2.1; +http://www.google.com/bot.html)&#8221;</p>
<p>Drop your .log files into the logs/apache or logs/nginx directory accordingly.</p>
<h3>3-Play</h3>
<p>Go back to <span style="text-decoration: underline;"><span style="color: #0000ff; text-decoration: underline;">http://DOCKER-IP:9000</span></span>. You should have figures and graphs. Congratulations!</p>
<p>&nbsp;</p>
<figure id="attachment_43217" aria-describedby="caption-attachment-43217" style="width: 624px" class="wp-caption aligncenter"><a href="https://www.relevance.com/wp-content/uploads/2016/06/3-open-source-log-analyzer.png"><img loading="lazy" decoding="async" class=" wp-image-43217" src="https://www.relevance.com/wp-content/uploads/2016/06/3-open-source-log-analyzer.png" alt="open source log analyzer" width="624" height="701" /></a><figcaption id="caption-attachment-43217" class="wp-caption-text">Image source: OnCrawl open source log analyzer</figcaption></figure>
<p>&nbsp;</p>
<p>You can now start using the free open source log analyzer and daily monitor your SEO performance.</p>
<p>To sum up, log file analysis is a powerful partner to increase your SEO performances and as a result, helps you drive more traffic and conversions to your website!</p>
<p>&nbsp;</p>
<p>[xyz-ihs snippet=&#8221;Hubspot-CTA-Leaderboard&#8221;]</p>
<p>The post <a href="https://www.relevance.com/a-guide-to-log-file-analysis-how-to-open-the-google-blackbox/">A Guide to Log File Analysis – How to Open the Google Blackbox</a> appeared first on <a href="https://www.relevance.com">Relevance</a>.</p>
]]></content:encoded>
					
					<wfw:commentRss>https://www.relevance.com/a-guide-to-log-file-analysis-how-to-open-the-google-blackbox/feed/</wfw:commentRss>
			<slash:comments>0</slash:comments>
		
		
			</item>
		<item>
		<title>How to Fix Duplicate Content and Improve your SEO</title>
		<link>https://www.relevance.com/how-to-fix-duplicate-content-and-improve-your-seo/</link>
					<comments>https://www.relevance.com/how-to-fix-duplicate-content-and-improve-your-seo/#respond</comments>
		
		<dc:creator><![CDATA[Emma Labrador]]></dc:creator>
		<pubDate>Wed, 17 Feb 2016 10:30:47 +0000</pubDate>
				<category><![CDATA[SEO]]></category>
		<guid isPermaLink="false">https://www.relevance.com/?p=41756</guid>

					<description><![CDATA[<p>Duplicate content is an SEO issue many SEOs or content marketers probably have experienced once a time in their daily routine. Content marketers who spend the time to create qualitative content strategies do not want to get penalized for duplication or near duplicates. In 2013, Matt Cutts stated that 25% of the web was duplicate&#8230;</p>
<p>The post <a href="https://www.relevance.com/how-to-fix-duplicate-content-and-improve-your-seo/">How to Fix Duplicate Content and Improve your SEO</a> appeared first on <a href="https://www.relevance.com">Relevance</a>.</p>
]]></description>
										<content:encoded><![CDATA[<p>Duplicate content is an SEO issue many SEOs or content marketers probably have experienced once a time in their daily routine. Content marketers who spend the time to create qualitative content strategies do not want to get penalized for duplication or near duplicates. In 2013, Matt Cutts stated that 25% of the web was duplicate content, you can find this number in the video at the end of the article.</p>
<p>Duplication refers to blocks of content that appear more than one time inside or outside a website or which are pretty similar. Duplicate content or near duplicates can lead to SEO penalties.</p>
<p>In fact, crawlers have trouble indexing the right content between different versions. Bots are then obliged to pick the content likely to be the best one, but it can lead to a loss of relevancy as it is not always simple to choose the right version. Moreover, bots will face difficulties to deliver the link metrics to the right page or share it between the different versions. Then, duplicate content also leads to the inability to rank the right version for a given query. At the end, you will face a drop in traffic.</p>
<p>The thing is, Google gives a lot of credits to user experience and focuses on delivering the best content possible to its audience. This is also why duplicate content is penalized.</p>
<p>While not all types of duplicate content can hurt your SEO, some of them need to be on the look out to avoid SEO penalties.</p>
<p><strong>This article will cover:</strong></p>
<ul>
<li>types of duplicate content</li>
<li>how to deal with duplicate content</li>
<li>tools to get rid of duplicate content</li>
</ul>
<h2>Types of Duplicate Content Leading to SEO Penalties</h2>
<p>There are different types of duplicate content you should avoid.</p>
<p><strong>Duplicate product forms</strong></p>
<p>E-commerce websites often use manufacturer’s item descriptions to describe the products they sell. The problem is that those products are often sold to different e-commerce websites. Then, the same description appears on different websites and creates duplicate content.</p>
<p><strong>Syndicated or copied content</strong></p>
<p>Many bloggers use content, quotes or comments from other websites to illustrate their articles. There is nothing wrong with that if you link back to the original one. However, Google can still consider this as a duplication and will poorly value those pieces of content.</p>
<p><strong>Sorting and multi-pages lists</strong></p>
<p>Large e-commerce websites have filter and category options that generate unique URLs. Product pages can appear in different categories and be ordered differently depending on how the list is sorted. For instance, if you range 45 products by price or by alphabetical order, you will end up with two pages containing the same content, but with different URLs.</p>
<p><strong>URL issues</strong></p>
<p>Google considers URLs in www, http, https, .com and .com/index.html as different ones even if they point to identical pages and will evaluate them as duplicate content.</p>
<p><strong>Session IDs</strong></p>
<p>Session IDs issues refer to different session IDs stored in the same URL that are assigned to a visitor when he or she comes to the website.</p>
<p><strong>Printer-friendly</strong></p>
<p>Printable versions of content can lead to duplicate content issues when different versions of a page are indexed.</p>
<h2>How to Avoid Duplicate Content?</h2>
<p>There are different best practices to avoid duplicate content issues. The main solution with content located in different URLs is to canonicalize the original one. You can use a 301 redirect, a rel=canonical or parameters handling tool from Google Webmaster Central.</p>
<p><strong>301 redirect</strong></p>
<p>The 301 redirect is great for URL issues leading to duplications. It informs search engines which version of a page is the original and it links the duplications to that original one. Plus, when different well-ranked pages are linked to a single one, they are not in competition anymore and they create an overall stronger and more popular signal.</p>
<p><strong>Rel=canonical</strong></p>
<p>It works quite the same as 301 redirects, but it is easier to set up. That tag is located in the HTML head section of your web page and looks like this:</p>
<p>&lt;link href=&#8221;http://www.mywebsite.com/canonical-version-of-page/&#8221; rel=&#8221;canonical&#8221; /&gt;</p>
<p>Then search engines know the above URL is a copy of the original one.</p>
<p>You can use it for content you integrated from other websites. It will inform search engines that you know the content is not from you and that the link metrics of that content should pass to the original one.</p>
<p><strong>NoIndex, NoFollow</strong></p>
<p>Use the noindex,nofollow meta tag to tell search engines not to index the content. Bots will be able to crawl the page, but won’t index it. Thus, you won’t be penalized for duplicate content.</p>
<p><strong>Preferred domain</strong></p>
<p>A quite simple operation is to set a preferred domain for search engines. It will inform whether a site must be displayed under ‘www’ or not in the SERPs.</p>
<p><strong>Unique product description</strong></p>
<p>As we said, product information on e-commerce websites can lead to duplicate content issues. Take the time to write unique ones or enrich your descriptions as it will help you rank above sites whose descriptions are duplicated.</p>
<h2>What tools can help me to detect duplicate content?</h2>
<p>In order to save time, you can use different qualitative tools to help you eradicate duplicate content. Here are three different ones, with some being totally free.</p>
<p><strong>Siteliner</strong></p>
<p>This tool detects any duplicate content on your website. You just need to add your website’s URL and it will draw a full report with your content performances.</p>
<p><a href="https://www.relevance.com/wp-content/uploads/2016/02/fix-duplicate-content.jpg"><img loading="lazy" decoding="async" class="aligncenter wp-image-41766 size-large" src="https://www.relevance.com/wp-content/uploads/2016/02/fix-duplicate-content-1024x329.jpg" alt="fix-duplicate-content" width="640" height="206" /></a></p>
<p><a href="https://www.relevance.com/wp-content/uploads/2016/02/fix-duplicate-content-2.jpg"><img loading="lazy" decoding="async" class="aligncenter wp-image-41765 size-large" src="https://www.relevance.com/wp-content/uploads/2016/02/fix-duplicate-content-2-1024x411.jpg" alt="fix-duplicate-content" width="640" height="257" /></a></p>
<p><strong>OnCrawl</strong></p>
<p>This onsite SEO semantic crawler also offers a duplicate and near duplicate detection feature. It shows you clusters of duplicates and near duplicates, types of duplication and clearly indicates which URLs are concerned.</p>
<p><a href="https://www.relevance.com/wp-content/uploads/2016/02/fix-duplicate-content-3.jpg"><img loading="lazy" decoding="async" class="aligncenter wp-image-41764 size-large" src="https://www.relevance.com/wp-content/uploads/2016/02/fix-duplicate-content-3-1024x434.jpg" alt="fix-duplicate-content" width="640" height="271" /></a></p>
<p><a href="https://www.relevance.com/wp-content/uploads/2016/02/fix-duplicate-content-4.jpg"><img loading="lazy" decoding="async" class="aligncenter wp-image-41763 size-large" src="https://www.relevance.com/wp-content/uploads/2016/02/fix-duplicate-content-4-1024x661.jpg" alt="fix-duplicate-content" width="640" height="413" /></a></p>
<p><a href="https://www.relevance.com/wp-content/uploads/2016/02/fix-duplicate-content-5.jpg"><img loading="lazy" decoding="async" class="aligncenter wp-image-41762 size-large" src="https://www.relevance.com/wp-content/uploads/2016/02/fix-duplicate-content-5-1024x487.jpg" alt="fix-duplicate-content" width="640" height="304" /></a></p>
<p><a href="https://www.relevance.com/wp-content/uploads/2016/02/fix-duplicate-content-6.jpg"><img loading="lazy" decoding="async" class="aligncenter wp-image-41761 size-large" src="https://www.relevance.com/wp-content/uploads/2016/02/fix-duplicate-content-6-1024x432.jpg" alt="fix-duplicate-content" width="640" height="270" /></a></p>
<p>The tool offers a 30 day-trial. If you want to enjoy all the functionalities, you just need to pick a plan and cancel your subscription before the end of the trial and you won’t pay anything.</p>
<p><strong>Copyscape</strong></p>
<p>Copyscape is a great partner because it detects duplicate content outside of your blog. You can thus easily know if someone has duplicated your content without your permission or without giving you credits.</p>
<p><a href="https://www.relevance.com/wp-content/uploads/2016/02/fix-duplicate-content-7.jpg"><img loading="lazy" decoding="async" class="aligncenter wp-image-41760 size-large" src="https://www.relevance.com/wp-content/uploads/2016/02/fix-duplicate-content-7-1024x554.jpg" alt="fix-duplicate-content" width="640" height="346" /></a></p>
<p>&nbsp;</p>
<p style="text-align: center;">[xyz-ihs snippet=&#8221;Hubspot-CTA-Leaderboard&#8221;]</p>
<p>The post <a href="https://www.relevance.com/how-to-fix-duplicate-content-and-improve-your-seo/">How to Fix Duplicate Content and Improve your SEO</a> appeared first on <a href="https://www.relevance.com">Relevance</a>.</p>
]]></content:encoded>
					
					<wfw:commentRss>https://www.relevance.com/how-to-fix-duplicate-content-and-improve-your-seo/feed/</wfw:commentRss>
			<slash:comments>0</slash:comments>
		
		
			</item>
	</channel>
</rss>
