You work hard, contact so many professional A list bloggers out there, and suggest your post for incoming link. If they provide you with a backlink, you will be happy. A high PageRank, unpaid link, given with the free will of the giver, from the most relevant page/category is going to be a million dollar vote to your page. It can itself get your page to skyrocket from SERP 400 to SERP 3.
Yesterday, I had a discussion in the Digital Point Forums about backlinks and their validity. It seems that most people are not knowledgeable about backlink validity analysis. People know only about DoFollow and NoFollow, and nothing above that. Here, we will see the importance of ‘robots.txt’ file in the backlink analysis. Robots.txt is a simple link invalidation secret several professional bloggers won’t share with you.
What Is Robots.txt
When a search crawler accesses a website for searching and indexing, the first thing it looks for is the Robots.txt file. If it doesn’t find one, it goes about normally indexing the page.
A robots.txt file is a small text file on the root directory of the domain of a web page that directs search engine crawlers as to which of the pages or sections of the website shouldn’t be indexed. Its general format is thus:
User-agent: Googlebot
Disallow: /links.html
The above code simply disallows the Google bot from accessing the links page of a website (whose relative path is /links.html). This means, if you are an advertiser and you exchanged links with this person, who has disallowed the links page, your backlink from that page holds no value whatever. But it will definitely look like a DoFollow backlink to the untrained eyes. The search bots will normally index any pages not mentioned in the Robots.txt file.
User-agent: *
Disallow: /page/categories.htm
This code disallows the page for all search bots, not only Google’s. The wildcard, ‘*’ is used to specify that all search bots are to follow the rule. So, none of the search bots will index the mentioned page.
A careful, clever advertiser thinks from the search bot point of view and first looks for the Robots.txt file of the pages he plans to purchase links from. He merely won’t purchase backlink from any disallowed page.
Checking Robots.txt File of Any Website
You can easily check the Robots.txt file of any website out there. It’s right there in the root directory of the domain. It is named that way, ‘robots.txt’. You may just access it through the browser. For instance, if you need to access the Robots.txt file of Microsoft.com, just go to the main home page of Microsoft, and put this on the address bar:
“http://www.microsoft.com/robots.txt”
Google: “www.google.com/robots.txt”
Simple? You will notice that these sites have disallowed a lot of internal pages from being indexed. This is why these internal backlinks do not show up in search results or hold any weight.
Through robots.txt file, you can even specify the sitemap of a website. If you check my robots.txt file, “http://cutewriting.blogspot.com/robots.txt”, you will see that a sitemap has been specified.
Conclusion
Before you purchase links from any website for SEO (if you are purchasing at all) or going for link exchange, first look for the robots.txt file and see if your links are going to get any weight at all. If not, that link exchange simply won’t work.
By the way, this article or this blog does not recommend purchasing backlinks for the purpose of SEO. If you are purchasing backlinks, it should be for traffic and the links should be Nofollow. Purchasing backlinks for SEO can get your site banned by Google easily.
In the next article of Link Analysis series, we will see another important thing to check for before exchanging links: The Robots Meta tag. Subscribe and enjoy.
Related Entries:
What is DoFollow? What is NoFollow (Make Your Blog DoFollow)
Ten Effective Link Building Techniques
Eleven Nasty Ways to Build Links
Copyright © Lenin Nair 2008
Important, Read This: Get your professional web hosting (unlimited all) for a full year at only 4.95 dollars! No strings attached. In order to get this, click here and use the coupon code, FREEJULY.
Backlink Analysis Part I: Robots.txt Can Invalidate Your Backlinks
Posted by
Lenin
at
9/03/2008 04:15:00 AM
Copyright © CuteWriting 2009. All rights reserved.
For republication, please contact us.
[This notice is void in case of contributions and posts in which the author is explicitly specified.]
Subscribe
For republication, please contact us.
[This notice is void in case of contributions and posts in which the author is explicitly specified.]
Subscribe
Subscribe to:
Post Comments (Atom)
Search
Topics
- Bestseller Lists (3)
- Blog Notification (34)
- Blogging Tidbits (5)
- Blogging Tips (76)
- Breaking News (12)
- Copyright and License (4)
- Copywriting (4)
- Creative Writing (49)
- CW Web Services (1)
- Detective Fiction (5)
- DoFollow vs. NoFollow (4)
- English Classics (27)
- English Language (14)
- Freelancing (12)
- Grammar and Style (45)
- Great Writers (23)
- Idioms and Usages Reference (23)
- Link Building (8)
- Literature Resources (13)
- Money Making and Affiliate Programs (9)
- Online Advertising (9)
- Online Writing (23)
- Plagiarism (5)
- Publication Help (30)
- Punctuation (10)
- Reader Contributions (3)
- SEO (45)
- SEO Tidbits (2)
- SEO Tools (4)
- Special Posts (42)
- Sponsored Entries (31)
- Tips and Help for Writers (50)
- Usage and Vocabulary (42)
- Web 2.0 and Social Media (27)
- Web 2.0 and Social Media Tidbits (3)
- Web 3.0 (1)
- Web Service Reviews (22)
- Writer's Block (3)
- Writing Site Reviews (12)
Blog Archive
-
▼
2009
(124)
-
►
April
(30)
- Windows 7 Test Release From Thursday
- ClickBank Ads Reloaded!
- Microsoft Vine: A New Contestant in Social Warfare...
- Foreword and Forward
- Literally, Practically, and Virtually
- How to Serve Third Party Ads
- Using Google Reader
- Artist and Artiste
- The Crystal Egg: Short Story By H G Wells
- Is Your Twitter Account Affected Too?
- Ain't: An Old Usage That Still Sustains
- What You Should Know About Link Rot
- Birth of Born and Borne
- Turning Firefox Into Internet Explorer
- A List of Medical Terms for Phobias (Fears)
- The Autograph Hunters by PG Wodehouse
- Which Is More Important? Content or Links?
- Landmark for Apple: One Billion App Downloads
- Five Tips for Generating More Page Views for Your ...
- Egg White Is Albumin or Albumen?
- Fake Google Adsense Reports and How to Detect Them...
- Five Ways to Lose Your PageRank Successfully?
- Software for Tracking Manuscript Submissions
- Set Up Your Google Alerts to Protect Your Constant...
- After Death or Anno Domini?
- The Burial of the Rats by Bram Stoker
- Examples of Comment Spamming in Blogger
- Word Tips: Engaged Men and Women
- What Went Wrong on the Web Yesterday?
- Google Aiming for Microsoft's Hostile Acquisition!...
-
►
March
(32)
- Wikipedia Defeats Microsoft's Encarta
- The Beautiful Suit by H G Wells
- What Is RSS? Avoid These RSS Feed Syndication Mist...
- Taking Control of Your Blog: Some Social Interacti...
- Freelancing At ODesk: A Global Opportunity
- World's Most Affordable Hosting Offer Will En...
- Use Google Trends for Instant Traffic
- The Benefits of Popular Posts Widget for Blogs
- Former Convicts as Friends in Facebook: Jail Offic...
- Thesis: The Best Professional WordPress Theme
- Five Unavoidable Firefox Add-ons for Search Engine...
- About Conjunctive Adverbs
- Important! Read About My Twitter Password Hijacker...
-
►
April
(30)



























8 Opinions:
Hai Lenin,
Being a blogger can we do any thing with robot.text file ?
Is there any control for blogger on this file ?
Let me know.
Suresh, Unless you are using Blogger self-hosting, there is no control over Robots.txt. Thanks for commenting.
thanks for sending me a note about this post! i really learned something new today!
What a readable, informative and educating post - always great to learn something new - thanks for sharing your knowledge!
This is certainly an informative article. I "dugg" it!!!
This morning I was researching robot.text file. The Google Tools, etc. website has lots of great info on this topic. Later I discovered this outstanding blog and it has helped me immensely.
I'm concerned about an odd backlink to some of my youtube videos. It's some kind of fake website about insurance and it has 2 thumbnails of my videos that have nothing to do with insurance. When I clicked the thumbnails I was brought to 2 videos that were not mine. In fact they belonged to 2 competitors in my field.
For some of my vids this is the only backlink, one backlink. I cannot seem to get my legitimate backlinks to connect to these vids. Do you think someone is screwing with robot.text files to block real backlinks from my vids? Why are these people using my thumbnails on a fake insurance website backlink to direct the next link to their videos?
I am looking at your blog to see how I can subscribe to this excellent content. Thank you, you have made my day!
Regards,
Chuck
Chuck, no I don think someone may be doing anything with robots.txt. It's only one backlink right? You can try building more links to your vids. Good to know you liked the resource and decided to subscribe. Since they are youtube videos, no one has access to the robots.txt file within youtube, so don't worry. I don't get why you say you can't connect your legitimate links to these videos. Can you give specifics so that I can help?
Lenin
thank's for writing an essential seo article, somthing which I, and probably other people have completely forgotten about; eg robots.txt file.
Post a Comment
Comments are very strictly moderated. If you don't want to waste your time, please read the comment policy at the blog usage policy. You will also find what to do if there is no comment form below...