Showing posts with label Google. Show all posts

Google Ranking Factors and Blog Research by Ch Zain Ishaq

For a long time my blogs have performed amazingly well with Google Blog Search. I always appear in the relevant results quickly, and the results I obtain have some reasonable longevity, even when I am not the original source of a story. Considering how much competition I often have for certain search terms which everyone seems to be writing about because of common interest, I must have been doing a number of things right. I am going to do a little bit of mix and match here, and inject my own commentary but my interpretation of the patent is actually slightly different to those that I have read so far. It should be noted I am working my way through the patent itself, and notrecompiling the summaries of others.

Relevancy & Quality – Blog | Blogpost

It should first of all be noted that in the patent Google doesn’t differentiate between individual blog posts and whole blogs.
The phrase “blog document,” as used hereinafter, is to be broadly interpreted to include a blog, a blog post, or both a blog and a blog post. It will be appreciated that the techniques described herein are equally applicable to blogs and blog posts.
Later on in the patent, they also mention that feeds are also included within the documents that are compared and rated.
two distinct sets of data are used to determine a score of a blog (or blog post) in response to a search query–the topical relevance of the blog (or blog post) to the terms in the search query and the quality of the blog (or blog post), which is independent of the query terms. The quality of the blog (or blog post) may positively or negatively affect the score of the blog (or blog post)

Relevancy –

this applies to the search term, thus Google will analyse the blog page, and they will also in some way determine the relevance to the whole blog.

Quality –

this is irrespective of the search term, so think about factors from outside your niche.

Google Blog Search – Positive Factors Affecting Search Quality | Relevancy

Popularity Of The Blog Document

A number of news aggregator sites (commonly called “news readers” or “feed readers”) exist where individuals can subscribe to a blog document (through its feed). Such aggregators store information describing how many individuals have subscribed to given blog documents. A blog document having a high number of subscriptions implies a higher quality for the blog document.
This patent was first of all applied for 13th September 2005, with Google Blog Search launched 13 September 2005. At the time they were logically not basing this on numbers available for Google Reader subscribers. The Google Reader blog was launched October 21, 2005 with a post saying they had been up and running for 2 weeks.
Maybe there is a coincidence between the 2 events.
Google Ranking factors
So which data were Google basing this part of their patent on? Some services such as Technorati and Bloglines do provide readership data, as does Feedburner, though most services report readership data as they are collecting new blog posts to a service like Feedburner, who aggregate the statistics.
It seems there might be some value is collecting Technorati favorites (my reciprocation policy might be well worth it) beyond limited bragging rights. Google of course through Google Reader now have access to lots of usage data, so maybe other sources will eventually be phased out.

Implied Popularity Of The Blog Document

This implied popularity may be identified by, for example, examining the click stream of search results. For example, if a certain blog document is clicked more than other blog documents when the blog document appears in result sets, this may be an indication that the blog document is popular and, thus, a positive indicator of the quality of the blog document.
Click data from search results, possible from Google Toolbar users.

Existence Of The Blog Document In Blogrolls

The existence of the blog document in blogrolls may be a positive indication of the quality of the blog document. It will be appreciated that blog documents often contain not only recent entries (i.e., posts), but also “blogrolls,” which are a dense collection of links to external sites (usually other blogs) in which the author/blogger is interested. A blogroll link to a blog document is an indication of popularity of that blog document, so aggregated blogroll links to a blog document can be counted and used to infer magnitude of popularity for the blog document.
Everything I have ever read has suggested that for normal search, blogroll links that are site wide carry diminishing value. Just because it is listed here as part of the calculation does not necessarily mean that everyone should start building up huge blogrolls… well unless they want to game Technorati and have a blog network.

Existence Of The Blog Document In A High Quality Blogroll

The existence of the blog document in a high quality blogroll may be a positive indication of the quality of the blog document. A high quality blogroll is a blogroll that links to well-known or trusted bloggers. Therefore, a high quality blogroll that also links to the blog document is a positive indicator of the quality of the blog document.
Another revelation, links on high quality pages are worth more than links on low quality pages. Remember that “blog document” can mean both blog page and blog site. Can blogroll just refer to a list of links on what is identified as a blog. Thus a column of links to related pages might also class as a blogroll, whether in the sidebar or below the content. Thus a list of links to related documents on the same site could be looked on as a blogroll on a blog document. Related links plugins are very powerful, especially if you also include them in content that gets syndicated by design, or by sploggers.

Tagging Of The Blog Document

Tagging of the blog document may be a positive indication of the quality of the blog document. Some existing sites allow users to add “tags” to (i.e., to “categorize”) a blog document. These custom categorizations are an indicator that an individual has evaluated the content of the blog document and determined that one or more categories appropriately describe its content, and as such are a positive indicator of the quality of the blog document.
Well some sites do allow you to tag in a meaningful way, maybe Google uses shared tags from Del.icio.us and other sites, but many of those use nofollow extensively. It is my own belief that self tagging content heavily with plugins such as Ultimate Tag Warrior helps a huge amount. I have given lots of examples before, but more recent examples include
  1. toolbar pagerank
  2. google reader feedburner
  3. feedburner google reader
  4. compete toolbar
  5. duplicate content supplemental results
Yes, I am just going down the inbound traffic results looking for likely candidates that rank well in both blog and normal search and aren’t totally obscure. These are subjects that sites in my niche have also talked about, with the keywords in the title, and which you would expect to rank higher than my own content.
This doesn’t just affect blogsearch, Google have been using it for some time with the main results as well.
Here are my observations regarding tagging from back in November, especially how they could relate to LSI calculations.

References To The Blog Document By Other Sources

Wow revelation again, god links are worth having either to pages or blog.

Pagerank Of The Blog Document

Pagerank is still relevant, who knows for how long and how much.
It will be appreciated that other indicators may also be used.
What seems to be missing, at least at time of application?
  • Domain age?
  • Trustrank?
  • Page Titles?
  • URLs?
  • Growth rate of link popularity
Plus lots more that also factor into it, but general search patents probably also cover blog search.

Google Blog Search – Negative Factors Affecting Search Relevancy | Quality

Frequency Of New Posts On The Blog Document

The frequency at which new posts are added to the blog document may be a negative indication of the quality of that blog document. Feeds typically include only the most recent posts from a blog document. Spammers often generate new posts in spurts (i.e., many new posts appear within a short time period) or at predictable intervals (one post every 10 minutes, or a post every 3 hours at 32 minutes past the hour). Both behaviors are correlated with malicious intent and can be used to identify possible spammers. Therefore, if the frequency at which new posts are added to the blog document matches a predictable pattern, this may be a negative indication of the quality of the blog document.
Make sure there is some variation when you publish your content for the day, especially with future dated posts.
Most spamming tools are actually fairly sophisticated, thus I am not sure this measurement is very accurate. It most likely indicated a blogger who is very organised these days.

The Content Of The Posts In The Blog Document

The content of the posts in the blog document may be a negative indication of the quality of that blog document. A feed typically contains some or all of the content of several posts from a given blog document. The blog document itself also includes the content of the posts. Spammers may put one version of content into a feed to improve their ranking in search results, while putting a different version on their blog document (e.g., links to irrelevant ads). This mismatch (between feed and blog document) can, therefore, be a negative indication of the quality of the blog document.
This is actually a very significant and interestingly worded item. Google are stating that they are comparing the content of a feed with the content on your pages to ensure it matches.
Based upon this:-
  • Don’t use a content spinner on your feeds to avoid duplicate content
  • Allow Google to index your feeds
  • If you use related links on your blog, make sure you use them in your feeds too.

Duplicate Content, Especially In Feeds

Also, in some instances, particular content may be duplicated in multiple posts in a blog document, resulting in multiple feeds containing the same content. Such duplication indicates the feed is low quality/spam and, thus, can be a negative indication of the quality of the blog document.
I can’t say I have noticed a problem having a lot of straggling RSS feeds on categories and tags. This could also be referring to things like the large footer I have on each post, though I haven’t seen a problem with that either.
After the last toolbar pagerank update I spent some time studying Matt Cutts’ blog, and also looking at how pagerank was being transferred around my own site. Pagerank is only slightly useful as a guide, and only immediately after an update. Rather than repeat myself, you can read about my organic garden approach to this site.

Collective Intelligence

The words/phrases used in the posts of a blog document may also be a negative indication of the quality of that blog document. For example, from a collection of blog documents and feeds that evaluators rate as spam, a list of words and phrases (bigrams, trigrams, etc.) that appear frequently in spam may be extracted. If a blog document contains a high percentage of words or phrases from the list, this can be a negative indication of quality of the blog document.
Google invest a lot of research analysing spam, detecting various word matching patterns, and use that to identify other documents.

A Size Of The Posts In The Blog Document

The size of the posts in a blog document may be a negative indication of quality of the blog document. Many automated post generators create numerous posts of identical or very similar length. As a result, the distribution of post sizes can be used as a reliable measure of spamminess. When a blog document includes numerous posts of identical or very similar length, this may be a negative indication of quality of the blog document.
This might be of special interest to those that use out-sourcing for articles, you need to ensure the article length changes.

A Link Distribution Of The Blog Document

A link distribution of the blog document may be a negative indication of quality of the blog document. As disclosed above, some posts are created to increase the pagerank of a particular blog document. In some cases, a high percentage of all links from the posts or from the blog document all point to ether a single web page, or to a single external site. If the number of links to any single external site exceeds a threshold, this can be a negative indication of quality of the blog document.
In some ways this debunks the benefits of blogrolls mentioned as a benefit, but as previously quoted, Google are using blog document in multiple context, and comparing the context, thus it could just refer to multiple spam links always pointing to a single domain within the content.

The Presence Of Ads In The Blog Document

The presence of ads in the blog document may be a negative indication of quality of the blog document. If a blog document contains a large number of ads, this may be a negative indication of the quality of the blog document.
Remember this is just a patent, and Google recently relaxed the rules about having ads from other networks along with Adsense. As long as a page is of a reasonable size to support the adverts, I don’t think there is a problem. If you just have a heading and 5 words, with 10 advertising blocks, you might want to add a few more words.
However they go on to say this
Moreover, blog documents typically contain three types of content: the content of recent posts, a blogroll, and blog metadata (e.g., author profile information and/or other information pertinent to the blog document or its author). Ads, if present, typically appear within the blog metadata section or near the blogroll. The presence of ads in the recent posts part of a blog document may be a negative indication of the quality of the blog document.
Thus if you are using blocks in the content for all your ads, you might not rank as well, especially if you use multiple networks. You can probably get away with 3 in the content, or maybe 1 or 2 per post.
It will be appreciated that other indicators may also be used
Conclusion
The feed stats information is very useful, and looking at the timing, my conclusion is that Google might have been using Bloglines and Technorati Favourites data, with Google Reader in its infancy, or maybe though less likely, when blog search was introduced, they weren’t using that part of the patent.
For me the most significant information was tagging, but just linking though to Technorati with your tags isn’t a great idea.
Remember that Google have their own blogging system, and they have archives and labels, and they are not going to create a system to generate duplicate content and then penalise you for it. Google wouldn’t have added such a system unless they intended to benefit from the enhanced data.
You don’t have to build your blogs in a 1990s era tree like structure to rank well.

Read More »

How to Easily Index your Website Faster in Google Search



How to Easily Index your Website Faster in Google Search

Today, having a website is not a challenging task anymore as it was considered before. In present online age people are more concern about getting their brand name or blog searchable on search engines to engage more valuable visitors. Many brand companies have already started to implement techniques that help their website to index faster in search engines. If you’re looking for some tips or techniques that could help your site to rank faster, then you are luckily at the right place. Today in this article, we will show you How to Easily Index your website faster in Google Search

Why Google Indexes new websites slowly?

We see a lot of users always complain about the fact that their new site takes a lot of time in indexing posts or articles. The main reasons of not getting your new site index are mentioned as below:

Less Authority: When you create a new site, it has almost ZERO authority while other old sites have better authority this is the reason why your site takes a while in indexing. The older your site would be, the more authority it would gain so it’s just a matter of getting your site a bit older.

Less Content: Google loves content, not just content, but fresh and quality content. New sites have less or no content, this makes almost impossible for search engines to get a blank page index (in some cases Google even indexes blank pages). However, write content that has quality and quantity at the same time.

Lack of Social Signals: Now days, search engines do count on social signals (the no of shares, tweets, likes the content is received from social networking websites) and on the basis of that they index or rank websites. For a new site, it is almost impossible to overnight get a lot of social followers.

Under the Scan: For those who don’t know, Google always takes a nice look at your site before it is indexed in the search engines so if there is any delay in indexation then your site is under the scan..

To find the number of pages indexed go to Google.com and search for site:mybloggerlab.com(Note: Do not forget to replace mybloggerlab.com with your website address). If you see no results, then you should implement the tips we have discussed in this article. Following screenshot shows the search results for the above query.

How to Easily Index your website faster in Google Search:

The techniques which we’ll be discussing in this article are applicable for both new and old websites. In short, we will be using some external websites to send search engine crawlers to our sites, forcing them to index our site faster than usual.

1. Submit sitemap to Google:

What is a sitemap? A sitemap is the list of posts or pages that you’ve published on your site which helps search engine crawlers to index your posts according to the its published dates. By submitting a sitemap to Google, Yahoo, Bing, AOL and etc. you’re inviting them to index your blog faster and efficiently.

If you haven't submit your sitemap to Google then do it today by going to the webmaster tool.

2. Commenting on Dofollow or CommentLuv Blogs:

After publishing a new article, comment on authority high traffic blogs. This not only provides some traffic, but at the same time search engine crawlers will follow your site if the attribution is set to Dofollow.

This is an old but working technique that every newbie blogger implements and if you haven’t then try it from today. You will start seeing the changes within a week or so depends on how much blogs you’re commenting on daily basis, I would prefer 20 to 30 comments a day.

3. Sharing on Social Communities: 

This is another method which is very effective for new blog owners. Facebook or even Google+ has active groups and communities that allows you to share your posts URLS on them. When you’re sharing those URLs on those groups, you are actually sending a social signal to your site. This means the more you share, the more social signals it will send back to your site.

We have also implemented the same technique in the early days of MyBloggerLab.com and results were pretty much amazing.

4. Ping Your Posts or Blog:

Ping service is another working way to alarm search engines to index your newly published posts. However, Blogger as well as WordPress automatically sends pings to search engines, but doing it efficiently by doing it yourself won’t hurt much.

Ping From this site: http://pingomatic.com/

5. Guest Posting:

People say Guest posting is no longer effective, but the reality is they have the same significance at the moment as well. If you have time write 5 to 10 quality articles and ask high authority blogs to publish them and provide a Dofollow link to you.

Try to guest posts on blogs that are related to your niche, are updated regularly and consists of the active audience. By doing this, you will get some healthy traffic and again in your authority.

Just to prove this thing really works, I applied these techniques on one of my new testing blog and results were pretty pleasing. The following screenshot shows a website that was indexed in Google within 10 minutes.
We hope this article has helped you in learning some of the unrevealed secrets of getting your site indexed faster in Google search engine. Now, since you know they go ahead and apply them for results that you’re looking for a lot of months or even years.

Read More »
--------------