maandag 28 oktober 2013

Historic Index Update

We have now updated our Historic Index with data for start of August 2013. Here are the new stats: Historic Index – Unique Pages crawled: 582,304,941,631 Unique URLs: 2,274,429,442,587 Date range: 27 Feb 2008 to 29 Sep 2013


The post Historic Index Update appeared first on Majestic SEO Blog.


via Majestic SEO Blog» English http://blog.majesticseo.com/index-updates/historic-index-update-21/


http://seocompanyadvice.com/historic-index-update-3/?utm_source=rss&utm_medium=rss&utm_campaign=historic-index-update-3

Search News from the Future

Posted by Reinhart


Citizens of Moz, I come to you at a most desperate hour. I’ve just returned from London, Distilled’s international headquarters, and I’ve been patiently awaiting this moment to share some potentially niche-shattering news with you all.


I don’t quite know where to begin, so I’ll just say it: You see, the stories are all true. Will Critchlow is a wizard. I know, it’s common knowledge that nearly all Englishmen are wizards, I’ve seen Harry Potter too. But Mr. Critchlow is a wizard with a most peculiar and exciting gift: that of clairvoyance. He can see the future!



And no, I don’t mean in a Steve Jobs/Carl Sagan/George Orwell futurist kind of way either. I mean he quite literally has a translucent, viridian ball of crystal sitting on his desk that divines that which has yet to transpire! I wouldn’t have thought anything of the object upon first glance, but one night I came back to the office to grab my misplaced jacket to low mutterings, frantic typing, and wisps of smoke coming from the other side of the room. I dove into the bean bag room so as not to draw his attention and waited patiently, shaking with dread but with a fully piqued curiosity.


I couldn’t make out what he was chanting and I don’t think I would have been able to translate the Latin anyway. After about 30 minutes of this I heard him pack up his things and leave. I’m normally more of the craven type when it comes to adventure, but something that night pressed me to snoop around my boss’s desk for the truth.


The smoke and emerald glow dissipated as I shuffled some papers around. The smell of ozone lingered in the air. Nothing looked too out of the ordinary: the latest issue of Inc. Magazine, a Post-it note with a hastily scrawled and circled “Fire Phil Nottingham: Oct 31“… wait… this news clip read… 2016? Maybe he was just tired and mistak— 2020?! What was I looking at here?!


What I’m about to reveal may shock or even scare some readers, but I believe it is essential that the Moz community hear it nevertheless. I may lose my job—nay, I may be turned into a toad with a dreadful cockney accent—but it will have all been worth it to bring this knowledge to you all. My interpretations may be shaky at best, but the headlines were as clear as day: These are digital marketing news items from the future!


You may never get a better chance to peek behind the tapestries of time as you do now. So read on, friends, and be brave.


Term “mobile” removed from Analytics, Google’s vocabulary


MOUNTAIN VIEW, CA — April 14, 2015 — A term commonly used by webmasters, digital marketers and industry analysts may not be so common after today. Over the weekend, Google removed the term “mobile” from all of its web products, including Webmaster Tools, Google Analytics, and the company’s AdWords tool set.


“Mobile has been a deprecated term for some time now,” the search giant explained in a corresponding blog post. “The lines between where and when we view our various screens have been blurred beyond parsability. All web-based content can be viewed on any device these days and thus it makes little sense to refer to all non-traditional desktop visitors as ‘mobile.’ ”


The web is very close to becoming truly device-neutral largely thanks to thoughtful webmasters, CMS development teams and device manufacturers who have all come together to deal with the issue of rendering content from multiple angles. Data on device type, screen size, and other metrics is still readily available throughout Google’s suite of webmaster tools. [...]


Voice searches now constitute 28% of all queries


AUSTIN, TX — June 29, 2020 — Search engine corporations such as Google, Microsoft, and Yahoo have traditionally held on tight to their data, offering limited info on global search trends, but a recent study conducted by the University of Texas has unearthed compelling evidence that shows almost a third of all search queries are now conducted via voice search.


The nation’s only “Professor of Search,” Dr. Pete Meyers of the University of Texas explains the results of his institution’s study:


“They called me mad back in 2013, but voice searches now constitute a huge chunk of the search pie. Several years ago we would have found it laughable to be walking down the street talking to our devices, let alone talking to our devices within our home. But with the advancement of voice recognition software and the nearly ubiquitous nature of the hardware to back it up, today we’re estimating that voice search makes up almost a third of all search queries, and that number seems to be on the rise.”


A few of the major contributing factors to the ascendancy of voice search include web-enabled automobiles, home appliances, [...]


Traditional television advertising revenues wane as new year begins; YouTube, Twitter and Facebook post record annual reports


NEW YORK, NY — January 1, 2019 — Google’s video platform, YouTube (GOOG) along with social networks Twitter (TWTR) and Facebook (FB), posted record gains in 2018 as social and video advertising revenues shattered forecasts and industry expectations. Analysts speculate that this was due in no small part to the transference of advertising spends on traditional television media. When FB and TWTR first hit the stock market, many buyers felt the social networks needed to prove themselves in the competitive world of media advertising, but as the multi-billion dollar industry of traditional television advertising continues to crumble amid stiff competition from a la carte alternatives like Netflix and Amazon, more marketing budgets are now trickling down to companies such as YouTube, Facebook, Twitter, and Google’s AdWords platform.


“Television is evolving and has been for some time,” says Will Critchlow, founder and president of the world’s foremost digital marketing agency, Distilled International. “Companies want to get their products in front of consumers, and those consumers are now watching television online. They’re doing everything online.” [...]


Netflix introduces video, text advertisements for streaming content


LOS GATOS, CA — January 16, 2019 — Earlier this month we saw reports that television advertising revenues were waning in the new year. Today we can report that some of those dollars will most certainly be spent on Netflix’s streaming video platform. The company issued a press release this morning indicating that the company, for the first time in its history, will now display advertisements before many of their most popular original programs such as Arrested Development, Orange is the New Black, and the much-anticipated final season of House of Cards. Advertisements will be similar to those seen on YouTube, Hulu, and other video sites.


“With the amount of quality content and general media access we’re collecting, we have no choice but to find revenue from other sources if we want to remain at the $9.99 price point we set in 2015,” Netflix CEO Reed Hastings said on an investor conference call yesterday. [...]


Google cracks down on fake, purchased +1s


MOUNTAIN VIEW, CA — February 1, 2017 — For the first several years of the social platform’s life, Google Plus seemed a joke to many. Comparisons were made to MySpace and other defunct social platforms, and G+ was often called “a graveyard” as it faced competition from the already-established Facebook. But since that time, the network has shown some real staying power with the full faith and credit of Google Inc. behind it. To that end, in late 2016 we reported on Google’s announcement that plus ones, Google’s own brand of “Likes,” would help determine the order in which documents appeared in its search engine results pages. This move forced webmasters everywhere, for big and small companies alike, to reconsider the social platform for conducting regular business. Since then, various scams have been created to generate fake or paid “+1s” for sites who want quick and easy exposure in Google’s search engine. While this practice has been effective for some, it is not sitting at all well with the search giant.


Today, Google announced a crack-down on those sites which it has determined to have been generating fake +1s. The process should be easy enough for Google as it has access to all of its users’ account data and history. One well known Google representative, speaking on condition of anonymity, cut straight to the point, asking “What were they thinking?” in reference to marketers who’ve been attempting to game Google’s algorithm. “As if we haven’t been aware of fake Google plus accounts since Plus’s inception?… [...]


Panda and Penguin now refresh daily


MOUNTAIN VIEW, CA — July 22, 2016 — Five years ago, Google launched a pair of systems designed to keep poor content out of its search engine’s results and combat questionable citation building tactics. The former is known as the “Panda” update and the latter, “Penguin.” Until this week, the two algorithms have been updated on unpredictable schedules based on when the massive amounts of data required to make proper determinations about the quality of a website and its internet-wide citations were parsed. This would result in periodic “refresh” days, where webmasters who engaged in deceptive marketing practices would brace themselves for potential losses of traffic to their webpages. These updates would traditionally occur four or five times per year. Last Monday, Google announced that they’ve dedicated additional resources to these systems and are now able to parse the same data sets many times faster than before, meaning that these updates will now essentially occur in real time.


“Our users don’t want clean and relevant answers three months from now, they need them immediately,” declared a Google representative at an industry conference in San Diego… [...]


Google removes “organic keywords” tab from analytics


MOUNTAIN VIEW, CA — October 31, 2020 — [...] and for those who don’t remember, Google used to give webmasters access to what was known as “keyword data,” allowing them to make better decisions about their sites’ development and what their users’ intent might be when visiting. In the fall of 2013, Google denied access to almost all of this organic data by encrypting all searches generated through Google.com. Today, Google took it a step further and completely removed the “Organic Keywords” tab from its popular web analytics program.


“We’ve been meaning to do this for some time now,” said a senior Google representative, speaking on condition of anonymity. “We’re very concerned about our users’ privacy, and that’s why we started to deny access to sensitive data such as search queries. We know our users don’t want webmasters knowing what they’re searching for, and we want to respect that.”


Later in the same conversation, the same Googler said with a grin and a chortle, “There’ll always be keyword data in AdWords.” Adopting a sing-song tone, one might have quoted Arrested Development‘s timeless one-liner, “There’s always money in the banana stand.”


In other news, Google added a new tab to analytics titled “ASL,” which includes less-sensitive data about users such as age, sex, location, weight, and sexual preference [...]


Cutts: Given today’s technology, page speed a “deprecated metric” for determining site quality


MOUNTAIN VIEW, CA — October 31, 2022 — For as long as there have been websites, there have been slow websites. In past generations, particularly in the 2000s and 2010s, internet users have been frustrated with delay times and unresponsive pages. Because of this, Google has made several attempts to help webmasters create more efficient sites, and has also taken measures to ensure that particularly slow sites do not register as frequently in their search results. But technology has come a long way since the days of 4G and 5G wireless networks.


Matt Cutts, the long-time head of Google’s anti-spam team, addressed a group of fledgling digital marketers this weekend at SMX 2022 saying, “Given the bandwidth speeds of today’s internet service providers, we’re no longer using page speed as an indicator of ‘site quality,’ as this metric is now almost completely deprecated. In the past it made sense to devalue a site that took 10 to 15 seconds to load, and thus provided a negative experience for Googlers. But with today’s 7G Quantum LTE-X technology, the difference between page load times are negligible, almost instant and ultimately irrelevant.


“Does anyone else remember 4G LTE? How about 2G? [Expletive], I’m old.”


So, there you have it. That’s all I found in Will’s office that fateful night. Leave a comment and wish me good tidings if you be so bold. Though as a soon-to-be-toad, I may have a difficult time responding to your queries.


Sign up for The Moz Top 10, a semimonthly mailer updating you on the top ten hottest pieces of SEO news, tips, and rad links uncovered by the Moz team. Think of it as your exclusive digest of stuff you don’t have time to hunt down but want to read!



via Moz Blog http://feedproxy.google.com/~r/MozBlog/~3/6Vi5EaivkLY/search-news-from-the-future


http://seocompanyadvice.com/search-news-from-the-future/?utm_source=rss&utm_medium=rss&utm_campaign=search-news-from-the-future

zaterdag 26 oktober 2013

Link Reclamation – Whiteboard Friday

Posted by RossHudgens


The one thing any good marketer appreciates more than a mention of their brand is a link back to their own domains. For a variety of reasons, some authors—no matter how well meaning they are—don’t include that link with the mention. With the right tools and a little diplomacy, these are some of the easiest opportunities to earn valuable links back to our own properties, and in today’s Whiteboard Friday, Ross Hudgens gives us several great places to start.






























For reference, here’s a still of this week’s whiteboard:



Video Transcription




Hey Moz fans, welcome to Whiteboard Friday. My name is Ross Hudgens, and I work for Siege Media, a content market agency/link building and link development agency. Today I’m going to talk a little bit about link reclamation, one of my favorite subjects.


Link reclamation, if you’re not familiar with it, is the task of finding opportunities on the web where you’ve been linked to or you’ve been mentioned, but haven’t been properly linked to for whatever reason. Maybe the webmaster messed something up, maybe they just didn’t find the right URL, maybe they misspelled your domain name, all kinds of reasons inform why someone might do that incorrectly.


For the purposes of hopefully getting more traffic and also getting those links that we like that potentially can add a lot of value to our domain and help us rank for keywords we want to, it makes sense to do link reclamation. And even more so, I really like it because the conversion is so high. Because someone has already mentioned you, you get a really high conversion on your request, because normally they already have a positive brand sentiment for you.


So one thing to think of in general for link reclamation, if you think about it as a main concept, is it most frequently occurs when your brand has an experience outside of the digital world or disconnected from your main domain. So it could still be digital, but disconnected. So for example, if you’re Target, you probably have a lot of experiences in store where people refer to that, and it doesn’t necessarily make sense to refer to your domain.


Similarly, if you have a YouTube video, it might not make as much sense. But there are still opportunities to ask for a citation back to your domain in those instances that you might not have totally realized before that. That is possible with these kinds of opportunities.


So there are a lot of ways to go about doing that. I’m going to dive into a few of them in this strategy section. A big one for big brands is brand misspellings. Frequently, people will have webmaster error. For whatever reason, they will misspell your domain name.


For example, if you’re a giant company, Pepsi or something like that, you could look for PEPS.com, and likely you’re going to find some instances where people have linked to that thinking it’s Pepsi. So you can go to them and say, “Hey, please fix your error on your site, how you’ve linked to us. It will help your users. It will help us. It will be a great thing all around.” They will generally do that.


So a good process for finding those and also seeing if there is actually link volume around different misspellings is detailed on this link by John Henry Scherck, who I might have mispronounced his name. But that URL, you can find that process, which I can’t really get into as much as I would like to here. But I definitely recommend you go to that post and check it out.


The most powerful ongoing process is simple brand monitoring. So Moz’s tool, Fresh Web Explorer is especially powerful for that. You can use advanced operators to see where your brand has been mentioned but you haven’t been linked to. I think it’s negative LD, or you can see in the advanced operators dropdown. But that’s extremely powerful to just monitor and see who’s mentioning you and all those things as an ongoing thing.


Similarly, Google by date, so Google has an advanced search setting where you can search by 24 hours, a week, a month, things like that, and it will give you the opportunity to see recent mentions that sometimes offer a nice supplement to the fresh web index. So it’s always good to get multiple looks at the web, Fresh Web Explorer I believe uses RSS feeds specifically. Google, by date has, of course, their own comprehensive index. So it’s nice to get a blend of both for finding mentions where people have talked about you, but not linked to you.


So another good one is allinURL/tag/brand. So insert brand. If you’re Pepsi, allin/tag/Pepsi. So these are instances where people think you’re significant enough to actually link to you, significant enough to actually mention you. But sometimes they haven’t linked to you. They’ve created a tag for you because they’ve talked about you in some way in a post. So there’s a lot of sometimes opportunity to get links on those kinds of pages.


So it’s kind of a cool way to easily use Google search engine to find pages that do those kinds of things. So you can just search by that, scrape the results, dump it into like a spreadsheet, and you can quickly find who hasn’t linked to you by doing that kind of process.


Short form text is a kind of a unique thing. It’s any kind of asset that is short, like a definition, a stat, anything like that, that might have been mentioned or stolen without attribution. So examples of that: One stat that is frequently referred to is every second of page load time is a seven percent dip in conversation rate.


So if you have that stat and you actually were the source of that stat, you could track it in Fresh Web Explorer and Google search, these same kind of things just like you do your brand, see who’s mentioned it, and say, “Please attribute us properly with that stat.”


Other examples I like pointing to frequently is Content Marketing Institute. They have a “what is the definition of content marketing,” and that has been stolen like 125,000 times. There is this huge opportunity where people are just taking that, not linking to it, not attributing it properly just because they are lazy or what have you. If you go out and reach out to those webmasters, you can easily get links back, because most of the people will panic in that moment and be like, “Oh, it was just a honest mistake,” and link to you as they should have.


So sometimes you might have that asset, sometimes you might not, but it’s also something to think about and have in the back of your mind when you do that kind of data analysis that might come with an interesting stat that people might want to take.


Reverse image search, so you have a logo or a set of logos, maybe you have interesting assets on your site. For example, if you’re National Geographic, you might have images that everyone takes. You can start monitoring those images, see who’s taken them without attribution, and get links back by doing requests of, “Please attribute properly.”


There are tools like Image Raider, which I know does that. I haven’t used it extensively, but it’s pretty good. Similarly you can use tools like TinEye and also just reverse image search on Google to find those mentions. Of course, an explicit and powerful one is your own logos. So you can see who has mentioned your logo, but not linked to you in the same kind of ways.


Also, just as a tangent from this, Screaming Frog is a really powerful tool. If you haven’t heard of it, it’s a good way to dump a lot of links in there, because sometimes you might see this as a negative process if you go to a lot of these links and you’ve already been linked to. So you can use Screaming Frog as a custom filter and find exactly who hasn’t linked to you by setting an exclude to your domain name. So it can make this process more efficient for you and less frustrating depending on the domain.


YouTube videos, so a lot of people make video assets that are hosted on YouTube or some other platforms, but they don’t necessarily get linked to for whatever reason. Again, it’s something separate from your main domain. So because of that disconnect, they don’t probably link to it a lot of time. Or they’re going to link to the YouTube video, but it doesn’t mean that they’re not willing to link to you as a business if you request it.


So a good way to find that is either dump the YouTube URL in Open Site Explorer or whatever your link management tool is and see what the data is behind that. Or just look at the dashboard. YouTube specifically has an embed dashboard, so you can see where people have embedded that video and not linked back to you, or hopefully they have already linked to you, of course. But you can capture that gap where they haven’t linked to you and they should have because it has a slight disconnect from your main domain as a YouTube video.


Links and tweets and +1s. This is more of an advanced thing. If you have a pretty powerful Twitter account or a Google+ account, you can actually take your archive from Twitter and create a spreadsheet essentially, dump all of your tweets in there. Use a tool like Screaming Frog or use the Moz API for example. Look at the data and see if any of those have been linked to and then see if there is an opportunity to actually reach out and say, “Hey, this tweet, you looked to my Twitter account. I’d really appreciate it if you link to my domain instead.”


A similar thing can be done for Google+. You have to do a site search for your Google+ URL, and they won’t get all of your URLs. But if you scrape that and do the same kind of process, you might find places where people have linked to maybe an interesting tweet by you or some interesting quote you gave or something like that, where you might have not been linked to that you might have wanted to and do a request like that.


So finally, moving links to primary domain. So what that means is, if you’re a big brand, sometimes you have multiple owned properties across the web, but not all of them are you primary KPI for SEO purposes. So I’ve worked with companies who have two main domains, because they can’t make up their mind really, and one is very clearly their SEO domain where they want to rank for stuff. It’s not totally clear which one in the mind of consumers is the primary one.


So something that you can do and that I did in that instance is you go to the one that they don’t really have reason to rank for anything and ask them to link to the other one because maybe you’re changing focus or that’s how you would evangelize it to them. Most of the time they’ll do it because they like you and they’re already linking to you and things like that. You’ll get the more direct link power from those kinds of links.


So, when you’re doing this process, it’s not as simple as asking every single person to link to you. There’s risk involved if you do this incorrectly. So it’s definitely be delicate and don’t step on the wrong toes, because when you do this, there’ll be sites if you’re a big brand that people cover you all the time. Occasionally people will write about you, but maybe in that single post they don’t link to you, but in the previous ten they linked to you.


So in those kinds of instances, just let it go. You don’t need a link in every single post that you get. Potentially it can be kind of put offish to that person that covers you all the time, and you don’t want to lose that good press by burning bridges by being over aggressive as a SEO. So in those kinds of instances, verify that you have links already, that they cover you all the time and just let it go if that’s the kind of instance where they just mistakenly didn’t link to you that one time.


Similarly, don’t step on PR. It’s kind of a similar idea. When you’re doing this process, reaching out to big people that are covering you, it’s frequently newspapers don’t link cite properly so you have to do that kind of outreach. This is where it’s a high value kind of campaign when you hit those big newspapers. But it’s also at risk with PR if you step on them and they don’t like what you’re doing. They hate that you are talking to their contacts directly. So verify that this is an okay thing before you start doing this outreach with your PR team


Ask for links where you should be linked. What I mean by that is sometimes you’ll get mentions in articles in jest. Maybe they’ll be talking about general soda trends, and they’ll randomly jitter off five brands that are sodas, Pepsi, Coca-Cola, Sprite, something like that, all in a sentence. So you could go and say, “Hey link to Pepsi,” but there are four other companies there that they would also have to link to.


You’re just kind of a one off thing in the article. You don’t totally make sense to be linked to in those kinds of situations. So it’s an example of a non-harmonious kind of event where you shouldn’t ask for a link because it doesn’t make sense necessarily. So in those kinds of situations, skip it. You want places where it gets a positive brand sentiment. You should have been linked to in the article, but they didn’t


So if you can really say in your mind it adds value to the article being linked here, then that’s when you should do outreach for this kind of link reclamation. If not, you potentially could burn bridges, step on people’s toes, and put yourself at risk for future coverage that might have been more powerful had you not actually ruined your relationship with that press person.


And finally, if you’re doing this at scale and you have a lot of people mentioning you, you’re a huge brand, you want to use tools that make this more efficient. So one of the problems is sometimes you’ll have no idea if someone has linked to you before, unless you have a process put in place.


So there are a lot of link management tools where you can dump all of your links into it, and it will automatically have a popup in the corner saying that you have a link from this domain. So if you have a lot of people doing outreach and doing link reclamation, you can see, “Hey, I’ve already gotten a link from this domain or multiple links from this domain. I don’t need to do this outreach again.”


Otherwise it’s kind of time intensive, trying to remember who’s linked to you, who hasn’t link to you, whether or not you should do that outreach, and all of those things. So doing that on top of all of these things I think is really powerful.


That’s pretty much it, but I hope you guys see this as a valuable thing that I do. It’s really powerful, especially the bigger your brand, the more powerful it’s going to be. But I think any business on the web today who’s hopefully building a brand, because that’s what it takes to rank in Google and today’s search results, is going to get occasional mentions in any of these instances that you can potentially capitalize on that were missed where people don’t link to you.


So I hope this was valuable and have a good one.




Video transcription by Speechpad.com


Sign up for The Moz Top 10, a semimonthly mailer updating you on the top ten hottest pieces of SEO news, tips, and rad links uncovered by the Moz team. Think of it as your exclusive digest of stuff you don’t have time to hunt down but want to read!


via Moz Blog http://moz.com/blog/link-reclamation-whiteboard-friday


http://seocompanyadvice.com/link-reclamation-whiteboard-friday-2/?utm_source=rss&utm_medium=rss&utm_campaign=link-reclamation-whiteboard-friday-2

Hummingbird Unleashed

Posted by gfiorelli1


Sometimes I think that us SEOs could be wonderful characters for a Woody Allen movie: We are stressed, nervous, paranoid, we have a tendency for sudden changes of mood…okay, maybe I am exaggerating a little bit, but that’s how we tend to (over)react whenever Google announces something.


Cases like this webmaster, who is desperately thinking he was penalized by Hummingbird, are not uncommon.



One thing that doesn’t help is the lack of clarity coming from Google, which not only never mentions Hummingbird in any official document (for example, in the post of its 15th anniversary), but has also shied away from details of this epochal update in the “off-the-record” declarations of Amit Singhal. In fact, in some ways those statements partly contributed to the confusion.


When Google announces an update—especially one like Hummingbird—the best thing to do is to avoid trying to immediately understand what it really is based on intuition alone. It is better to wait until the dust falls to the ground, recover the original documents, examine those related to them (and any variants), take the time to see the update in action, calmly investigate, and then after all that try to find the most plausible answers.


This method is not scientific (and therefore the answers can’t be defined as “surely correct”), it is philological, and when it comes to Google and its updates, I consider it a great method to use.


The original documents are the story for the press of the event during which Google announced Hummingbird, and the FAQ that Danny Sullivan published immediately after the event, which makes direct reference to what Amit Singhal said.


Related documents are the patents that probably underlie Hummingbird, and the observations that experts like Bill Slawski, Ammon Johns, Rand Fishkin, Aaron Bradley and others have derived.


This post is the result of my study of those documents and field observations.


Why did Amit Singhal mix apples with oranges?


When announcing Hummingbird, Amit Singhal said that it wasn’t since Caffeine in 2010 that the Google Algorithm was updated so deeply.


The problem is that Caffeine wasn’t an algorithmic change; it was an infrastructural change.


Caffeine’s purpose, in fact, was to optimize the indexation of the billions of Internet documents Google crawls, presenting a richer, bigger, and fresher pool of results to the users.


Instead, Hummingbird’s objective is not a newer optimization of the indexation process, but to better understand the users’ intent when searching, thereby offering the most relevant results to them.


Nevertheless, we can affirm that Hummingbird is also an infrastructural update, as it governs the more than 200 elements that make up Google’s algorithm.



The (maybe unconscious) association Amit Singhal created between Caffeine and Hummingbird should tell us:



  • That Hummingbird would not be here if Caffeine wasn’t deployed in 2010, and hence it should be considered an evolution of Google Search, and not a revolution.

  • Moreover, that Hummingbird should be considered Google’s most ambitious attempt to solve all the algorithmic issues that Caffeine caused.


Let me explain this last point.


Caffeine, quitting the so-called “Sand Box,” caused the SERPs to be flooded with poor-quality results.


Google reacted by creating “patches” like Panda, Penguin, and the exact-match domain (EMD) updates, among others.


But these updates, so effective in what we define as middle- and head-tail queries, were not so effective for a type of query that—mainly because of the fast adoption of mobile search by the users—more and more people have begun to use: conversational long tail queries, or those that Amit Singhal has defined as “verbose queries.”


The evolution of natural language recognition by Google, the improved ability to disambiguate entities and concepts through technology inherited from Metaweb and improved with Knowledge Graph, and the huge improvements made in the SERPs’ personalized customization have given Google the theoretical and practical tools not only for solving the problem of long-tail queries, but also for giving a fresh start to the evolution of Google Search.


That is the backstory that explains what Amit Singhal told about Hummingbird, paraphrased here by Danny Sullivan:



[Hummingbird] Gave us an opportunity [...] to take synonyms and knowledge graph and other things Google has been doing to understand meaning to rethink how we can use the power of all these things to combine meaning and predict how to match your query to the document in terms of what the query is really wanting and are the connections available in the documents. and not just random coincidence that could be the case in early search engines.



How does Hummingbird work?


“To take synonyms and knowledge graph and other things…”


Google has been working with synonyms for a long time. If we look at the timeline Google itself shared in its 15th anniversary post, it has used them since 2002, even though we can also tell that disambiguation (meant as orthographic analysis of the queries) has been applied since 2001.


Image from Fifteen years on—and we’re just getting started by Amit Singhal on Inside Search blog


Last year Vanessa Fox wrote “Is Google’s Synonym Matching Increasing?…” on Search Engine Land.


Reading that post and seeing the examples presented, it is clear that synonyms were already used by Google—in connection with the user intent underlying the query—in order to broaden the query and rewrite it to offer the best results to the users.


That same post, though, shows us why only using a thesaurus of synonyms or relying on the knowledge of the highly ranked queries was not enough to assure relevant SERPs (see how Vanessa points out how Google doesn’t consider “dogs” pets in the query “pet adoption,” but does consider “cats”).


Amit Singhal, in this old patent, was also conscious that only relying on synonyms was not a perfect solution, because two words may be synonyms and may not be so depending on the context they are used (i.e.: “coche” and “automóvil” both mean “car” in Spanish, but “carro” only means “car” in Latin American Spanish, meaning “wagon” in Spain).


Therefore, in order to deliver the best results possible using semantic search, what Google needed to understand better, easier, and faster was context. Hummingbird is how Google solved that need.


Image from "The Google Hummingbird Patent?" by Bill Slawki on SEO by the Sea.


Synonyms remain essential; Amit Singhal confirmed that in the post-event talk with Danny Sullivan. How they are used now has been described by Bill Slawski in this post, where he dissects the Synonym identification based on co-occurring terms patent.


That patent, then is also based on the concept of “search entities,” which I described in my last post here on Moz, when talking about personalized search.


Speaking literally, words are not “things” themselves but the verbal representation of things, and search entities are how Google objectifies words into concepts. An object may have a relationship with others that may change depending on the context in which they are used together. In this sense, words are treated like people, cities, books, and all the other named entities usually related to the Knowledge Graph.


The mechanisms Google uses in identifying search entities are especially important in disambiguating the different potential meanings of a word, and thereby refining the information retrieval accordingly to a “probability score.”


This technique is not so different from what the Knowledge Graph does when disambiguating, for instance, Saint Peter the Apostle from Saint Peter the Basilica or Saint Peter the city in Minnesota.


Finally, there is a third concept playing an explicit role in what could be the “Hummingbird patent:” co-occurrences.


Integrating these three elements, Google now is (in theory) able:



  1. To better understand the intent of a query;

  2. To broaden the pool of web documents that may answer that query;

  3. To simplify how it delivers information, because if query A, query B, and query C substantively mean the same thing, Google doesn’t need to propose three different SERPs, but just one;

  4. To offer a better search experience, because expanding the query and better understanding the relationships between search entities (also based on direct/indirect personalization elements), Google can now offer results that have a higher probability of satisfying the needs of the user.

  5. As a consequence, Google may present better SERPs also in terms of better ads, because in 99% of the cases, verbose queries were not presenting ads in their SERPs before Hummingbird.



Maybe Hummingbird could have solved Fred Astaire and Ginger Rogers speaking issues…


90% of the queries affected, seriously?


Many SEOs have questioned the fact that Hummingbird has affected the 90% of all queries for the simple reason they didn’t notice any change in traffic and rankings.


Apart from the fact that the SERPs were in constant turmoil between the end of August and the first half of September, during which time Hummingbird first saw the light (though it could just be a coincidence, quite an opportune one indeed), the typical query that Hummingbird targets is the conversational one (e.g.: “What is the best pizzeria to eat at close to Piazza del Popolo e via del Corso?”), a query that usually is not tracked by us SEOs (well, apart from Dr. Pete, maybe).


Moreover, Hummingbird is about queries, not keywords (much less long-tail ones), as was so well explained by Ammon Johns in his post “Hummingbird – The opposite of long-tail search.” For that reason, tracking long-tail rankings as a metric of the impact of Hummingbird is totally wrong.


Finally, Hummingbird has not meant the extinction of all the classic ranking factors, but is instead a new framework set upon them. If a site was both authoritative and relevant for a query, it still will be ranking as well as it was before Hummingbird.


So, which sites got hit? Probably those sites that were relying just on very long tail keyword-optimized pages, but had no or very low authority. Therefore, as Rand said in his latest Whiteboard Friday, now it is far more convenient to create better linkable/shareable content, which also semantically relates to long-tail keywords, than it is to create thousands of long tail-based pages with poor or no quality or utility.


If Hummingbird is a shift to semantic SEO, does that mean that using Schema.org will make my site rank better?


One of the myths that spread very fast when Hummingbird was announced was that it is heavily using structured data as a main factor.


Although it is true that for some months now Google has stressed the importance of structured data (for example, dedicating a section to it in Google Webmaster Tools), considering Schema.org as the magic solution is not correct. It is an example of how us SEOs sometimes confuse the means with the purpose.


Google Data Highlighter is a good alternative to Schema.org, even though not such potent


What we need to do is offer Google easily understandable context for the topics around which we have created a page, and structured data are helpful in this respect. By themselves, however, they are not enough. As mentioned before, if a page is not considered authoritative (thanks to external links and mentions), it most likely will not have enough strength for ranking well, especially now that long-tail queries are simplified by Hummingbird.


Is Hummingbird related to the increased presence of the Knowledge Graph and Answers Cards?


Many people came up with the idea that Hummingbird is the translation of the Knowledge Graph to the classic Google Search, and that it has a direct connection with the proliferation of the Answer Cards. This theory led to some very angry posts ranting against the “scraper” nature of Google.


This is most likely due to the fact that Hummingbird was announced alongside new features of Knowledge Graph, but there is no evident relationship between Hummingbird and Knowledge Graph.


What many have thought as being a cause (Hummingbird causing more Knowledge Graph and Answer Cards, hence being the same) is most probably a simple correlation.


Hummingbird substantially simplified verbose queries into less verbose ones, the latter of which are sometimes complemented with the constantly expanding Knowledge Graph. For that reason, we see a greater number of SERPs presenting Knowledge Graph elements and Answer Cards.


That said, the philosophy behind Hummingbird and the Knowledge Graph is the same, moving from strings to things.


Is Hummingbird strongly based on the Knowledge Base?


The Knowledge Base is potent and pervasive in how Google works, but reducing Hummingbird to just the Knowledge Base would be simplistic.



As we saw, Hummingbird relies on several elements, the Knowledge Base being one of them, especially in all queries with personalization (which should be considered a pervasive layer that affects the algorithm).


If Hummingbird was heavily relying on the Knowledge Base, without complementing it with other factors, we could fall into the issues that Amit Singhal was struggling with in the earlier patent about synonyms.


Does Hummingbird mean the end of the link graph?


No. PageRank and link-related elements of the algorithm are still alive and kicking. I would also dare to say that links are even more important now.


In fact, without the authority a good link profile grants to a site, a web page will have even more difficulty ranking now (see what I wrote just above about the fate of low-authority pages).


What is even more important now is the context in which the link is present. We already learned this with Penguin, but Hummingbird reaffirms how inbound links from topically irrelevant contexts are bad links.


That said, Google still has to improve on the link front, as Danny Sullivan said well in this tweet:





At the same time, though (again because of context and entity recognition), brand co-occurrences and co-citations assume an even more important role with Hummingbird.


Is Hummingbird related to 100% (not provided)?


The fact that Hummingbird and 100% (not provided) were rolled out at almost the same time seems to be more than just a coincidence.


If Hummingbird is more about search entities, better information retrieval, and query expansion—an update where keywords by themselves have lost part of the omnipresent value they had—then relying on keyword data alone is not enough anymore.


We should stop focusing only on keyword optimization and start thinking about topical optimization.


This obliges us to think about great content, and not just about “content.” Things like “SEO copywriting” will end up being the same as “amazing copywriting.”


For that, as SEOs, we should start understanding how search entities work, and not simply become human thesauruses of synonyms.


If Hummingbird is a giant step toward Semantic SEO, then as SEOs, our job “is not about optimizing for strings, or for things, but for the connections between things,” as brilliantly says Aaron Bradley in this post and deck for SMX East.



Semantic SEO – The Shift From Strings To Things by Aaron Bradley #SMX
from Search Marketing Expo – SMX


What must we do to be Hummingbird-friendly?


Let me ask you few questions, and try to answer them sincerely:



  1. When creating/optimizing a site, are you doing it with a clear audience in your mind?

  2. When performing on-page optimization for your site, are you following at least these SEO best practices?

    1. Using a clear and not overly complex information architecture;

    2. Avoiding canonicalization issues;

    3. Avoiding thin-content issues;

    4. Creating a semantic content model;

    5. Topically optimizing the content of the site on a page-by-page basis, using natural and semantically rich language and with a landing page-centric strategy in mind;

    6. Creating useful content using several formats, that you yourself would like to share with your friends and link to;

    7. Implementing Schema.org, Open Graph and semantic mark-ups.



  3. Are your link-building objectives:

    1. Better brand visibility?

    2. Gaining referral traffic?

    3. Enhancing the sense of thought leadership of your brand?

    4. Topically related sites and/or topically related sections of a more generalist site (i.e.: News site)?



  4. As an SEO, is social media offering these advantages?

    1. Wider brand visibility;

    2. Social echo;

    3. Increased mentions/links in the form of derivatives, co-occurrences, and co-citation in others’ web sites;

    4. Organic traffic and brand ambassadors’ growth.




If you answered yes to all these questions, you don’t have to do anything but keep up the good work, refine it, and be creative and engaging. You were likely already seeing your site ranking well and gaining traffic thanks to the more holistic vision of SEO you have.


If you answered no to few of them, then you have just to correct the things you’re doing wrong and follow the so-called SEO best practices (and the 2013 Moz Ranking Factors are a good list of best practices).


If you sincerely answered no to many of them, then you were having problems even before Hummingbird was unleashed, and things won’t get better with it if you don’t radically change your mindset.


Hummingbird is not asking us to rethink SEO or to reinvent the wheel. It is simply asking us to not do crappy SEO… but that is something we should know already, shouldn’t we?


Sign up for The Moz Top 10, a semimonthly mailer updating you on the top ten hottest pieces of SEO news, tips, and rad links uncovered by the Moz team. Think of it as your exclusive digest of stuff you don’t have time to hunt down but want to read!


via Moz Blog http://moz.com/blog/hummingbird-unleashed


http://seocompanyadvice.com/hummingbird-unleashed-2/?utm_source=rss&utm_medium=rss&utm_campaign=hummingbird-unleashed-2

Take the 2013 Moz Industry Survey: Share Your Voice!

Posted by Cyrus-Shepard


We’re very excited to announce the 2013 Moz Industry Survey is ready to take. This is the fourth edition of the survey, which started in 2008 as the SEO Industry Survey and only ran every two years. So much has changed since the last survey that we thought it was important to run it anew in order to gain fresh insights. Some of what we hope to learn and share:



  • Who works in inbound marketing and SEO today?

  • What tactics and tools are most popular?

  • Where are marketers spending their dollars?

  • What does the future of the industry hold?


This year’s survey was redesigned to be easier and only take 5-10 minutes. When the results are in we’ll share the data freely with you, our partners, and the rest of the world.



Prizes


It wouldn’t be the Industry Survey without a few excellent prizes thrown in as an added incentive.


This year we’ve upped the game with prizes we feel are both exciting and perfect for the busy inbound marketer. To see the full sweepstakes terms and rules, go to our sweepstakes rules page. The winners will be announced by June 4th. Follow us on Twitter to stay up to date.


Grand Prize: Attend MozCon 2014 in Seattle


+ Flight

+ Hotel

+ Lunch with an industry expert



Come see us Mozzers in Seattle! The Grand Prize includes one ticket to MozCon 2014 plus airfare and accommodations. We’ll also arrange a one-on-one lunch for you with an industry expert.


2 First Prizes: iPad 2


We’re giving away two separate iPad 2s.



10 Second Prizes: $100 Amazon.com gift cards


Yep, 10 lucky people will win $100 Amazon.com gift cards. Why not buy yourself a nice book?




Why the survey is important


By comparing answers and predictions from one year to the next, we can spot trends and gain insight not easily reported through any other source. This is our best chance to understand exactly where the future of our industry is headed. Some of the things we hope to learn:



  • Demographics: Who is practicing inbound marketing and SEO today? Where do we work and live?

  • Agencies vs. in-house vs. other: How are agencies growing? What’s the average size? Who is doing inbound marketing on their own?

  • Tactics and strategies: What’s working for people today? How have strategies and tactics evolved?

  • Tools and technology: What are marketers using to discover opportunities, promote themselves, and measure the results?

  • Budget and spending: What tools and platforms are marketers investing in?


Every year the Industry Survey delivers new insights and surprises. For example, the chart below (from the 2012 survey) lists average reported salary by role. Will it change in 2013?



2012 SEO Industry Survey



Thanks to our partners


Huge thanks to our partners who are helping to spread the word and encouraging their audience to participate in the survey. We’d especially like to give special recognition to Search Engine Land, Buffer, aimClear, SEOverflow, CopyBlogger, Econsultancy, Content Marketing Institute, TopRank Marketing, MarketingProfs, HootSuite and Entreprenuer.com, Distilled and Hubspot.




Sharing is caring


The number of people who take the survey is very important! The more people who take the survey, the better and more accurate the data will be, and the more insight we can share with the industry.


So please share with your co-workers. Share on social media. Share with your email lists. You can use the buttons below this post to get you started, but remember to take the survey first!



Sign up for The Moz Top 10, a semimonthly mailer updating you on the top ten hottest pieces of SEO news, tips, and rad links uncovered by the Moz team. Think of it as your exclusive digest of stuff you don’t have time to hunt down but want to read!


via Moz Blog http://moz.com/blog/take-2013-industry-survey


http://seocompanyadvice.com/take-the-2013-moz-industry-survey-share-your-voice-2/?utm_source=rss&utm_medium=rss&utm_campaign=take-the-2013-moz-industry-survey-share-your-voice-2

Say Hello to Fresh Alerts: New Mentions and Link Notifications in Your Inbox

Posted by Cyrus-Shepard


Imagine a product similar to Google Alerts, only much better. It’s built specifically for marketers and SEOs. This product not only finds mentions of your keywords and brand, but also reports new links to any website or URL you choose. It comes equipped with advanced search operators to discover new opportunities, and its exportable metrics are sortable by both date and Feed Authority.


To top it all off, it now alerts you via email whenever it finds something new.


Announcing Fresh Alerts from Fresh Web Explorer


For the past few months we’ve enjoyed using Fresh Web Explorer, which has quickly become my favorite new marketing tool. Since then, our engineers and developers have been working to add email alerts to the mix to vastly improve its value.



Starting today, when you use Fresh Web Explorer, you can now set up alerts for up to 10 queries of your choice. The emails are sent daily whenever anything new is discovered. Because Fresh Web Explorer refreshes its index every 8 hours, this means you can be notified of new links and mentions literally within hours after they appear on the web.


When you run a query in Fresh Web Explorer, you have the opportunity to create an alert based on that search.



One key feature is the ability to set your timezone. This helps tailor the reporting specific to your area of the world, so the alerts are more relevant to you.


Fresh Alerts for SEO and inbound marketing


I’ve had the opportunity to beta-test Fresh Alerts for two months, and I can say without hesitation that it’s literally changed the way I do SEO and inbound marketing. We also tested the product with 1,000 Moz beta users, and the feedback has showcased the variety of ways folks are using Fresh Alerts.


1. Link building


While we built Fresh Alerts as a mentions tool, it does a remarkably good job at helping to build links through the process of link reclamation. By using the built-in search operators, you can set your alerts to find non-linking mentions of your brand or keywords on the web.


For example, if I want to search for folks who mention MozRank (a Moz branded term) but don’t include a link to Moz, I’d set up my Fresh Alert like this:



mozrank –rd:moz.com (mentions of Mozrank that don’t link to the root domain moz.com)




With this alert set, every day I would get a new Fresh Alert in my inbox with a list of mentions. Looking at the number of non-linking mentions above, I’d better get link building!


2. Reputation management


Using Fresh Alerts, you can easily be notified whenever anyone mentions you name or brand on the web. Hopefully the information is positive which gives you the opportunity to open a relationship or simply stay on top of the information. If negative, you can reach out and try to mitigate the damage.


Here’s a Fresh Alert email set up for mentions of Rand Fishkin. (In this case, only included mentions that don’t link to moz.com are included.)



You could also use reputation-based alerts to send daily emails to your clients and monitor the conversation about your brand across the web.


3. Competitive intelligence


You can easily set up Fresh Alerts to notify you when your competition is in the news. Better yet, use the search operators to notify you when specific media outlets mention your competition.


In the example below, FEW shows us whenever “Amazon” is mentioned specifically on TechCrunch.



You can also monitor when and where your competitors earn new links. For example, if you wanted to set up a link alert for yourcompetition.com, simply use the Root Domain search operator, like so:



rd:yourcompetition.com (alerts for all new links to the root domain)



By understanding how your competitors earn links and mentions, you may discover new opportunities that are easy to replicate.


4. Reporting and content performance


This is a tip I don’t hear people talk about, so I thought I’d share it. Whenever we publish a big piece of content here at Moz, I set up a Fresh Alert to notify me whenever someone mentions it.


For example, we recently published the 2013 Search Engine Ranking Factors. Because this was an important piece of content for us, I set up 2 different Fresh Alerts:



  • One Fresh Alert notified me whenever people mentioned “Search Engine Ranking Factors” but didn’t link to Moz


  • Another alert to tell me when people linked to the report itself




In the first example, I can reach out to those people who mentioned us without linking to see if I can start a relationship and possibly earn a link.


In the second example, as seen in the graph below, I can monitor our link-building efforts.



5. Discover publishing and guest-post opportunities


Fresh Alerts has to be one of the easiest ways to find distribution, publishing and guest-post opportunities for your content. Yes, high-quality guest posting, when combined with quality content and smart placement, remains a powerful tactic when integrated with other marketing opportunities.


For example, let’s say your subject is “dragons” and you want to find blogs that have posted guest posts in the past few days. You can simply create an alert for “dragons” AND “guest post”.



This alert will notify you whenever a new post is published mentioning both “guest post” and “dragons”.


This technique isn’t limited to guest posting, either. Getting creative, you could find other publishing opportunities for your specific niche.


The details


Starting now, we’ve made Fresh Alerts available to subscribers of Moz Analytics. If you’re not a PRO member, you can sign up for a 30-day trial to give them a spin if you’d like, which also includes access to our new Moz Analytics and full suite of inbound marketing tools.


Here’s what you need to know about Fresh Alerts:



  • Activate up to 10 Alerts per Moz Analytics account

  • When Fresh Web Explorer finds new mentions or links, you receive an email within 24 hours

  • Alerts are sorted by Feed Authority, our new metric created specifically for FWE

  • All advanced operators used by Fresh Web Explorer are available with Fresh Alerts



Have you tried Fresh Web Explorer already? If so, let us know your best ideas for Fresh Alerts in the comments below.


Sign up for The Moz Top 10, a semimonthly mailer updating you on the top ten hottest pieces of SEO news, tips, and rad links uncovered by the Moz team. Think of it as your exclusive digest of stuff you don’t have time to hunt down but want to read!


via Moz Blog http://moz.com/blog/say-hello-to-fresh-alerts


http://seocompanyadvice.com/say-hello-to-fresh-alerts-new-mentions-and-link-notifications-in-your-inbox-2/?utm_source=rss&utm_medium=rss&utm_campaign=say-hello-to-fresh-alerts-new-mentions-and-link-notifications-in-your-inbox-2

A [Poorly] Illustrated Guide to Google’s Algorithm

Posted by Dr-Pete


Like all great literature, this post started as a bad joke on Twitter on a Friday night:



If you know me, then this kind of behavior hardly surprises you (and I probably owe you an apology or two). What’s surprising is that Google’s Matt Cutts replied, and fairly seriously:



Matt’s concern that even my painfully stupid joke could be misinterpreted demonstrates just how confused many people are about the algorithm. This tweet actually led to a handful of very productive conversations, including one with Danny Sullivan about the nature of Google’s “Hummingbird” update.


These conversations got me thinking about how much we oversimplify what “the algorithm” really is. This post is a journey in pictures, from the most basic conception of the algorithm to something that I hope reflects the major concepts Google is built on as we head into 2014.


The Google algorithm


There’s really no such thing as “the” algorithm, but that’s how we think about it—as some kind of monolithic block of code that Google occasionally tweaks. In our collective SEO consciousness, it looks something like this:



So, naturally, when Google announces an “update”, all we see are shades of blue. We hear about a major algorithm update ever month or two, and yet Google confirmed 665 updates (technically, they used the word “launches”) in 2012—obviously, there’s something more going on here than just changing a few lines of code in some mega-program.


Inputs and outputs


Of course, the algorithm has to do something, so we need inputs and outputs. In the case of search, the most fundamental input is Google’s index of the worldwide web, and the output is search engine result pages (SERPs):



Simple enough, right? Web pages go in, [something happens], search results come out. Well, maybe it’s not quite that simple. Obviously, the algorithm itself is incredibly complicated (and we’ll get to that in a minute), but even the inputs aren’t as straightforward as you might imagine.


First of all, the index is really roughly a dozen data centers distributed across the world, and each data center is a miniature city unto itself, linked by one of the most impressive global fiber optic networks ever built. So, let’s at least add some color and say it looks something more like this:



Each block in that index illustration is a cloud of thousands of machines and an incredible array of hardware, software and people, but if we dive deep into that, this post will never end. It’s important to realize, though, that the index isn’t the only major input into the algorithm. To oversimplify, the system probably looks more like this:



The link graph, local and maps data, the social graph (predominantly Google+) and the Knowledge Graph—essentially, a collection of entity databases—all comprise major inputs that exist beyond Google’s core index of the worldwide web. Again, this is just a conceptualization (I don’t claim to know how each of these are actually structured as physical data), but each of these inputs are unique and important pieces of the search puzzle.


For the purposes of this post, I’m going to leave out personalization, which has its own inputs (like your search history and location). Personalization is undoubtedly important, but it impacts many areas of this illustration and is more of a layer than a single piece of the puzzle.


Relevance, ranking and re-ranking


As SEOs, we’re mostly concerned (i.e. obsessed) with ranking, but we forget that ranking is really only part of the algorithm’s job. I think it’s useful to split the process into two steps: (1) relevance, and (2) ranking. For a page to rank in Google, it first has to make the cut and be included in the list. Let’s draw it something like this:



In other words, first Google has to pick which pages match the search, and then they pick which order those pages are displayed in. Step (1) relies on relevance—a page can have all the links, +1s, and citations in the world, but if it’s not a match to the query, it’s not going to rank. The Wikipedia page for Millard Fillmore is never going to rank for “best iPhone cases,” no matter how much authority Wikipedia has. Once Wikipedia clears the relevance bar, though, that authority kicks in and the page will often rank well.


Interestingly, this is one reason that our large-scale correlation studies show fairly low correlations for on-page factors. Our correlation studies only measure how well a page ranks once it’s passed the relevance threshold. In 2013, it’s likely that on-page factors are still necessary for relevance, but they’re not sufficient for top rankings. In other words, your page has to clearly be about a topic to show up in results, but just being about that topic doesn’t mean that it’s going to rank well.


Even ranking isn’t a single process. I’m going to try to cover an incredibly complicated topic in just a few sentences, a topic that I’ll call “re-ranking.” Essentially, Google determines a core ranking and what we might call a “pure” organic result. Then, secondary ranking algorithms kick in—these include local results, social results, and vertical results (like news and images). These secondary algorithms rewrite or re-rank the original results:



To see this in action, check out my post on how Google counts local results. Using the methodology in that post, you can clearly see how Google determines a base set of rankings, and then the local algorithm kicks in and not only adds new features but re-ranks the original results. This diagram is only the tip of the iceberg—Bill Slawski has an excellent three-part series on re-ranking that covers 40 different ways Google may re-rank results.


Special inputs: penalties and disavowals


There are also special inputs (for lack of a better term). For example, if Google issues a manual penalty against a site, that has to be flagged somewhere and fed into the system. This may be part of the index, but since this process is managed manually and tied to Google Webmaster Tools, I think it’s useful to view it as a separate concept.


Likewise, Google’s disavow tool is a separate input, in this case one partially controlled by webmasters. This data must be periodically processed and then fed back into the algorithm and/or link graph. Presumably, there’s a semi-automated editorial process involved to verify and clean this user-submitted data. So, that gives us something like this:



Of course, there are many inputs that feed other parts of the system. For example, XML sitemaps in Google Webmaster Tools help shape the index. My goal it to give you a flavor for the major concepts. As you can see, even the “simple” version is quickly getting complicated.


Updates: Panda, Penguin and Hummingbird


Finally, we have the algorithm updates we all know and love. In many cases, an update really is just a change or addition to some small part of Google’s code. In the past couple of years, though, algorithm updates have gotten a bit more tricky.


Let’s start with Panda, originally launched in February of 2011. The Panda update was more than just a tweak to the code—it was (and probably still is) a sub-algorithm with its own data structures, living outside of the core algorithm (conceptually speaking). Every month or so, the Panda algorithm would be re-run, Panda data would be updated, and that data would feed what you might call a Panda ranking factor back into the core algorithm. It’s likely that Penguin operates similarly, in that it’s a sub-algorithm and separate data set. We’ll put them outside of the big, blue oval:



I don’t mean to imply that Panda and Penguin are the same—they operate in very different ways. I’m simply suggesting that both of these algorithm updates rely on their own code and data sources and are only periodically fed back into the system.


Why didn’t Google just re-write the algorithm to account for the Panda and/or Penguin intent? Part of it is computational—the resources required to process this data are beyond what the real-time infrastructure can probably handle. As Google gets faster and more powerful, these sub-algorithms may become fully integrated (and Panda is probably more integrated than it once was). The other reason may involve testing and mitigating impact. It’s likely that Google only updates Penguin periodically because of the large impact that the first Penguin update had. This may not be a process that they simply want to let loose in real-time.


So, what about the recent Hummingbird update? There’s still a lot we don’t know, but Google has made it pretty clear that Hummingbird is a fundamental rewrite of how the core algorithm works. I don’t think we’ve seen the full impact of Hummingbird yet, personally, and the potential of this new code may be realized over months or even years, but now we’re talking about the core algorithm(s). That leads us to our final image:




Image credit for hummingbird silhouette: Michele Tobias at Experimental Craft.


The end result surprised even me as I created it. This was the most basic illustration I could make that didn’t feel misleading or simplistic. The reality of Google today far surpasses this diagram—every piece is dozens of smaller pieces. I hope, though, that this gives you a sense for what the algorithm really is and does.


Additional resources


If you’re new to the algorithm and would like to learn more, Google’s own “How Search Works” resource is actually pretty interesting (check out the sub-sections, not just the scroller). I’d also highly recommend Chapter 1 of our Beginner’s Guide: “How Search Engines Operate.” If you just want to know more about how Google operates, Steven Levy’s book “In The Plex” is an amazing read.


Special bonus nonsense!


While writing this post, the team and I kept thinking there must be some way to make it more dynamic, but all of our attempts ended badly. Finally, I just gave up and turned the post into an animated GIF. If you like that sort of thing, then here you go…



Sign up for The Moz Top 10, a semimonthly mailer updating you on the top ten hottest pieces of SEO news, tips, and rad links uncovered by the Moz team. Think of it as your exclusive digest of stuff you don’t have time to hunt down but want to read!


via Moz Blog http://moz.com/blog/a-poorly-illustrated-guide-to-googles-algorithm


http://seocompanyadvice.com/a-poorly-illustrated-guide-to-googles-algorithm-2/?utm_source=rss&utm_medium=rss&utm_campaign=a-poorly-illustrated-guide-to-googles-algorithm-2