An inside look at Google's news-ranking algorithm
Patent application seeks to refine algorithm for third time since 2003
Computerworld - A patent application filed by Google last year provides a detailed look at some of the metrics the company considers when ranking news stories and deciding how prominently to display them on its Google News page.
The application, filed in February 2012 and published last July, seeks to build on a patent Google was issued in 2009 titled "Systems and Methods for Improving the Ranking of News Articles." Computerworld found the document while conducting an unrelated patent search on the United States Patent Office's website.
A Google spokesman had no comment on the specifics of the application. "We file patent applications on a variety of ideas that our employees come up with," he said via email. " Some of those ideas later mature into real products or services some don't. Prospective product announcements should not necessarily be inferred from our patent applications."
The 2012 patent application offers details on more than a dozen separate metrics the company uses to rank news stories created by other Websites. How Google evaluates stories has been a point of contention with various media companies who have in the past said the company is infringing on their work. Many have also complained that Google can effectively turn on or off a spigot of visitors to a Web site by prominently displaying, or downplaying, a story. Google's decisions also affect what stories readers see, potentially shaping their view of news events.
Since it was launched in 2002, Google News has become the largest aggregator of news content on the Web. The site, which is completely computer-generated, collects and display headlines from thousands of news sources around the world.
The metrics cited in the patent application include: the number of articles produced by a news organization during a given time period; the average length of an article from a news source; and the importance of coverage from the news source.
Other metrics include a breaking news score, usage patterns, human opinion, circulation statistics and the size of the staff associated with a particular news operation.
Also factored in are the number of news bureaus a news source has, the number of original named entities used in stories, breadth of coverage, international diversity and even writing style.
The patent application provides some much needed visibility into how companies like Google select and rank online content, said Sree Sreenivasan, professor of professional practice at Columbia University's Journalism School and the university's first chief digital officer.
"In the world of technology, so much is opaque. It's nice to have some clarity around this stuff," Sreenivasan said. He noted that some of the metrics Google appears to be using to judge the quality of a news source are the same kind of metrics editors would use in deciding whether to trust a publication or not.
He pointed to metrics like staff size and audience diversity as examples. Even Google's use of story length is a good metric, Sreenivasan said. At first blush, it would appear that Google is emphasizing quantity over quality, he said. But the reality is that many high-quality media organizations now generate more content than they used to. So using story lengths and word counts is valid, he said.
- Best iPhone, iPad Business Apps for 2014
- 14 Tech Conventions You Should Attend in 2014
- 10 Desktop Apps to Power Your Windows PC
- How to Add New Job Skills Without Going Back to School
- Slideshow: 7 security mistakes people make with their mobile device
- iOS vs. Android: Which is more secure?
- 11 sure signs you've been hacked
- Four Myths of High-Productivity App Dev Debunked Debunk the main myths surrounding high-productivity application development and how both platforms have overcome them.
- IDC: Dual Perspectives on ITaaS This global survey by IDG Research Services of more than 350 IT and BU directors at enterprises of 1,000 employees or more reveals...
- ESG Whitepaper: Integrated Computing Platforms: Infrastructure Builds Tomorrow's Data Center Thought Leadership Report: 'Integrated Computing Platforms: Infrastructure Builds for Tomorrow's Data Center.' Highlights the challenges customers face when planning to deploy a Cloud...
- Analyst Report-Mixed All Flash Arrays Delivers Safer Higher Performance What is the impact of an all-flash array with enterprise features and reliability on the mainstream data center? In the mainstream environment, storage...
- Live Webcast Increasing the Value of Your Reports and Dashboards Learn how incorporating other analytical capabilities such as predictive modeling and visualization can increase the value of your reports and dashboards by providing...
On-Demand Webcast: 7 Reasons to Choose VoIP
Thinking about a new phone system for your business?
Be sure to watch this informative webcast. Steve Strauss, small business columnist for USA...
- Top 8 Communications Tools for Small Businesses Powerful technology is available to help your small business improve its communications with customers, employees and suppliers. View this free On-Demand Webcast produced... All Business Intelligence/Analytics White Papers | Webcasts
By Rob F. Walker, Ph.D.
In the previous installment, we looked at and discussed strategies for business simulation and the infrastructure needed to make such initiatives successful. Now, we¿re ready to discuss some practical examples of business simulation. Imagine a mail order company selling products together with the necessary financing. more