Skip the navigation
)
News

Yahoo spinoff unveils Hadoop data analysis system

Yahoo's spin off of its Hadoop expertise will release its first distribution

By Joab Jackson
November 1, 2011 08:04 AM ET

IDG News Service - Yahoo spinoff company Hortonworks has released a preview edition of what will be a fully open source distribution of the Apache Hadoop data analysis platform, called the Hortonworks Data Platform, the company announced Tuesday.

The company has also started offering commercial support and training programs, designed for system integrators and independent software vendors planning to deploy Hadoop on behalf of their customers.

"We'll use our deep-domain knowledge to empower our partners to solve their customers' problems with Hadoop," said Eric Baldeschwieler, CEO of Hortonworks. He likened Hortonworks business model to Linux system software provider Red Hat, which maintains open source versions of all its software and charges for maintenance and support.

The current pre-production version of the Hortonworks Data Platform (HDP) is now being distributed through the company's Technology Preview Program, available to a select number of users. Hortonworks plans to publicly release the pre-production version of the software within the first three months of 2012.

HDP will be a fully open source distribution, including all the components used for a typical Hadoop deployment, including Hadoop Distributed File System (HDFS), MapReduce, Pig, Hive, HBase and Zookeeper. The package will also include a number of newer enterprise-ready components as well, all open source. One is HCatalog, a metadata management service that can aid in connecting Hadoop with other enterprise information systems. The package will also include the Ambari open source installation and management system for Hadoop clusters.

Last June, Yahoo, along with Benchmark Capital, set up Hortonworks as a stand-alone company dedicated to facilitating the enterprise use of Hadoop. Yahoo played a pivotal role in the early development of Hadoop, hiring Hadoop creator Doug Cutting in order to refine the software, which it used for large-scale data analysis.

About 20 Yahoo engineers, all with considerable Hadoop skills, were shifted to Hortonworks when the company was created (though Cutting himself is now at Cloudera). The company is hoping that its in-house expertise will make its services a viable choice alongside earlier entrants into the field of Hadoop support and software, including Cloudera and IBM.

Also, unlike their competitors, Hortonworks will keep its distribution fully open source, Baldeschwieler said. Other companies "have proprietary components where they replaced parts of the Hadoop stack with their own proprietary components," Baldeschwieler said. Hortonworks plans to fill in the missing pieces of the Hadoop stack, such as management tools, with fully open source components.

Hortonworks also plans to aggressively pursue partnerships with other software and system companies. Last month, the company teamed with Microsoft to bring Hadoop to Windows Server and Microsoft's Azure cloud service. "We're developing to build an ecosystem of partners," Baldeschwieler said.

Joab Jackson covers enterprise software and general technology breaking news for The IDG News Service. Follow Joab on Twitter at @Joab_Jackson. Joab's e-mail address is Joab_Jackson@idg.com

Reprinted with permission from IDG.net. Story copyright 2012 International Data Group. All rights reserved.
What is Tech Briefcase?
TechBriefcase is a new, free service where IT Professionals can Search, Store and Share IT white papers and content like this. Learn more
Bookmark content
Speed up your research efforts with content across the web.
Search and Store
Find the white papers you need. Create folders for any topic.
View Anywhere
Open your briefcase on your iPhone, tablet or desktop. Share with colleagues.
Don't have an account yet?
Additional Resources
Security KnowledgeVault
WHITE PAPER
Security is not an option. This KnowledgeVault Series offers professional advice how to be proactive in the fight against cybercrimes and multi-layered security threats; how to adopt a holistic approach to protecting and managing data; and how to hire a qualified security assessor. Make security your Number 1 priority.

Read now.

Cut Communications Costs Once and for All
WHITE PAPER
New IP-based communications systems are being deployed by small and midsized businesses at a rapid rate. Learn how these organizations are enabling faster responsiveness, creating better customer experiences, speeding office or mobile interactions, and dramatically reducing existing communications costs.

Read now.

BI and Analytics White Papers
Thinking Outside The Data Warehouse
This high level, business problem focused eBook uses 5 customer scenarios to show how people and organizations are tackling real issues using IBM...
Using BD for Smarter Decision Making
This paper looks at new developments in business analytics and discusses the benefits analyzing big data bring to the business.
Measuring the Business Value of CI in the Data Center
One of the key strategies that IT teams are pursuing to reduce capital costs while boosting asset utilization and employee productivity is the...
Switching Schedulers - Not As Complicated As You Think
Changing or consolidating job schedulers may seem daunting. However, the benefits of switching to enterprise workload automation outweigh the risks. Read how BMC...
Capture-Enabled Business Process Management
Organizations today must deal with a vast amount of incoming information from many different sources. Efficient, automated business processes are critical to managing...
All BI and Analytics White Papers
BI and Analytics Webcasts
InfoSphere Warehouse Packs Demo
These flash modules make warehousing more tangible and relevant to business users through detailed explanations of the InfoSphere Warehouse Packs.
Delivery Management -- Extending Lifecycle Management
Date: Wednesday, June 20, 2012, 1:00 PM EDT

Siloed organizations continue doing the wrong things and doing things wrong, leading to increased costs,...
Leverage automation today to reduce IT complexity
Date: Tuesday, June 5, 2012, 2:00 PM EDT

Whether your B2B complexity is caused by multiple technologies due to M&A, business or application specific...
BMC Control-M - Single Point of Control Demo
With BMC Control-M, you schedule and manage everything - down to the very last platform and application - from one simple interface. It's...
BMC Control-M - Single Point of Control Demo
With BMC Control-M, you schedule and manage everything - down to the very last platform and application - from one simple interface. It's...
All BI and Analytics Webcasts
Newsletter Sign-Up

Receive the latest news test, reviews and trends on your favorite technology topics

Choose a newsletter
  1. View all newsletters | Privacy Policy
IT Jobs