Skip the navigation
)
News

EMC unveils Hadoop appliance, BI software

Joins with quiet start-up MapR Technologies to create free and enterprise versions of Hadoop-based data analytics software

May 9, 2011 03:37 PM ET

Computerworld - LAS VEGAS - EMC Monday unveiled a purpose-built appliance for processing both structured and unstructured data sets for business analytics tasks.

EMC today also announced the availability of two new business intelligence software products -- the Hadoop-based EMC Greenplum HD Community and Enterprise Editions -- at its EMC World user conference here.

Service contracts for both software products includes installation, training and global technical support.

The Greenplum HD Community Edition is a fully-certified downloadable free software stack. The software is based on Hadoop, an Apache data management software, and is optimized to run on virtual machines.

Greenplum HD Enterprise Edition is tailored for corporate data centers, with capabilities like fault tolerance through automated node failure detection and notification, multi-site management and data management features such as snapshots and wide area replication.

It also offers simple data loading from databases and access via a native Network File System (NFS) interface.

EMC claims its version of Hadoop delivers two to five times the performance over the standard packaged versions of Apache Hadoop.

EMC had signed an agreement with Cloudera last fall to use its Hadoop-based data management software and services.

However, Scott Yara, vice president of products with EMC's Data Computing Division and co-founder of Greenplum, said today that the company is moving in a "new direction" with partner MapR Technologies, a start-up in development mode for the past two years.

MapR built a proprietary replacement for the Hadoop Distributed File System (HDFS) that can replace existing installations of the Hadoop file system.

John Schroeder, CEO of MapR, said his company's version of map reduce technology returns far faster data analytics results, and can manage larger data sets on fewer machines than current Hadoop iterations.

"We can reduce the size of a cluster," he told a gathering of reporters and analysts gathered at the show. "That's a tremendous TCO savings."

Luke Lonergan, CTO of EMC's Data Computing Division and another co-founder of Greenplum, added that EMC is working with dozens of resellers to get the MapR Hadoop software to customers. The distribution channels should go live later the second quarter of this year.

No pricing has yet been released for the Enterprise-edition.

"Hadoop has played a leading role in the transformation from traditional data warehousing to Big Data Analytics," said John Webster, a senior partner with research firm the Evaluator Group, in a statement. "EMC's Hadoop commercialization strategy is aimed at streamlining and bulletproofing Hadoop for enterprise users, making Hadoop more of a must-have real-time analytics tool for the enterprise."

EMC's Hadoop appliance

Along with its new software products, EMC introduced an updated Greenplum Data Computing Appliance that runs Hadoop for easy installation of the business intelligence technology.

The new Greenplum HD Data Computing Appliance is built on top of Intel X86 servers and it uses both a structured database built by Greenplum, which EMC acquired last year , and the Apache open-source version of Hadoop. The older version of the appliance is based on Sun Fire x64-based servers.

According to Yara, administrators can read and write files in parallel from Greenplum to HDFS, enabling rapid data sharing. Cross-platform analysis can be performed using Greenplum SQL and advanced analytic functions accessing data on HDFS.

"We're here to build a big data analytics stack," Yara said. "It's a unified stack whether it's for structured data in Greenplum's database, or through a data computing appliance."

The new Hadoop appliance is expected to be able to scale to a large number of nodes, but EMC did not disclose details.

The appliance is due out in the third quarter of this year, Yara said.

Lucas Mearian covers storage, disaster recovery and business continuity, financial services infrastructure and health care IT for Computerworld. Follow Lucas on Twitter at Twitter@lucasmearian, or subscribe to Lucas's RSS feed Mearian RSS. His e-mail address is lmearian@computerworld.com.

Read more about BI and Analytics in Computerworld's BI and Analytics Topic Center.



What is Tech Briefcase?
TechBriefcase is a new, free service where IT Professionals can Search, Store and Share IT white papers and content like this. Learn more
Bookmark content
Speed up your research efforts with content across the web.
Search and Store
Find the white papers you need. Create folders for any topic.
View Anywhere
Open your briefcase on your iPhone, tablet or desktop. Share with colleagues.
Don't have an account yet?
Additional Resources
Security KnowledgeVault
WHITE PAPER
Security is not an option. This KnowledgeVault Series offers professional advice how to be proactive in the fight against cybercrimes and multi-layered security threats; how to adopt a holistic approach to protecting and managing data; and how to hire a qualified security assessor. Make security your Number 1 priority.

Read now.

Cut Communications Costs Once and for All
WHITE PAPER
New IP-based communications systems are being deployed by small and midsized businesses at a rapid rate. Learn how these organizations are enabling faster responsiveness, creating better customer experiences, speeding office or mobile interactions, and dramatically reducing existing communications costs.

Read now.

BI and Analytics White Papers
Thinking Outside The Data Warehouse
This high level, business problem focused eBook uses 5 customer scenarios to show how people and organizations are tackling real issues using IBM...
Using BD for Smarter Decision Making
This paper looks at new developments in business analytics and discusses the benefits analyzing big data bring to the business.
Measuring the Business Value of CI in the Data Center
One of the key strategies that IT teams are pursuing to reduce capital costs while boosting asset utilization and employee productivity is the...
Switching Schedulers - Not As Complicated As You Think
Changing or consolidating job schedulers may seem daunting. However, the benefits of switching to enterprise workload automation outweigh the risks. Read how BMC...
Capture-Enabled Business Process Management
Organizations today must deal with a vast amount of incoming information from many different sources. Efficient, automated business processes are critical to managing...
All BI and Analytics White Papers
BI and Analytics Webcasts
InfoSphere Warehouse Packs Demo
These flash modules make warehousing more tangible and relevant to business users through detailed explanations of the InfoSphere Warehouse Packs.
Delivery Management -- Extending Lifecycle Management
Date: Wednesday, June 20, 2012, 1:00 PM EDT

Siloed organizations continue doing the wrong things and doing things wrong, leading to increased costs,...
Leverage automation today to reduce IT complexity
Date: Tuesday, June 5, 2012, 2:00 PM EDT

Whether your B2B complexity is caused by multiple technologies due to M&A, business or application specific...
BMC Control-M - Single Point of Control Demo
With BMC Control-M, you schedule and manage everything - down to the very last platform and application - from one simple interface. It's...
BMC Control-M - Single Point of Control Demo
With BMC Control-M, you schedule and manage everything - down to the very last platform and application - from one simple interface. It's...
All BI and Analytics Webcasts
Newsletter Sign-Up

Receive the latest news test, reviews and trends on your favorite technology topics

Choose a newsletter
  1. View all newsletters | Privacy Policy
IT Jobs