Ex-Google, Yahoo staffers release Hadoop distribution

By Chris Kanaracus, IDG News Service |  Open Source, Cloudera, Hadoop Add a new comment

A startup called Cloudera on Monday publicly released its distribution of the open-source Hadoop distributed computing framework, hoping to sell enterprise users on the system employed by Google, Yahoo and others to process large data sets.

Cloudera, which was launched by former Google, Yahoo, Oracle and Facebook employees last year, has been providing its initial customers with support for Hadoop.

"One of the repeating themes we have heard while working with our customers and the community is that Hadoop configuration and deployment is a pain," Cloudera employee and former Googler Christophe Bisciglia said in a blog post. "In order for Big Data to truly disrupt the enterprise, Hadoop needs to be just as easy to configure, deploy and manage as any other piece of software."

That is why Cloudera has decided to release its distribution, which is available as an RPM bundle for systems running Red Hat Linux, as well as an image for Amazon's Elastic Compute Cloud (EC2).

The distribution is available at no charge, under the Apache 2 license. By releasing the package, Cloudera is no doubt hoping more enterprises take a look at using Hadoop and subsequently tap Cloudera's support services, for which pricing information was not immediately available Monday.

Cloudera's distribution has three components: the Hadoop distributed file system, which can run on commodity machines; an implementation of the MapReduce framework originally developed by Google for parallel processing of large data sets; and Hive, a data warehousing layer that uses the SQL-based HQL language for querying.

Also Monday, Cloudera said it had secured US$5 million in funding from Accel Partners along with other investors, including former MySQL and Sun executive Marten Mickos, LinkedIn president Jeff Weiner and Gideon Yu, chief financial officer at Facebook.

    Add a comment

    Post a comment using one of these accounts
    Or join now
    At least 6 characters

    Note: Comment will appear soon after you have activated your account.
    Obscene/spam comments will be removed and accounts suspended.
    The information you submit is subject to our Privacy Policy and Terms of Service.

    ITworld LIVE

    Open SourceWhite Papers & Webcasts

    White Paper

    Consolidating SAP Applications to Linux on Power by IDC

    IDC studied a group of enterprises that had deployed SAP applications on IBM Power Systems servers running Linux server operating environments and had been working with those systems for several years. Learn about the results...

    White Paper

    An Interactive eGuide: Open Source

    By now, enterprises are well aware of the benefits of open-source software, which boasts a clean design, reliability, and maintainability, as well as support for standards and community values. But perhaps the biggest benefit is quality; since open-source software users have access to source code, bug fixes and enhancements come from multiple sources, often resulting in superior software.

    See more White Papers | Webcasts

    Ask a question

    Ask a Question