Welcome!

Linux Authors: Jim Kaskade, Carmen Gonzalez, Nikita Ivanov, Pat Romanski, Victoria Livschitz

News Feed Item

Mu Sigma Helps Companies Accelerate Big Data Analysis With First Packaged MapReduce Algorithms for Hadoop

New Offering Helps Accelerate Big Data Analysis Projects; in Testing, Mu Sigma's Offering Consistently Outperformed Commercial Software Tools

CHICAGO, IL -- (Marketwired) -- 06/27/13 -- Mu Sigma (http://www.mu-sigma.com), the largest pure-play provider of decision sciences and analytics solutions for global enterprise customers, launched a new addition to its series of analytical products. muHPC™ (for High Performance Computing) is a library of popular statistical algorithms written in MapReduce, designed for enterprise-class Big Data analysis in Hadoop environments. As with Mu Sigma's other products, muHPC was successfully and extensively used within Mu Sigma on many client engagements before the company brought it to market.

Traditionally, enterprises that wanted to leverage R and Hadoop for Big Data analysis have had to write their own algorithms, or rely on open-source options that had not been widely used or tested. Quality varied, and it was a challenge for companies to acquire talent with relevant skills and competencies in order to code their own algorithms. Mu Sigma's offering enables enterprises to accelerate their R and Hadoop initiatives, and their overall Big Data analysis programs. In testing, muHPC packages consistently outperformed a leading commercial software in equivalent procedures in terms of execution time on large data sets -- in fact, muHPC algorithms proved to be 2-4 times faster while achieving the same results.

"We talk with so many large enterprises that want to leverage open-source tools such as R and Hadoop but simply cannot find staff with the requisite skills," said Zubin Dowlaty, Head of Innovation and Development at Mu Sigma -- the group responsible for developing new technology solutions for Mu Sigma's internal use and for eventual launch to the public. "muHPC directly addresses that market need by providing a packaged set of the most common R-based algorithms that can be used in a Hadoop environment right out of the box. muHPC is a breakthrough concept that removes significant barriers to Big Data analysis."

muHPC consists of three packages currently:

  • muGLM: Offers easy-to-use R functions for building a wide variety of generalized linear models (OLS, Logistic, Poisson, Negative Binomial, Gamma etc.) on Big Data.
  • muEDA: Offers easy-to-use R functions for performing exploratory analysis on Big Data.
  • muKMeans: Offers easy-to-use R functions for data clustering on Big Data using the K-means algorithm.

Mu Sigma leveraged technology from Cloudera and Revolution Analytics to build muHPC. "Cloudera's Distribution including Apache Hadoop is a great platform for organizations to store and analyze data," said Tim Stevens, vice president, Business Development, Cloudera. "Mu Sigma developing and certifying their new solution on top of Cloudera ensures that their approach to solving big data challenges is both innovative and effective."

muHPC algorithms have been written using components from Revolution Analytics' open-source RHadoop project. Hadoop integration is based on the rmr2 package, which provides Hadoop MapReduce functionality in R, and has been implemented and tested with Cloudera's distribution of Hadoop and Revolution R Enterprise. "Mu Sigma's initiative to utilize open-source R packages for commercial implementations provides a great impetus to R's popularity and it being adopted as the standard environment for mainstream analytical analysis," said Greg Fuller, Vice President, Partners and Channels at Revolution Analytics. "We are very pleased that our Alliance partners are building differentiated solutions based on Revolution Analytics solutions. We look forward to doing even more with our partners as we further develop Revolution R Enterprise Platform-as-a-Service."

muHPC is available now, and comes with an annual subscription license on a per-cluster basis with an incremental price per-node. To learn more, visit www.mu-sigma.com/muhpc.

About Mu Sigma
Mu Sigma, one of the world's largest pure-play decision sciences and analytics firms, helps companies institutionalize data-driven decision making and harness Big Data. Mu Sigma solves high-impact business problems in the areas of Marketing, Risk and Supply Chain across 10 industry verticals. With over 2500 decision science professionals and more than 75 Fortune 500 clients, Mu Sigma has driven disruptive innovation in the analytics industry with its interdisciplinary approach combining business, math and technology, and its integrated decision support ecosystem comprised of technology platforms, processes, methodologies and people. Visit www.mu-sigma.com.

Mu Sigma is a trademark of Mu Sigma, Inc. in the United States and other countries. All other trademarks contained herein are the property of their respective owners.

Add to Digg Bookmark with del.icio.us Add to Newsvine

For information, contact:
Michelle G. Faulkner
Big Swing Communications
+1 617-510-6998
Email Contact

More Stories By Marketwired .

Copyright © 2009 Marketwired. All rights reserved. All the news releases provided by Marketwired are copyrighted. Any forms of copying other than an individual user's personal reference without express written permission is prohibited. Further distribution of these materials is strictly forbidden, including but not limited to, posting, emailing, faxing, archiving in a public database, redistributing via a computer network or in a printed form.

@ThingsExpo Stories
The definition of IoT is not new, in fact it’s been around for over a decade. What has changed is the public's awareness that the technology we use on a daily basis has caught up on the vision of an always on, always connected world. If you look into the details of what comprises the IoT, you’ll see that it includes everything from cloud computing, Big Data analytics, “Things,” Web communication, applications, network, storage, etc. It is essentially including everything connected online from hardware to software, or as we like to say, it’s an Internet of many different things. The difference ...
The 3rd International @ThingsExpo, co-located with the 16th International Cloud Expo - to be held June 9-11, 2015, at the Javits Center in New York City, NY - announces that it is now accepting Keynote Proposals. The Internet of Things (IoT) is the most profound change in personal and enterprise IT since the creation of the Worldwide Web more than 20 years ago. All major researchers estimate there will be tens of billions devices - computers, smartphones, tablets, and sensors - connected to the Internet by 2020. This number will continue to grow at a rapid pace for the next several decades.
The Internet of Things will greatly expand the opportunities for data collection and new business models driven off of that data. In her session at @ThingsExpo, Esmeralda Swartz, CMO of MetraTech, discussed how for this to be effective you not only need to have infrastructure and operational models capable of utilizing this new phenomenon, but increasingly service providers will need to convince a skeptical public to participate. Get ready to show them the money!
The 3rd International Internet of @ThingsExpo, co-located with the 16th International Cloud Expo - to be held June 9-11, 2015, at the Javits Center in New York City, NY - announces that its Call for Papers is now open. The Internet of Things (IoT) is the biggest idea since the creation of the Worldwide Web more than 20 years ago.
We are reaching the end of the beginning with WebRTC, and real systems using this technology have begun to appear. One challenge that faces every WebRTC deployment (in some form or another) is identity management. For example, if you have an existing service – possibly built on a variety of different PaaS/SaaS offerings – and you want to add real-time communications you are faced with a challenge relating to user management, authentication, authorization, and validation. Service providers will want to use their existing identities, but these will have credentials already that are (hopefully) i...
The Internet of Things is tied together with a thin strand that is known as time. Coincidentally, at the core of nearly all data analytics is a timestamp. When working with time series data there are a few core principles that everyone should consider, especially across datasets where time is the common boundary. In his session at Internet of @ThingsExpo, Jim Scott, Director of Enterprise Strategy & Architecture at MapR Technologies, discussed single-value, geo-spatial, and log time series data. By focusing on enterprise applications and the data center, he will use OpenTSDB as an example t...
SYS-CON Events announced today that Gridstore™, the leader in hyper-converged infrastructure purpose-built to optimize Microsoft workloads, will exhibit at SYS-CON's 16th International Cloud Expo®, which will take place on June 9-11, 2015, at the Javits Center in New York City, NY. Gridstore™ is the leader in hyper-converged infrastructure purpose-built for Microsoft workloads and designed to accelerate applications in virtualized environments. Gridstore’s hyper-converged infrastructure is the industry’s first all flash version of HyperConverged Appliances that include both compute and storag...
"There is a natural synchronization between the business models, the IoT is there to support ,” explained Brendan O'Brien, Co-founder and Chief Architect of Aria Systems, in this SYS-CON.tv interview at the 15th International Cloud Expo®, held Nov 4–6, 2014, at the Santa Clara Convention Center in Santa Clara, CA.
The Internet of Things promises to transform businesses (and lives), but navigating the business and technical path to success can be difficult to understand. In his session at @ThingsExpo, Sean Lorenz, Technical Product Manager for Xively at LogMeIn, demonstrated how to approach creating broadly successful connected customer solutions using real world business transformation studies including New England BioLabs and more.
WebRTC defines no default signaling protocol, causing fragmentation between WebRTC silos. SIP and XMPP provide possibilities, but come with considerable complexity and are not designed for use in a web environment. In his session at @ThingsExpo, Matthew Hodgson, technical co-founder of the Matrix.org, discussed how Matrix is a new non-profit Open Source Project that defines both a new HTTP-based standard for VoIP & IM signaling and provides reference implementations.
How do APIs and IoT relate? The answer is not as simple as merely adding an API on top of a dumb device, but rather about understanding the architectural patterns for implementing an IoT fabric. There are typically two or three trends: Exposing the device to a management framework Exposing that management framework to a business centric logic Exposing that business layer and data to end users. This last trend is the IoT stack, which involves a new shift in the separation of what stuff happens, where data lives and where the interface lies. For instance, it's a mix of architectural styles ...
An entirely new security model is needed for the Internet of Things, or is it? Can we save some old and tested controls for this new and different environment? In his session at @ThingsExpo, New York's at the Javits Center, Davi Ottenheimer, EMC Senior Director of Trust, reviewed hands-on lessons with IoT devices and reveal a new risk balance you might not expect. Davi Ottenheimer, EMC Senior Director of Trust, has more than nineteen years' experience managing global security operations and assessments, including a decade of leading incident response and digital forensics. He is co-author of t...
DevOps Summit 2015 New York, co-located with the 16th International Cloud Expo - to be held June 9-11, 2015, at the Javits Center in New York City, NY - announces that it is now accepting Keynote Proposals. The widespread success of cloud computing is driving the DevOps revolution in enterprise IT. Now as never before, development teams must communicate and collaborate in a dynamic, 24/7/365 environment. There is no time to wait for long development cycles that produce software that is obsolete at launch. DevOps may be disruptive, but it is essential.
The Internet of Things will put IT to its ultimate test by creating infinite new opportunities to digitize products and services, generate and analyze new data to improve customer satisfaction, and discover new ways to gain a competitive advantage across nearly every industry. In order to help corporate business units to capitalize on the rapidly evolving IoT opportunities, IT must stand up to a new set of challenges. In his session at @ThingsExpo, Jeff Kaplan, Managing Director of THINKstrategies, will examine why IT must finally fulfill its role in support of its SBUs or face a new round of...
Scott Jenson leads a project called The Physical Web within the Chrome team at Google. Project members are working to take the scalability and openness of the web and use it to talk to the exponentially exploding range of smart devices. Nearly every company today working on the IoT comes up with the same basic solution: use my server and you'll be fine. But if we really believe there will be trillions of these devices, that just can't scale. We need a system that is open a scalable and by using the URL as a basic building block, we open this up and get the same resilience that the web enjoys.
The security devil is always in the details of the attack: the ones you've endured, the ones you prepare yourself to fend off, and the ones that, you fear, will catch you completely unaware and defenseless. The Internet of Things (IoT) is nothing if not an endless proliferation of details. It's the vision of a world in which continuous Internet connectivity and addressability is embedded into a growing range of human artifacts, into the natural world, and even into our smartphones, appliances, and physical persons. In the IoT vision, every new "thing" - sensor, actuator, data source, data con...
Connected devices and the Internet of Things are getting significant momentum in 2014. In his session at Internet of @ThingsExpo, Jim Hunter, Chief Scientist & Technology Evangelist at Greenwave Systems, examined three key elements that together will drive mass adoption of the IoT before the end of 2015. The first element is the recent advent of robust open source protocols (like AllJoyn and WebRTC) that facilitate M2M communication. The second is broad availability of flexible, cost-effective storage designed to handle the massive surge in back-end data in a world where timely analytics is e...
The 3rd International Internet of @ThingsExpo, co-located with the 16th International Cloud Expo - to be held June 9-11, 2015, at the Javits Center in New York City, NY - announces that its Call for Papers is now open. The Internet of Things (IoT) is the biggest idea since the creation of the Worldwide Web more than 20 years ago.
Explosive growth in connected devices. Enormous amounts of data for collection and analysis. Critical use of data for split-second decision making and actionable information. All three are factors in making the Internet of Things a reality. Yet, any one factor would have an IT organization pondering its infrastructure strategy. How should your organization enhance its IT framework to enable an Internet of Things implementation? In his session at Internet of @ThingsExpo, James Kirkland, Chief Architect for the Internet of Things and Intelligent Systems at Red Hat, described how to revolutioniz...
P2P RTC will impact the landscape of communications, shifting from traditional telephony style communications models to OTT (Over-The-Top) cloud assisted & PaaS (Platform as a Service) communication services. The P2P shift will impact many areas of our lives, from mobile communication, human interactive web services, RTC and telephony infrastructure, user federation, security and privacy implications, business costs, and scalability. In his session at @ThingsExpo, Robin Raymond, Chief Architect at Hookflash, will walk through the shifting landscape of traditional telephone and voice services ...