Click here to close now.

Welcome!

Linux Containers Authors: Sanjeev Sharma, Liz McMillan, Carmen Gonzalez, Elizabeth White, Pat Romanski

Related Topics: Linux Containers

Linux Containers: Article

Sandia's CRF Team and Penguin Computing Case Study

Sandia's CRF team and Penguin Computing put their heads together and harnessed the power of several Penguin Altus Opteron server

The minimal intervention required to manage a Scyld cluster significantly enhanced our productivity but in a cost effective manner.

-Joe Oefelein, senior member, technical staff
Sandia National Laboratory's Combustion Research Facility

Even though Sandia National Laboratory's Combustion Research Facility (CRF) was doing science for the Department of Energy (DOE) with real-world pocketbook impact, they were often limited by a lack of supercomputer resources required to conduct numerical simulations and analyze data. Knowing they had a tight budget and needed massive compute power, Sandia's CRF team and Penguin Computing put their heads together and harnessed the power of several Penguin Altus Opteron servers using Sycld Software's Beowulf cluster software. This single departmental cluster gave them over five million compute hours per year and a fivefold increase in performance. The result was a dramatic increase in the amount of research they could complete. It also saved CRF $150,000 a year in administration overhead. Being brilliant scientists, they were smart enough to invest that "found money" in more compute nodes so they can crank out even more research!

The Challenge
The DOE funds Sandia to find high-efficiency, low-emission solutions to complex combustion problems. As part of this mission, the CRF conducts complex simulations and analysis of turbulent reacting flows to study interactions between fluid dynamics and combustion chemistry that affect the performance and emissions of combustion devices.

Simulations done by CRF are so complex that a typical baseline case takes one to two weeks. More sophisticated jobs need up to eight weeks of compute time. Nothing in-house was powerful enough to do that sort of work so CRF had to rely on off-site supercomputer grants to complete their research.

Because demand for supercomputing time always exceeds supply, the CRF had to compete for resources at the handful of supercomputing centers nationwide against other government facilities, academic researchers, and other applicants. In the end, CRF got only a small fraction of the total compute hours per year needed to perform the required calculations.

The Solution
Two members of CRF's technical staff, senior member Joe Oefelein and distinguished member Jackie Chen, realized high-performance technology had evolved to the point where CRF might be able to cost-effectively create a departmental scale Linux cluster that would allow them to run many calculations. With their own powerful cluster to turn to as a first line of research, precious supercomputer time could be reserved for larger simulations that required substantial system support.

"We purchased hardware from Penguin Computing [because] their system engineering appeared to be the most robust," Oefelein said. "We chose Scyld Beowulf because it is easier to use [than other options] and offered us a turn-key solution. The Scyld BeoMaster interface emulates a workstation."

Within a day of arrival, the new Penguin/Scyld system was in place. It turned out that the biggest implementation challenge to CRF was the time the staff needed to get the computer lab prepared and the power in place. Since then, the cluster has been in continuous operation. Their new departmental cluster gave them over five-million compute hours per year and a five-time increase in performance. The result was a dramatic increase in the amount of research they could complete.

The Scyld Beowulf clustering software chosen by CRF also dramatically simplifies ease of deployment and manageability. This means three CRF principal investigators can access the Scyld Beowulf cluster from their workstations just by using an informal queue to manage shared use of the cluster. A single process ID space for the entire cluster on the master node also means that the cluster seems like one computer so CRF can even scale up incremental without redesign or administrative effort. The ease of use of the cluster also saved CRF $150,000 a year in administration overhead.

"It's been clear as we've emerged from the shake-down phase that [the new server and cluster is] performing as promised. The time I need to manage the cluster is truly minimal," said Oefelein.

"Cluster technology provides an affordable system with fairly significant capability dedicated to our problems," concluded Oefelein. "We expect to buy more clusters. Other groups at Sandia also recognize this and are currently evaluating Linux clusters."

The Installation
Penguin Computing created a 72-node Linux cluster with 144 processors for CRF. The Altus master node, back-up battery system, GigaBit Ethernet switch, and Infiniband are on one rack. Two other racks house Altus compute nodes with motherboards and communication hardware.

CRF invested in Infiniband because they knew that they would likely add another 72 nodes later if the strong performance they expected materialized. The experiment results were even better than expected so CRF is adding more nodes.

Advice from Sandia

  1. Do your homework as to how your specific applications perform on the clusters being evaluated.
  2. Reliability is key. Cluster technology isn't perfect and depending on which products you choose, you may need supplemental cluster management expertise in-house.
  3. Buy in logical increments. Anticipate what you'll need in the future especially if you work with uncertain budgets.
About Sandia National Laboratory Combustion Research Facility
(www.ca.sandia.gov/CRF)
The CRF is an internationally recognized Department of Energy Office of Science user facility. The CRF is home to about 100 scientists, engineers, and technologists who conduct basic and applied research aimed at improving our nation's ability to use and control combustion processes. The need for a thorough and basic understanding of combustion and combustion-related processes lies at the heart of CRF research.

The CRF is an Office of Science user facility for broad-based research in energy science and technology. Using the facility's unique laser diagnostic capabilities, staff researchers and visiting investigators explore fundamental chemical reactivity and dynamics problems, as well as conduct applied studies that support industry's needs in areas such as engines and materials processing.

More Stories By Joseph C. Oefelein

Joseph C. Oefelein is a senior member of the technical staff at the Sandia National Laboratories Combustion Research Facility. He received a doctorate in mechanical engineering from Pennsylvania State University in May 1997, an MS in mechanical engineering from Penn State in December 1992, and a BS in mechanical engineering (with highest honors) from Rutgers University in May 1989.

Comments (2) View Comments

Share your thoughts on this story.

Add your comment
You must be signed in to add a comment. Sign-in | Register

In accordance with our Comment Policy, we encourage comments that are on topic, relevant and to-the-point. We will remove comments that include profanity, personal attacks, racial slurs, threats of violence, or other inappropriate material that violates our Terms and Conditions, and will block users who make repeated violations. We ask all readers to expect diversity of opinion and to treat one another with dignity and respect.


Most Recent Comments
Enterprise Open Source Magazine News Desk 11/10/05 11:34:01 AM EST

LinuxWorld Feature: Sandia's CRF Team and Penguin Computing Case Study. The DOE funds Sandia to find high-efficiency, low-emission solutions to complex combustion problems. As part of this mission, the CRF conducts complex simulations and analysis of turbulent reacting flows to study interactions between fluid dynamics and combustion chemistry that affect the performance and emissions of combustion devices.

LinuxWorld News Desk 11/10/05 11:00:48 AM EST

LinuxWorld Feature: Sandia's CRF Team and Penguin Computing Case Study. The DOE funds Sandia to find high-efficiency, low-emission solutions to complex combustion problems. As part of this mission, the CRF conducts complex simulations and analysis of turbulent reacting flows to study interactions between fluid dynamics and combustion chemistry that affect the performance and emissions of combustion devices.

@ThingsExpo Stories
Internet of Things (IoT) will be a hybrid ecosystem of diverse devices and sensors collaborating with operational and enterprise systems to create the next big application. In their session at @ThingsExpo, Bramh Gupta, founder and CEO of robomq.io, and Fred Yatzeck, principal architect leading product development at robomq.io, discussed how choosing the right middleware and integration strategy from the get-go will enable IoT solution developers to adapt and grow with the industry, while at the same time reduce Time to Market (TTM) by using plug and play capabilities offered by a robust IoT ...
17th Cloud Expo, taking place Nov 3-5, 2015, at the Santa Clara Convention Center in Santa Clara, CA, will feature technical sessions from a rock star conference faculty and the leading industry players in the world. Cloud computing is now being embraced by a majority of enterprises of all sizes. Yesterday's debate about public vs. private has transformed into the reality of hybrid cloud: a recent survey shows that 74% of enterprises have a hybrid cloud strategy. Meanwhile, 94% of enterprises are using some form of XaaS – software, platform, and infrastructure as a service.
SYS-CON Events announced today that Secure Infrastructure & Services will exhibit at SYS-CON's 17th International Cloud Expo®, which will take place on November 3–5, 2015, at the Santa Clara Convention Center in Santa Clara, CA. Secure Infrastructure & Services (SIAS) is a managed services provider of cloud computing solutions for the IBM Power Systems market. The company helps mid-market firms built on IBM hardware platforms to deploy new levels of reliable and cost-effective computing and high availability solutions, leveraging the cloud and the benefits of Infrastructure-as-a-Service (IaaS...
The 17th International Cloud Expo has announced that its Call for Papers is open. 17th International Cloud Expo, to be held November 3-5, 2015, at the Santa Clara Convention Center in Santa Clara, CA, brings together Cloud Computing, APM, APIs, Microservices, Security, Big Data, Internet of Things, DevOps and WebRTC to one location. With cloud computing driving a higher percentage of enterprise IT budgets every year, it becomes increasingly important to plant your flag in this fast-expanding business opportunity. Submit your speaking proposal today!
"We have a tagline - "Power in the API Economy." What that means is everything that is built in applications and connected applications is done through APIs," explained Roberto Medrano, Executive Vice President at Akana, in this SYS-CON.tv interview at 16th Cloud Expo, held June 9-11, 2015, at the Javits Center in New York City.
The 5th International DevOps Summit, co-located with 17th International Cloud Expo – being held November 3-5, 2015, at the Santa Clara Convention Center in Santa Clara, CA – announces that its Call for Papers is open. Born out of proven success in agile development, cloud computing, and process automation, DevOps is a macro trend you cannot afford to miss. From showcase success stories from early adopters and web-scale businesses, DevOps is expanding to organizations of all sizes, including the world's largest enterprises – and delivering real results. Among the proven benefits, DevOps is corr...
The basic integration architecture, as defined by ESBs, hasn’t changed for more than a decade. Most cloud integration providers still rely on an ESB architecture and their proprietary connectors. As a result, enterprise integration projects suffer from constraints of availability and reliability of these connectors that are not re-usable across other integration vendors. However, the rapid adoption of APIs and almost ubiquitous availability of APIs amongst most SaaS and Cloud applications are rapidly redefining traditional integration approaches and their reliance on proprietary connectors. ...
The Internet of Things is not only adding billions of sensors and billions of terabytes to the Internet. It is also forcing a fundamental change in the way we envision Information Technology. For the first time, more data is being created by devices at the edge of the Internet rather than from centralized systems. What does this mean for today's IT professional? In this Power Panel at @ThingsExpo, moderated by Conference Chair Roger Strukhoff, panelists addressed this very serious issue of profound change in the industry.
Internet of Things is moving from being a hype to a reality. Experts estimate that internet connected cars will grow to 152 million, while over 100 million internet connected wireless light bulbs and lamps will be operational by 2020. These and many other intriguing statistics highlight the importance of Internet powered devices and how market penetration is going to multiply many times over in the next few years.
Today air travel is a minefield of delays, hassles and customer disappointment. Airlines struggle to revitalize the experience. GE and M2Mi will demonstrate practical examples of how IoT solutions are helping airlines bring back personalization, reduce trip time and improve reliability. In their session at @ThingsExpo, Shyam Varan Nath, Principal Architect with GE, and Dr. Sarah Cooper, M2Mi’s VP Business Development and Engineering, will explore the IoT cloud-based platform technologies driving this change including privacy controls, data transparency and integration of real time context wi...
WebRTC converts the entire network into a ubiquitous communications cloud thereby connecting anytime, anywhere through any point. In his session at WebRTC Summit,, Mark Castleman, EIR at Bell Labs and Head of Future X Labs, will discuss how the transformational nature of communications is achieved through the democratizing force of WebRTC. WebRTC is doing for voice what HTML did for web content.
To many people, IoT is a buzzword whose value is not understood. Many people think IoT is all about wearables and home automation. In his session at @ThingsExpo, Mike Kavis, Vice President & Principal Cloud Architect at Cloud Technology Partners, discussed some incredible game-changing use cases and how they are transforming industries like agriculture, manufacturing, health care, and smart cities. He will discuss cool technologies like smart dust, robotics, smart labels, and much more. Prepare to be blown away with a glimpse of the future.
Explosive growth in connected devices. Enormous amounts of data for collection and analysis. Critical use of data for split-second decision making and actionable information. All three are factors in making the Internet of Things a reality. Yet, any one factor would have an IT organization pondering its infrastructure strategy. How should your organization enhance its IT framework to enable an Internet of Things implementation? In his session at @ThingsExpo, James Kirkland, Red Hat's Chief Architect for the Internet of Things and Intelligent Systems, described how to revolutionize your archit...
It is one thing to build single industrial IoT applications, but what will it take to build the Smart Cities and truly society-changing applications of the future? The technology won’t be the problem, it will be the number of parties that need to work together and be aligned in their motivation to succeed. In his session at @ThingsExpo, Jason Mondanaro, Director, Product Management at Metanga, discussed how you can plan to cooperate, partner, and form lasting all-star teams to change the world and it starts with business models and monetization strategies.
SYS-CON Events announced today that BMC will exhibit at SYS-CON's 16th International Cloud Expo®, which will take place on June 9-11, 2015, at the Javits Center in New York City, NY. BMC delivers software solutions that help IT transform digital enterprises for the ultimate competitive business advantage. BMC has worked with thousands of leading companies to create and deliver powerful IT management services. From mainframe to cloud to mobile, BMC pairs high-speed digital innovation with robust IT industrialization – allowing customers to provide amazing user experiences with optimized IT per...
There will be 150 billion connected devices by 2020. New digital businesses have already disrupted value chains across every industry. APIs are at the center of the digital business. You need to understand what assets you have that can be exposed digitally, what their digital value chain is, and how to create an effective business model around that value chain to compete in this economy. No enterprise can be complacent and not engage in the digital economy. Learn how to be the disruptor and not the disruptee.
The Internet of Things is not only adding billions of sensors and billions of terabytes to the Internet. It is also forcing a fundamental change in the way we envision Information Technology. For the first time, more data is being created by devices at the edge of the Internet rather than from centralized systems. What does this mean for today's IT professional? In this Power Panel at @ThingsExpo, moderated by Conference Chair Roger Strukhoff, panelists will addresses this very serious issue of profound change in the industry.
Business as usual for IT is evolving into a "Make or Buy" decision on a service-by-service conversation with input from the LOBs. How does your organization move forward with cloud? In his general session at 16th Cloud Expo, Paul Maravei, Regional Sales Manager, Hybrid Cloud and Managed Services at Cisco, discusses how Cisco and its partners offer a market-leading portfolio and ecosystem of cloud infrastructure and application services that allow you to uniquely and securely combine cloud business applications and services across multiple cloud delivery models.
In his General Session at 16th Cloud Expo, David Shacochis, host of The Hybrid IT Files podcast and Vice President at CenturyLink, investigated three key trends of the “gigabit economy" though the story of a Fortune 500 communications company in transformation. Narrating how multi-modal hybrid IT, service automation, and agile delivery all intersect, he will cover the role of storytelling and empathy in achieving strategic alignment between the enterprise and its information technology.
Buzzword alert: Microservices and IoT at a DevOps conference? What could possibly go wrong? In this Power Panel at DevOps Summit, moderated by Jason Bloomberg, the leading expert on architecting agility for the enterprise and president of Intellyx, panelists peeled away the buzz and discuss the important architectural principles behind implementing IoT solutions for the enterprise. As remote IoT devices and sensors become increasingly intelligent, they become part of our distributed cloud environment, and we must architect and code accordingly. At the very least, you'll have no problem fillin...