Welcome!

Linux Containers Authors: Carmen Gonzalez, Mehdi Daoudi, Mano Marks, Pat Romanski, Elizabeth White

Related Topics: Linux Containers

Linux Containers: Article

Sandia's CRF Team and Penguin Computing Case Study

Sandia's CRF team and Penguin Computing put their heads together and harnessed the power of several Penguin Altus Opteron server

The minimal intervention required to manage a Scyld cluster significantly enhanced our productivity but in a cost effective manner.

-Joe Oefelein, senior member, technical staff
Sandia National Laboratory's Combustion Research Facility

Even though Sandia National Laboratory's Combustion Research Facility (CRF) was doing science for the Department of Energy (DOE) with real-world pocketbook impact, they were often limited by a lack of supercomputer resources required to conduct numerical simulations and analyze data. Knowing they had a tight budget and needed massive compute power, Sandia's CRF team and Penguin Computing put their heads together and harnessed the power of several Penguin Altus Opteron servers using Sycld Software's Beowulf cluster software. This single departmental cluster gave them over five million compute hours per year and a fivefold increase in performance. The result was a dramatic increase in the amount of research they could complete. It also saved CRF $150,000 a year in administration overhead. Being brilliant scientists, they were smart enough to invest that "found money" in more compute nodes so they can crank out even more research!

The Challenge
The DOE funds Sandia to find high-efficiency, low-emission solutions to complex combustion problems. As part of this mission, the CRF conducts complex simulations and analysis of turbulent reacting flows to study interactions between fluid dynamics and combustion chemistry that affect the performance and emissions of combustion devices.

Simulations done by CRF are so complex that a typical baseline case takes one to two weeks. More sophisticated jobs need up to eight weeks of compute time. Nothing in-house was powerful enough to do that sort of work so CRF had to rely on off-site supercomputer grants to complete their research.

Because demand for supercomputing time always exceeds supply, the CRF had to compete for resources at the handful of supercomputing centers nationwide against other government facilities, academic researchers, and other applicants. In the end, CRF got only a small fraction of the total compute hours per year needed to perform the required calculations.

The Solution
Two members of CRF's technical staff, senior member Joe Oefelein and distinguished member Jackie Chen, realized high-performance technology had evolved to the point where CRF might be able to cost-effectively create a departmental scale Linux cluster that would allow them to run many calculations. With their own powerful cluster to turn to as a first line of research, precious supercomputer time could be reserved for larger simulations that required substantial system support.

"We purchased hardware from Penguin Computing [because] their system engineering appeared to be the most robust," Oefelein said. "We chose Scyld Beowulf because it is easier to use [than other options] and offered us a turn-key solution. The Scyld BeoMaster interface emulates a workstation."

Within a day of arrival, the new Penguin/Scyld system was in place. It turned out that the biggest implementation challenge to CRF was the time the staff needed to get the computer lab prepared and the power in place. Since then, the cluster has been in continuous operation. Their new departmental cluster gave them over five-million compute hours per year and a five-time increase in performance. The result was a dramatic increase in the amount of research they could complete.

The Scyld Beowulf clustering software chosen by CRF also dramatically simplifies ease of deployment and manageability. This means three CRF principal investigators can access the Scyld Beowulf cluster from their workstations just by using an informal queue to manage shared use of the cluster. A single process ID space for the entire cluster on the master node also means that the cluster seems like one computer so CRF can even scale up incremental without redesign or administrative effort. The ease of use of the cluster also saved CRF $150,000 a year in administration overhead.

"It's been clear as we've emerged from the shake-down phase that [the new server and cluster is] performing as promised. The time I need to manage the cluster is truly minimal," said Oefelein.

"Cluster technology provides an affordable system with fairly significant capability dedicated to our problems," concluded Oefelein. "We expect to buy more clusters. Other groups at Sandia also recognize this and are currently evaluating Linux clusters."

The Installation
Penguin Computing created a 72-node Linux cluster with 144 processors for CRF. The Altus master node, back-up battery system, GigaBit Ethernet switch, and Infiniband are on one rack. Two other racks house Altus compute nodes with motherboards and communication hardware.

CRF invested in Infiniband because they knew that they would likely add another 72 nodes later if the strong performance they expected materialized. The experiment results were even better than expected so CRF is adding more nodes.

Advice from Sandia

  1. Do your homework as to how your specific applications perform on the clusters being evaluated.
  2. Reliability is key. Cluster technology isn't perfect and depending on which products you choose, you may need supplemental cluster management expertise in-house.
  3. Buy in logical increments. Anticipate what you'll need in the future especially if you work with uncertain budgets.
About Sandia National Laboratory Combustion Research Facility
(www.ca.sandia.gov/CRF)
The CRF is an internationally recognized Department of Energy Office of Science user facility. The CRF is home to about 100 scientists, engineers, and technologists who conduct basic and applied research aimed at improving our nation's ability to use and control combustion processes. The need for a thorough and basic understanding of combustion and combustion-related processes lies at the heart of CRF research.

The CRF is an Office of Science user facility for broad-based research in energy science and technology. Using the facility's unique laser diagnostic capabilities, staff researchers and visiting investigators explore fundamental chemical reactivity and dynamics problems, as well as conduct applied studies that support industry's needs in areas such as engines and materials processing.

More Stories By Joseph C. Oefelein

Joseph C. Oefelein is a senior member of the technical staff at the Sandia National Laboratories Combustion Research Facility. He received a doctorate in mechanical engineering from Pennsylvania State University in May 1997, an MS in mechanical engineering from Penn State in December 1992, and a BS in mechanical engineering (with highest honors) from Rutgers University in May 1989.

Comments (2) View Comments

Share your thoughts on this story.

Add your comment
You must be signed in to add a comment. Sign-in | Register

In accordance with our Comment Policy, we encourage comments that are on topic, relevant and to-the-point. We will remove comments that include profanity, personal attacks, racial slurs, threats of violence, or other inappropriate material that violates our Terms and Conditions, and will block users who make repeated violations. We ask all readers to expect diversity of opinion and to treat one another with dignity and respect.


Most Recent Comments
Enterprise Open Source Magazine News Desk 11/10/05 11:34:01 AM EST

LinuxWorld Feature: Sandia's CRF Team and Penguin Computing Case Study. The DOE funds Sandia to find high-efficiency, low-emission solutions to complex combustion problems. As part of this mission, the CRF conducts complex simulations and analysis of turbulent reacting flows to study interactions between fluid dynamics and combustion chemistry that affect the performance and emissions of combustion devices.

LinuxWorld News Desk 11/10/05 11:00:48 AM EST

LinuxWorld Feature: Sandia's CRF Team and Penguin Computing Case Study. The DOE funds Sandia to find high-efficiency, low-emission solutions to complex combustion problems. As part of this mission, the CRF conducts complex simulations and analysis of turbulent reacting flows to study interactions between fluid dynamics and combustion chemistry that affect the performance and emissions of combustion devices.

@ThingsExpo Stories
Internet of @ThingsExpo, taking place June 6-8, 2017 at the Javits Center in New York City, New York, is co-located with the 20th International Cloud Expo and will feature technical sessions from a rock star conference faculty and the leading industry players in the world. @ThingsExpo New York Call for Papers is now open.
WebRTC sits at the intersection between VoIP and the Web. As such, it poses some interesting challenges for those developing services on top of it, but also for those who need to test and monitor these services. In his session at WebRTC Summit, Tsahi Levent-Levi, co-founder of testRTC, reviewed the various challenges posed by WebRTC when it comes to testing and monitoring and on ways to overcome them.
DevOps is being widely accepted (if not fully adopted) as essential in enterprise IT. But as Enterprise DevOps gains maturity, expands scope, and increases velocity, the need for data-driven decisions across teams becomes more acute. DevOps teams in any modern business must wrangle the ‘digital exhaust’ from the delivery toolchain, "pervasive" and "cognitive" computing, APIs and services, mobile devices and applications, the Internet of Things, and now even blockchain. In this power panel at @...
WebRTC services have already permeated corporate communications in the form of videoconferencing solutions. However, WebRTC has the potential of going beyond and catalyzing a new class of services providing more than calls with capabilities such as mass-scale real-time media broadcasting, enriched and augmented video, person-to-machine and machine-to-machine communications. In his session at @ThingsExpo, Luis Lopez, CEO of Kurento, introduced the technologies required for implementing these idea...
Buzzword alert: Microservices and IoT at a DevOps conference? What could possibly go wrong? In this Power Panel at DevOps Summit, moderated by Jason Bloomberg, the leading expert on architecting agility for the enterprise and president of Intellyx, panelists peeled away the buzz and discuss the important architectural principles behind implementing IoT solutions for the enterprise. As remote IoT devices and sensors become increasingly intelligent, they become part of our distributed cloud enviro...
"A lot of times people will come to us and have a very diverse set of requirements or very customized need and we'll help them to implement it in a fashion that you can't just buy off of the shelf," explained Nick Rose, CTO of Enzu, in this SYS-CON.tv interview at 18th Cloud Expo, held June 7-9, 2016, at the Javits Center in New York City, NY.
The WebRTC Summit New York, to be held June 6-8, 2017, at the Javits Center in New York City, NY, announces that its Call for Papers is now open. Topics include all aspects of improving IT delivery by eliminating waste through automated business models leveraging cloud technologies. WebRTC Summit is co-located with 20th International Cloud Expo and @ThingsExpo. WebRTC is the future of browser-to-browser communications, and continues to make inroads into the traditional, difficult, plug-in web co...
In his keynote at @ThingsExpo, Chris Matthieu, Director of IoT Engineering at Citrix and co-founder and CTO of Octoblu, focused on building an IoT platform and company. He provided a behind-the-scenes look at Octoblu’s platform, business, and pivots along the way (including the Citrix acquisition of Octoblu).
For basic one-to-one voice or video calling solutions, WebRTC has proven to be a very powerful technology. Although WebRTC’s core functionality is to provide secure, real-time p2p media streaming, leveraging native platform features and server-side components brings up new communication capabilities for web and native mobile applications, allowing for advanced multi-user use cases such as video broadcasting, conferencing, and media recording.
Web Real-Time Communication APIs have quickly revolutionized what browsers are capable of. In addition to video and audio streams, we can now bi-directionally send arbitrary data over WebRTC's PeerConnection Data Channels. With the advent of Progressive Web Apps and new hardware APIs such as WebBluetooh and WebUSB, we can finally enable users to stitch together the Internet of Things directly from their browsers while communicating privately and securely in a decentralized way.
WebRTC is about the data channel as much as about video and audio conferencing. However, basically all commercial WebRTC applications have been built with a focus on audio and video. The handling of “data” has been limited to text chat and file download – all other data sharing seems to end with screensharing. What is holding back a more intensive use of peer-to-peer data? In her session at @ThingsExpo, Dr Silvia Pfeiffer, WebRTC Applications Team Lead at National ICT Australia, looked at differ...
The security needs of IoT environments require a strong, proven approach to maintain security, trust and privacy in their ecosystem. Assurance and protection of device identity, secure data encryption and authentication are the key security challenges organizations are trying to address when integrating IoT devices. This holds true for IoT applications in a wide range of industries, for example, healthcare, consumer devices, and manufacturing. In his session at @ThingsExpo, Lancen LaChance, vic...
With all the incredible momentum behind the Internet of Things (IoT) industry, it is easy to forget that not a single CEO wakes up and wonders if “my IoT is broken.” What they wonder is if they are making the right decisions to do all they can to increase revenue, decrease costs, and improve customer experience – effectively the same challenges they have always had in growing their business. The exciting thing about the IoT industry is now these decisions can be better, faster, and smarter. Now ...
Who are you? How do you introduce yourself? Do you use a name, or do you greet a friend by the last four digits of his social security number? Assuming you don’t, why are we content to associate our identity with 10 random digits assigned by our phone company? Identity is an issue that affects everyone, but as individuals we don’t spend a lot of time thinking about it. In his session at @ThingsExpo, Ben Klang, Founder & President of Mojo Lingo, discussed the impact of technology on identity. Sho...
Fact is, enterprises have significant legacy voice infrastructure that’s costly to replace with pure IP solutions. How can we bring this analog infrastructure into our shiny new cloud applications? There are proven methods to bind both legacy voice applications and traditional PSTN audio into cloud-based applications and services at a carrier scale. Some of the most successful implementations leverage WebRTC, WebSockets, SIP and other open source technologies. In his session at @ThingsExpo, Da...
A critical component of any IoT project is what to do with all the data being generated. This data needs to be captured, processed, structured, and stored in a way to facilitate different kinds of queries. Traditional data warehouse and analytical systems are mature technologies that can be used to handle certain kinds of queries, but they are not always well suited to many problems, particularly when there is a need for real-time insights.
You think you know what’s in your data. But do you? Most organizations are now aware of the business intelligence represented by their data. Data science stands to take this to a level you never thought of – literally. The techniques of data science, when used with the capabilities of Big Data technologies, can make connections you had not yet imagined, helping you discover new insights and ask new questions of your data. In his session at @ThingsExpo, Sarbjit Sarkaria, data science team lead ...
WebRTC has had a real tough three or four years, and so have those working with it. Only a few short years ago, the development world were excited about WebRTC and proclaiming how awesome it was. You might have played with the technology a couple of years ago, only to find the extra infrastructure requirements were painful to implement and poorly documented. This probably left a bitter taste in your mouth, especially when things went wrong.
WebRTC is bringing significant change to the communications landscape that will bridge the worlds of web and telephony, making the Internet the new standard for communications. Cloud9 took the road less traveled and used WebRTC to create a downloadable enterprise-grade communications platform that is changing the communication dynamic in the financial sector. In his session at @ThingsExpo, Leo Papadopoulos, CTO of Cloud9, discussed the importance of WebRTC and how it enables companies to focus o...
Providing secure, mobile access to sensitive data sets is a critical element in realizing the full potential of cloud computing. However, large data caches remain inaccessible to edge devices for reasons of security, size, format or limited viewing capabilities. Medical imaging, computer aided design and seismic interpretation are just a few examples of industries facing this challenge. Rather than fighting for incremental gains by pulling these datasets to edge devices, we need to embrace the i...