
How Leading Experts Leverage HPC to Solve Complex Problems
Join us for an engaging panel discussion on the power of high performance computing (HPC) and the software used by leading experts in the field. Our panelists will share their experiences and insights on how they utilize HPC to solve complex problems in various industries. They will discuss the benefits and challenges of using HPC, as well as the different software tools they use to optimize their computing workflows. Whether you are a seasoned HPC user or just getting started, this webinar will provide valuable insights on how to harness the power of HPC to accelerate your research and innovation.
Webinar Synopsis:
Speakers:
-
Zane Hamilton, Vice President - Sales Engineering, CIQ
-
Rose Stein: Solutions Engineer, CIQ
-
Gary Jung: HPC General Manager, UC Berkeley
-
Brain Phan, Solutions Architect, CIQ
-
Brock Taylor, VP of HPC, CIQ
-
Forrest Burt, High Performance Computing Systems Engineer, CIQ
Note: This transcript was created using speech recognition software. While it has been reviewed by human transcribers, it may contain errors.
Full Webinar Transcript:
Zane Hamilton:
Good morning, good afternoon, and good evening wherever you are. Thank you for joining CIQ. We're focused on powering the next generation of software infrastructure, leveraging the capabilities of Cloud, Hyperscale, and HPC. From research to the enterprise, our customers rely on us for the ultimate Rocky Linux, Warewulf, and Apptainer support escalation. We provide deep development capabilities and solutions, all delivered in the collaborative spirit of open source.
Hello everyone, and welcome to another CIQ webinar. Hello, Rose.
Rose Stein:
Hello.
Zane Hamilton:
Welcome. How are you?
Rose Stein:
Great, thank you.
Zane Hamilton:
What are we talking about this week?
Rose Stein:
Ooh, so this is actually a really interesting topic and conversation. We're going to be talking about HPC, but in a broader context, like how are experts in the field actually utilizing HPC resources to solve the complex problems that we're facing?
Zane Hamilton:
Very nice. Who do we have this week? Let's bring in our panel. Gary, Brian, Brock, Forrest. Forrest, I forced you to come. I expected you to be here. How's it going, everyone?
Gary Jung:
Good.
Zane Hamilton:
Good to see everybody. We'll go around, and do introductions real quick. As you know, Rose and myself, we'll start with Gary.
Gary Jung:
Hi, my name is Gary Jung. I manage the institutional HPC for Lawrence Berkeley National Laboratory, and I also manage the HPC program down at UC Berkeley.
Zane Hamilton:
Thank you, Gary. Brian?
Brain Phan:
Everyone. Brian Phan here. I'm a Solutions Architect here at CIQ. My background in HPC administration architecture, and I've experienced supporting HPC users across various domains of science. I'm excited to be back and discussing this topic today.
Zane Hamilton:
Thank you very much, Brian. Brock, I hope you're having nice weather. It's getting to be that time of year, right?
Brock Taylor:
It is. It's definitely very nice. Has not hit that boiling hot phase yet, you can go outside and enjoy the weather. I'm Brock, a long time HPCer with a deep background, a couple of decades' worth in HPC, primarily in clustering.
Zane Hamilton:
Thank you, Brock. And Forrest.
Forrest Burt:
Hi, can everyone hear me all right?
Zane Hamilton:
We got you.
Forrest Burt:
Good morning everyone. My name is Forrest Burt. I'm an HPC Systems Engineer here at CIQ. My background is in the academic and national lab HPC space. I was previously a student system administrator at a university here in the States, but have since started working at CIQ, on the systems engineering side of things. Very excited to be on and excited to discuss HPC.
Importance of HPC [7:51]
Zane Hamilton:
Thank you, Forrest. Rose, I'm going to let Brock kind of set the stage here and let him tell us a brief overview of the importance of HPC if you don't mind.
Brock Taylor:
Absolutely. It's actually underpinning so many things that we interact with on a daily basis. I think sometimes it's overlooked because when you say HPC, a lot of people will think of the largest systems on the planet. But the technology's actually used across many different industries. And again, in the cars we drive, definitely if you get on an airplane, but just think about as you are driving and seeing road construction crews, the machines they're using, farmers, the tractors that they're driving. Even things that are in our local hardware stores that are helping build the decks for our houses or the houses themselves. High Performance Computing has been used to help create a lot of those products. Over the past 20, or 30 years, the advances in HPC have actually enabled many different improvements.
One example that I would bring up is weather forecasting. If you go back 25 years, the capabilities to project out one week or two weeks, and be semi-accurate was more consulting almanac and trends and things that you could actually try and predict with. But now, because computing has become powerful, simulations are able to actually predict much more accurately where/how things are going to occur and weather patterns. Again, it's still very, very difficult. There's a lot of research going on in weather forecasting, but if you think about how accurate we can get in predicting where a hurricane is going to hit, and what that means to be able to warn people, to evacuate people, to prepare, and actually save lives in the process.
A lot of that is because of improvements in high performance computing and its usage, across different industries. The fun things in life, the new Avatar movie, would not be possible without high performance computing. The first Avatar movie would not be possible, but the technologies that have gone into making that film, and who is involved, they've won awards for the technological improvements. You can see what they're able to do. And even simulating and actually rendering water is a highly difficult thing to do. But making it so realistic, again, is high performance computing. It's very broad. It's hours and hours in any one subject. But there are so many different things that you can use high performance computing to do. And I think what we'll see is even more industries, more everyday products that are going to be able to tap into very powerful technologies.
What is HPC Computing? [11:23]
Rose Stein:
I'm sorry, Brock, I'm just going to have to ask you if you could please backtrack a little bit. What is high performance computing? And if you're not high performance computing, what are you?
Brock Taylor:
That's a great and loaded question, and I'll start, and probably, the others will want to chime in as well. It's an argument to describe or concretely define high performance computing, what it is. To some, it is only the very high-end, the most powerful supercomputers on the planet. That's what high performance computing is. Don't talk to me about anything else. If you think of high performance and computing, you can argue that our phones are, in fact, high performance computing machines, and in some ways, they definitely are. They have the power of supercomputers from 20 years ago. To try and generalize it is taking a problem in science that you can describe by functions and formulas, and converting it from an analog world to a digital world, which is tough.
And trying to solve a very complex problem that involves a lot of, I'll call it, moving parts. Basically, a lot of science says that, if you do it by hand, it takes you a very long time. If you do it by a single computer, it will take you a lot less time, but you're still very much limited to high performance computing, and clustering, in particular, helps spread out computation to make it go faster, but at the cost of complexity and difficulty. That's where CIQ and the communities that we are engaged with come in, helping pull all these technologies together so that more people can actually use them easily. Again, a scientist that is researching bone growth does not have to learn how to build a cluster or learn how different pieces fit together. They're focusing on improving their science.
Types of Software Used on HPC Systems [14:01]
Zane Hamilton:
You mentioned earlier that you talked about rendering water, and I heard that they had to go create software to do that. I think when we look at HPC, we think about clusters, and we think about big machines. But I think a big part of this that you touched on is the software itself. I'll go to Gary next. What are the different types of software? I know that everything about it is very specific for what you're trying to accomplish or what you're trying to do, but tell me about the different types of software that people run on an HPC system.
Gary Jung:
Well, you mean like in terms of applications, maybe? Is that what you're going for?
Zane Hamilton:
Absolutely. Whenever a scientist is actually doing research, I mean, it depends on obviously the discipline of what you're trying to do. Or if you are an engineer and you're trying to render something, there are a lot of different pieces of software that people are using out there to do all kinds of different things.
Gary Jung:
Maybe I'm not the best person to answer this question, but people run a lot of different software. A lot of it is written by scientists. I'm talking in a researcher context, but a lot of times, people are trying to model something which only they know about. They'll write the software for that. Then maybe it becomes open source, and other people doing that type of research will use that software. There's software like the VASP, which a lot of material science people use for electronic structure calculations at the atomic level of materials, that can calculate things like the phases and band gaps, stuff like that.
There's software for doing imaging. A lot of people use clusters for imaging and imaging reconstruction. There's a lot of software written for that. There's software for every different type of domain of science. Really, the thing that distinguishes the software is that it can run on a large system, and it can either run multiple instances of it, or it can split up a problem. It can run on a lot of systems. That's the way I would characterize it.
HPC Today For Customers [16:39]
Zane Hamilton:
Thank you, Gary. Brian, and Forrest, I know we throw a lot at you guys from a software perspective, and talking to customers and going to prove that software works in an HPC environment, or at least show kind of what that looks like. What are you guys running across from a software perspective at HPC today, from a customer perspective? Start with Brian. Oh, Forrest came off mute first, but Brian's off now. Go ahead, Brian.
Brain Phan:
From the CFD perspective, at the end of the day, when you're running simulations, you set up a simulation problem, and you're basically setting up a system of differential equations. You're basically trying to find the solution to that. With CFD software, they're basically writing solvers to solve these differential equations. Different software vendors basically add their secret sauce to these solvers to make it go faster. Depending on your needs, you might go with one vendor versus another. That's from my experience in automotive and aerospace.
Zane Hamilton:
Thank you, Brian. Forrest?
Forrest Burt:
At the moment, the stuff I've seen most is the academic and AI side of HPC. From the academic side of things, a lot of what people are doing and using is the same type of stuff they've been using for a while. Molecular dynamics work within GROWMACS, and LAMMPS, that type of stuff is alive and well. The usual licensed use cases are Mathematica, and Matlab. I haven't heard of Mathematica for a bit, but Matlab, that type of thing, is still of interest to people. More and more, there's interest in dedicated data science platforms that give a whole all in one interface for doing a lot of different data analytics, machine learning, and those types of tasks. I mean, different things, QMCPACK, Quantum ESPRESSO, a lot of that stuff is pretty standard out of that.
On the AI side of things, we're really seeing a lot of interesting stuff going on. A lot of the work in that field is centered around a few major programming languages and a few major frameworks for those programming languages. We're kind of seeing, sort of like what Brian alluded to with the CFD industry, and we're seeing in AI these large scale frameworks, these large scale pre-trained models, that type of thing that is out there at the moment. Being leveraged widely by people and being fine-tuned by researchers and companies and places like that. There's a very strong culture at the moment of using what's out there. Because so much work has gone into developing these AI frameworks like JAX, TensorFlow, PyTorch, that type of thing.
There's so much that they're already capable of. Additionally, for example, with the GPT space, there's so much pre-trained material out there that saves so much time and effort for people that want to tinker with it that we're seeing a lot of use of these pre-trained models and people taking those. I think Brian's words are adding their own secret sauce to it. Then being able to go and have something more customized for their own purposes. All different use cases are coming out of my side of things. It's pretty interesting.
Zane Hamilton:
Thank you, Forrest.
Is Everything High Performance Computing? [20:02]
Rose Stein:
Brock, I'm going to, it's possible that you did not answer this part on purpose, but I would love to know if it's not high performance computing; what is it? Is there a name for it, or is it all computing? I mean, you mentioned the phone. Which, I mean, 40 years ago, this would've been like, oh my God, it's the high performance, you couldn't even imagine. I couldn't even imagine it. Is everything high performance computing, or is there something else that is happening?
Brock Taylor:
There are things that I would not call high performance computing. Not to disparage that, but take a web server, for instance, that's just feeding us the pages we're looking at. That's not really high performance computing. It's not computationally intensive work. But it doesn't mean that it's not connected to some elements of high performance computing. It's a tough question, and people live in different camps, but if you think of computation, mathematical computation, number crunching, if you will. Almost anything that is doing that can be classified as high performance computing workloads or applications. Again, in today's world, with the capabilities that we have, it's not just a single domain or application. It's many different high performance computing applications working together to solve a very complex problem that you wouldn't have been able to solve 10, 20, or 30 years ago.
Now problems that were being solved with high performance computing, but as an example, if you take a bulldozer, there are actually a lot of different systems that go into making a bulldozer. There's the actual structure, how much earth or material it can move, and how much force it has. There's the hydraulics that lift the bucket and things of that nature. All of those elements are simulated before they ever build that first bulldozer. Same thing with the car. If you're as old as I am, you remember crash test dummies were in all kinds of commercials 30 years ago. They still exist, crash test dummies still exist, but they're not used nearly as much, and physical car crashes are done pretty much to prove a design operates the way it should.
High performance computing is the reason simulation is the reason that you don't need to crash cars as much, or they're really proving that the design of the card does actually do what it's supposed to. I am kind of dancing around concretely answering it because it can cause an allergic reaction to some people. Again, have the debate, is AI high performance computing? I am in the camp that it absolutely is because it's computationally intensive. It is different from traditional computational models that are doing; sorry to get a little technical floating point operations, but it's more integer work. It exercises silicon in a different way. The algorithms are different, but they are computationally intensive. That's kind of my bedrock definition of HPC or high performance computing, is if something is computationally intensive, it's doing, as Brian said, solving differential equations for some physical science, that is HPC.
Rose Stein:
Thank you for that. I could tell you were going around it, and that was something that people ask, what do you mean by HPC? And it's like, Ooh. And everyone's got a little bit of a different perspective in the way that they understand it. That was very delicate and well said.
Gary Jung:
I was just going to say the way that we usually do because we have a lot of scientists and they have a lot of laboratories. They can't do it on their own with their laptop, the most powerful workstation it can put together in their lab, and whatever either the amount of data they're gathering or the computation, they can't do it on their own as high performance computing. They have to go somewhere else to get it done.
Forrest Burt:
That's similar, just to chime in, as to how I've always kind of defined HPC. It's essentially the use of some dedicated resources beyond what you already have been running on to do some work. If you're finding that your laptop isn't working, it's not running fast enough on that, and you even go get one cloud instance that gives you more capability, and you're running on that, that's HPC. If you're trying to improve the performance of Edge devices beyond these little CPU-based things, and you've maybe got one of the old GPU boards or something like that. You are looking for resources that, as I said, are beyond what you already have available. You kind of have to look at it relative to what already existed there. It could be that when you need to move something off your laptop, you have one of these world-class massive systems available, and that's definitely HPC. But it could also be that you're in a small lab, and you're just moving to a cloud instance that has an A100 on it or something like that. That's still HPC because you are going in search of something that is better than what you have and that can improve the efficiency of your computation.
High-Throughput Computing [26:03]
Zane Hamilton:
How does high-throughput computing play into this? Get a reaction from Brock?
Brock Taylor:
This may not be the answer you necessarily want to hear. Again, you'd have to define high-throughput computing. It can be HPC again, when we talk about parallelizing an application, it's splitting that workload across multiple resources. That can be within a single processor. It can be across multiple processors, it can be across multiple machines. How you parallelize what the algorithm is doing can actually sometimes break up a problem where each individual resource doesn't need to know anything about what any other resource is doing. That's essentially called embarrassingly parallel, meaning nobody has to talk to each other. Everybody can just go off, do their part of the problem, and at the end, the results get gathered. That in itself could be thought of as high-throughput. Each individual resource is then how much it can do.
You can spread that out in many different ways. You can spread that across different types of resources. and it's all about how fast you can get results out of that resource. Something that is tightly coupled, meaning the computation of an individual resource has to actually talk to other resources to do the computation, becomes a different problem. And throughput in that respect is the overall problem. What all resources combined can actually do in a given time period. How much throughput you can have? You can also have throughput defined as the outputs of individual parts of a workflow. You can think of throughput, as I have a week to come up with the design. I'm going to simulate as many different design parameters as I can. That's the throughput of design parameters, and design parameter exploration. Now step back, high throughput could be, again, how many web pages can I serve? That wouldn't necessarily be high performance computing. Again, it's kind of a loaded answer, but that's what I would say.
Models Being Used Throughout Different Disciplines of Science [28:57]
Zane Hamilton:
That's great. Whenever we talk about embarrassing parallels or the other types, what are some examples of the different disciplines within science that are using those different types of models or different types of software? I know Gary alluded to earlier, the scientists are having to write them themselves, so they have to have some sort of understanding of what they're trying to accomplish. What type of methodology they're going to use, and which algorithms they're going to use? Is there something like chemistry typically uses this, and engineering, and CAD things use this?
Brock Taylor:
Absolutely. Gary, go ahead. I don't want to step on it.
Forrest Burt:
I was just going to give two real quick ones. In bioscience, a really common task here that's embarrassingly parallel is gene sequencing. If you have a whole bunch of samples that you've taken from some type of sampling machine, like tissue samples, something like that in bioscience. And you're trying to determine exactly what genes are in that sample, taking those hundred thousand samples you might have and aligning them against a genome to figure out what genes each one of those samples is pretty much, as I understand it, an embarrassingly parallel task. Another good example I heard of in my past HPC experience was space telescopes that take large amounts of images. If you've got 20,000 images coming down from Hubble, you need to be able to do some type of specific processing across all those images.
That's going to be something embarrassingly parallel because you're going to do the same processing to all 20,000 of those images. The execution of that processing on any one of those images doesn't depend upon any other finishing. You tend to find that embarrassingly parallel architecture maps really broadly to a wide variety of different workloads. Even outside of HPC, there's a lot of things that essentially go do the same bit of processing to all of these files. I won't get too ahead of myself, but the MPI side of things is where you tend to get much more complex and much more like dedicated HPC-type applications because of the much, much higher barrier to entry that there is to write MPI code, that's doing like core to core communication on a machine. Versus writing some Python scripts that can easily be run over 20,000 files, through like a swarm job or something similar. It's really quite a broad architecture, but you see it really widely in HPC.
Rose Stein:
Brian, what are some things that you've noticed? Not noticed necessarily, but where have you seen the coolest simulations?
Brain Phan:
On the CFD side, I've seen, typically, the simulations are rotten, like through MPI. For example, if you're running a simulation on a car, the car model is basically broken up across all of the different servers. Each server works on solving the differential equations for a particular time step. Then once you're done computing that part of the model, all the servers coordinate back with each other to combine all the results back, and then you move on to the next time step. If you do enough of these time steps, maybe you could simulate, for example, a crash test.
Visualization of Simulation [32:38]
Zane Hamilton:
Brian, what does it look like from an end user perspective as this thing is going through, and some of these clusters get very, very large? Can you watch these things taking place in real-time? You could almost watch a simulation taking place. Is there visualization behind this?
Brain Phan:
Typically, while the simulation is running, you could take visualization software and attach it to or watch the directory that the simulation is running in. And if the correct files are being generated, you should be able to render the visualization within the software. As your simulation is running, you can kind of watch it to see like your results are what you actually expect, and if there's something that's kind of off, you can just stop it, checkpoint your simulation, modify your model, and then just restart it.
Zane Hamilton:
Is that something you see typically more on the engineering side? People doing crash tests, people trying to simulate wind tunnels, people working on wind generators, those types of things, but they're watching something in real-time. Is it staying there? Or if you go back into genomics, or are you watching a model being built? Is that something that they care about as much?
Brain Phan:
I think with genomics, you don't care as much. This is just from my experience, I don't think you would care as much while the computation is actually running, but you're more interested in what type of results are generated. I think the result that people usually look for are VCF files, which are basically a genetic diff to a source, genomic sequence, for example.
Social Use of HPC [34:16]
Rose Stein:
You guys, I just was totally tripping out. So you can do things like weather, car crashes and I imagine bombs and things like that. You can probably simulate any physical thing in the world, and people just have either curiosity or need or whatever will do these sorts of things either at universities or companies or wherever. What about social constructs? Like social, like life, could you take all the data and all of the information from all the cameras all across the world and be like, okay, well, if we were to do this behavior or have this thing or build this building or destroy this city or whatever, could you simulate out then what would happen? Has anyone done that? They probably wouldn't share it with us anyway.
Gary Jung:
Are you asking about social, like how we could use this in a social context, HPC? Is that what you're asking?
Rose Stein:
Yes, absolutely. I mean, if you can do it to predict the weather based on past patterns. You can use high performance computing for that. Couldn't you do it socially? Like, hey, is this the right move for our society, for our country, for our world? Could you, is that maybe what they did with the, what do you call it, climate change? That's a similar type of thing, but like for people.
Brock Taylor:
They're definitely doing it now. Climate simulation is becoming very important, obviously. The example that actually comes to mind when you say social, it's a little bit of the AI space. But again, to me, that's HPC. There's a researcher and efforts that are going on at the University of Illinois, where she actually used data mining to go back through digitized articles and recreate the history of African American women in the US. It's an element of if you try to do that without computing, you have to go read articles physically, you have to dig and find the articles, and you, you essentially have to get lucky. You have to pull up the right element, the right book with the power of computing. Now the algorithms are going out and searching.
Finding things, and then there is vetting whether or not it's a good match. And continually improve whether or not it's a good match to an actual event that's significant. What this researcher was able to do is actually go out and find and piece together all this information that can collectively recreate the history that you wouldn't otherwise be able to do. Because either you have to employ too many people to go out and read and dig through libraries manually, and again, get lucky to stumble across something and know that there are three other people that have connected pieces. Computation can put it all together, and present it and allow the researcher to actually start to look at that, vet it, and bring it forward as a collective piece. That's a huge power in society.
The legal system is obviously using this quite a bit as well. If you think about law and how cases are presented, precedence is a huge element of the law. Being able to go back through all law cases and pull together things that may be related, and you can do it in a way that's time sensitive. You don't have infinite amounts of time, but you can go out and find these pieces. That's where as well, you have the ramifications of the accuracy of what is being done, and AI is still in the early phases. If it's non-consequential and it's wrong, it's okay. But if it's mission-critical, like an autonomous driving system, and AI makes a mistake because it doesn't have the data, it can actually be very damaging. Again, it's how the technology will be used. How do you ensure that it's correct? How do you go from, Hey, I'm confident this is the right answer to The answer is credible. That's a big issue that many high performance computing domains face.
Zane Hamilton:
That's interesting, Brock. I would've never thought about recreating history or actually trying to get an accurate history from that aspect. That's very interesting. I was thinking of Rosa's question more from a perspective of, I feel like advertisers are doing this with Facebook, Twitter, TikTok, and all the data that's coming from there. I mean, somebody's using that for something, and the only thing I can ever go back to is probably doing it for an advertising purpose. To Rose's point, are they using the data for other stuff, for other social predictions? Gary, I think you were going to.
Gary Jung:
They actually have like two examples for climate change and one with social science. We are down at UC Berkeley, one of the PIs that runs in our system is Saul Shung, and he's with the Goldman School of Public Policy. His work is in using high performance computing to model how public policy affects climate change or how different policies would affect climate change, for example. From the science perspective on climate change, we have somebody else here, his name is David Roms, and he's a PI here, and he works on the cloud resolution models. In a lot of the climate change models that are used to model climate the side of the cell is like a hundred kilometers, and which is actually really, really large. But one of the things that they're not able to model with that is the formation of clouds, which happens on a one-kilometer scale. His work is in doing cloud-resolving models so that this way, you can run that in conjunction with the climate change model and get a better picture of what's going on at a smaller level.
Then one other one that I'll toss in here because what made me think of this is Rose thinking about this data that we get from cameras and stuff. One of our astrophysicists, here at Berkeley Laboratory, his work is in looking at telescopes and looking at the sky. He had come up with this idea like, well, gee, let's what happens if we return this around and we can look at the ground? So he's using high resolution cameras to look for fire detection. With that, he's built this training model, which essentially can do like early smoke detection and can notify Cal fire. He's done this a few years ago, so this isn't that new, but it was just interesting that he was using these cameras to do early detection of forest fires and using our HPC systems to do the training of the model to do that. They just run that all the time on our systems.
Where Does the Data Come From? [42:30]
Zane Hamilton:
The question I was going to ask Gary, is I noticed earlier David DeBonis had made a comment about data, and it seems like data is the key in this whole thing. You've got to have data to actually build the models and look at them. In your example, for the fire and everything, where are they getting that data from? Are they taking satellite images?
Gary Jung:
They were using some satellite images, and then at one point, there was some sensitivity about getting, like super high, access to government images that are super high resolution. Then all of a sudden, the data became sensitive. It used a lot, a lot of data.
Zane Hamilton:
Interesting. It's always interesting stuff like that where things come from. If they're getting thermal imaging, if they're getting airplanes flying over, taking pictures that are just grabbing pictures they go by, it's always interesting to see where the data comes from.
Gary Jung:
We've had other projects where people are doing, we did one project with Homeland Security, and their thing was like trying to detect radioactive sources. They have a flying platform, lidar and high speed cameras, and gamma ray detection. Essentially what they're doing is the flying platform is just flying and building the slight, you think, like a super rich version of Google Maps where it has all this extra information. Then hopefully, what'll happen is that or what they were doing was they were using it to spot radioactive sources or like bombs and stuff like that. That was really interesting because then the algorithms essentially, if you're flying at a hundred miles an hour, you have to be able to triangulate and figure out like where you spotted that stuff. It becomes a complex mathematical problem.
Zane Hamilton:
Very interesting. Forrest, did you have somebody you wanted to add?
Forrest Burt:
I just had one random thought about the social side of HPC. One use case that I've heard that's really big and emerging in the space, the more people are messing with it, is graph analytics. I've heard at a couple of conferences I was at last year I didn't hear a ton about it, but there were definitely a few questions asked about how graph analytics played into a few things. These are things like being able to take the web of connections between friends on Facebook. The canonical example is taking a sample of people from Facebook, looking at all of their connections and how they're all friends with each other and stuff. Then doing whatever analytics. I'm sure Facebook is very into doing that type of thing for advertising, as we've noted, and stuff like that. But it's becoming a bigger and bigger generalized HBC use case. Being able to look at these massive connection graphs and do meaningful analytics on them from social media sites or what have you.
Brock Taylor:
I would throw city planning as another area. If you just think, planning where cell phone towers go or where those access points are so that there's adequate signal and coverage, you've got acoustics of how sound can be amplified inside a city or dampened inside of a city. I mean, there are just all kinds of ways it can be used and is used, and it's probably underused in a lot of areas that are impacted to our daily lives.
Zane Hamilton:
5G and microcells have made that a whole other thing, putting a microcell in every bus stop because they run fiber right next to it. Remember looking at those maps for a while myself.
Rose Stein:
Speaking of social and Facebook and high performance computing, my son showed me this. I think it was my cousin. She had posted this picture of herself right next to her boyfriend, and then it said, I'll love you forever, or something like that. Then it went, do do do do, do, do, and it aged him like 50 years. There's like her standing next to him, and you know what I mean? I was like, that is so weird. I was just thinking that might be a good motivation for people where it's like, hey, if you were to quit smoking, I mean, say there's a smoker or whatever, like do but some kind of habit. Right now, this is how you would age if you continued smoking, this is how you would age.
There are so many little different things, and what I was thinking about is these last few years have been rough and not to get into any, like, controversy about people's feelings and perceptions about that. There were some consequences to the actions that we took regardless. There's like a lot of human beings who just don't do very well when they are isolated from each other and afraid of each other. Even if it was for a good reason or whatever, what if we could simulate that stuff? The pros and cons and this action may be worth the potential consequences. Can we project that out? There are a lot of different questions right here.
Brock Taylor:
I'll chime in first. I definitely think we can. I get back to the question of how do you trust the answer that comes out. That's the human element of high performance computing. We have to have the researchers that are aware that they have to vet the answers. If you look at academia and that model my opinion, I'll say it is an opinion; we get into trouble when people take an answer and just throw it out as fact, without having it vetted. That's where academia has the peer review system. You can put your ideas out there, but other people have to be able to recreate and agree with the answer and come to a consensus. You're never going to get a hundred percent, but you try to get those answers to make sure that you're credible in having a model.
Again, we are talking about a digitized element of an analog world where you don't have an exact answer. What comes out of a simulation is a simulation. It is not exact. You can always find a flaw somewhere. An example is I recently saw a researcher who is working as a material scientist designing and simulating the material, then actually manufacturing it. What is manufactured has flaws because of the manufacturing process. It's different from what was simulated. Now, you have to look at the simulation as a pure model, and the real world object is not, so the two are different. You are actually now dealing with two different things. You have to figure out what's an acceptable element, and what's an acceptable error bar. Those are the things we could have simulated what a policy might have meant and done, but you also have to make sure that you don't put data out there that may not be that accurate. It came up, I'm not going to say what it was, but an agency, basically he said, we now believe this data point to be true with a low confidence level. What does that mean? It means, yes, they simulated it, and it could be true, but we're not very confident in the answer. It's feasible. It may not be true. They do not have a confidence level in how they got to that answer. But they chose to put it out there. It's a dilemma. It's what we face. There's no right or wrong answer necessarily, but it's tough.
Rose Stein:
It really sounds like it's going in the direction of, sorry.
Zane Hamilton:
No, go ahead. Go ahead.
Rose Stein:
Of the high performance computing, and it's exciting, and it's still new, and there are lots of potentials. This idea is that the software that is giving whatever scientists or people or someone with an idea access to the compute resources that ability to then share the data and share the information and have it be reproducible. I mean, that sounds like it's really going in the direction of what we are doing at CIQ is making that accessible to the people who want to use it.
Zane Hamilton:
Brock, I was going to ask on the one where they created the material, did they go back and retrain the model and input what they had learned so it would be more accurate the next time? Or does that seem like it'd be a good idea to do?
Brock Taylor:
This is not a retrain, this is you have to build that model for the simulation, and it's actually a huge problem. Because how does this researcher create that model that has all these individual imperfections at the microscopic level by hand? You would essentially have to be able to take that material, and have something that can scan it like cryo-EM does to create that computational model that can then be simulated. These are the problems that these researchers are, in fact, working on trying to help solve. I can't say exactly what it is because it's a private conference. But this research is working on bone growth. Literally,in a sense, that could help people across the planet heal after major trauma. Either from surgery or from accidents. You have broken bones, and instead of having titanium implants screwed into the bone, there could be materials injected that actually grow the boat faster.
Differences Between Enterprises and HPC [53:45]
Zane Hamilton:
Very interesting. That's cool stuff. Forrest, I know you wanted to talk about the differences between enterprise and HPC, and I think we're starting to see them come together, but I know you have thoughts on this topic.
Forrest Burt:
I just wanted to touch on this a little bit. We talked a little bit about websites, that type of stuff as something that's done in computing. I tend to class websites, large databases, software-as-a-service deployments, and that type of thing in the categorization of enterprise computing. These are large scale workloads that are serving some resource out to a large number of people at once. But for the most part, the resources that are being used are fairly generic. Something that we have to consider here is the hyperscale versus HPC because a lot of the time, these enterprise computing deployments will take social media sites, as we've discussed as an example. A lot of times, the enterprise computing that they're doing requires such an incredible scale that it can be hard to differentiate it from high performance computing.
One thing that I look at there is with the hyper scale. You are essentially trying to serve one resource out to thousands of users at once. With HPC, you're trying to use all those resources to create something for a much smaller group of people. Hypothetically, aren't thousands of people waiting on the simulation? It's probably one small group's work that's being done there. In general, that's a hyperscale back to enterprise computing a little bit. In general, one of the biggest emergencies between enterprise computing and HPC is in the tooling and stuff like the automation and platforms and things that are available to them. We see an enterprise that has been really common for a long time now. There's Ansible, containerization, and CICD platforms, there are all these different tools that are used ubiquitously to manage these massive deployments.
They are everywhere. There's Kubernetes, container orchestration, and that type of stuff. All these major sites. Most enterprise computing relies on these technologies, but for the most part, these technologies are unheard of in HPC. On the converse, the scale of some of the enterprise computing, like around, I don't want to call enterprise computing AI here, but the AI computing that they are doing in the enterprise is reaching a scale that they're starting to need HPC class systems to be able to work with that. The end result of this is that we have an enterprise wanting HPC class resources. We have HPC wanting enterprise-class tooling, and there's a bit of a divide between these two things. At CIQ, our answer to this is our platform working on Fuzzball, which aims to bring enterprise and high performance computing together in a way that allows HPC to use container orchestration and CIC platforms, container or cluster deployments that rely on things like Ansible, et cetera.
It's a very, very interesting space at the moment. Like I said, we touched on social media sites' websites, so I wanted to mention that for a moment. There's a very interesting divide that has been growing between enterprise and HPC that they're both doing the thing that the other one wants. It'll be very interesting to see how they come back together. As I said, one of the biggest things that we work on at CIQ is bringing that divergence, to a meaningful resolution for everybody involved. Very interesting space.
Future of HPC [57:18]
Rose Stein:
Shameless plug for us. I love it. That's what we're talking about because it's exciting, and there is a need to bring the two together. That was where I was going with Brock like, Hey, if it's a non-HPC, then what is it? Well, we are kind of blending, so it's less and less distinctive. Thank you for bringing that up, Forrest, and it really kind of ties into the next question that I've been thinking about here. Where are we going, right? We talked a lot about what it is like and where we are now. What is the future of HPC as you see it? We'll start with Brock.
Brock Taylor:
I'm going to reference what Gary brought up earlier. 20 years ago, it definitely was the researchers themselves writing codes to solve these highly complex problems. You see that less and less, and that's a good thing because these researchers are usually not computer scientists. They do not know the details of hardware architectures, or cash architectures. They wouldn't know what InfiniBand is. They may not even know what a Linux platform is. That's that element of, as we move forward and make this more accessible, the domain specialists stay in their domain. That's really important because these researchers are specialists looking at these highly complex systems to come up with solutions. And the more that they have to learn something else to do that research, the more they are delayed from coming up with their answers. I see the future as that distraction and more elements of software layers, allowing them to do more in their specialty areas.
Rose Stein:
Thank you. Brian, what's your, what's your vision? Where do you see HPC going?
Brain Phan:
Adding on to what Brock was talking about, in the future, I see just better tools and better hardware coming out. With these two things, users will be able to get results a lot faster, and that's what they care about. If you're able to get results faster, it means you can make design decisions faster. If you could do that, you could probably get to market a lot faster before everyone else. As a result of that, we'll just see cooler new products coming out that will hopefully improve our lives.
Rose Stein:
I like it. Bigger, better, faster, more. Awesome. The slogan of this year's super computing conference is "I am HPC." Claiming the affirmation. I love it. Thanks, David. Gary, do you have a vision of where you see HPC going?
Gary Jung:
I think we're just getting started. I really think we're just getting started and maybe Rose the what you're saying about like how it could be used. I think what would be nice is that we could be used in a way that people don't have to think about it. For example, that app where somebody could just click on this thing and see how they age, I could see HPC just behind the scenes is just becoming part of the infrastructure where people just utilize it, and they're able to help them make better decisions, hopefully, make better decisions by using it. I think there's a lot of work to get there. I agree with Brock in that the domain specialists, I can see now, a lot of them don't want to do the coding and develop their own codes. But there's a little bit of a dilemma there because the people who have done that have been critical to like, making HPC successful.
I'm not sure what the model is going to be in order to continue to achieve what we have. That's going to be tough. As far as just HPC and enterprise, not that long ago with just APC was just the computation, and then the next big thing was data. Data's kind of where people are talking about it as a first class citizen, but I don't know that that's actually true yet, and a lot of HPC centers. But just merging all of that with AI I think there are still a lot of interesting things to be invented and put into practice.
Rose Stein:
Thanks guys. That's all the questions I have. I don't know if there's anything else that you want to bring up Zane or Forrest. You got what you wanted to see?
Forrest Burt:
I was just going to sum up essentially what I had said there at the moment. HPC is going nowhere but up. HPC is everything, as we've discussed here, from really high level simulations to the weather modeling that tells public utilities how much power they're going to have available from renewables and stuff like that. As we've noted, HPC touches so many aspects of life, and the use cases that are supporting all of this technology and civilization stuff are only getting more complex. At the moment, HPC is going nowhere but up. There's a complete Renaissance in hardware and software at the moment with new chips, new devices, new appliances, in software, and new paradigms. Like I touched on, the combining of enterprise and HPC together, that ultimately allow users to do more work faster and focus more so on their domain science and less so on system internals and that type of thing. Ultimately, to kind of reiterate what everyone else has said here, it's a very great time for HPC, and it's very interesting to see where it's going on all fronts.
Zane Hamilton:
I think it's a great way to wrap up, Forrest. I agree with you, it is very interesting to see where things are headed. We are up on time. Really appreciate your time, Gary, it's always good to see you. Thank you for joining us. Brian, Brock, Forrest. Thank you, Rose. It's been great. I appreciate it. If you don't mind going and liking and subscribing to the CIQ channel, we would really appreciate it, and we will see you again next week. Thank you.
Rose Stein:
Awesome. Thanks, guys.
Transcript
good morning good afternoon and good evening wherever you are thank you for joining at ciq we're focused on powering the next generation of software infrastructure leveraging the capabilities of cloud hyperscale and HPC from research to the Enterprise our customers rely on us for the ultimate Rocky Linux werewolf and abtainer support escalation we provide deep development capabilities and solutions all delivered in the collaborative Spirit of Open Source hello everyone and welcome to another ciq webinar hello Rose hello and welcome how are you great thank you so what are you talking about this week ooh so this is actually really interesting topic in conversation we're going to
be talking about HPC but in a broader context like how are experts in the field actually utilizing HPC resources to solve the complex problems that we're facing very nice so who do we have this week let's bring in our panel Gary Brian Brock's Forest Forest I force you to come so I expected you to be here how's it going everyone good to see everybody so we'll go around do introductions real quick uh as you know Rose and myself let's start with Gary hi my name is Gary Jung I manage the institutional HPC for Lawrence Berkeley National Laboratory and I also managed HP program down at
EC Berkeley thank you Gary Brian everyone Brian Fenn here I'm a Solutions architect here at ciq my background's in HPC Administration architecture and I've I've experienced supporting HPC users across various domains of science so I'm excited to be back and discussing this topic today thank you very much Brian Brock I hope you're having nice weather it's going to be that time of year right yeah it is uh yeah it's it's definitely very nice uh has not hit that boiling hot phase so if you can go outside and enjoy the weather yeah a long time HP yeah exactly long time HBC are in deep background a
couple decades worth in HPC primarily in clustering thank you Brock and Forest hi can everyone hear me all right we got you fantastic okay good morning everyone my name is Forrest Burt I'm an HBC systems engineer here at ciq uh my background is in the academic and National Lab HPC space I was previously a student system administrator at uh University here in the states uh but it's since started working at ciq on the systems engineering side of things so very excited to be on and excited to discuss HPC thank you Forrest so Rose I'm going to let Brock kind of set the stage here and
View full transcriptHide full transcript
let him tell us kind of a brief overview of the importance of HPC if you don't mind yeah absolutely I you know it's it's actually underpinning so many things that we interact with on a daily basis I think sometimes it's overlooked because when you say HPC a lot of people will think of you know the largest systems on the planet but the Technologies actually used across many different Industries and again in you know the cars we drive uh you know definitely if you get on an airplane um but just think about as your you know driving and seeing road construction Crews the machines they're using
uh Farmers the the tractors that they're driving um even things that are in our local hardware stores that are helping you know build the decks for our houses or the houses themselves I performs Computing has been used to help create a lot of those products um over the past 20 30 years the advances in HBC have actually enabled many different improvements and you know one example that that I would bring up is weather forecasting if you go back 25 years the capabilities to project out you know one week or two week two weeks and be semi-accurate was more Consulting almanacs and and Trends and things
that that you know you could you could actually try and predict with but now because Computing has become powerful uh simulations are able to actually predict much more accurately where things you know how things are going to occur in weather patterns and again it's still very very difficult there's a lot of research going on in weather forecasting but if you think you know how accurate we can get in predicting where a hurricane is going to hit and what that means to be able to warn people to evacuate people uh to you know prepare and actually save lives in the process a lot of that is
because of improvements in high performance Computing and its usage across different Industries um yeah the fun things in life uh the the new Avatar movie would not be possible without high performance Computing the first Avatar movie would not be possible but uh it's already the technologies that have gone into making that film uh who is involved they've won awards for the technology improvements um you can see what they're able to do and even simulating and actually rendering water is is a highly difficult thing to do but making it so realistic again is high performance Computing so it's very broad uh you know it's it's hours
and hours in any one subject but there's so many different things that that you can use high performance Computing to do and I think what we'll see is even more Industries more everyday products that are going to be able to tap into very powerful Technologies thank you Brock so I'm sorry Brock I'm just gonna have to ask you if you could please backtrack a little bit and what is high performance Computing and if you're not high performance Computing what are you yeah and that's uh that's a a great and loaded question and I'll start it in probably uh the others will want to chime in
as well it's it's an argument to describe or concretely Define high performance Computing what it is uh to again some it is only the very high end the most powerful super computers on the planet that's what high performance Computing is don't talk to me about anything else um if you think of high performance and Computing you can argue that our phones are in fact high performance Computing machines and in some ways they definitely are they have the power of super computers from 20 years ago um but you know to try and generalize it it is it is taking a a problem in science that you
know you can describe by functions and formulas and converting it from an analog world to a digital world which is tough and trying to solve a very complex problem that involves a lot of I'll call it moving Parts but basically a lot of science that if you do it by hand it takes you a very long time if you do it by a single computer it will take you a lot less time but you're still very much limited into high performance Computing and clustering in particular helps spread out computation to make it go faster but at at a cost of complexity and difficulty and you
know that's where ciq and the the communities that we are engaged with come in helping pull all these Technologies together so that more people can actually use them easily and again a scientist that is is researching bone growth does not have to learn how to build a cluster or learn how different pieces fit together they're focusing on improving their science you mentioned earlier and you kind of talked about rendering water and I heard that they had to go create software to do that so I think when we look at HPC we think about clusters and we think about big machines but I think a big
part of this that you kind of touched on is the software itself so I'll go to Gary next what are the different types of software or I mean I know everything about it is very specific for what you're trying to accomplish or what you're trying to do but tell me about the different types of software that people run on an HPC system boy you mean like in terms of applications maybe is that what you're going for absolutely so whenever a scientist is actually doing research I mean it depends on obviously the discipline of what you're trying to do or if you're an engineer and you're
trying to render something there's a lot of different pieces of software that people are using out there to do all kinds of different things yeah you know I you personally taking but there's people run a lot of different software software a lot of it is written by the scientists because they're trying you know I'm talking in a researcher uh context but a lot of times people are trying to model something and which only they know about and so they'll write the software for that and then and then maybe it becomes open source and other other um other people doing that type of research will will
use that software so uh there's software like the vasp which is uh like a lot of uh material signs people use that for electronic structure calculations atomic at the atomic level of materials uh it can calculate things like the phases and band gaps stuff like that there's um uh there's image there's software for doing Imaging a lot of people do use clusters for Imaging and imaging reconstruction so there's a lot of software written for that um there's uh their software for every different type of of domain of Science and uh really the thing that uh distinguishes uh this the software is that it can run
on a large system and that can either run multiple instances of it or it can split up a problem so it can run on a lot of systems um and so that's that's the way I would characterize it I can carry some Brian Force I know we throw a lot at you guys from a software perspective and talking to customers and going to prove that that software Works in an HPC environment or at least show kind of what that looks like what are you guys running across from a software perspective on HPC today from a customer perspective start with Brian oh Forest came off mute
first but Brian's off now go ahead Brian um I'm a cfd perspective like uh at the end of the day Engineers are when you're running simulations you're basically when you set up a simulation problem you're basically setting up a system of differential equations and you're basically trying to find the solution to that uh with uh cfd software uh they're basically writing solvers to solve these differential equations and with different uh software vendors they basically add their secret sauce to these solvers to make it go faster and depending on your needs you might go with one vendor versus another uh but yeah that's from my experience
in um Automotive in Aerospace thank you Brian of course so at the moment the stuff I've been seeing most is kind of the academic and AI side of HPC um from the academic side of things uh a lot of what people are doing and using is the same type of stuff that they've been using for a while molecular Dynamics work within grommax lamps that type of stuff is alive and well um kind of the usual licensed use cases Mathematica Matlab um I've heard mathematics for a bit like Matlab that type of thing is still of interest to people um more and more there's interest around
kind of dedicated data science platforms um that give like a whole all-in-one interface for doing a lot of different data analytics machine learning those type of tasks um I mean different things qmc pack Quantum espresso you know a lot of that stuff is pretty standard out of that and the AI side of things we're really seeing a lot of interesting stuff going on um a lot of the work in that field is centered around a few major programming languages a few major programming languages and a few major Frameworks for those programming languages so we're kind of seeing um sort of like what Brian alluded to
with the cfd industry where uh we're kind of seeing an AI of these large-scale Frameworks these large-scale pre-trained models that type of thing that are out there at the moment being leveraged widely by people and being fine-tuned by researchers and companies and places like that there's a very strong culture at the moment of kind of using what's out there because so much work has gone in developing these AI Frameworks like Jacks tensorflow pytorch that type of thing there's so much that they're already capable of and additionally um kind of for example with the GPT space there's so much pre-trained material out there that saves so
much time and effort for people that want to Tinker with it but we're seeing a lot of use of kind of these pre-trained models and people taking those and then uh I think Brian's words are adding their own secret sauce to it um and then being able to go and uh have something more customized for their own purposes so all kinds of different use cases coming out of my side of things it's pretty interesting hmm thank you Forrest that is awesome Brock I'm gonna I'm gonna uh it's possible that you did not answer this part on purpose but I would love to know if it's
not high performance Computing what is it is there a name for or or is all Computing I mean you kind of mentioned the phone which I was thinking about I mean you know 40 years ago this was like oh my God high performance imagine I couldn't even imagine that so is everything high performance Computing or is there something else that is happening uh there are there are things that I would not call high performance Computing and uh you know not to disparage that but you know take a web server for instance that's you know just feeding us our the pages we're looking at that's that's
not really high performance Computing um it's not computationally intensive work um but it doesn't mean that it it's not connected to some elements of high performance computing all right so I view it again it's it's a it's a tough question and people live in different camps but if you think of computation you know mathematical computation uh number crunching if you will almost anything that is doing that can be classified as high performance Computing workloads or applications and again in in today's world with the capabilities that we have it's not just a single domain or application it's many different high performance Computing applications working together to
solve a very complex problem that you wouldn't have been able to to actually solve 10 20 30 years ago okay um yeah problems that that were being solved with high performance Computing but as an example you know if you take a bulldozer there's actually a lot of different systems that go into making a bulldozer there's there's the actual structure how much uh earther material it can move how much force it has there's the Hydraulics that lift the bucket and things of that nature all of those elements are simulated before they ever build that first bulldozer right same thing with the car if you you know
if if you're as old as I am you know you remember Crash Test Dummies were in all kinds of commercials 30 years ago right um they're they still exist Crash Test Dummies still exists but they're you know they're not used nearly as much in physical car crashes are done pretty much to prove a design operates the way it should high performance Computing is the reason simulation is the reason that you don't need to crash cars as much or they're really proving that it does actually the design of the card does actually do what it's supposed to right so I do I'm kind of dancing around
concretely answering it because it can cause an allergic reaction to some people uh again have the debate is AI high performance computing I am in the camp that it absolutely is because it's computationally intensive it is different than traditional computational models that are doing uh it's hard to get a little technical floating pulling out operations it's more integer work right it exercises silicon in a different way the algorithms are different but they are computationally intensive right and so that's that's kind of my bedrock definition of HPC or high performance Computing is if something is computationally intensive it's doing as Brian said you know solving differential
equations for some physical science that is HPC yeah and thank you for that I can kind of tell you were going around it is something that you know people what you mean by HPC and it's like and everyone's got a little bit of a different perspective and the way that they understand it and so that was very delicate I was just going to say the way that we usually because we have a lot of scientists and they have a lot of Laboratories um you know they can't do it on their own with their laptop or whatever the most powerful workstation it can put together in
your lab and whatever their the data either the amount of data they're Gathering or the computation just can't they can't do it on their own then to us as high performance computing yeah yeah go somewhere else to get it done that's kind of similar just to chime in as to how I've always kind of defined HPC it's essentially the use of some dedicated resources beyond what you already have been running on to do some work so if you're finding that you know your laptop isn't working it's not running fast enough on that and you even go get you know one Cloud instance that gives you
more capability and you're running on that that's HPC if you're trying to improve you know the performance of edge devices Beyond you know these little CPU based things that you've maybe got you know one of the little GPU boards or something like that you're you're uh you're looking for resources that as I said are beyond what you already had available and so you kind of have to look at it relative to what already existed there because yeah it could be that um you know when you need to move something off your laptop you have one of these world-class massive systems available and that's definitely HPC
but it could also be that you know you're in a small lab and you're just moving to you know a cloud instance that has an a100 on it or something like that um that's still HBC because you're going you're going in search um something that is better than what you have and that can improve the efficiency of your computation so how does high throughput Computing play into this let's even get a reaction from Brock may not be the answer you you necessarily want to hear um again it's it's you'd have to Define High throughput Computing uh it can be HPC again think um when we
talk about parallelizing an application it's splitting that workload across you know multiple resources and that can be within a single processor it can be across multiple processors it can be across multiple machines right how you parallelize what the algorithm is doing can actually sometimes break up a problem where each individual resource doesn't need to know anything about what any other resource is doing right and that's that's essentially called embarrassingly parallel meaning nobody has to talk to each other everybody can just go off do their part of the problem and at the end the results get gathered that in itself could be thought of as high
throughput you know itch each individual resource is then how much can it do and you can spread that out in many different ways you can spread that across different types of resources um and it's all about you know how much result how fast can you get results out of that resource and something that is tightly coupled meaning the the computation of an individual resource has to actually talk to other resources to do computation that becomes a different problem and throughput in that respect is the overall problem what all resources combine can actually do in a given time period how much throughput you can have you
can also have throughput defined as the outputs of individual parts of a workflow right and and you can think of throughput as I have a week to come up with the design I'm going to simulate as many different design parameters as I can that's throughput of design parameters right design parameter exploration now step back you know High throughput could be again how many web pages can I serve that wouldn't necessarily be high performance computing so again it's kind of a loaded answer but that's that's what I would say oh that's great and I whenever we talk about embarrassing parallel or the other types what are
some examples of the different disciplines within science that are using those different types of models or different types of software I know Gary alluded to earlier the scientists are having to write them there themselves so they have to have some sort of understanding of what they're trying to accomplish and what what type of methodology they're going to use and which algorithms they're going to use so is there something like chemistry typically uses this and Engineering CAD things use this absolutely yeah go ahead I don't want to step on I was just going to give two real quick ones um in bioscience so I think there's
a bit of a delay on someone's um in bioscience a really common task here that's embarrassingly parallel is Gene sequencing so if you have a whole bunch of samples that you've taken from some type of sampling machine like tissue samples something like that and bioscience and you're trying to determine exactly what genes are in that sample taking those you know 100 000 samples you might have and aligning them against a genome to figure out what genes each one of those samples is is pretty much as I understand it embarrassingly parallel task um doing something like another good example I heard of kind of in my
past HKC experience um space telescopes that are taking large amounts of images if you've got you know 20 000 images coming down from Hubble and you need to be able to do some type of specific processing across all those images that's going to be something embarrassing and parallel because you're going to do the same processing to all twenty thousand of those images and the execution of that processing on any one of those images doesn't depend upon any other one finishing um so you tend to find that embarrassingly parallel architecture uh Maps really broadly to a wide variety of different workloads even outside of HBC there's
a lot of things that are essentially go to the same bit of processing to all of these files um I won't get too ahead of myself but the MPI side of things is where you tend to get much more complexity and much more like dedicated HPC type applications because of the much much higher um just barrier to entry that there is to write NPI code that's doing like quarter core communication on a machine versus writing um you know some Python scripts that can easily be run over 20 000 files through like a swarm job or something similar so it's really quite a broad architecture but
you see it um really really widely in HD hmm um Brian what are some things that you've noticed not noticed necessarily but yeah like where have you seen like the coolest simulations or uh so on the cfd side uh let's see what type of so on the cfd side like I've seen typically the simulations are run like through MPI and how like for example if you're running a simulation on a car uh the car model is basically broken up across all of the different the various servers each uh each server works on solving the differential equations for a particular time step and then once you're
done Computing on that part of the model uh all of them coordinate with all the servers coordinate back with each other to kind of combine all the results back and then you move on to the next time step and you know if you do enough of these time steps maybe you could simulate for example a crash test yeah what's with that Brian what does it look like from a an end user perspective as this thing is going through and some of these clusters get very very large can you watch these things taking place in real time so you could almost watch a simulation taking place
is there visualization behind this yes uh typically uh while the simulation is running you could take visualization software and attach it to or watch the directory that the uh simulation is running in and if the correct files are being generated you should be able to render the visualization within the software so as your simulation is running you can kind of watch it to see like your results are what you actually expect and if there's something that's kind of off you can just stop it checkpoint your simulation modify your model and then just restart it that's something you see typically in the more the engineering side
so people doing crash tests people trying to simulate wind tunnels people working on wind generators those type of things but they're watching something in real time or that is it staying there or if you go back into genomics are you watching a model being built is that something that they care about as much uh I think with genomics you don't care as much a wall well this is just from my experience I don't think you would care as much like while the computation is actually running but you're more interested in what type of results are generated I think the result that people usually look for
are VCF files which are basically a genetic diff to a like Source um genomic sequence for example yeah you guys I just was like totally tripping out like can you okay so you can do like weather right and car crashes and I imagine like you know bombs and things like that right like you can probably simulate any physical thing in the world and people just had either curiosity or need or whatever will do these sorts of things either at universities or companies or wherever like what about social constructs right like social like life like could you take all the data and all of the information
from like all the cameras all across the world and be like okay well if we were to do this Behavior have this thing or build this building or destroy this city or whatever like could you simulate out then what would happen like has anyone done that they probably wouldn't share it with us anyway but are you asking about like social like how we could use this in a social context HPC is that what you're asking yes absolutely I mean if you can do it to predict the weather based on past patterns right and you can use high performance Computing for that like couldn't you do
it for yes like socially like hey is this the right move for our our society for our country for our world like could you is that maybe what they did with the uh what do you call it climate change like that's similar type of thing but like for they're definitely doing it now uh you know climate simulation is becoming you know very important obviously um the example that actually comes to mind when you say social is um it's a little bit of of the AI space um but again to me that's HPC uh there's a a researcher and efforts that are going on the University
of Illinois um where she she actually used uh data mining to go back through digitized articles and recreate uh history of African-American women in the U.S right and it's an element of if you try to do that without Computing you have to physically go read articles you have to dig and find the articles and you you essentially have to get lucky right you have to pull up the right the right element the right book with the power of computing now algorithms are going out and searching finding things and then there is vet whether or not it's a good match right and continually improve whether or
not it's a good match to uh an actual event that's significant and so what this researcher was able to do is actually go out and find and piece together all this information that can collectively recreate the history that you wouldn't otherwise be able to do because either you have to employ too many people to go out and read and dig through libraries manually and get lucky to stumble across something and know that there's three other people that have connected pieces computation can put it all together right and present it and allow the researcher to actually start to look at that that it and and bring
it forward as a collective piece and that's that's a huge power in society um the legal system is obviously using this quite a bit as well if you think about law and how cases are presented precedence is is a huge element of law so being able to go back through all law cases and pull together things that may be related and do it in a way that's that's time sensitive you don't have infinite amounts of time but you could go out and find these pieces and that's where as well you have the ramifications of the accuracy of what is being done right and and AI
is still in the early phases uh if it's if it's non-consequential and it's wrong it's okay um but if it's if it's Mission critical like an autonomous driving system and and AI makes a mistake because it doesn't have the data it can actually be very damaging right and again it's it's how will the technology be used how do you ensure that it's correct uh how do you go from hey I'm confident this is the right answer to the answer is credible that's a big issue that that many high performance Computing domains face it's interesting Brockett I would have never thought about recreating history or actually
trying to get an accurate history from that aspect that's very interesting I was thinking of Rosa's question more from a perspective of I feel like advertisers are doing this with Facebook Twitter Tick Tock all the data that's coming from there I mean somebody's using that for something and the only the only thing I can ever go back to is probably doing it for an advertising purpose but to roses point are they using the data for other stuff for other social predictions Gary I think you were gonna yeah yeah I actually have like two examples for climate change and one with social science and so we
have down at UC Berkeley uh one of the pis that runs on our system is uh Saul Sean and he's with the Goldman School of public policy and his work is in using high performance Computing to model how public policy affects climate change or how different policies would affect climate change for example and then and from the science perspective on climate change we have somebody else here uh his name is David romps and he's a pi here and he works on the cloud resolution models and so a lot of the climate change models uh that uh model climate you know the side of the cell
is like 100 kilometers which is and which is actually really really large but one of the things that that they're not able to model what that is like the res formation of clouds which happens on a one kilometer scale so his work is in doing uh Cloud resolving models so that this way you can run that in conjunction with the climate change model and you get a better picture of like what's going on at a smaller level um Target and then one other one that I'll toss in here where is uh what made me think of this is Rose think about this data that we
get from cameras and stuff um one of our uh uh astrophysicists uh here at Berkeley laboratory you know his work is in you know looking at telescopes and looking at the sky and uh he came up with this idea like well Gee let's what happens if we return this around and we can look at the ground and so um so he's using high resolution cameras uh to look to do fire detection and with that you know he's built this training model which uh essentially uh you know can do like early smoke detection in it and can notify Cal Fire and so he had this is
he's done this a few years ago so this isn't that new but um it was just kind of interesting that um he was using these um these cameras to do early detection of forest fires and using our HPC systems to do the training of the model to to do that and they just run that all the time on our systems so the question I was going to ask Gary is I noticed earlier uh David bonus had made a comment about data and it seems like data is kind of the key in this whole thing you've got to have data to actually build the models and
and look at in your example for the fire and everything where are they getting that data from are they taking satellite images is it yeah yeah and so you know there there's um uh they were using some satellite images and then at one point there was some sensitivity about getting like super high uh access to you know government images that are super high resolution and so then all of a sudden the data became sensitive so um they're uh but yeah yeah it's used a lot a lot of data interesting it's always interesting on stuff like that where things come from if they're getting thermal imaging
if they're getting uh airplanes flying over taking pictures that are just grabbing pictures they go by it's always interesting to see where the data comes from yeah yeah we we've had other projects where people are doing uh we did one project with Homeland Security and uh you know their thing was like trying to detect radioactive sources and so they have a flying platform it has lidar and high-speed cameras and um you know gamma ray detection so essentially what they're doing is they're the flying platform is just flying and building this like you think like a super rich version of Google Maps where it has all
this extra information and then and then hopefully what will happen is that you know or what their what they were doing was they were um easing to uh spot radioactive sources uh for like bombs and stuff like that but you know that that was really interesting because then the um uh the uh algorithms essentially if you're flying at 100 miles an hour you have to be able to train triangulate and figure out like where you spot that stuff so it becomes a complex mathematical problem very interesting Forest did you have somebody wanted to add graph now I just said one random uh thought about the
social side of HPC one thing that I've uh come to use case that I've heard that's really big and kind of emerging in the space the more people are messing with is graph analytics I've heard a couple conferences I was at last year uh it was kind of I didn't hear a ton about it but there was definitely a few questions asked about um you know kind of how graph analytics played into a few things so this is thing or these are things like being able to take you know the web of uh connections between like friends on Facebook the canonical example is taking like
a sample of people from Facebook looking at all of their connections and how they're all friends with each other and stuff and then doing you know whatever analytics I'm sure Facebook is very into doing that type of thing for advertising as we've noted and stuff like that um but it's becoming kind of a bigger and bigger generalized HBC use case like I said being able to look at these massive connection graphs and um do meaningful analytics on them from the social media sites or what have you yeah I would throw City Planning City Planning is another area and if you just think uh planning where
cell phone towers go or where those those uh access points are so that there's adequate signal and coverage um you've got Acoustics you know how sound can be Amplified inside a city or dampened inside of a city you know lots I mean there's this all kinds of ways it can be used and is used it's probably underused in in a lot of areas that are you know impact to our daily lives yeah 5G and microcells have made that a whole nother thing putting a MicroCell in every bus stop because they've run fiber right next to it I'm sorry whatever looking at those maps for a
while myself social and and Facebook and high performance Computing I I've just showed me this so I think it was my cousin she had posted this picture of herself right next to her boyfriend and then it like said I'll love you forever or something like that and then I went and it aged him like 50 years so there's like her standing next to him and you know what I mean and it I was like that is so weird right and I was just thinking like that might be a good motivation for people where it's like hey if you were to quit smoking I mean say
there's you know smoker or whatever like doing but some because some kind of have it right now like this is how you would age if you continue smoking this is how you would Age right like there's there's so many little different things and like what I was thinking about is you know these last few years have been rough and not to get into any like controversy about you know people's feelings uh you know perceptions about that but there were some consequences to the actions that we took regardless right there's like a lot of um like people human beings just don't do very well when we
are isolated from each other and like afraid of each other right like even though if it was a good reason or whatever but like is like what if we could like simulate that kind of stuff like like the the pros and cons and like is this action maybe worth the potential consequences um and can we project that out there's like there's like a lot of different questions right here but I would I'll chime in first I definitely think we can uh I get back to the question of uh how do you trust the answer that comes out and that takes I mean that's that's the
the human element of of high performance Computing we have to have the researchers that are aware you know that they have to vet the answers and you know if you look at Academia and that model um we we my opinion I'll say it is an opinion we get into trouble when people uh take an answer just throw it out as fact without having it vetted and that's where Academia has the peer review system um you you can put your ideas out there but other people have to be able to recreate and agree with the answer and come to a consensus and and you know you're
never going to get a hundred percent but you you you try to vet those answers to make sure that you're credible in having a model because again we are talking about a digitized element of an analog world where you don't have an exact answer right what comes out of a simulation is a simulation it is not exact you can always find a flaw somewhere and I yeah an example is I recently saw a researcher who is working on you know his material scientist designed you know simulated the material then actually manufactured it what is manufactured has flaws because of the manufacturing process so it's different
than what what was simulated and now you have to look at the simulation is a pure model and the the real the real world object is not so the two are different right you are actually now with dealing with two different things right and you have to figure out what's an acceptable element what's that acceptable error bar right and so those are the things of yeah we could have simulated what a policy might have meant and done but you also have to make sure that you don't put data out there that may not be that accurate and there was three it it came up I'm
not gonna I'm not gonna say what it was but an agency basically said we now believe this uh this data point to be true with a low confidence level what does that mean it means it means yes they simulated it and it could be true but we're not very confident in the answer so you know it's feasible it may not be true we they do not have the confidence level in how they got to that answer um but you know they chose to put it out there so it's it's a dilemma it's what we face um there's no right or wrong answer necessarily but it's
tough so three times going in the direction of sorry yeah of the high performance Computing and the the the it's exciting and it's still new and there's lots of potential um and this idea that the that the software that kind of is is giving whatever scientists or people or someone with an idea access to the compute resources that that ability to then share the data and share the information and have it be reproducible I mean that that sounds like it's really going in the direction of what we are doing at ciq is making that accessible to the people who want to use it so Mark
I was going to ask on the on the one where they created a material did they go back and retrain the model and input what they had learned so it would be more accurate the next time or is that um seems like it'd be a good idea to do well it's it this is not a retrain this is you have to build that model for the simulation and it's it's actually uh a huge problem because how does this researcher create that model that has all these individual imperfections at the the microscopic level by here right um you would essentially have to be able to take
that material have something that can scan it like you know like cryo em does uh to create that computational model that can then be simulated and so the you know these are the problems that these researchers are in fact working on trying to help solve and uh while I can't I can't say exactly what it is because it was it's a private conference but this researcher is working on bone growth right literally on a sense that could help people across the planet heal after major major trauma either from surgery or from accidents your broken bones instead of having titanium implants screwed into the bone there
could be materials injected that actually grow the boat faster very interesting it's cool stuff so Forrest I know you wanted to talk about the differences between Enterprise and HPC and I think we're starting to see them come together but I I know you have thoughts on this topic yeah so I just wanted to touch on this a little bit we talked a little bit about websites um that type of stuff is something that's done in Computing I attend a class uh you know websites large databases software as a service deployments that type of thing and the categorization of Enterprise Computing these are large-scale workloads that
are serving some resource onto a large amount of people at once but for the most part I mean versus that are being used are fairly generic um something that we have to consider here is like the hyperscale versus HPC because a lot of the time these Enterprise Computing deployments will take social media sites as we've discussed them as an example A lot of times the Enterprise Computing that they're doing requires such an incredible scale that it can be hard to differentiate it from high performance computing one thing that I kind of look at there is you know with the hyperscale you are essentially trying to
serve one resource out the thousands of users at once with HPC you're trying to use all those resources to create something for a much much smaller group of people there aren't a thousand they're hypothetically aren't you know thousands of people waiting on the simulation it's probably you know one small group's work that's being done there um in general so that's kind of a hyperscale back to Enterprise Computing a little bit in general one of the biggest divergences between Enterprise Computing and HPC is in the tooling and stuff like uh kind of the Automation and platforms and things like that that are available that are available
to them we see an Enterprise that really commonly for a long time now there's um you know ansible uh containerization CI CD platforms there's all these different tools that are used ubiquitously to manage these massive deployments um and that are just they they are everywhere those kubernetes container orchestration that type of stuff all these major sites most Enterprise Computing relies on these Technologies but for the most part these Technologies are unheard of in HPC on the converse the scale of some of the Enterprise Computing like around I don't want to call Enterprise Computing AI here but the AI Computing that they are doing in Enterprise
is reaching a scale that they are starting to need HPC class systems to be able to work with that and so the end result of this is that we have Enterprise wanting HPC class resources we have HPC wanting Enterprise class tooling and there's kind of a um there's a bit of you know a divide between these two things um at ciq our answer to this is our platform working on fuzzball which aims to bring Enterprise and high performance Computing together in a way that allows HPC to use container orchestration and cice platforms uh container or cluster deployments that rely on things like ansible Etc um
and so it's a very very interesting space at the moment like I said we touched on social media sites website so I wanted to mention that for a moment there's a very very interesting divide that has been growing between Enterprise and HPC that you know they're kind of both doing the thing that the other one wants uh and so it'll be very interesting to see kind of how those come back together and like I said one of the biggest things that we work on at ciq is bringing that Divergence uh to a meaningful resolution for everybody involved so yeah very interesting space yeah Shameless plug
for us I love it I mean but it is that's about it because it's and there is a need to to bring the two together and that was kind of where I was going out with uh Brock like hey if it's an on HP scene what is it it's like well we are kind of blending so it's it's less and less distinctive um so thank you for bringing that up for us and it really kind of ties into the next question that I've been thinking about here like where are we going right so we talked a lot about like what it is and and where
we are now and what is the future of HPC as you see it you start with Brock well I I think it um I'm gonna reference what Gary brought up earlier uh you know 20 years ago it definitely was the researchers themselves writing codes to solve these highly complex problems um you see that less and less and that's a good thing because these researchers are usually not computer scientists they do not know the details of Hardware architectures cash architectures uh they they wouldn't know what infiniband is they may not even know what a Linux platform is and that's that element of as we move forward
and make this more accessible the domain Specialists stay in their domain and that's really important because these researchers are Specialists looking at these highly complex systems to come up with Solutions and the more that they have to learn something else to do that research the more they are delayed from coming up with their answers so I see the future as that abstraction and more element of software layers allowing them to do more in their specialty areas awesome thank you Brian what's your what's your vision where you see HPC going uh kind of adding on to what Brock was talking about like in the future I
I see just better tools and better Hardware uh coming out and uh with these two things uh users will be able to get results a lot faster and that's what they care about and uh if you're able to get results faster it means you you can make design decisions faster and if you can do that you could probably get to Market a lot faster before everyone else and as a result of that we'll just see cooler new products coming out that will hopefully improve our lives yeah I like it bigger better faster more [Laughter] of this year's Super Computing conference is I am HPC okay
claiming the affirmation I love it thanks David Gary do you have a vision where you see HBC going um yeah you know I think we're just getting started I I really think we're just getting started and maybe Rose the you know what you're saying about like how it could be used you know I I think what would be nice is if it could be used in a way that people don't have to think about it so they you know like for example that app where somebody could just click on this thing and see how they age uh you know I I could see HPC you
know just behind the scenes is just becoming part of the infrastructure where people just utilize it and they're able to help them make better decisions by hopefully making better decisions by using it so um I think there's a lot a lot of work to get there I agree with Brock in that uh you know the domain Specialists I can see it now a lot of them don't want to do the coding and develop their own codes and but there's a little bit of a dilemma there because the people who have done that have been critical to like making HPC successful so I'm not sure what
the model is going to be in order to achieve uh you know achieve to continue to achieve what we have and um so that's that's going to be tough and as far as just HPC and Enterprise um you know not that long ago with just ABC with just the computation and then the next big thing was data and you know data is kind of you know we're people are talking about it as a first-class citizen but I don't know that that's actually true yet in a lot of HPC centers and um but just merging all of that with AI I think is there's still a
lot of interesting things to be invented and put into practice cool thanks guys that's all the questions I have I don't know if there's anything else that you want to bring up saying or Force you got you got you wanna Vision I was just gonna sum up you know essentially what I had said there at the moment HBC is going nothing better or it's going nowhere but I'm sorry um going nowhere but uh HPC is everything you know as we've discussed here from really high level simulations to you know the weather modeling that tells Public Utilities how much power they're going to have available for
maneuverables and stuff like that so you know as we've noted HPC touches so many aspects of Life uh and the use cases that are supporting you know all of this technology and civilization and stuff are only getting more complex so you know at the moment HBC is going nowhere but up there's a complete Renaissance and hardware and software at the moment with new Chips new devices uh new appliances um in software new paradigms like I touched on the combining of Enterprise and HPC together um that ultimately allow users to do more work faster and focus more so on their domain signs and less so on
you know system internals and that type of thing um so you know ultimately to kind of reiterate what everyone else has said here it's uh it's a very great time for HBC and it's very interesting to see where it's going on all fronts I think it's a great way to wrap up for us it's I agree with you it is very interesting to see where things are headed guys we are up on time really appreciate your time Gary it's always good to see you thank you for joining us Brian brockforce thank you Rose it's been great I appreciate it you guys if you don't mind
going and liking and subscribing to the ciq channel we would really appreciate it and we will see you again next week thank you awesome thanks guys [Music]
Built for scale. Chosen by the world’s best.
2.75M+
Rocky Linux instances
Being used world wide
90%
Of fortune 100 companies
Use CIQ supported technologies
250k
Avg. monthly downloads
Rocky Linux
Have questions about your infrastructure?
Talk to a CIQ engineer about Rocky Linux, HPC, and AI infrastructure.