HPC

This article is more than 1 year old

Oak Ridge goes gaga for Nvidia GPUs

Fermi chases Cell for HPC dough

Thu 1 Oct 2009 // 22:06 UTC

Hybrid futures

It would not be at all surprising to see a hybrid architecture for the future Oak Ridge machine that uses PCI-Express 2.0 links to hook Fermi GPUs into Opteron server nodes, just like IBM is using PCI-Express 1.0 links to hook Cell boards into the Opteron nodes with Roadrunner.

Jeff Nichols, associate lab director for computing and computational sciences at Oak Ridge, said in a statement that the Fermi GPUs, which have eight times the double precision floating point performance as the Teslas, at around 500 gigaflops, would enable "substantial scientific breakthroughs that would be impossible without the new technology."

Working with the future Tesla GPU coprocessors and their successors, Oak Ridge is hoping to push up into the exaflops barrier within ten years. Getting to 10 petaflops next year with a parallel super that uses the Fermi GPUs is just a down payment.

The important thing about the Fermi GPUs is that the CUDA programming environment from nVidia supports not just C, but C++ as well. When Fortran compilers can see and dispatch work to the GPUs, the combination of decent double-precision performance and C++ and Fortran support will truly push GPU co-processors into the mainstream. This is exactly what Nvidia, AMD (with its FireStream GPUs), and Intel (with its Larrabee GPUs) are all hoping for.

Cell multiplication

The question now is, what will Big Blue do to counter these moves onto its hybrid supercomputing turf?

Several years back, when the Cell chips were first being commercialized and offered terrible double-precision floating-point performance (like 42 gigaflops versus 460 gigaflops for a two-socket Cell blade), Big Blue's roadmap called for a Cell board with two sockets that could deliver 460 gigaflops of single-precision and 217 gigaflops of double-precision math. We know this blade server as the BladeCenter QS22.

The roadmap also called for a BladeCenter QS2Z, which would have Cell chips that in turn had two Power cores and a whopping 32 vector processors each, using a next-generation memory and interconnection technology; the QS2Z blade would sport 2 teraflops per blade at single precision and 1 teraflops per blade at double precision.

That's about twice the oomph in a Cell chip compared to the forthcoming Fermi GPUs. Oak Ridge knew that, of course, but maybe this future Cell chip never made it out of the concept stage, as it was in early 2007.

IBM is mum on its Cell roadmap plans at this point, but this future Cell chip was slated for delivery in the first half of 2010, more or less concurrent with the Fermi GPU co-processors. ®

Topics

Special Features

Vendor Voice

Resources

HPC

Oak Ridge goes gaga for Nvidia GPUs

Fermi chases Cell for HPC dough

Hybrid futures

Cell multiplication

More about

More about

Narrower topics

Broader topics

More about

More about

More about

Narrower topics

Broader topics

TIP US OFF

Other stories you might like

Intel Gaudi's third and final hurrah is an AI accelerator built to best Nvidia's H100

AI cloud startup TensorWave bets AMD can beat Nvidia

Los Alamos Lab powers up Nvidia-laden Venado supercomputer

Industrial systems integrating digitalisation

With Run:ai acquisition, Nvidia aims to manage your AI kubes

Banned Nvidia GPUs sneak into sanction-busting Chinese servers

China scientists talk of powering hypersonic weapon with cheap Nvidia chip

Microsoft foresees a new type of AI PC: A Surface designed with help from machines

Lambda borrows half a billion bucks to grow its GPU cloud

AI bubble or not, Nvidia is betting everything on a GPU-accelerated future

Intel's neuromorphic 'owl brain' swoops into Sandia labs

China's mega-telcos are spending billions on AI servers

About Us

Our Websites

Your Privacy