AMD tries to spoil Nvidia's week by teasing high-end accelerators, Epyc chips with 3D L3 cache, and more

Microsoft cloud first to privately preview Milan-X parts

AMD teased a bunch of enterprise-class developments today, from claiming the AI and HPC accelerator crown to word of 3D-stacked cache, all perhaps to spoil rival Nvidia's forthcoming GPU conference.

Here's a summary of what AMD said on Monday ahead of Nv's GTC this week:

3D-stacked Level 3 cache

AMD said it has crafted a server microprocessor code-named Milan-X, which is a third-generation Epyc with a layer of SRAM stacked on top of dies within the IC package. It's likely Milan-X will, as far as customers are concerned, be a handful of SKUs extending today's 7nm Epyc 7003s, and will sport up to 64 cores per socket without any architectural changes.

The added layer of memory in the Milan-X is essentially a vertically positioned L3 cache that AMD calls the V-Cache. The chip designer gave us a glimpse of this approach in June albeit with the RAM stacked in a Ryzen prototype part. AMD co-designed the 3D structure with TSMC, which manufactures the components.

The additional layer of L3 cache sits right over the L3 cache on the CPU dies, adding 64MB of SRAM to the 32MB below resulting in a roughly 10 per cent increase in overall cache latency. AMD said the additional cache will increase performance of the processor for certain engineering and scientific workloads, at least.

The Milan-X can be a drop-in replacement for third-gen Epycs, with some BIOS updates required to make use of the added technology, according to AMD. Microsoft's Azure cloud is the first to offer access to these stacked processors in a private preview, and it is expected to widen this service over the next few weeks.

If you want to get your hands on the physical silicon, be aware it is not due to launch until the first quarter of next year. Cisco, Dell, Lenovo, HPE, Supermicro, and others are expected to sell data-center systems featuring the chips.

For a top-end Milan-X SKU, you're looking at up to 768 MB of L3 cache total per socket, or 1.5GB per dual-socket system.

You can find more info and analysis of the Milan-X here, and what it means for Microsoft here, by our friends at The Next Platform.

AMD claims accelerator crown

AMD launched its MI200 series of accelerators for AI and HPC systems, and claimed they are "the most advanced" of their kind in the world. AMD reckons its devices can outdo the competition (ie, Nvidia) in terms of FP64 performance in supercomputer applications, and top 380 tera-FLOPS of peak theoretical half-precision (FP16) performance for machine-learning applications.

The US Dept of Energy's Oak Ridge Laboratory is going to use the hardware, along with third-gen Epyc processors, in a 1.5 exa-FLOPS HPE-built super called Frontier, according to AMD. Frontier is due to power up next year.

The MI200 series uses AMD's CDNA 2 architecture and multiple GPU dies within a package, and will come in two form factors: the MI250X and MI250 as OCP accelerator modules, and the MI210 as a PCIe card. The given specifications are as follows:

Table of the MI200-series specifications

AMD's specifications for the MI200 series

The MI250X is available now if you're in the market for an HPE Cray EX Supercomputer; otherwise, you'll have to wait until Q1 2022 to get the add-ons from a system builder.

Zen 4 roadmap teased

AMD also didn't want you to forget about its upcoming Zen 4 processor families. We're told to look out for a 96-core 5nm Zen 4-based server chip code-named Genoa. This will support DDR5 memory, CXL interconnects, and PCIe 5 devices, and will launch next year for data-center-class environments, according to AMD.

Then there's the Bergamo, a 5nm 128-core Zen 4c component due to arrive in 2023. The c in 4c means it's optimized for cloud providers – higher core density and performance per socket.

You can find more from AMD here, and more analysis coming up on The Next Platform here. ®

Similar topics

Narrower topics

Other stories you might like

  • Lonestar plans to put datacenters in the Moon's lava tubes
    How? Founder tells The Register 'Robots… lots of robots'

    Imagine a future where racks of computer servers hum quietly in darkness below the surface of the Moon.

    Here is where some of the most important data is stored, to be left untouched for as long as can be. The idea sounds like something from science-fiction, but one startup that recently emerged from stealth is trying to turn it into a reality. Lonestar Data Holdings has a unique mission unlike any other cloud provider: to build datacenters on the Moon backing up the world's data.

    "It's inconceivable to me that we are keeping our most precious assets, our knowledge and our data, on Earth, where we're setting off bombs and burning things," Christopher Stott, founder and CEO of Lonestar, told The Register. "We need to put our assets in place off our planet, where we can keep it safe."

    Continue reading
  • Conti: Russian-backed rulers of Costa Rican hacktocracy?
    Also, Chinese IT admin jailed for deleting database, and the NSA promises no more backdoors

    In brief The notorious Russian-aligned Conti ransomware gang has upped the ante in its attack against Costa Rica, threatening to overthrow the government if it doesn't pay a $20 million ransom. 

    Costa Rican president Rodrigo Chaves said that the country is effectively at war with the gang, who in April infiltrated the government's computer systems, gaining a foothold in 27 agencies at various government levels. The US State Department has offered a $15 million reward leading to the capture of Conti's leaders, who it said have made more than $150 million from 1,000+ victims.

    Conti claimed this week that it has insiders in the Costa Rican government, the AP reported, warning that "We are determined to overthrow the government by means of a cyber attack, we have already shown you all the strength and power, you have introduced an emergency." 

    Continue reading
  • China-linked Twisted Panda caught spying on Russian defense R&D
    Because Beijing isn't above covert ops to accomplish its five-year goals

    Chinese cyberspies targeted two Russian defense institutes and possibly another research facility in Belarus, according to Check Point Research.

    The new campaign, dubbed Twisted Panda, is part of a larger, state-sponsored espionage operation that has been ongoing for several months, if not nearly a year, according to the security shop.

    In a technical analysis, the researchers detail the various malicious stages and payloads of the campaign that used sanctions-related phishing emails to attack Russian entities, which are part of the state-owned defense conglomerate Rostec Corporation.

    Continue reading
  • FTC signals crackdown on ed-tech harvesting kid's data
    Trade watchdog, and President, reminds that COPPA can ban ya

    The US Federal Trade Commission on Thursday said it intends to take action against educational technology companies that unlawfully collect data from children using online educational services.

    In a policy statement, the agency said, "Children should not have to needlessly hand over their data and forfeit their privacy in order to do their schoolwork or participate in remote learning, especially given the wide and increasing adoption of ed tech tools."

    The agency says it will scrutinize educational service providers to ensure that they are meeting their legal obligations under COPPA, the Children's Online Privacy Protection Act.

    Continue reading
  • Mysterious firm seeks to buy majority stake in Arm China
    Chinese joint venture's ousted CEO tries to hang on - who will get control?

    The saga surrounding Arm's joint venture in China just took another intriguing turn: a mysterious firm named Lotcap Group claims it has signed a letter of intent to buy a 51 percent stake in Arm China from existing investors in the country.

    In a Chinese-language press release posted Wednesday, Lotcap said it has formed a subsidiary, Lotcap Fund, to buy a majority stake in the joint venture. However, reporting by one newspaper suggested that the investment firm still needs the approval of one significant investor to gain 51 percent control of Arm China.

    The development comes a couple of weeks after Arm China said that its former CEO, Allen Wu, was refusing once again to step down from his position, despite the company's board voting in late April to replace Wu with two co-chief executives. SoftBank Group, which owns 49 percent of the Chinese venture, has been trying to unentangle Arm China from Wu as the Japanese tech investment giant plans for an initial public offering of the British parent company.

    Continue reading
  • SmartNICs power the cloud, are enterprise datacenters next?
    High pricing, lack of software make smartNICs a tough sell, despite offload potential

    SmartNICs have the potential to accelerate enterprise workloads, but don't expect to see them bring hyperscale-class efficiency to most datacenters anytime soon, ZK Research's Zeus Kerravala told The Register.

    SmartNICs are widely deployed in cloud and hyperscale datacenters as a means to offload input/output (I/O) intensive network, security, and storage operations from the CPU, freeing it up to run revenue generating tenant workloads. Some more advanced chips even offload the hypervisor to further separate the infrastructure management layer from the rest of the server.

    Despite relative success in the cloud and a flurry of innovation from the still-limited vendor SmartNIC ecosystem, including Mellanox (Nvidia), Intel, Marvell, and Xilinx (AMD), Kerravala argues that the use cases for enterprise datacenters are unlikely to resemble those of the major hyperscalers, at least in the near term.

    Continue reading

Biting the hand that feeds IT © 1998–2022