Former Intel AI boss Naveen Rao is now counting the cost of machine learning, literally

MosaicMLdelving into the details


A former head of artificial intelligence products at Intel has started a company to help companies cut overhead costs on AI systems.

Naveen Rao, CEO and co-founder of MosaicML, previously led Nervana Systems, which was acquired by Intel for $350m. But like many Intel acquisitions, the marriage didn't pan out, and Intel killed the Nervana AI chip last year, after which Rao left the company.

MosaicML's open source tools focus on implementing AI systems based on cost, training time, or speed-to-results. They do so by analyzing an AI problem relative to the neural net settings and hardware, which then paves an efficient path to generate optimal settings while reducing electric costs.

One component is Composer, which provides the building blocks on which AI applications can be efficiently trained. MosaicML developed these methods after months of researching common settings in computer vision models that include ResNets and natural language processing models like Transformer and GPT.

But developers will ultimately need to chose the best approach. That is where the second component, Explorer, steps in. The tool has a visual interface that provides fine-grained details on parameters that include better results, training time or cost, and users can filter results by the hardware type, cloud and technique.

"We change the learning algorithms themselves to make them use less compute to arrive at the result," Rao told The Register.

AI systems can be inefficient and costly, and more thought needs to be put into economizing machine learning, Rao said. "We find Nvidia GPUs give us the fastest and easiest way to get going. We plan on adding support for other chips in the future," he explained.

The library works within PyTorch right now, and support for Tensorflow will be added later, Rao said.

AI isn't a one-size fits all approach, and inefficiencies in both software and hardware are considerations when accounting for total cost of ownership, said Dan Hutcheson, analyst a VLSI Research.

"The amount of computation required to train the largest models is estimated to be growing >5x every year, yet hardware performance per dollar is growing at only a fraction of that rate," MosaicML said in a blog, citing a 2018 study by OpenAI.

OpenAI in a study last year said algorithmic progress have shown further AI speedups than hardware efficiency.

Many systems use racks of power-hungry Nvidia GPUs for machine learning. Rao is a proponent of this distributed approach, with AI processing split over a network of cheaper chips and components that include low-cost DDR memory and PCI-Express interconnects.

In a tweetstorm last week, he took a jab at monolithic AI chips like the WSE-2 chip produced by Cerebras Systems Inc. being inefficient in AI relative to performance-per-dollar.

The distributed approach reflects a fundamental flaw in understanding the cost of doing AI at a chip level and scaling performance, Cerebras CEO Andrew Feldman told The Register.

"The real waste - it's got nothing to do with the individual chip level. To get 10x the performance you're going to spend 100 or 1,000 times the power," Feldman said.

Feldman invoked Moore's Law, saying "What Intel showed over decades is that you can build a great business if you can keep your prices flat and doubling performance every three to four years."

One angel investor in MosaicML, Steve Jurvetson of Future Ventures, in a tweet floated the idea of a "Mosaic's Law" corollary to measuring advances in algorithms per dollar spent.

Venture capitalists have poured $37m into MosaicML, with other investors also including Lux Capital, DCVC Playground Global, AME, Correlation and E14. ®

Similar topics


Other stories you might like

  • Google sours on legacy G Suite freeloaders, demands fee or flee

    Free incarnation of online app package, which became Workplace, is going away

    Google has served eviction notices to its legacy G Suite squatters: the free service will no longer be available in four months and existing users can either pay for a Google Workspace subscription or export their data and take their not particularly valuable businesses elsewhere.

    "If you have the G Suite legacy free edition, you need to upgrade to a paid Google Workspace subscription to keep your services," the company said in a recently revised support document. "The G Suite legacy free edition will no longer be available starting May 1, 2022."

    Continue reading
  • SpaceX Starlink sat streaks now present in nearly a fifth of all astronomical images snapped by Caltech telescope

    Annoying, maybe – but totally ruining this science, maybe not

    SpaceX’s Starlink satellites appear in about a fifth of all images snapped by the Zwicky Transient Facility (ZTF), a camera attached to the Samuel Oschin Telescope in California, which is used by astronomers to study supernovae, gamma ray bursts, asteroids, and suchlike.

    A study led by Przemek Mróz, a former postdoctoral scholar at the California Institute of Technology (Caltech) and now a researcher at the University of Warsaw in Poland, analysed the current and future effects of Starlink satellites on the ZTF. The telescope and camera are housed at the Palomar Observatory, which is operated by Caltech.

    The team of astronomers found 5,301 streaks leftover from the moving satellites in images taken by the instrument between November 2019 and September 2021, according to their paper on the subject, published in the Astrophysical Journal Letters this week.

    Continue reading
  • AI tool finds hundreds of genes related to human motor neuron disease

    Breakthrough could lead to development of drugs to target illness

    A machine-learning algorithm has helped scientists find 690 human genes associated with a higher risk of developing motor neuron disease, according to research published in Cell this week.

    Neuronal cells in the central nervous system and brain break down and die in people with motor neuron disease, like amyotrophic lateral sclerosis (ALS) more commonly known as Lou Gehrig's disease, named after the baseball player who developed it. They lose control over their bodies, and as the disease progresses patients become completely paralyzed. There is currently no verified cure for ALS.

    Motor neuron disease typically affects people in old age and its causes are unknown. Johnathan Cooper-Knock, a clinical lecturer at the University of Sheffield in England and leader of Project MinE, an ambitious effort to perform whole genome sequencing of ALS, believes that understanding how genes affect cellular function could help scientists develop new drugs to treat the disease.

    Continue reading

Biting the hand that feeds IT © 1998–2022