Dreaming of building an AI R&D lab? You'll need deep pockets: Tax filing reveals millions of bucks OpenAI alone spent on cloud ML compute

No wonder it split off its non-profit arm to make money

The latest tax return forms filed by neural-network boffinry center OpenAI this year shows just how expensive it is to run an independent AI research institute without the financial backing of giant tech orgs.

The San Francisco-based lab recently effectively overturned its non-profit status to kickstart a for-profit spinoff to compete with the likes of Google, Facebook, Amazon, Apple, and so on. OpenAI is now split into two parts: OpenAI LP, the bigger money-making side, and OpenAI Nonprofit, the smaller portion of the outfit focused on things like its OpenAI Scholars educational program. The ultimate goal is to develop artificial general intelligence (AGI), which is no easy task.

Research and development in modern AI is costly. Neural networks typically perform well on specific tasks like image recognition or language translation after being trained on heaps of data, and then you usually have to start all over again for another task. Researchers often have to retrain their models several times to tweak its performance. All of this requires heavy computing resources, and renting out CPUs, GPUs, TPUs, whatever, you name it, is pricey. Buying the hardware and space for it all outright is a non-starter for independent or fledgling labs.

OpenAI’s income tax exemption form, filed [PDF] in March for 2017 when it was still a non-profit, revealed that it splashed out on a whopping $7.95m on cloud computing expenses that year. That’s more than double compared to the $2.33m spent in 2016. We were alerted to the filings this week by an industry source.

Some of OpenAI’s largest projects such as OpenAI Five, the Dota-2 playing bots and GPT-2 and the giant language model, together enlisted more than 100,000 CPU cores and hundreds of GPUs and TPUs.

There’s also another money sucking component in AI research, too. Talent is hard to come by and companies compete with one another to win over the best researchers in the field. They offer premium packages, complete with excellent health benefits and eye-popping salaries.


OpenAI retires its Dota-2 playing bots after crushing e-sport pros one last time


Ilya Sutskever, research director, remained the highest-paid OpenAI employee. In 2017, he topped the list at $748,908, whilst Greg Brockman, co-founder and CTO banked $264,201. John Schulman, a senior research was second with $718,728, and Pieter Abbeel was next with $658,077. Other notable names include Andrej Karpathy and Diederik Kingma, who were paid $424,231 and $545,833. The total wages dished out to the rest of the company was just over $13.2m for the year.

Those salaries are significantly greater compared to the previous year. Sutskever was the only exception. He was paid a hefty $900,000 base wage on top of a $1m signing bonus. Abbeel, Karpathy, and Kingma have left OpenAI since, which gives you an idea of staff turnover at these AI research labs and Bay Area tech companies in general.

There are also other costs, such as rent and equipment that contributed to the total expenses cost of over $28.6m. Considering, OpenAI was given about $33.2m in grants and a $3m loan from Sam Altman, its CEO, the research lab burns a lot of its money through expenses, leaving it with just over $4.5m to spare.

OpenAI declined to comment on the loan and its salaries. Obviously, the likes of Google and Facebook are spending orders of magnitude more on machine-learning technology, though they have buckets of ad cash and other revenue to throw at AI-based products, whereas OpenAI is independently focused on achieving AGI.

“We’ve experienced firsthand that the most dramatic AI systems use the most computational power in addition to algorithmic innovations, and decided to scale much faster than we’d planned when starting OpenAI,” the organization previously said. “We’ll need to invest billions of dollars in upcoming years into large-scale cloud compute, attracting and retaining talented people, and building AI supercomputers.”

It appears the latest tax filing was rejected by California’s bureaucrats because the outfit had not yet completed a required independent financial audit. “The state did not deem our submission complete because our third-party financial audit is still wrapping up,” an OpenAI spokesperson told The Register on Thursday. “We expect to finish it in the next few months, at which point we will submit those additional filings.” ®

Similar topics

Other stories you might like

  • Lonestar plans to put datacenters in the Moon's lava tubes
    How? Founder tells The Register 'Robots… lots of robots'

    Imagine a future where racks of computer servers hum quietly in darkness below the surface of the Moon.

    Here is where some of the most important data is stored, to be left untouched for as long as can be. The idea sounds like something from science-fiction, but one startup that recently emerged from stealth is trying to turn it into a reality. Lonestar Data Holdings has a unique mission unlike any other cloud provider: to build datacenters on the Moon backing up the world's data.

    "It's inconceivable to me that we are keeping our most precious assets, our knowledge and our data, on Earth, where we're setting off bombs and burning things," Christopher Stott, founder and CEO of Lonestar, told The Register. "We need to put our assets in place off our planet, where we can keep it safe."

    Continue reading
  • Conti: Russian-backed rulers of Costa Rican hacktocracy?
    Also, Chinese IT admin jailed for deleting database, and the NSA promises no more backdoors

    In brief The notorious Russian-aligned Conti ransomware gang has upped the ante in its attack against Costa Rica, threatening to overthrow the government if it doesn't pay a $20 million ransom. 

    Costa Rican president Rodrigo Chaves said that the country is effectively at war with the gang, who in April infiltrated the government's computer systems, gaining a foothold in 27 agencies at various government levels. The US State Department has offered a $15 million reward leading to the capture of Conti's leaders, who it said have made more than $150 million from 1,000+ victims.

    Conti claimed this week that it has insiders in the Costa Rican government, the AP reported, warning that "We are determined to overthrow the government by means of a cyber attack, we have already shown you all the strength and power, you have introduced an emergency." 

    Continue reading
  • China-linked Twisted Panda caught spying on Russian defense R&D
    Because Beijing isn't above covert ops to accomplish its five-year goals

    Chinese cyberspies targeted two Russian defense institutes and possibly another research facility in Belarus, according to Check Point Research.

    The new campaign, dubbed Twisted Panda, is part of a larger, state-sponsored espionage operation that has been ongoing for several months, if not nearly a year, according to the security shop.

    In a technical analysis, the researchers detail the various malicious stages and payloads of the campaign that used sanctions-related phishing emails to attack Russian entities, which are part of the state-owned defense conglomerate Rostec Corporation.

    Continue reading
  • FTC signals crackdown on ed-tech harvesting kid's data
    Trade watchdog, and President, reminds that COPPA can ban ya

    The US Federal Trade Commission on Thursday said it intends to take action against educational technology companies that unlawfully collect data from children using online educational services.

    In a policy statement, the agency said, "Children should not have to needlessly hand over their data and forfeit their privacy in order to do their schoolwork or participate in remote learning, especially given the wide and increasing adoption of ed tech tools."

    The agency says it will scrutinize educational service providers to ensure that they are meeting their legal obligations under COPPA, the Children's Online Privacy Protection Act.

    Continue reading
  • Mysterious firm seeks to buy majority stake in Arm China
    Chinese joint venture's ousted CEO tries to hang on - who will get control?

    The saga surrounding Arm's joint venture in China just took another intriguing turn: a mysterious firm named Lotcap Group claims it has signed a letter of intent to buy a 51 percent stake in Arm China from existing investors in the country.

    In a Chinese-language press release posted Wednesday, Lotcap said it has formed a subsidiary, Lotcap Fund, to buy a majority stake in the joint venture. However, reporting by one newspaper suggested that the investment firm still needs the approval of one significant investor to gain 51 percent control of Arm China.

    The development comes a couple of weeks after Arm China said that its former CEO, Allen Wu, was refusing once again to step down from his position, despite the company's board voting in late April to replace Wu with two co-chief executives. SoftBank Group, which owns 49 percent of the Chinese venture, has been trying to unentangle Arm China from Wu as the Japanese tech investment giant plans for an initial public offering of the British parent company.

    Continue reading
  • SmartNICs power the cloud, are enterprise datacenters next?
    High pricing, lack of software make smartNICs a tough sell, despite offload potential

    SmartNICs have the potential to accelerate enterprise workloads, but don't expect to see them bring hyperscale-class efficiency to most datacenters anytime soon, ZK Research's Zeus Kerravala told The Register.

    SmartNICs are widely deployed in cloud and hyperscale datacenters as a means to offload input/output (I/O) intensive network, security, and storage operations from the CPU, freeing it up to run revenue generating tenant workloads. Some more advanced chips even offload the hypervisor to further separate the infrastructure management layer from the rest of the server.

    Despite relative success in the cloud and a flurry of innovation from the still-limited vendor SmartNIC ecosystem, including Mellanox (Nvidia), Intel, Marvell, and Xilinx (AMD), Kerravala argues that the use cases for enterprise datacenters are unlikely to resemble those of the major hyperscalers, at least in the near term.

    Continue reading

Biting the hand that feeds IT © 1998–2022