Close Menu
TechCentralTechCentral

    Subscribe to the newsletter

    Get the best South African technology news and analysis delivered to your e-mail inbox every morning.

    Facebook X (Twitter) YouTube LinkedIn
    WhatsApp Facebook X (Twitter) LinkedIn YouTube
    TechCentralTechCentral
    • News
      ANC's attack on Solly Malatsi shows how BEE dogma trumps economic reality

      ANC’s attack on Solly Malatsi shows how BEE dogma trumps economic reality

      14 December 2025
      Political war erupts over BEE in the ICT sector - Solly Malatsi

      Political war erupts over BEE in the ICT sector

      13 December 2025
      Icasa told to align on BEE in move that will favour Starlink - Solly Malatsi

      Icasa told to align on BEE in move that will favour Starlink

      12 December 2025
      South African solar industry faces a reality check

      South African solar industry faces a reality check

      12 December 2025
      OpenAI launches GPT-5.2 after 'code red' push to counter Google. Shelby Tauber/Reuters

      OpenAI launches GPT-5.2 after ‘code red’ push to counter Google

      12 December 2025
    • World
      Oracle’s AI ambitions face scrutiny on earnings miss

      Oracle’s AI ambitions face scrutiny on earnings miss

      11 December 2025
      China will get Nvidia H200 chips - but not without paying Washington first

      China will get Nvidia H200 chips – but not without paying Washington first

      9 December 2025
      IBM reportedly close to $11-billion deal to buy Confluent - Arvind Krishna

      IBM reportedly close to $11-billion deal to buy Confluent

      8 December 2025
      Amazon and Google launch multi-cloud service for faster connectivity

      Amazon and Google launch multi-cloud service for faster connectivity

      1 December 2025
      Google makes final court plea to stop US breakup

      Google makes final court plea to stop US breakup

      21 November 2025
    • In-depth
      Black Friday goes digital in South Africa as online spending surges to record high

      Black Friday goes digital in South Africa as online spending surges to record high

      4 December 2025
      Canal+ plays hardball - and DStv viewers feel the pain

      Canal+ plays hardball – and DStv viewers feel the pain

      3 December 2025
      Jensen Huang Nvidia

      So, will China really win the AI race?

      14 November 2025
      Valve's Linux console takes aim at Microsoft's gaming empire

      Valve’s Linux console takes aim at Microsoft’s gaming empire

      13 November 2025
      iOCO's extraordinary comeback plan - Rhys Summerton

      iOCO’s extraordinary comeback plan

      28 October 2025
    • TCS
      TCS+ | Africa's digital transformation - unlocking AI through cloud and culture - Cliff de Wit Accelera Digital Group

      TCS+ | Cloud without culture won’t deliver AI: Accelera’s Cliff de Wit

      12 December 2025
      TCS+ | How Cloud on Demand helps partners thrive in the AWS ecosystem - Odwa Ndyaluvane and Xenia Rhode

      TCS+ | How Cloud On Demand helps partners thrive in the AWS ecosystem

      4 December 2025
      TCS | MTN Group CEO Ralph Mupita on competition, AI and the future of mobile

      TCS | Ralph Mupita on competition, AI and the future of mobile

      28 November 2025
      TCS | Dominic Cull on fixing South Africa's ICT policy bottlenecks

      TCS | Dominic Cull on fixing South Africa’s ICT policy bottlenecks

      21 November 2025
      TCS | BMW CEO Peter van Binsbergen on the future of South Africa's automotive industry

      TCS | BMW CEO Peter van Binsbergen on the future of South Africa’s automotive industry

      6 November 2025
    • Opinion
      Netflix, Warner Bros deal raises fresh headaches for MultiChoice - Duncan McLeod

      Netflix, Warner Bros deal raises fresh headaches for MultiChoice

      5 December 2025
      BIN scans, DDoS and the next cybercrime wave hitting South Africa's banks - Entersekt Gerhard Oosthuizen

      BIN scans, DDoS and the next cybercrime wave hitting South Africa’s banks

      3 December 2025
      Your data, your hardware: the DIY AI revolution is coming - Duncan McLeod

      Your data, your hardware: the DIY AI revolution is coming

      20 November 2025
      Zero Carbon Charge founder Joubert Roux

      The energy revolution South Africa can’t afford to miss

      20 November 2025
      It's time for a new approach to government IT spend in South Africa - Richard Firth

      It’s time for a new approach to government IT spend in South Africa

      19 November 2025
    • Company Hubs
      • Africa Data Centres
      • AfriGIS
      • Altron Digital Business
      • Altron Document Solutions
      • Altron Group
      • Arctic Wolf
      • AvertITD
      • Braintree
      • CallMiner
      • CambriLearn
      • CYBER1 Solutions
      • Digicloud Africa
      • Digimune
      • Domains.co.za
      • ESET
      • Euphoria Telecom
      • Incredible Business
      • iONLINE
      • IQbusiness
      • Iris Network Systems
      • LSD Open
      • NEC XON
      • Netstar
      • Network Platforms
      • Next DLP
      • Ovations
      • Paracon
      • Paratus
      • Q-KON
      • SevenC
      • SkyWire
      • Solid8 Technologies
      • Telit Cinterion
      • Tenable
      • Vertiv
      • Videri Digital
      • Vodacom Business
      • Wipro
      • Workday
      • XLink
    • Sections
      • AI and machine learning
      • Banking
      • Broadcasting and Media
      • Cloud services
      • Contact centres and CX
      • Cryptocurrencies
      • Education and skills
      • Electronics and hardware
      • Energy and sustainability
      • Enterprise software
      • Financial services
      • Information security
      • Internet and connectivity
      • Internet of Things
      • Investment
      • IT services
      • Lifestyle
      • Motoring
      • Public sector
      • Retail and e-commerce
      • Satellite communications
      • Science
      • SMEs and start-ups
      • Social media
      • Talent and leadership
      • Telecoms
    • Events
    • Advertise
    TechCentralTechCentral
    Home » Sections » AI and machine learning » The insanely powerful supercomputer Microsoft built for AI workloads

    The insanely powerful supercomputer Microsoft built for AI workloads

    When Microsoft invested $1-billion in OpenAI in 2019, it agreed to build a cutting-edge supercomputer for the AI research start-up.
    By Dina Bass13 March 2023
    Twitter LinkedIn Facebook WhatsApp Email Telegram Copy Link
    News Alerts
    WhatsApp

    When Microsoft invested US$1-billion in OpenAI in 2019, it agreed to build a massive, cutting-edge supercomputer for the artificial intelligence research start-up. The only problem: Microsoft didn’t have anything like what OpenAI needed and wasn’t totally sure it could build something that big in its Azure cloud service without it breaking.

    OpenAI was trying to train an increasingly large set of AI programs called models, which were ingesting greater volumes of data and learning more and more parameters, the variables the AI system has sussed out through training and retraining. That meant OpenAI needed access to powerful cloud computing services for long periods of time.

    To meet that challenge, Microsoft had to find ways to string together tens of thousands of Nvidia’s A100 graphics chips — the workhorse for training AI models — and change how it positions servers on racks to prevent power outages. Scott Guthrie, the Microsoft executive vice president who oversees cloud and AI, wouldn’t give a specific cost for the project, but said “it’s probably larger” than several hundred million dollars.

    We built a system architecture that could operate and be reliable at a very large scale.

    “We built a system architecture that could operate and be reliable at a very large scale. That’s what resulted in ChatGPT being possible,” said Nidhi Chappell, Microsoft GM of Azure AI infrastructure. “That’s one model that came out of of it. There’s going to be many, many others.”

    The technology allowed OpenAI to release ChatGPT, the viral chatbot that attracted more than a million users within days of going public in November and is now getting pulled into other companies’ business models. As generative AI tools such as ChatGPT gain interest from businesses and consumers, more pressure will be put on cloud services providers such as Microsoft, Amazon.com and Google to ensure their data centres can provided the enormous computing power needed.

    Now Microsoft uses that same set of resources it built for OpenAI to train and run its own large AI models, including the new Bing search bot introduced last month. It also sells the system to other customers. The software giant is already at work on the next generation of the AI supercomputer, part of an expanded deal with OpenAI in which Microsoft added $10-billion to its investment.

    Better for AI

    “We didn’t build them a custom thing — it started off as a custom thing, but we always built it in a way to generalise it so that anyone that wants to train a large language model can leverage the same improvements,” said Guthrie in an interview. “That’s really helped us become a better cloud for AI broadly.”

    Training a massive AI model requires a large pool of connected graphics processing units in one place like the AI supercomputer Microsoft assembled. Once a model is in use, answering all the queries users pose — called inference — requires a slightly different setup. Microsoft also deploys graphics chips for inference but those processors — hundreds of thousands of them — are geographically dispersed throughout the company’s more than 60 regions of data centres. Now the company is adding the latest Nvidia graphics chip for AI workloads — the H100 — and the newest version of Nvidia’s Infiniband networking technology to share data even faster, Microsoft said on Monday in a blog post.

    The new Bing is still in preview with Microsoft gradually adding more users from a waitlist. Guthrie’s team holds a daily meeting with about two dozen employees they’ve dubbed the “pit crew”, after the group of mechanics that tune race cars in the middle of the race. The group’s job is to figure out how to bring greater amounts of computing capacity online quickly, as well as fix problems that crop up.

    “It’s very much a kind of a huddle, where it’s like, ‘Hey, anyone has a good idea, let’s put it on the table today, and let’s discuss it and let’s figure out, okay, can we shave a few minutes here? Can we shave a few hours? A few days?’” Guthrie said.

    A cloud service depends on thousands of different parts and items — the individual pieces of servers, pipes, concrete for the buildings, different metals and minerals — and a delay or short supply of any one component, no matter how tiny, can throw everything off. Recently, the pit crew had to deal with a shortage of cable trays — the basket-like contraptions that hold the cables coming off the machines. So they designed a new cable tray that Microsoft could manufacture itself or find somewhere to buy. They’ve also worked on ways to squish as many servers as possible in existing data centres around the world so they don’t have to wait for new buildings, Guthrie said.

    When OpenAI or Microsoft is training a large AI model, the work happens at one time. It’s divided across all the GPUs and at certain points, the units need to talk to each other to share the work they’ve done. For the AI supercomputer, Microsoft had make sure the networking gear that handles the communication among all the chips could handle that load, and it had to develop software that gets the best use out of the GPUs and the networking equipment. The company has now come up with software that lets it train models with tens of trillions of parameters.

    Because all the machines fire up at once, Microsoft had to think about where they were placed and where the power supplies were located. Otherwise you end up with the data centre version of what happens when you turn on a microwave, toaster and vacuum cleaner at the same time in the kitchen, Guthrie said.

    Read: Microsoft is infusing AI into business apps, including Teams

    The company also had to make sure it could cool off all of those machines and chips, and uses evaporation, outside air in cooler climates and high-tech swamp coolers in hot ones, said Alistair Speirs, director of Azure global infrastructure.

    Microsoft is going to keep working on customised server and chip designs and ways to optimise its supply chain in order to wring any speed gains, efficiency and cost-savings it can, Guthrie said.

    Read: Microsoft to bake Bing AI into Windows 11

    “The model that is wowing the world right now is built on the supercomputer we started building a couple of years ago. The new models will be built on the new supercomputer we’re training now, which is much bigger and will enable even more sophistication,” he said.  — Reported with Max Chafkin and Ian King, (c) 2023 Bloomberg LP

    Get TechCentral’s daily newsletter



    ChatGPT Microsoft OpenAI
    Subscribe to TechCentral Subscribe to TechCentral
    Share. Facebook Twitter LinkedIn WhatsApp Telegram Email Copy Link
    Previous ArticleAbsa adds to growing chorus of alarm over load shedding
    Next Article MTN takes R695-million hit from load shedding

    Related Posts

    OpenAI launches GPT-5.2 after 'code red' push to counter Google. Shelby Tauber/Reuters

    OpenAI launches GPT-5.2 after ‘code red’ push to counter Google

    12 December 2025
    OpenAI warns new models pose high cybersecurity risk

    OpenAI warns new models pose high cybersecurity risk

    11 December 2025
    Big Microsoft 365 price increases coming next year

    Big Microsoft price increases coming next year

    5 December 2025
    Company News
    When the physical world goes online: the new front line of cyber risk - Snode Technologies

    When the physical world goes online: the new front line of cyber risk

    12 December 2025
    Endless possibilities with Adapt IT Telecoms' unified VAS platform - Matthew Seabrook

    Endless possibilities with Adapt IT Telecoms’ unified VAS platform

    11 December 2025
    Securing IoT connectivity: how MSB Micro Systems keeps devices in check

    Securing IoT connectivity: how MSB Micro Systems keeps devices in check

    11 December 2025
    Opinion
    Netflix, Warner Bros deal raises fresh headaches for MultiChoice - Duncan McLeod

    Netflix, Warner Bros deal raises fresh headaches for MultiChoice

    5 December 2025
    BIN scans, DDoS and the next cybercrime wave hitting South Africa's banks - Entersekt Gerhard Oosthuizen

    BIN scans, DDoS and the next cybercrime wave hitting South Africa’s banks

    3 December 2025
    Your data, your hardware: the DIY AI revolution is coming - Duncan McLeod

    Your data, your hardware: the DIY AI revolution is coming

    20 November 2025

    Subscribe to Updates

    Get the best South African technology news and analysis delivered to your e-mail inbox every morning.

    Latest Posts
    ANC's attack on Solly Malatsi shows how BEE dogma trumps economic reality

    ANC’s attack on Solly Malatsi shows how BEE dogma trumps economic reality

    14 December 2025
    Political war erupts over BEE in the ICT sector - Solly Malatsi

    Political war erupts over BEE in the ICT sector

    13 December 2025
    Icasa told to align on BEE in move that will favour Starlink - Solly Malatsi

    Icasa told to align on BEE in move that will favour Starlink

    12 December 2025
    South African solar industry faces a reality check

    South African solar industry faces a reality check

    12 December 2025
    © 2009 - 2025 NewsCentral Media
    • Cookie policy (ZA)
    • TechCentral – privacy and Popia

    Type above and press Enter to search. Press Esc to cancel.

    Manage consent

    TechCentral uses cookies to enhance its offerings. Consenting to these technologies allows us to serve you better. Not consenting or withdrawing consent may adversely affect certain features and functions of the website.

    Functional Always active
    The technical storage or access is strictly necessary for the legitimate purpose of enabling the use of a specific service explicitly requested by the subscriber or user, or for the sole purpose of carrying out the transmission of a communication over an electronic communications network.
    Preferences
    The technical storage or access is necessary for the legitimate purpose of storing preferences that are not requested by the subscriber or user.
    Statistics
    The technical storage or access that is used exclusively for statistical purposes. The technical storage or access that is used exclusively for anonymous statistical purposes. Without a subpoena, voluntary compliance on the part of your Internet Service Provider, or additional records from a third party, information stored or retrieved for this purpose alone cannot usually be used to identify you.
    Marketing
    The technical storage or access is required to create user profiles to send advertising, or to track the user on a website or across several websites for similar marketing purposes.
    • Manage options
    • Manage services
    • Manage {vendor_count} vendors
    • Read more about these purposes
    View preferences
    • {title}
    • {title}
    • {title}