Industry News

With more than 4,000 chips strung together, Google says its supercomputers are faster and more energy efficient than Nvidia's

Views : 88
Update time : 2023-04-18 12:00:50
        Alphabet Inc.'s Google Inc. on Tuesday unveiled new details about its supercomputers used to train artificial intelligence models, saying the systems are faster and more power-efficient than their Nvidia counterparts.
        Google has designed its own chips, called Tensor Processing Units (TPUs), to train artificial intelligence models, which are used in more than 90 percent of the company's AI training efforts, for tasks such as answering questions in human language or generating images. Google's TPUs are now in their fourth generation. Google published a scientific paper on Tuesday detailing how they use their own custom-developed optical switches to connect more than 4,000 chips in series into a supercomputer.
 
        Improving these connections has become a key point of competition between companies building AI supercomputers, as the size of the so-called large language models that power technologies like Google's Bard or OpenAI's ChatGPT has exploded, meaning they are too big to be stored on a single chip. 
        These models have to be partitioned into thousands of chips, which then have to work in tandem for weeks or more to train the models. Google's PaLM model - its largest publicly disclosed language model to date - was trained over 50 days by spreading it across two supercomputers with 4,000 chips. 
        Google says its supercomputers can easily reconfigure the connections between the chips in real time, helping to avoid problems and improve performance. 
        In a blog post about the system, Google researcher Norm Jouppi and Google Distinguished Engineer David Patterson wrote: "Circuit switching made it easy for us to bypass faulty components. This flexibility even allows us to change the topology of the supercomputer interconnect to accelerate the performance of ML (machine learning) models." 
        Although Google is only now announcing details of its supercomputer, it is already coming online internally in 2020, running in a data centre in Mayes County, Oklahoma (USA). Google said the startup Midjourney used the system to train its model, which can generate images after inputting text. 
        In its paper, Google said its supercomputer was 1.7 times faster and 1.9 times more energy efficient than a system based on the Nvidia A100 chip for a system of the same size. Google said it did not compare its fourth-generation product to Nvidia's current flagship H100 chip because the H100 came to market after Google's chip and was built with newer technology. Google hinted that they may be working on a new TPU to compete with the Nvidia H100.


 
Related News
Read More >>
LDK220 LDO Voltage Regulators Specifications, Features, Pinout, and Applications LDK220 LDO Voltage Regulators Specifications, Features, Pinout, and Applications
Feb .02.2026
The LDK220 series of low-dropout linear regulators (LDOs) is a high-performance device designed specifically for low-power consumption and high-precision voltage regulation, widely used in scenarios such as consumer electronics, industrial control, and po
Xilinx Spartan®-7 FPGA Family: A High-Performance and Energy-Efficient Solution for Mid-Range FPGAs Xilinx Spartan®-7 FPGA Family: A High-Performance and Energy-Efficient Solution for Mid-Range FPGAs
Jan .20.2026
Xilinx Spartan®-7 FPGA Family stands as a defining solution in the mid-range FPGA landscape, blending high performance, energy efficiency, and cost-effectiveness to redefine versatility for industrial, IoT, and consumer electronics applications. Built on
Altera FLEX Series: Architecture, Innovation, and Application Across Four Generations Altera FLEX Series: Architecture, Innovation, and Application Across Four Generations
Jan .07.2026
The Altera FLEX series was more than a lineup of FPGAs—it was a blueprint for how programmable logic devices could evolve to meet diverse market needs. The FLEX 8000 laid the architectural groundwork, the FLEX 10K redefined functionality with embedded mem
LM4765 vs. LM4766: A Comprehensive Comparison of Dual-Channel Audio Power Amplifiers LM4765 vs. LM4766: A Comprehensive Comparison of Dual-Channel Audio Power Amplifiers
Dec .16.2025
Among TI standout offerings, the LM4765 and LM4766 are dual-channel amplifiers designed to cater to diverse audio needs—from compact setups to high-fidelity systems. While sharing the same product lineage, these chips differ significantly in power output,