Industry News

NVIDIA Releases H200, World's Most Powerful AI Chip: Performance Soars 90%, Doubles Llama 2 Inference Speed

Views : 155
Update time : 2023-12-09 10:24:38
        November 13, 2011 - NVIDIA today unveiled the next generation of AI supercomputer chips that will play an important role in deep learning and large-scale language models (LLMs), such as OpenAI's GPT-4. The new chips represent a significant leap forward from the previous generation, and will be used in data centers and supercomputers to handle tasks such as weather and climate prediction, drug discovery, quantum computing, and more. and other tasks.
        The key product release is the HGX H200 GPU, based on NVIDIA's "Hopper" architecture, which is the successor to the H100 GPU and the company's first chip to use HBM3e memory, which is faster and larger, and therefore better suited for large language models. According to NVIDIA, "With HBM3e, the NVIDIA H200 delivers 141GB of memory at 4.8 TB per second, nearly twice the capacity and a 2.4x increase in bandwidth compared to the A100." On the AI side, NVIDIA says the HGX H200 doubles the speed of inference on Llama 2 (70 billion parameter LLM) compared to the H100.The HGX H200 will be available in 4-way and 8-way configurations that are compatible with the software and hardware in the H100 system. It will be available for every type of data center (local, cloud, hybrid cloud and edge) and deployed by Amazon Web Services, Google Cloud, Microsoft Azure and Oracle Cloud Infrastructure, among others, and will be available in the second quarter of 2024.
        Another key product announcement from NVIDIA was the GH200 Grace Hopper "superchip," which combines the HGX H200 GPU with the Arm-based NVIDIA Grace CPU via the company's NVLink-C2C interconnect, and is officially designed for supercomputers. Designed specifically for supercomputers, it allows "scientists and researchers to solve the world's most challenging problems by accelerating complex AI and HPC applications running terabytes of data. The GH200 will be used in "more than 40 AI supercomputers at research centers, system manufacturers and cloud providers around the world," including Dell, Eviden, Hewlett-Packard Enterprise (HPE), Lenovo, QCT, and Supermicro. Notably, HPE's Cray EX2500 supercomputer will use the quad GH200, which scales up to tens of thousands of Grace Hopper superchip nodes. Perhaps the largest Grace Hopper supercomputer will be the GH200. Perhaps the largest Grace Hopper supercomputer is JUPITER at the Jülich facility in Germany, which will be "the world's most powerful AI system" when installed in 2024. It uses a liquid-cooled architecture, and its enhancement module consists of nearly 24,000 NVIDIA GH200 supercomputers interconnected by NVIDIA's Quantum-2 InfiniBand networking platform.
        NVIDIA says JUPITER will contribute to scientific breakthroughs in a number of areas, including climate and weather prediction, generating high-resolution climate and weather simulations with interactive visualizations. It will also be used in drug discovery, quantum computing and industrial engineering, many of which use customized NVIDIA software solutions that simplify development but also make supercomputing teams dependent on NVIDIA hardware. Last quarter, NVIDIA achieved record revenues of $10.32 billion ($13.51 billion total) in AI and data center alone, up 171 percent from a year ago, and NVIDIA is no doubt hoping that the new GPUs and supercomputing chips will help it continue that trend.
 
Related News
Read More >>
LDK220 LDO Voltage Regulators Specifications, Features, Pinout, and Applications LDK220 LDO Voltage Regulators Specifications, Features, Pinout, and Applications
Feb .02.2026
The LDK220 series of low-dropout linear regulators (LDOs) is a high-performance device designed specifically for low-power consumption and high-precision voltage regulation, widely used in scenarios such as consumer electronics, industrial control, and po
Xilinx Spartan®-7 FPGA Family: A High-Performance and Energy-Efficient Solution for Mid-Range FPGAs Xilinx Spartan®-7 FPGA Family: A High-Performance and Energy-Efficient Solution for Mid-Range FPGAs
Jan .20.2026
Xilinx Spartan®-7 FPGA Family stands as a defining solution in the mid-range FPGA landscape, blending high performance, energy efficiency, and cost-effectiveness to redefine versatility for industrial, IoT, and consumer electronics applications. Built on
Altera FLEX Series: Architecture, Innovation, and Application Across Four Generations Altera FLEX Series: Architecture, Innovation, and Application Across Four Generations
Jan .07.2026
The Altera FLEX series was more than a lineup of FPGAs—it was a blueprint for how programmable logic devices could evolve to meet diverse market needs. The FLEX 8000 laid the architectural groundwork, the FLEX 10K redefined functionality with embedded mem
LM4765 vs. LM4766: A Comprehensive Comparison of Dual-Channel Audio Power Amplifiers LM4765 vs. LM4766: A Comprehensive Comparison of Dual-Channel Audio Power Amplifiers
Dec .16.2025
Among TI standout offerings, the LM4765 and LM4766 are dual-channel amplifiers designed to cater to diverse audio needs—from compact setups to high-fidelity systems. While sharing the same product lineage, these chips differ significantly in power output,