Sony Pixel Power calrec Sony

NVIDIA Takes Inference to New Heights Across MLPerf Tests

05/04/2023

MLPerf remains the definitive measurement for AI performance as an independent, third-party benchmark. NVIDIA's AI platform has consistently shown leadership across both training and inference since the inception of MLPerf, including the MLPerf Inference 3.0 benchmarks released today.

Three years ago when we introduced A100, the AI world was dominated by computer vision. Generative AI has arrived, said NVIDIA founder and CEO Jensen Huang.

This is exactly why we built Hopper, specifically optimized for GPT with the Transformer Engine. Today's MLPerf 3.0 highlights Hopper delivering 4x more performance than A100.

The next level of Generative AI requires new AI infrastructure to train large language models with great energy efficiency. Customers are ramping Hopper at scale, building AI infrastructure with tens of thousands of Hopper GPUs connected by NVIDIA NVLink and InfiniBand.

The industry is working hard on new advances in safe and trustworthy Generative AI. Hopper is enabling this essential work, he said.

The latest MLPerf results show NVIDIA taking AI inference to new levels of performance and efficiency from the cloud to the edge.

Specifically, NVIDIA H100 Tensor Core GPUs running in DGX H100 systems delivered the highest performance in every test of AI inference, the job of running neural networks in production. Thanks to software optimizations, the GPUs delivered up to 54% performance gains from their debut in September.

In healthcare, H100 GPUs delivered a 31% performance increase since September on 3D-UNet, the MLPerf benchmark for medical imaging.

Powered by its Transformer Engine, the H100 GPU, based on the Hopper architecture, excelled on BERT, a transformer-based large language model that paved the way for today's broad use of generative AI.

Generative AI lets users quickly create text, images, 3D models and more. It's a capability companies from startups to cloud service providers are rapidly adopting to enable new business models and accelerate existing ones.

Hundreds of millions of people are now using generative AI tools like ChatGPT - also a transformer model - expecting instant responses.

At this iPhone moment of AI, performance on inference is vital. Deep learning is now being deployed nearly everywhere, driving an insatiable need for inference performance from factory floors to online recommendation systems.

L4 GPUs Speed Out of the Gate NVIDIA L4 Tensor Core GPUs made their debut in the MLPerf tests at over 3x the speed of prior-generation T4 GPUs. Packaged in a low-profile form factor, these accelerators are designed to deliver high throughput and low latency in almost any server.

L4 GPUs ran all MLPerf workloads. Thanks to their support for the key FP8 format, their results were particularly stunning on the performance-hungry BERT model.

In addition to stellar AI performance, L4 GPUs deliver up to 10x faster image decode, up to 3.2x faster video processing and over 4x faster graphics and real-time rendering performance.

Announced two weeks ago at GTC, these accelerators are already available from major systems makers and cloud service providers. L4 GPUs are the latest addition to NVIDIA's portfolio of AI inference platforms launched at GTC.

Software, Networks Shine in System Test NVIDIA's full-stack AI platform showed its leadership in a new MLPerf test.

The so-called network-division benchmark streams data to a remote inference server. It reflects the popular scenario of enterprise users running AI jobs in the cloud with data stored behind corporate firewalls.

On BERT, remote NVIDIA DGX A100 systems delivered up to 96% of their maximum local performance, slowed in part because they needed to wait for CPUs to complete some tasks. On the ResNet-50 test for computer vision, handled solely by GPUs, they hit the full 100%.

Both results are thanks, in large part, to NVIDIA Quantum Infiniband networking, NVIDIA ConnectX SmartNICs and software such as NVIDIA GPUDirect.

Orin Shows 3.2x Gains at the Edge Separately, the NVIDIA Jetson AGX Orin system-on-module delivered gains of up to 63% in energy efficiency and 81% in performance compared with its results a year ago. Jetson AGX Orin supplies inference when AI is needed in confined spaces at low power levels, including on systems powered by batteries.

For applications needing even smaller modules drawing less power, the Jetson Orin NX 16G shined in its debut in the benchmarks. It delivered up to 3.2x the performance of the prior-generation Jetson Xavier NX processor.

A Broad NVIDIA AI Ecosystem The MLPerf results show NVIDIA AI is backed by the industry's broadest ecosystem in machine learning.

Ten companies submitted results on the NVIDIA platform in this round. They came from the Microsoft Azure cloud service and system makers including ASUS, Dell Technologies, GIGABYTE, H3C, Lenovo, Nettrix, Supermicro and xFusion.

Their work shows users can get great performance with NVIDIA AI both in the cloud and in servers running in their own data centers.

NVIDIA partners participate in MLPerf because they know it's a valuable tool for customers evaluating AI platforms and vendors. Results in the latest round demonstrate that the performance they deliver today will grow with the NVIDIA platform.

Users Need Versatile Performance NVIDIA AI is the only platform to run all MLPerf inference workloads and scenarios in data center and edge computing. Its versatile performance and efficiency make users the real winners.

Real-world applications typically employ many neural networks of different kinds that often need to deliver answers in real time.

For example, an AI application may need to understand a user's spoken request, classify an image, make a recommendation and then deliver a response as a spoken message in a human-sounding voice. Each step requires a different type
LINK: https://blogs.nvidia.com/blog/2023/04/05/inference-mlperf-ai/...
See more stories from nvidia

Most recent headlines

13/03/2025

FCC Chairman Carr Launches Massive Deregulation Initiative

WASHINGTON FCC Chairman Brendan Carr announced that the agency has launched a massive, new deregulatory initiative that could potentially subject virtually all ...

13/03/2025

SDVI Integrates Rally With Spectra Vail

SUNNYVALE, Calif. SDVI has announced that it has integrated its Rally media supply chain management platform with the Spectra Vail multi-cloud data management s...

13/03/2025

FCC Approves Gray's Acquisition of KXLT

ATLANTA Gray Media has announced that the Federal Communications Commission (FCC) has granted a waiver of its local ownership rules to permit Gray Media to acqu...

13/03/2025

NAB Show 2025 Exhibitor Insight: Blackmagic Design

TV Tech: What do you anticipate will be the most significant technology trends at the 2025 NAB Show?...

13/03/2025

Watch Student Naomi Soleils Folk Pop Performance on The Voice

Watch Student Naomi Soleils Folk Pop Performance on The Voice The songwriting major sang Stars by Grace Potter and the Nocturnals during the blind auditions. ...

13/03/2025

Relive the Magic as GeForce NOW Brings More Blizzard Gaming to the Cloud

Bundle up - GeForce NOW is bringing a flurry of Blizzard titles to its ever-expanding library. Prepare to weather epic gameplay in the cloud, tackling the genr...

13/03/2025

Gaming Goodness: NVIDIA Reveals Latest Neural Rendering and AI Advancements Supercharging Game Development at GDC 2025

AI is leveling up the world's most beloved games, as the latest advancements...

13/03/2025

Drop It Like It's Mod: Breathing New Life Into Classic Games With AI in NVIDIA RTX Remix

PC game modding is massive, with over 5 billion mods downloaded annually. Mods p...

12/03/2025

Como o Impacto Cultural e Financeiro da Indstria Musical Define Seu Sucesso em 2025

O Spotify acaba de lan ar o relat rio Loud & Clear deste ano, uma vis o transpar...

12/03/2025

Cmo el impacto cultural y financiero de la industria musical define su xito en 2025

Spotify acaba de presentar el informe Loud & Clear de este a o, una mirada trans...

12/03/2025

Nourish Mind, Body, and Soul This Ramadan With Spotify

Ramadan, a period of profound spiritual significance for Muslims worldwide, is a time for fasting, prayer, reflection, and community. Enrich your experience thi...

12/03/2025

How the Music Industry's Cultural and Financial Impact Define Its Success in 2025

Spotify has just unveiled this year's Loud & Clear report, a transparent loo...

12/03/2025

FAST now transcends back catalog content

More than 70% of FAST programming has been produced since 2010, according to new Gracenote report NEW YORK March 12, 2025 Gracenote, the content data busin...

12/03/2025

Viamedia Adopts ShowSeeker Pilot To Streamline Ad Workflows

GEONA, Nev. Independent digital and linear advertising rep firm Viamedia will adopt cloud-based ShowSeeker Pilot as its primary ad campaign and order management...

12/03/2025

Mediagenix Appoints Wael Yasin as Sales Director Central Europe

BRUSSELS Mediagenix has announced that Wael Yasin has joined the company as sales director Central Europe....

12/03/2025

Omdia: Global Online Consumer Expenditure to Hit $6.6 Trillion by 2029

LONDON A new study highlights opportunities for shoppable TV and the massive impact online consumer spending is having on the economy, with Omdia predicting tha...

12/03/2025

Comcast Technology Solutions Upgrades to Interra Systems BATON Version 9

CUPERTINO, Calif. Interra Systems has announced that Comcast Technology Solutions has integrated recent updates to BATON Version 9 into its operations....

12/03/2025

Nevion announces new 400G addition to its eMerge SDN media fabric offering

Nevion announces new 400G addition to its eMerge SDN media fabric offering Brie Clayton March 12, 2025 0 Comments High-capacity switch enhances existi...

12/03/2025

DaVinci Resolve Studio Delivers Cinematic Sound for Adam Bol

DaVinci Resolve Studio Delivers Cinematic Sound for Adam Bol Brie Clayton March 12, 2025 0 Comments Feature film relies on DaVinci Resolve Studio for ...

12/03/2025

SVT Leverages Ateliere Live to Pioneer 100% Software-Defined Production at Rally Sweden 2025

SVT Leverages Ateliere Live to Pioneer 100% Software-Defined Production at Rally...

12/03/2025

Berklee Awards Fenway Neighborhood Improvement Grant to Four Organizations

Berklee Awards Fenway Neighborhood Improvement Grant to Four Organizations A total of $17,000 will be distributed among the Boston-based nonprofits. By Madd...

12/03/2025

Offering a broad umbrella for IP standards

TVBEuropes Jenny Priestley sits down with new SMPTE president Richard Welsh to discuss his aims for the organisation going forward, its efforts to attract a you...

12/03/2025

Warner Bros Belgium takes a new approach to its post production workflows

The new process has significantly enhanced operational efficiencies, with automation freeing up creators to concentrate on higher value tasks By Matthew Corrig...

12/03/2025

Another VFX company suspends operations

Jellyfish Pictures, which has offices in London and Sheffield, said it has been battling hard in the face of strong headwinds over the last 12 months By Jenny ...

12/03/2025

Bitcentral to Showcase AI-Powered Innovations for Smarter...

Visit Booth #W2213 to Experience the Latest in AI-Driven Content Discovery and Workflow Automation Bitcentral, a leader in media workflow solutions, is set to ...

12/03/2025

Jeff Lilly Named WGN-TV Director Of Technology

CHICAGO Jeff Lilly has been named WGN-TV director of technology effective March 17, 2025, according to Ric Harris, WGN-TV vice president and general manager....

12/03/2025

Survey: Consumers Want More Shoppable TV Experiences

MOUNTAIN VIEW, Calif. A new study from LG Ad Solutions indicates that consumers want more features that would allow them to shop for products on the connected T...

12/03/2025

IP Showcase To Focus On IP, Broadcast Tech Convergence at 2025 NAB Show

BOTHELL, Wash. The Alliance for IP Media Solutions (AIMS), Advanced Media Workflow Association (AMWA) and the Video Services Forum (VSF) will once again present...

12/03/2025

Comcast Boosts Internet Speeds for More Than 20 Million Customers

PHILADELPHIA Comcast announced that it has upgraded Xfinity Internet speeds for more than 20 million customers for no additional cost....

12/03/2025

Create with Maxon: Cinema 4D Fundamentals Workshop - March 12-14

Create with Maxon: Cinema 4D Fundamentals Workshop - March 12-14 Brie Clayton March 11, 2025 0 Comments Makin' Waffles with Elly Wade During Marc...

12/03/2025

Ross Video Strengthens Cloud Leadership with Appointment of Aaron Tunnell

Ottawa, Canada - March 12, 2025 - Ross Video is pleased to announce the appointment of Aaron Tunnell as Business Development Director, Cloud Solutions, reinforc...

12/03/2025

VEON to release 4Q 2024 trading update on 20 March 2025

12 Mar 2025 VEON to release 4Q 2024 trading update on 20 March 2025 Dubai, 12 March 2025 - VEON Ltd. (NASDAQ: VEON), a global digital operator, today confirms ...

12/03/2025

Getting Set for the Winter Olympics: Inside SVT's Ongoing Software-Based Production Evolution

Getting set for the Winter Olympics: Inside SVT's ongoing software-based pro...

12/03/2025

Seamless SMPTE 2110: A Discussion on Making the Move to IP Less Painful

Seamless SMPTE 2110: A Discussion on Making the Move to IP Less Painful Leaders from Netflix, Monumental Sports & Entertainment, LinkedI, and Megapixel address ...

12/03/2025

Opening Doors of Perception: Meta's Ajit Ninan on Rethinking AR/MR Perceptual Displays

Opening Doors of Perception: Meta's Ajit Ninan on Rethinking AR/MR Perceptua...

12/03/2025

Sky Sports Unveils Exciting Plans for 2025 Formula 1 Coverage

With the 2025 Formula 1 season set to begin in Melbourne on March 16, Sky Sports is gearing up to deliver its most comprehensive F1 coverage yet. Following a th...

12/03/2025

Unicanal and Trece TV revolutionize Digital TV in Paraguay with Rohde & Schwarz transmission solutions

Unicanal and Trece TV revolutionize Digital TV in Paraguay with Rohde & Schwarz ...

12/03/2025

Honoring Service, Celebrating Sport: Rohde & Schwarz UK Sponsors NavyFit Rugby Double-Header

Honoring Service, Celebrating Sport: Rohde & Schwarz UK Sponsors NavyFit Rugby D...

12/03/2025

FOR-A America Introduces Top of the Line HANABI Series Switcher at NAB 2025 Exhibition

New HVS-Q12 Represents a Quantum Leap in Switcher Design...

12/03/2025

Telespazio and ABS Expand Partnership to Deliver Tailored C-Band Connectivity Solutions for the Brazilian Market

Telespazio and ABS Expand Partnership to Deliver Tailored C-Band Connectivity So...

12/03/2025

Congratulations to all our clients who have achieved TAG certification through our independent validation

ABC conducts independent audits across the TAG programs. In the March 2025 TAG s...

12/03/2025

Moonlight: Thales Alenia Space to develop the space segment of the navigation system orbiting around the Moon

Facebook Twitter LinkedIn Moonlight: ESA program to create a satellite con...

12/03/2025

Thales among the world's top 100 most innovative companies for the 12th year, according to Clarivate

Facebook Twitter LinkedIn In 2025, Thales ranks among the top 0.01% of the...

12/03/2025

CCPC sponsored TV series The Complaints Bureau returns to RT One

Consumer show back following hugely successful first series 10 March 2025 Following a hugely successful first series, The Complaints Bureau returns to RT ...

11/03/2025

Give Me the Backstory: Get to Know Amel Guellaty, the Writer-Director of Where the Wind Comes From

By Lucy Spicer One of the most exciting things about the Sundance Film Festival...

11/03/2025

Salsa's Rhythmic Revival: A New Generation Discovers the Genre's Timeless Appeal

Salsa is making a comeback, captivating new listeners with its infectious energy...

11/03/2025

Canada's Own Rupi Kaur Curates the Latest Edition of Local Spot, Revealing the Songs and Artists That Shaped Her

For many people, music can serve as a reflection of their roots and upbringing. ...

11/03/2025

ST Engineering iDirect's Public Safety Solution Wins MSUA Satellite Mobile Innovation Award 2025

The solution delivers reliable, scalable and secure connectivity for critical pu...

11/03/2025

Weigel Broadcasting Taps Harmonic for Cloud-Based Playout-to-Delivery

SAN JOSE, Calif. Harmonic has announced that Weigel Broadcasting has deployed Harmonics VOS Media Software, which offers playout-to-delivery capabilities, inclu...