Sony Pixel Power calrec Sony

SLMming Down Latency: How NVIDIA's First On-Device Small Language Model Makes Digital Humans More Lifelike

21/08/2024

Editor's note: This post is part of the AI Decoded series, which demystifies AI by making the technology more accessible, and showcases new hardware, software, tools and accelerations for RTX PC and workstation users.

At Gamescom this week, NVIDIA announced that NVIDIA ACE - a suite of technologies for bringing digital humans to life with generative AI - now includes the company's first on-device small language model (SLM), powered locally by RTX AI.

The model, called Nemotron-4 4B Instruct, provides better role-play, retrieval-augmented generation and function-calling capabilities, so game characters can more intuitively comprehend player instructions, respond to gamers, and perform more accurate and relevant actions.

Available as an NVIDIA NIM microservice for cloud and on-device deployment by game developers, the model is optimized for low memory usage, offering faster response times and providing developers a way to take advantage of over 100 million GeForce RTX-powered PCs and laptops and NVIDIA RTX-powered workstations.

The SLM Advantage An AI model's accuracy and performance depends on the size and quality of the dataset used for training. Large language models are trained on vast amounts of data, but are typically general-purpose and contain excess information for most uses.

SLMs, on the other hand, focus on specific use cases. So even with less data, they're capable of delivering more accurate responses, more quickly - critical elements for conversing naturally with digital humans.

Nemotron-4 4B was first distilled from the larger Nemotron-4 15B LLM. This process requires the smaller model, called a student, to mimic the outputs of the larger model, appropriately called a teacher. During this process, noncritical outputs of the student model are pruned or removed to reduce the parameter size of the model. Then, the SLM is quantized, which reduces the precision of the model's weights.

With fewer parameters and less precision, Nemotron-4 4B has a lower memory footprint and faster time to first token - how quickly a response begins - than the larger Nemotron-4 LLM while still maintaining a high level of accuracy due to distillation. Its smaller memory footprint also means games and apps that integrate the NIM microservice can run locally on more of the GeForce RTX AI PCs and laptops and NVIDIA RTX AI workstations that consumers own today.

This new, optimized SLM is also purpose-built with instruction tuning, a technique for fine-tuning models on instructional prompts to better perform specific tasks. This can be seen in Mecha BREAK, a video game in which players can converse with a mechanic game character and instruct it to switch and customize mechs.

ACEs Up ACE NIM microservices allow developers to deploy state-of-the-art generative AI models through the cloud or on RTX AI PCs and workstations to bring AI to their games and applications. With ACE NIM microservices, non-playable characters (NPCs) can dynamically interact and converse with players in the game in real time.

ACE consists of key AI models for speech-to-text, language, text-to-speech and facial animation. It's also modular, allowing developers to choose the NIM microservice needed for each element in their particular process.

NVIDIA Riva automatic speech recognition (ASR) processes a user's spoken language and uses AI to deliver a highly accurate transcription in real time. The technology builds fully customizable conversational AI pipelines using GPU-accelerated multilingual speech and translation microservices. Other supported ASRs include OpenAI's Whisper, a open-source neural net that approaches human-level robustness and accuracy on English speech recognition.

Once translated to digital text, the transcription goes into an LLM - such as Google's Gemma, Meta's Llama 3 or now NVIDIA Nemotron-4 4B - to start generating a response to the user's original voice input.

Next, another piece of Riva technology - text-to-speech - generates an audio response. ElevenLabs' proprietary AI speech and voice technology is also supported and has been demoed as part of ACE, as seen in the above demo.

Finally, NVIDIA Audio2Face (A2F) generates facial expressions that can be synced to dialogue in many languages. With the microservice, digital avatars can display dynamic, realistic emotions streamed live or baked in during post-processing.

The AI network automatically animates face, eyes, mouth, tongue and head motions to match the selected emotional range and level of intensity. And A2F can automatically infer emotion directly from an audio clip.

Finally, the full character or digital human is animated in a renderer, like Unreal Engine or the NVIDIA Omniverse platform.

AI That's NIMble In addition to its modular support for various NVIDIA-powered and third-party AI models, ACE allows developers to run inference for each model in the cloud or locally on RTX AI PCs and workstations.

The NVIDIA AI Inference Manager software development kit allows for hybrid inference based on various needs such as experience, workload and costs. It streamlines AI model deployment and integration for PC application developers by preconfiguring the PC with the necessary AI models, engines and dependencies. Apps and games can then orchestrate inference seamlessly across a PC or workstation to the cloud.

ACE NIM microservices run locally on RTX AI PCs and workstations, as well as in the cloud. Current microservices running locally include Audio2Face, in the Covert Protocol tech demo, and the new Nemotron-4 4B Instruct and Whisper ASR in Mecha BREAK.

To Infinity and Beyond Digital humans go far beyond NPCs in games. At last month's SIGGRAPH conference, NVIDIA previewed James, an interactive digital human that can connect with people using emotions, humor and more. James is based on
LINK: https://blogs.nvidia.com/blog/ai-decoded-gamescom-ace-nemotron-instruc...
See more stories from nvidia

Most recent headlines

09/12/2024

Dalet Named an IDC Innovator in Media and Entertainment

Dalet, a leading technology and service provider for media-rich organizations, today announced that it has been named an IDC Innovator in the IDC Innovators: ...

09/11/2024

Dalet Expands Leadership Team to Fuel Next Stage of Growth

Dalet, a leading technology and service provider for media-rich organizations, today announced three new members of its executive team. Tara Bryant joins as Chi...

18/09/2024

Line up revealed for World's Most Dangerous Roads S6

Johnny Vegas & Lucy Beaumont and Babatunde Al sh & Kae Kurd arethecomedians taking part in brand new episodes of World's Most Dangerous Roads, which return...

18/09/2024

TV channel Gold's Absolutely Fabulous: Inside Out is joined by celebrities Emma Bunton, Meera Syal, Ruby Wax and more

The nation's favourite comedy channel, Gold, is set to take a reflective loo...

18/09/2024

Press Release: ToolsOnAir Wins 2024 OEM & Developer Award from Blackmagic Design at IBC

Press Release: ToolsOnAir Wins 2024 OEM & Developer Award from Blackmagic Design...

18/09/2024

A Different Man Is a Triumph That Lingers Long After Credits Role

Warning: This feature contains spoilers about the film. By Bailey Pennick Aaron Schimberg kept it brief in his introduction before A Different Man had its wor...

18/09/2024

Mi Primer Escenario' ofrece a los artistas emergentes en Mxico la oportunidad de tomar el escenario en MEXCLA Spotify

Apoyar a los artistas emergentes es parte fundamental del ADN de Spotify. Ahora ...

18/09/2024

Mi Primer Escenario' Offers Emerging Mexican Artists a Chance To Perform at MEXCLA Spotify'

Supporting emerging artists is a fundamental part of Spotify's DNA, and we&#...

18/09/2024

Karmic lessons and laughs with new SBS podcast Comedy Karma'

Karmic lessons and laughs with new SBS podcast Comedy Karma' 18 September, 2024 Media releases Join the ever-curious stand-up comedian Aditya Gautam a...

18/09/2024

Clear-Com Supports the Next Generation of Media Professionals at Rise Academy Summer...

eds3_5_jq(document).ready(function($) { $(#eds_sliderM519).chameleonSlider_2_1({...

18/09/2024

Nielsen launches Advanced Audiences, enhancing digital campaign precision, reach and effectiveness across Australia and New Zealand

Sydney, September 18, 2024 - Nielsen today announced the launch of Advanced Audi...

18/09/2024

Heartland Video Systems Welcomes Industry Veteran Dan Whe...

Heartland Video Systems (HVS), a premier video systems integration and consulting firm, is proud to announce the appointment of Dan Whealy as Director of Busine...

18/09/2024

Clear-Com Supports the Next Generation of Media Professio...

Clear-Com is proud to have participated in the Rise Academy Summer School 2024, a transformative experience aimed at introducing young people to the dynamic wor...

18/09/2024

Amplify Berklee Honors the Legacy and Celebrates the Future of Berklee City Music

Amplify Berklee Honors the Legacy and Celebrates the Future of Berklee City Musi...

18/09/2024

Riedel Expands Its Range of NSA Network Stream Adapters

Riedel Communications today announced the launch of two new additions to its acclaimed Network Stream Adapter (NSA) series: the NSA-003A and NSA-006A. Unveiled ...

18/09/2024

WDR Relies on Riedel Backbone for Remote Production of UE...

The German regional public broadcaster Westdeutscher Rundfunk (WDR) has implemented a Riedel backbone for communications and signal distribution for the ARD bro...

18/09/2024

Amagi Expands Footprint in Latin America with Mexico Entr...

Amagi, the global leader in cloud-based SaaS technology for broadcast and connected TV (CTV), today announced that it is entering the Mexican market. This strat...

18/09/2024

ITV to Modernize Its Media Supply Chain With Cloud-Native...

SDVI, the leading platform provider for cloud-native media supply chains, today announced that U.K. broadcaster ITV is deploying the SDVI Rally platform as part...

18/09/2024

Triveni Digital SCTE TechExpo24 Exhibitor Preview

ATSC 3.0 is gaining momentum across the U.S., and for cable operators, ensuring top-notch service quality is more important than ever. Monitoring the performanc...

18/09/2024

Amagi and BuyDRM Partner to Secure Streaming Video on Pla...

Amagi, the global leader in cloud-based SaaS technology for broadcast and connected TV (CTV), today announced a partnership with BuyDRM, a leading content secur...

18/09/2024

Charter Bumps Up Broadband Speeds, Unveils New Bundle Pricing

STAMFORD, Conn. Spectrum has made a series of announcements that include a new simplified pricing strategy, increased broadband speeds and new customer service ...

18/09/2024

Lawo Doubles the Number of HOME Apps

Lawo has announced that its HOME Apps platform now hosts nine essential processing apps, effectively doubling the previous offering with more apps to follow in ...

18/09/2024

Gray, New Orleans Pelicans Announce New Sports Network

Gray Media is partnering with the New Orleans Pelicans to create a new network that promises to bring every non-national Pelicans NBA game to its viewers....

18/09/2024

Fox Weather Expands Distribution to DirecTV

NEW YORK Fox News Media has announced that Fox Weather, a free ad-supported streaming television service (FAST), is now available to DirecTV customers....

18/09/2024

NAB Show New York Exhibitor Insight TAG Video Systems

TV TECH: What do you anticipate will be the most significant technology trends at the 2024 NAB Show New York?...

18/09/2024

Key Conversations With News, Sports Panels on Tap at NAB Show New York

Broadcast, media and entertainment leaders will gather once again at NAB Show New York in October to explore key innovations and strategies reshaping how conten...

18/09/2024

SCTE Foundation Rebrands

EXTON, Pa. The SCTE Foundation has announced a comprehensive relaunch of its efforts. The relaunch includes new branding that will be on display at the Georgia ...

18/09/2024

Viant Technology Launches New Programmatic Ad Solution, ViantAI

IRVINE, Calif. Viant Technology Inc. has launched ViantAI, an advanced AI-powered platform that it says will reshape how programmatic advertising is planned, pu...

18/09/2024

Blackmagic Design Announces Pricing for Blackmagic URSA Cine 17K 65

Blackmagic Design Announces Pricing for Blackmagic URSA Cine 17K 65 Brie Clayton September 17, 2024 0 Comments Revolutionary large format digital film...

18/09/2024

Avid | Stream IO ingest & playout solution now supports SMPTE 2110

Avid | Stream IO ingest & playout solution now supports SMPTE 2110 Brie Clayton September 17, 2024 0 Comments Avid's next-gen software-based produ...

18/09/2024

Telemundo's Celebrando Todo Lo Que Somos' Marks Hispanic Heritage Month

NBCUniversal's Telemundo Enterprises is marking Hispanic Heritage Month with year four of its multiplatform initiative with the slogan Celebrando Todo Lo Q...

18/09/2024

Hispano Soy' Marks Hispanic Heritage Month on Warner Bros. Discovery U.S. Hispanic Nets

To celebrate Hispanic Heritage Month, Warner Bros. Discovery U.S. Hispanic is la...

18/09/2024

MeTV Toons Plans Spooky Sundays in October

As Halloween approaches, MeTV Toons will feature frightful Sunday programming. On October 13, it's a scary Flintstones block 1-3 p.m., then Scooby-Doo! 3-5 ...

18/09/2024

TMZ Investigates Matthew Perry and His Death in Sept. 16 Special on Fox

TMZ takes a close look at actor Matthew Perry and the drug network that led to his death when TMZ Investigates: Matthew Perry & the Secret Celebrity Drug Ring a...

18/09/2024

Olympics Boost Broadcast, Peacock Viewing in August

The Paris Summer Olympics gave a boost to broadcast and popped on Peacock in August, according to Nielsen....

18/09/2024

Gray Forms Gulf Coast Network To Broadcast Pelicans NBA Games

Gray Media is forming a new venture called Gulf Coast Sports & Entertainment Network that will carry all of the locally televised games of the NBA's New Orl...

18/09/2024

LG Study: Asian-American Viewers Prefer Streaming Ads

LG Ad Solutions shared a new report The Inclusive Screen: Asian Americans, which found that 70% of Asian Americans feel that streaming TV ads are more relevant...

18/09/2024

VideoAmp Says It Serves as Currency in $1 Billion of Media Deals

Measurement company VideoAmp said that $1 billion in media buys have already been guaranteed using its measurement as a currency this year....

18/09/2024

Kenny Smith Will Host Season 3 of Harlem Globetrotters: Play It Forward'

Kenny Smith, former NBA star and current TNT Sports analyst, will be the host of Harlem Globetrotters: Play It Forward for Season 3, which premieres Saturday, O...

18/09/2024

ESPN's Jimmy Pitaro Reflects on ESPN Turning 45, Launch of a New App in 2025, and the Role of AI, RSNs

ESPN's Jimmy Pitaro Reflects on ESPN Turning 45, Launch of a New App in 2025...

18/09/2024

Chicago Sports Network Agrees to OTA-TV Deal With Millennial Telecommunications' WJYS-TV

Chicago Sports Network Agrees to OTA-TV Deal With Millennial Telecommunications&...

18/09/2024

New Orleans Pelicans Ink OTA-TV Distribution Deal With Gray Media; Raycom Sports To Produce All Games

New Orleans Pelicans Ink OTA-TV Distribution Deal With Gray Media; Raycom Sports...

18/09/2024

Seattle Kraken Officially Launch Kraken Hockey Network, Promise Biggest Regional Production in the NHL'

Seattle Kraken Officially Launch Kraken Hockey Network, Promise Biggest Regiona...

18/09/2024

New Victory+ Free, DTC Streaming Service From Dallas Stars and APMC Goes Live

New Victory Free, DTC Streaming Service From Dallas Stars and APMC Goes Live In addition to all Stars games, Victory will carry Anaheim Ducks games this seaso...

18/09/2024

ESPN Looking to Hire Drone Pilot

ESPN Looking to Hire Drone Pilot By Ken Kerschbaumer Wednesday, September 18, 2024 - 4:30 pm Print This Story | Subscribe Story Highlights Have the ski...

18/09/2024

Only on Netflix! With Stars like Rodrigo Santoro and Rafael Vitti, Four Brazilian Films Are Starting Production From Bahia to Bariloche

Back to All News Only on Netflix! With Stars like Rodrigo Santoro and Rafael Vi...

18/09/2024

Haivision Wins Prestigious IBC Innovation Award for its Live Video Contribution Solutions over Private 5G Networks

Haivision Wins Prestigious IBC Innovation Award for its Live Video Contribution ...

18/09/2024

Eutelsat Group Secures Additional Launches in New Agreement with Mitsubishi Heavy Industries

Photo credit: Mitsubishi Heavy Industries, Ltd Press release - 18 September 20...

18/09/2024

2024-09-18

iPadOS 18 makes the iPad experience more versatile and intelligent than ever, and is available today as a free software update. iPadOS 18 brings incredible new ...

18/09/2024

Telespazio and ABS Partner to Deliver Enhanced Connectivity Services for the Brazilian Air Force System

Telespazio and ABS Partner to Deliver Enhanced Connectivity Services for the Bra...