
NVIDIA Research's latest AI model is a prodigy among generative adversarial networks. Using a fraction of the study material needed by a typical GAN, it can learn skills as complex as emulating renowned painters and recreating images of cancer tissue.
By applying a breakthrough neural network training technique to the popular NVIDIA StyleGAN2 model, NVIDIA researchers reimagined artwork based on fewer than 1,500 images from the Metropolitan Museum of Art. Using NVIDIA DGX systems to accelerate training, they generated new AI art inspired by the historical portraits.
The technique - called adaptive discriminator augmentation, or ADA - reduces the number of training images by 10-20x while still getting great results. The same method could someday have a significant impact in healthcare, for example by creating cancer histology images to help train other AI models.
These results mean people can use GANs to tackle problems where vast quantities of data are too time-consuming or difficult to obtain, said David Luebke, vice president of graphics research at NVIDIA. I can't wait to see what artists, medical experts and researchers use it for.
The research paper behind this project is being presented this week at the annual Conference on Neural Information Processing Systems, known as NeurIPS. It's one of a record 28 NVIDIA Research papers accepted to the prestigious conference.
This new method is the latest in a legacy of GAN innovation by NVIDIA researchers, who've developed groundbreaking GAN-based models for the AI painting app GauGAN, the game engine mimicker GameGAN, and the pet photo transformer GANimal. All are available on the NVIDIA AI Playground.
The Training Data Dilemma Like most neural networks, GANs have long followed a basic principle: the more training data, the better the model. That's because each GAN consists of two cooperating networks - a generator, which creates synthetic images, and a discriminator, which learns what realistic images should look like based on training data.
The discriminator coaches the generator, giving pixel-by-pixel feedback to help it improve the realism of its synthetic images. But with limited training data to learn from, a discriminator won't be able to help the generator reach its full potential - like a rookie coach who's experienced far fewer games than a seasoned expert.
It typically takes 50,000 to 100,000 training images to train a high-quality GAN. But in many cases, researchers simply don't have tens or hundreds of thousands of sample images at their disposal.
With just a couple thousand images for training, many GANs would falter at producing realistic results. This problem, called overfitting, occurs when the discriminator simply memorizes the training images and fails to provide useful feedback to the generator.
In image classification tasks, researchers get around overfitting with data augmentation, a technique that expands smaller datasets using copies of existing images that are randomly distorted by processes like rotating, cropping or flipping - forcing the model to generalize better.
But previous attempts to apply augmentation to GAN training images resulted in a generator that learned to mimic those distortions, rather than creating believable synthetic images.
A GAN on a Mission NVIDIA Research's ADA method applies data augmentations adaptively, meaning the amount of data augmentation is adjusted at different points in the training process to avoid overfitting. This enables models like StyleGAN2 to achieve equally amazing results using an order of magnitude fewer training images.
As a result, researchers can apply GANs to previously impractical applications where examples are too scarce, too hard to obtain or too time-consuming to gather into a large dataset.
Different editions of StyleGAN have been used by artists to create stunning exhibits and produce a new manga based on the style of legendary illustrator Osamu Tezuka. It's even been adopted by Adobe to power Photoshop's new AI tool, Neural Filters.
With less training data required to get started, StyleGAN2 with ADA could be applied to rare art, such as the work by Paris-based AI art collective Obvious on African Kota masks.
Another promising application lies in healthcare, where medical images of rare diseases can be few and far between because most tests come back normal. Amassing a useful dataset of abnormal pathology slides would require many hours of painstaking labeling by medical experts.
Synthetic images created with a GAN using ADA could fill that gap, generating training data for another AI model that helps pathologists or radiologists spot rare conditions on pathology images or MRI studies. An added bonus: With AI-generated data, there are no patient data or privacy concerns, making it easier for healthcare institutions to share datasets.
NVIDIA Research at NeurIPS The NVIDIA Research team consists of more than 200 scientists around the globe, focusing on areas including AI, computer vision, self-driving cars, robotics and graphics. Over two dozen papers authored by NVIDIA researchers will be highlighted at NeurIPS, the year's largest AI research conference, taking place virtually from Dec. 6-12.
Check out the full lineup of NVIDIA Research papers at NeurIPS.
Main images generated by StyleGAN2 with ADA, trained on a dataset of fewer than 1,500 images from the Metropolitan Museum of Art Collection API.
Most recent headlines
13/03/2025
TAG Video Systems and Harmonic Partner to Deliver Enhanced Real-Time Monitoring ...
13/03/2025
DigitalGlue Invites NAB Visitors to Experience New Features in Managed Storage P...
13/03/2025
FOR-A America Theme - Connecting the Present, Building the Future - Comes to Lif...
13/03/2025
CTIA, the wireless industry association, has named former FCC Chairman Aji Pai as its President and Chief Executive Officer, effective April 1. He replaces Mere...
13/03/2025
he National Association of Broadcasters is urging the FCC to end its investigation of that controversial CBS interview with Kamala Harris stating there is no ...
13/03/2025
MIAMI CBS News & Stations continues to expand its use of augmented and virtual reality technologies with the planned launch of an augmented reality/virtual real...
13/03/2025
Srividhya Srinivasan, co-founder and chief customer success & innovation Officer at Amagi, tells TVBEurope how staying ahead of the latest trends is essential f...
13/03/2025
WASHINGTON FCC Chairman Brendan Carr announced that the agency has launched a massive, new deregulatory initiative that could potentially subject virtually all ...
13/03/2025
SUNNYVALE, Calif. SDVI has announced that it has integrated its Rally media supply chain management platform with the Spectra Vail multi-cloud data management s...
13/03/2025
ATLANTA Gray Media has announced that the Federal Communications Commission (FCC) has granted a waiver of its local ownership rules to permit Gray Media to acqu...
13/03/2025
TV Tech: What do you anticipate will be the most significant technology trends at the 2025 NAB Show?...
13/03/2025
Watch Student Naomi Soleils Folk Pop Performance on The Voice The songwriting major sang Stars by Grace Potter and the Nocturnals during the blind auditions.
...
13/03/2025
Here is your host, Patrick Kielty!
Shake your Shamrocks, Patrick Kielty will be...
13/03/2025
This St. Patrick's weekend, RT invites you to celebrate all things Irish, showcasing the very best of Irish sport, entertainment and live coverage from the...
13/03/2025
What's next in AI is at GTC 2025. Not only the technology, but the people and ideas that are pushing AI forward - creating new opportunities, novel solution...
13/03/2025
Facebook
Twitter
LinkedIn
By combining T-Mobile's robust network, Thal...
13/03/2025
Bundle up - GeForce NOW is bringing a flurry of Blizzard titles to its ever-expanding library.
Prepare to weather epic gameplay in the cloud, tackling the genr...
13/03/2025
AI is leveling up the world's most beloved games, as the latest advancements...
13/03/2025
PC game modding is massive, with over 5 billion mods downloaded annually. Mods p...
12/03/2025
O Spotify acaba de lan ar o relat rio Loud & Clear deste ano, uma vis o transpar...
12/03/2025
Spotify acaba de presentar el informe Loud & Clear de este a o, una mirada trans...
12/03/2025
Ramadan, a period of profound spiritual significance for Muslims worldwide, is a time for fasting, prayer, reflection, and community. Enrich your experience thi...
12/03/2025
Spotify has just unveiled this year's Loud & Clear report, a transparent loo...
12/03/2025
More than 70% of FAST programming has been produced since 2010, according to new Gracenote report
NEW YORK March 12, 2025 Gracenote, the content data busin...
12/03/2025
GEONA, Nev. Independent digital and linear advertising rep firm Viamedia will adopt cloud-based ShowSeeker Pilot as its primary ad campaign and order management...
12/03/2025
BRUSSELS Mediagenix has announced that Wael Yasin has joined the company as sales director Central Europe....
12/03/2025
LONDON A new study highlights opportunities for shoppable TV and the massive impact online consumer spending is having on the economy, with Omdia predicting tha...
12/03/2025
CUPERTINO, Calif. Interra Systems has announced that Comcast Technology Solutions has integrated recent updates to BATON Version 9 into its operations....
12/03/2025
Nevion announces new 400G addition to its eMerge SDN media fabric offering
Brie Clayton March 12, 2025
0 Comments
High-capacity switch enhances existi...
12/03/2025
DaVinci Resolve Studio Delivers Cinematic Sound for Adam Bol
Brie Clayton March 12, 2025
0 Comments
Feature film relies on DaVinci Resolve Studio for ...
12/03/2025
SVT Leverages Ateliere Live to Pioneer 100% Software-Defined Production at Rally...
12/03/2025
Berklee Awards Fenway Neighborhood Improvement Grant to Four Organizations A total of $17,000 will be distributed among the Boston-based nonprofits.
By
Madd...
12/03/2025
TVBEuropes Jenny Priestley sits down with new SMPTE president Richard Welsh to discuss his aims for the organisation going forward, its efforts to attract a you...
12/03/2025
The new process has significantly enhanced operational efficiencies, with automation freeing up creators to concentrate on higher value tasks
By Matthew Corrig...
12/03/2025
Jellyfish Pictures, which has offices in London and Sheffield, said it has been battling hard in the face of strong headwinds over the last 12 months
By Jenny ...
12/03/2025
Visit Booth #W2213 to Experience the Latest in AI-Driven Content Discovery and Workflow Automation
Bitcentral, a leader in media workflow solutions, is set to ...
12/03/2025
CHICAGO Jeff Lilly has been named WGN-TV director of technology effective March 17, 2025, according to Ric Harris, WGN-TV vice president and general manager....
12/03/2025
MOUNTAIN VIEW, Calif. A new study from LG Ad Solutions indicates that consumers want more features that would allow them to shop for products on the connected T...
12/03/2025
BOTHELL, Wash. The Alliance for IP Media Solutions (AIMS), Advanced Media Workflow Association (AMWA) and the Video Services Forum (VSF) will once again present...
12/03/2025
PHILADELPHIA Comcast announced that it has upgraded Xfinity Internet speeds for more than 20 million customers for no additional cost....
12/03/2025
Create with Maxon: Cinema 4D Fundamentals Workshop - March 12-14
Brie Clayton March 11, 2025
0 Comments
Makin' Waffles with Elly Wade
During Marc...
12/03/2025
Ottawa, Canada - March 12, 2025 - Ross Video is pleased to announce the appointment of Aaron Tunnell as Business Development Director, Cloud Solutions, reinforc...
12/03/2025
12 Mar 2025
VEON to release 4Q 2024 trading update on 20 March 2025 Dubai, 12 March 2025 - VEON Ltd. (NASDAQ: VEON), a global digital operator, today confirms ...
12/03/2025
Getting set for the Winter Olympics: Inside SVT's ongoing software-based pro...
12/03/2025
THE PLAYERS Championship 2025: AI Commentary, Shared-Reality Viewing at COSM, La...
12/03/2025
Seamless SMPTE 2110: A Discussion on Making the Move to IP Less Painful Leaders from Netflix, Monumental Sports & Entertainment, LinkedI, and Megapixel address ...
12/03/2025
Opening Doors of Perception: Meta's Ajit Ninan on Rethinking AR/MR Perceptua...
12/03/2025
With the 2025 Formula 1 season set to begin in Melbourne on March 16, Sky Sports is gearing up to deliver its most comprehensive F1 coverage yet. Following a th...
12/03/2025
Unicanal and Trece TV revolutionize Digital TV in Paraguay with Rohde & Schwarz ...
12/03/2025
Honoring Service, Celebrating Sport: Rohde & Schwarz UK Sponsors NavyFit Rugby D...