Sony Pixel Power calrec Sony

Mori Speech AI Model Helps Preserve and Promote New Zealand Indigenous Language

16/01/2024

Indigenous languages are under threat. Some 3,000 - three-quarters of the total - could disappear before the end of the century, or one every two weeks, according to UNESCO.

As part of a movement to protect such languages, New Zealand's Te Hiku Media, a broadcaster focused on the M ori people's indigenous language known as te reo, is using trustworthy AI to help preserve and revitalize the tongue.

Using ethical, transparent methods of speech data collection and analysis to maintain data sovereignty for the M ori people, Te Hiku Media is developing automatic speech recognition (ASR) models for te reo, which is a Polynesian language.

Built using the open-source NVIDIA NeMo toolkit for ASR and NVIDIA A100 Tensor Core GPUs, the speech-to-text models transcribe te reo with 92% accuracy. It can also transcribe bilingual speech using English and te reo with 82% accuracy. They're pivotal tools, made by and for the M ori people, that are helping preserve and amplify their stories.

There's immense value in using NVIDIA's open-source technologies to build the tools we need to ultimately achieve our mission, which is the preservation, promotion and revitalization of te reo M ori, said Keoni Mahelona, chief technology officer at Te Hiku Media, who leads a team of data scientists and developers, as well as M ori language experts and data curators, working on the project.

We're also helping guide the industry on ethical ways of using data and technologies to ensure they're used for the empowerment of marginalized communities, added Mahelona, a Native Hawaiian now living in New Zealand.

Building a House of Speech' Te Hiku Media began more than three decades ago as a radio station aiming to ensure te reo had space on the airwaves. Over the years, the organization incorporated television broadcasting and, with the rise of the internet, it convened a meeting in 2013 with the community's elders to form a strategy for sharing content in the digital era.

The elders agreed that we should make the stories accessible online for our community members - rather than just keeping our archives on cassettes in boxes - but once we had that objective, the challenge was how to do this correctly, in alignment with our strong roots in valuing sovereignty, Mahelona said.

Instead of uploading its video and audio sources to popular, global platforms - which, in their terms and conditions of use, require signing over certain rights related to the content - Te Hiku Media decided to build its own content distribution platform.

Called Whare K rero - meaning house of speech - the platform now holds more than 30 years' worth of digitized, archival material featuring about 1,000 hours of te reo native speakers, some of whom were born in the late 19th century, as well as more recent content from second-language learners and bilingual M ori people.

Now, around 20 M ori radio stations use and upload their content to Whare K rero. Community members can access the content through an app.

It's an invaluable resource of acoustic data, Mahelona said.

Turning to Trustworthy AI Such a trove held incredible value for those working to revitalize the language, the Te Hiku Media team quickly realized, but manual transcription required pulling lots of time and effort from limited resources. So began the organization's trustworthy AI efforts, in 2016, to accelerate its work using ASR.

No one would have a clue that there are eight NVIDIA A100 GPUs in our derelict, rundown, musky-smelling building in the far north of New Zealand - training and building M ori language models, Mahelona said. But the work has been game-changing for us.

To collect speech data in a transparent, ethically compliant, community-oriented way, Te Hiku Media began by explaining its cause to elders, garnering their support and asking them to come to the station to read phrases aloud.

It was really important that we had the support of the elders and that we recorded their voices, because that's the sort of content we want to transcribe, Mahelona said. But eventually these efforts didn't scale - we needed second-language learners, kids, middle-aged people and a lot more speech data in general.

So, the organization ran a crowdsourcing campaign, K rero M ori, to collect highly labeled speech samples according to the Kaitiakitanga license, which ensures Te Hiku Media uses the data only for the benefit of the M ori people.

In just 10 days, more than 2,500 signed up to read 200,000+ phrases, providing over 300 hours of labeled speech data, which was used to build and train the te reo M ori ASR models.

In addition to other open-source trustworthy AI tools, Te Hiku Media now uses the NVIDIA NeMo toolkit's ASR module for speech AI throughout its entire pipeline. The NeMo toolkit comprises building blocks called neural modules and includes pretrained models for language model development.

It's been absolutely amazing - NVIDIA's open-source NeMo enabled our ASR models to be bilingual and added automatic punctuation to our transcriptions, Mahelona said.

Te Hiku Media's ASR models are the engines running behind Kaituhi, a te reo M ori transcription service now available online.

The efforts have spurred similar ASR projects now underway by Native Hawaiians and the Mohawk people in southeastern Canada.

It's indigenous-led work in trustworthy AI that's inspiring other indigenous groups to think: If they can do it, we can do it, too,' Mahelona said.

Learn more about NVIDIA-powered trustworthy AI, the NVIDIA NeMo toolkit and how it enabled a Telugu language speech AI breakthrough.
LINK: https://blogs.nvidia.com/blog/te-hiku-media-maori-speech-ai/...
See more stories from nvidia

Most recent headlines

04/09/2025

Monumental Sports & Entertainment and Dalet Win Prestigious 2025 NAB Show Project of the Year Award

Monumental Sports & Entertainment (MSE), in collaboration with Dalet, has been a...

19/04/2025

SDVI Earns Both Product and Project of the Year Awards at...

SDVI, the leading platform provider for cloud-native media supply chains, today announced that the company earned multiple awards at the 2025 NAB Show, with two...

19/04/2025

Ateliere Announces CEO Transition

Ateliere Creative Technologies, a leading GenAI media software solutions company, today announced that Dan Goman has stepped down as CEO and David Bortis, Ateli...

19/04/2025

Marshall CV574 Miniature 4K UHD Camera Revolutionizes Off...

As Director of Media and Aerial Production at Terrible Herbst Motorsports, Bryan Moore is setting new standards in off-road racing media coverage thanks to his ...

19/04/2025

Lightware Launches Dual Screen Extended Desktop Taurus

A next-generation collaboration device that redefines connectivity for meeting environments Lightware, an industry-leading manufacturer of signal management so...

19/04/2025

Calrec Wins 2025 NAB Show Product of the Year Award for N...

Calrec is today announcing that its True Control 2.0 is a Remote Production winner in the 2025 NAB Show Product of the Year Awards. This official awards program...

19/04/2025

Appear and NBCUniversal Win Project of the Year at NAB Sh...

Appear, a global leader in live production technology, proudly announces it has been recognised alongside NBCUniversal with the prestigious NAB Show Delivery Pr...

19/04/2025

Deity Microphones Announces The Deity THEOS DIFB at NAB 2...

Deity Microphones, a leader in innovative audio equipment, is proud to announce the expected release of our Ultra-Wide Band IFB to the market. The THEOS DIFB wi...

19/04/2025

LiveU IQ Technology Delivers Breakthrough Network Perform...

A world renowned broadcaster and long-standing LiveU customer has successfully completed a series of live connectivity tests using LiveU's revolutionary, aw...

19/04/2025

BitFire Wins 2025 NAB Show Project of the Year and Produc...

BitFire (bitfire.tv), a longtime leader in live video transport, today announced dual NAB Show award wins at the 2025 NAB Show in Las Vegas. The company's M...

19/04/2025

BitFire Wins Three Top Awards at 2025 NAB Show

BitFire (bitfire.tv), a longtime leader in live video transport, today announced three major award wins at the 2025 NAB Show, April 5-9, in Las Vegas. The compa...

19/04/2025

Moments Lab and Satisfaction Group Form Unique Strategic...

AI video discovery company Moments Lab and Satisfaction Group, a leading independent unscripted television production company, are proud to announce a unique st...

19/04/2025

The Infrastructure of Creative Flow DigitalGlues creative...

As the media industry navigates the triple challenge of AI-driven production, distributed teams, and skyrocketing content demand, DigitalGlue s creative.space h...

19/04/2025

Miri Technologies Wins Two Prestigious Awards for X510 Bo...

Network technology startup Miri Technologies Inc. capped off its tremendously successful NAB Show debut by winning two prestigious industry awards for its cutti...

19/04/2025

Tablo Adds 45 New FAST Channels From Warner Bros. Discovery

CINCINNATI Scripp's Nuvyyo USA has concluded a deal with Warner Bros. Discovery to bring 45 FAST channels to Nuvyyo's Tablo TV device....

19/04/2025

Court Overrules FCC's $57 Million Fine Against AT&T

In a ruling that could have broader implications on the legality of regulatory agencies levying fines through administrative proceedings, the 5th U.S. Circuit C...

19/04/2025

FCC Chair Carr Blasts Comcast Over MSNBC Coverage

WASHINGTON Federal Communications Commission chair Brendan Carr has blasted Comcast over MSNBC's coverage of the deportation of Kilmar Abrego Garcia in a so...

19/04/2025

Berklee NYC and NYC Media Launch Season 3 of Inside Power Station @BerkleeNYC

Berklee NYC and NYC Media Launch Season 3 of Inside Power Station @BerkleeNYC This season features faculty member Arun Pandian as the new host and interviews ...

18/04/2025

Everyone Is Cordially Invited to Celebrate Queer Joy in The Wedding Banquet

Director Andrew Ahn, alongside actors Youn Yuh-jung and Joan Chen, takes a photo of the audience after the premiere of his film The Wedding Banquet at Eccles ...

18/04/2025

U.S. Judge Rules Google Illegally Monopolized Ad Technologies

In a ruling that could have a major impact on the digital advertising market, a federal judge has ruled that Google has monopolized some types of advertising te...

18/04/2025

TV News Outlets See March Spike in Social Media Usage

Broadcast and cable TV news outlets saw strong social media growth in March, according to new data from the social video analytics company Tubular Labs ....

18/04/2025

Berklee Student Yukai Yang Named 2025 Yamaha Young Performing Artist

Berklee Student Yukai Yang Named 2025 Yamaha Young Performing Artist The drummer secured a spot among the elite winners in this years competition. By Maddie...

18/04/2025

Boston Conservatory Alums Bring Real Women Have Curves to Broadway

Boston Conservatory Alums Bring Real Women Have Curves to Broadway The Latin American immigrant community takes center stage in a new musical featuring Tatian...

18/04/2025

UPDATED: Broadcasters Urge FCC to Hit the Delete Button on Antiquated Regs

WASHINGTON The FCC's call for public comments and suggestions on outdated regulations that it should be eliminated, has prompted a slew of fillings from bro...

18/04/2025

Federal Judge Rules Google Illegally Monopolized Ad Technologies

In a ruling that could have a major impact on the digital advertising market, a federal judge has ruled that Google has monopolized some types of advertising te...

18/04/2025

AMS, VideoAmp Collaborate on Cross-Channel Targeting and Measurement

PEARL RIVER, N.Y. Global media solutions company Active Media Services (AMS) has formed a new relationship with VideoAmp, a measurement company for linear TV, c...

18/04/2025

Netflix Reports Strong Q1 Revenue, Operating Income

Netflix reported generally positive results for first-quarter 2025, with revenue up 13% year-over-year to $10.543 billion and operating income growing by 27% to...

18/04/2025

NHL Playoffs 2025: TNT Sports Hits the Road for Onsite Productions With Mobile Units from NEP Group, Game Creek Video

NHL Playoffs 2025: TNT Sports Hits the Road for Onsite Productions With Mobile U...

18/04/2025

EVS's Sbastien Verlaine on U.S. Expansion, Next-Generation Products

EVSs S bastien Verlaine on U.S. Expansion, Next-Generation Products Beyond replay, offerings also target asset management and media infrastructure By Ken Kersc...

18/04/2025

ESPN Unleashes 4DREPLAY as NCAA Women's Gymnastics Championships Hit ABC

ESPN Unleashes 4DREPLAY as NCAA Women's Gymnastics Championships Hit ABC Men's championships to follow Saturday night on ESPN2 By Brandon Costa, Direct...

18/04/2025

Visualizing Victory: The Latest in AR, XR, and Virtual Production in Live Sports

Visualizing Victory: The Latest in AR, XR, and Virtual Production in Live Sports This panel discussion featured leaders from ESPN, CBS Sports, Warner Bros. Disc...

18/04/2025

NHL Playoffs 2025: With 16 Games in First Six Days, ESPN Deploys Variety of Remote-Production Models in U.S., Canada

NHL Playoffs 2025: With 16 Games in First Six Days, ESPN Deploys Variety of Remo...

17/04/2025

The Ugly Stepsister: A Cinderella Body Horror Story That Will Leave a Crowd in Shambles

Emilie Blichfeldt attends the 2025 Sundance Film Festival premiere of The Ugly ...

17/04/2025

Why Resilient GPS (R-GPS) Matters for US Military Superiority: We Must Address GPS Vulnerabilities

R-GPS gives warfighters a decisive battlefield advantage by punching through adv...

17/04/2025

What NAB told us about the future of media tech

This year's NAB Show in Las Vegas marked a noticeable shift in the priorities of media and broadcast organisations. Gone are the days of chasing flashy, or ...

17/04/2025

Changing Sustainable Production in Wales and Beyond

class=attachment-thumbnail size-thumbnail f-align-center alt= decoding=async data-lazy-srcset=https://www.antonbauer.com/wp-content/uploads/2024/12/Amy-Daniel-1...

17/04/2025

Roku to Collaborate with Adobe on Real-Time Customer Data

SAN JOSE, Calif. Roku and Adobe have announced that they are collaborating on a real time data platform made possible by a a new integration of the Roku Data C...

17/04/2025

IAB: Digital Ad Revenue Surges 14.9% YoY to $259 Billion in 2024

NEW YORK Internet advertising revenues demonstrated strong growth in 2024, increasing 14.9% year-over-year to $258.6 billion, according to the IAB Internet Adv...

17/04/2025

SDVI Earns Both Product and Project of the Year Awards at 2025 NAB Show

SDVI Earns Both Product and Project of the Year Awards at 2025 NAB Show Brie Clayton April 17, 2025 0 Comments Left to right, Geoff Stedman, CMO, SDVI...

17/04/2025

Singapore Polytechnic Readies Aspiring AV Professionals for Live IP Productions with AJA

Singapore Polytechnic Readies Aspiring AV Professionals for Live IP Productions ...

17/04/2025

Calrec Wins 2025 NAB Show Product of the Year Award for True Control 2.0

Calrec Wins 2025 NAB Show Product of the Year Award for True Control 2.0 Brie Clayton April 17, 2025 0 Comments Image: The Calrec True Control 2.o on ...

17/04/2025

In Return to Berklee, Lucius Looks Back and Moves Forward

In Return to Berklee, Lucius Looks Back and Moves Forward From mood boards to live demos, the alumni band gave students an exclusive look at the process behin...

17/04/2025

MyFree DirecTV Adds 8 NBCU Channels

DirecTV's free streaming service MyFree DirecTV has just added another eight channels from NBCUniversal....

17/04/2025

GameChanger Launches in the U.S.

LOS ANGELES The virtual production company GameChanger has announced that it is expanding its global footprint by bringing its virtual production technology to ...

17/04/2025

IBCAP Launches Automated VOD Monitoring and Takedown System

DENVER The International Broadcaster Coalition Against Piracy (IBCAP) has announced that it has developed a proprietary, automated software-based system to iden...

17/04/2025

Pixalate: Roku Continues to Dominate U.S. CTV Device Market

Pixalate's new CTV Device Market Share report for Q1 2025 shows that Roku has the highest open programmatic CTV device market share in the United States, wi...

17/04/2025

Edward J. Lewis III Named Senior Vice President of Institutional Advancement

Edward J. Lewis III Named Senior Vice President of Institutional Advancement Lewis has more than 20 years of industry experience, leading fundraising initiati...

17/04/2025

The Curling Group Puts On Inaugural Curling All-Star Game in Nashville

The Curling Group Puts On Inaugural Curling All-Star Game in Nashville The location in Music City is intended to broaden the sport's appeal By Dan Daley, ...

17/04/2025

Tribeca Festival 2025 Announces TV and NOW Lineup

April 17th, 2025 Press Materials Available Here Tribeca Festival 2025 Announces TV & NOW Lineup World Premieres and Exclusive Cast Panels with Apple TV '...

17/04/2025

SVG Sit-Down: Cisco's Bryan Bedford on Providing End-to-End Support for Clients, How Industry Trends Impact Workflows

SVG Sit-Down: Cisco's Bryan Bedford on Providing End-to-End Support for Clie...