Skip to content

Breaking News

WaterStone takes Salem Media private

RCS promotes Alissa Pollack to president of Mediabase

Inovonics opens bookings for Aaron 656

RedTech Magazine July/August 2026 drives growth through connections

U.K. gets new DAB radio broadcast platform

U.K. radio reaches record 51.1 million listeners in Q2

Broadcast group likes Australian reform package

U.S. Radio Hall of Fame announces 2026 Legends class

EMF to acquire Beasley stations in Charlotte and Las Vegas

SBE releases U.S. engineer salary survey

Monday August 24, 2026
Partners
Newsletter
Contact us
About
RedTech RedTech
  • News & Business
  • Strategy & Views
  • Technology
  • Products
  • All stories
  • Contact
  • Advertise
Telos Alliance Shares Omnia Volt Webinar Recording
Trending
Telos Alliance Shares Omnia Volt Webinar Recording

Featured

Nostalgie expands visual radio programming on LN24

Nostalgie’s weekday and weekend radio programs are now on Belgium’s French-language news TV channel

Beasley Media Group, WYUU, United States, concerts, marketing, promotions
Featured

Beasley Tampa schedules “Maxima” concert for September

Hispanic/Latin-flavored music to be on tap, Sept. 27

Featured Strategy & Views

Inside Podcasting: Telling the story where it leads

Proper investigative audio demands structure, collaboration and risk

Featured

Genelec names recipients of 2026 AES scholarships

The two $5,000 grants support emerging work in assistive listening technology and audio education

Featured News & Business

WaterStone takes Salem Media private

Salem plans further investment under private ownership

swXtch.io, IBC 2026, AoIP, SRT
Featured IBC2026 Products

swXtch.io promises upgraded network at IBC2026

Also to announce multicasting capability

  • Contact
  • About RedTech
RedTech RedTech
  • News & Business
  • Strategy & Views
    • Strategy & Views
    • Videos
  • Technology
    • Tech Focus
  • Products
  • Events
    • RedTech Summit 2026
    • Previous RedTech Summits
      • RedTech Summit 2025
      • RedTech Summit 2024
      • RedTech Summit 2023
      • RedTech Summit 2022
    • RadioWeek 2026
      • RadioWeek 2025
      • RadioWeek 2024
      • RadioWeek 2023
    • Global Online Content Series 2024
    • Events
      • IBC2026
      • 2026 NAB Show
      • World Radio Day 2026
      • IBC2025
      • 2025 NAB Show
      • IBC2024
      • 2024 NAB Show
      • IBC2023
      • 2023 NAB Show
      • IBC2022
    • Events Calendar
  • Publications
  • News & Business
  • Strategy & Views
    • Strategy & Views
    • Videos
  • Technology
    • Tech Focus
  • Products
  • Events
    • RedTech Summit 2026
    • Previous RedTech Summits
      • RedTech Summit 2025
      • RedTech Summit 2024
      • RedTech Summit 2023
      • RedTech Summit 2022
    • RadioWeek 2026
      • RadioWeek 2025
      • RadioWeek 2024
      • RadioWeek 2023
    • Global Online Content Series 2024
    • Events
      • IBC2026
      • 2026 NAB Show
      • World Radio Day 2026
      • IBC2025
      • 2025 NAB Show
      • IBC2024
      • 2024 NAB Show
      • IBC2023
      • 2023 NAB Show
      • IBC2022
    • Events Calendar
  • Publications

Click Here to Subscribe to RedTech's Newsletter

RedTech RedTech
  • News & Business
  • Strategy & Views
    • Strategy & Views
    • Videos
  • Technology
    • Tech Focus
  • Products
  • Events
    • RedTech Summit 2026
    • Previous RedTech Summits
      • RedTech Summit 2025
      • RedTech Summit 2024
      • RedTech Summit 2023
      • RedTech Summit 2022
    • RadioWeek 2026
      • RadioWeek 2025
      • RadioWeek 2024
      • RadioWeek 2023
    • Global Online Content Series 2024
    • Events
      • IBC2026
      • 2026 NAB Show
      • World Radio Day 2026
      • IBC2025
      • 2025 NAB Show
      • IBC2024
      • 2024 NAB Show
      • IBC2023
      • 2023 NAB Show
      • IBC2022
    • Events Calendar
  • Publications

Click Here to Subscribe to RedTech's Newsletter

Featured Strategy & Views

When a golden ear meets a neural net

by Davide Moro February 10, 2026 12 min read
 When a golden ear meets a neural net
Pedro Leite, left, and Luiz Fernando Kruszielski in a Globo post production room. Photo: Carlos Eduardo Rocha Miranda
Print Friendly, PDF & Email

RIO DE JANEIRO — It began like many good stories: Two colleagues with entirely different backgrounds, fired by a common, genuine passion for hands-on research, unexpectedly collaborating on an abstract concept. The one, Pedro Leite, machine learning engineer and AI researcher at Grupo Globo, has spent years exploring generative audio systems. The other, Luiz Fernando Kruszielski, is an innovation technologies specialist at Globo and a veteran sound engineer with a reputation for a “golden ear” — the ability to instinctively hear what others miss. 

Kruszielski and Leite were intrigued by early AI technology that enabled singers to create different voices and wondered whether it could produce something suitable for radio broadcasting.

The AI speech transformation process changes a performer’s voice into a target voice using information about timbre, pitch and spectrum — a sort of “voice DNA” — from a voice model. Emotions come from the performer’s voice — the source. This way, “the source plays with the target (voice), and the result is a very reliable voice,” Kruszielski said. “In broadcast applications, listeners shouldn’t be able to perceive the resultant voice as something unnatural.” 

Not all sounds have a fundamental frequency, but in voiced speech, the vocal folds generate a fundamental frequency — the basic rate at which they vibrate — and this determines the perceived pitch of the voice. The harmonic structure built on this base frequency shapes the timbre, the tonal quality that makes one voice sound different from another. Because every speaker has a unique combination of fundamental frequency and harmonic patterns, AI-assisted speech-to-speech systems must analyze these elements to recreate a speaker’s characteristic timbre while preserving the timing, rhythm, melody and emotional contours of the original performance. 

Perhaps the most impactful application for high-volume productions is converting low-quality recordings into studio-grade audio. 

Transforming into a wolf

In the beginning, Kruszielski and Leite explored the technology simply out of interest, running quick tests on open source models with hardware available in their lab. It was interesting but not obviously useful. The time to push their research a step further came when they encountered a particular creative challenge. A key scene of a Globo drama series “Vermelho Sangue” (“Blood Red”) required a girl to transform into a werewolf, and her voice had to transform accordingly. The writing team wanted accuracy — not a movie-style monster or a generic animal sound, but the vocalization of a specific species. They tried every traditional technique, including layering recordings, shifting pitch and formants (the resonant frequencies of the vocal tract that shape the characteristic timbre of a voice or vowel sound, independent of pitch), blending organic and synthetic tones, and using early voice-to-voice models. Yet nothing felt authentic enough.

Promotional poster for Globo’s “Vermelho Sangue” series. Photo: Grupo Globo

Kruszielski and Leite realized that if they wanted authenticity, they had to start with an authentic source. It would become the defining insight of the entire project. That meant real wolf vocalizations — scientifically documented recordings.

So, their next meeting was with a wildlife biologist specializing in animal vocalizations who had an extensive archive of wolf recordings. Unfortunately, the first outputs based on those samples sounded synthetic and unrealistic. Kruszielski and Leite didn’t give up. The field recordings included layers of environmental noise, such as wind, rustling vegetation and distant birds. To train a model capable of producing realistic transformations, they needed isolated vocalizations. So, they carefully cleaned each file, separating harmonic content from wildlife ambience and removing contamination without damaging the integrity of the wolf’s “voice.”

As soon as the refined dataset was fed into the AI model, everything changed. The voice of the transformed girl carried the texture, tension and resonance of a real animal. When they played the result for the director, he stood up from his chair. AI speech-to-speech was no longer an experiment. It was ready for production.

Built for sound engineers

The backbone of the system Kruszielski and Leite designed is the open-source RVC Project AI algorithm. Although powerful and flexible, it is not intended for everyday workflows in a sound department. It required command-line interfaces, cryptic flags, hidden configuration files and robust IT skills.

The purpose-designed graphic interface allows sound engineers to interact with the RVC processing engine in a familiar way.
Photo: Carlos Eduardo Rocha Miranda

Kruszielski and Leite designed a custom GUI specifically for audio specialists, with familiar controls. By reframing the AI system as studio-grade software rather than a technical experiment, they made it accessible to colleagues across the production team. The entire processing runs on a consumer-grade, gamer-level graphics card from the Nvidia RTX 40 family, which retails at a price well within the reach of any production studio and capable of faster-than-real-time processing. What started as a two-person side project became part of Globo’s broader audio production workflow.

The team applied lessons learned from the wolf transformation to everyday production challenges, such as correcting minor dialog mistakes without recalling actors for costly retakes. They have also used the system to modify accents. In one case, an actress needed to perform with a Yiddish accent that proved challenging during the shoot. With a reference sample and a carefully tuned AI audio model, they were able to shift her accent in post while preserving the shape, emotion and timing of her original performance. The result was seamless and expressive.

The risk of collapsing an illusion

Perhaps the most impactful application for high-volume productions is converting low-quality recordings into studio-grade audio. Actors or talents who are traveling or unavailable for booth time can record a line on a phone, in a hotel room, or anywhere convenient. The AI system reperforms the lines using the actor’s or talent’s vocal identity, producing audio that sounds as if it were recorded in ideal studio conditions.

For teams producing large volumes of scripted content, this flexibility can be transformative, improving the quality of everyday productions while saving time.

Kruszielski believes the technology does not support the idea AI might soon replace actors wholesale. While AI models can reproduce tone, timbre and certain expressive gestures, they cannot comprehend the subtle patterns that make a human performance unique. “An actor is defined not just by the sound of their voice but by their micropauses, breathing rhythm, nuanced hesitations, emotional timing and the way tension rises and releases across a line,” he explained.

Current AI models can approximate fragments of this but not sustain these characteristics across a long monolog without drifting into something increasingly unnatural. “As soon as a listener senses that something is off, the immersive experience breaks. The illusion collapses,” Kruszielski warned. For that reason, the team insists the technology is best understood not as a replacement for performers but as a tool that enhances their work, offers flexibility, and preserves creative intent.

After receiving a Master of Science in Engineering, the author worked for Telecom Italia and the Italian public broadcaster, Rai. Based in Bergamo, Italy, he now spends his time as a broadcast consultant for radio stations and equipment manufacturers, specializing in project management, network design and field measurement.

This article first appeared in the January/February 2026 edition of RedTech Magazine. You can read or download this edition for free here. You can access past editions of RedTech Magazine, also for free, here.

You might be interested in these stories

Super Hi-Fi and Connoisseur Media partner on AI

Saudi Media Forum to spotlight the kingdom’s media ambitions

DAB+ expansion gains momentum across Europe

Tags: AI AI Audio Grupo Globo RedTech Magazine January/February 2026
Previous post
Next post

Davide Moro

contributor


Most Recent
Featured

Nostalgie expands visual radio programming on LN24

August 24, 2026
Featured

Beasley Tampa schedules “Maxima” concert for September

August 24, 2026
Featured

Inside Podcasting: Telling the story where it leads

August 24, 2026
Latest Newsletters

13 Aug 2026 – New RedTech Magazine! | Inspiring Tomorrow’s Broadcasters | U.K. Radio Celebrates

10 Aug 2026 – RedTech Magazine July/August Is Now Available!

6 Aug 2026 – Canadian Radio Leads | NAB Joins Radio Ready | DAB In-Car Growth

30 July 2026 – French Radio Steps Up | Belgium’s In-Car Radio Stand | ARIAS Seeks producers

23 July 2026 – Intimate Spatial Sound | Tracking European Listening | Real-Time Relevance

16 July 2026 – Interchangeable Studios | Live Earbud Podcasting | Italy’s Local Radio Milestone

9 July 2026 – Broadcast Beats Streaming | Audio Academy Applications | Greatest Hits Coup

2 July 2026 – Switching Off FM | Radio News Slipping | Add Radio, Grow Profits

25 June 2026 – Summit’s Future Focus | MBC Powers Saudi Audio | Radio Outranks Spotify

18 June 2026 – RedTech Summit Kicks Off | In-Car Radio Strong | Delivering Dylan’s Voice

11 June 2026 – Radio Holds In France | New Creative Radio Resource | DRM For Wearables?

4 June 2026 – Power of Live Music | FM Switch-Off Plan | Improving Radio Journalism

1 June 2026 – RedTech Magazine May/June champions intelligent adaptation

28 May 2026 – TOPradio Tests AI DJ | Italy Radio’s Local Advantage | EBU Issues In-Car Position

21 May 2026 – Swedish Regional Licenses for NRJ | Capitalizing On Nostalgia | U.K. Radio Resilient

14 May 2026 – Flanders DAB+ Plan | Securing In-Car Radio | Good to Great Radio #10

7 May 2026 – Human Voice or AI Fake? | The Dashboard Fight | Unpredictable Programming

30 April 2026 – Innovations Reshaping Radio | Audio Without Boundaries | Cabsat Reschedules

22 April 2026 – Insights From the NAB Show | Media Convergence Scores | Cumulus Secures Approval

16 April 2026 – NAB Show Primer | K-pop Radio | Canadian Ad Decline

9 April 2026 – ARN Shifts Focus | Bauer Embraces Android | Belgium Plans Crisis Radio

6 April 2026 – RedTech Special Edition ‘The Innovators 2026’ Is Now Available

2 April 2026 – Radio’s Next Phase | New BBC Chief | Creating Tune-In

30 March 2026 – RedTech Magazine March/April is Here!

26 March 2026 – Celebrating Radio Luxembourg | RCS Names New Chief | Radiodays Riga Recap

Related Stories for you

Guest Commentary: Authenticity vs. AI — Embracing radio’s real competitive advantage

by Dan McQuillin August 17, 2026 8 min read

AI tools in radio can help broadcasters do what AI cannot

Jutel to connect radio workflows at IBC2026

by RedTech Staff August 13, 2026 4 min read

The lineup connects automation, mobile contribution, audio production and AI-powered monitoring

Telos Alliance adds AI-powered source separation to AudioTools Server

by RedTech Staff July 15, 2026 4 min read

New AudioShake module enables broadcasters to extract dialogue, music and effects from completed audio mixes

RedTech RedTech

RedTech International SAS
250 bis boulevard Saint-Germain
75007 Paris, France

contact@redtech.pro

Subscribe to our newsletter

About

About Us
Work With Us
Contact Us

Advertising

Advertise

Useful Links

Partners
Newsletter

more

Terms and Conditions
Privacy Policy

latest news

Featured

Nostalgie expands visual radio programming on LN24

Beasley Media Group, WYUU, United States, concerts, marketing, promotions
Featured

Beasley Tampa schedules “Maxima” concert for September

Featured

Inside Podcasting: Telling the story where it

Featured

Genelec names recipients of 2026 AES scholarships

Featured

WaterStone takes Salem Media private

Follow us:

Copyright RedTech International 2026. All Rights Reserved