AXD Website

HRTF spatial upsampling in the spherical harmonics domain employing a generative adversarial network

Xuyi Hu, Jian Li, Lorenzo Picinali, Aidan O. T. Hogg

September 3, 2024

Proceedings of the 27th International Conference on Digital Audio Effects (DAFx24)

Volume

Issue

Pages

396-403

A Head-Related Transfer Function (HRTF) is able to capture alterations a sound wave undergoes from its source before it reaches the entrances of a listener's left and right ear canals, and is imperative for creating immersive experiences in virtual and augmented reality (VR/AR). Nevertheless, creating personalized HRTFs demands sophisticated equipment and is hindered by time-consuming data acquisition processes. To counteract these challenges, various techniques for HRTF interpolation and up-sampling have been proposed. This paper illustrates how Generative Adversarial Networks (GANs) can be applied to HRTF data upsampling in the spherical harmonics domain. We propose using Autoencoding Generative Adversarial Networks (AE-GAN) to upsample low-degree spherical harmonics coefficients and get a more accurate representation of the full HRTF set. The proposed method is bench-marked against two baselines: barycentric interpolation and HRTF selection. Results from log-spectral distortion (LSD) evaluation suggest that the proposed AE-GAN has significant potential for upsampling very sparse HRTFs, achieving 17% improvement over baseline methods.

Visit Publication

Authors from the Audio Experience Design Team

Related Project

SONICOM

Related Tools & Devices

No items found.

AUDIO EXPERIENCE DESIGN

Imperial College London

About

Selected Publications

Projects

Tools & Devices

Facilities

News

Team

Summer Science