You are currently logged in as an
Institutional Subscriber.
If you would like to logout,
please click on the button below.
Home / Publications / E-library page
Only AES members and Institutional Journal Subscribers can download
In this work, we introduce a Neural 3D Audio Renderer (N3DAR) - a conceptual solution for creating acoustic digital twins of arbitrary spaces. We propose a workflow that consists of several stages including:;1. Simulation of high-fidelity Spatial Room Impulse Responses (SRIR) based on the 3D model of a digitalized space,
2. Building an ML-based model of this space for interpolation and reconstruction of SRIRs,
3. Development of a real-time 3D audio renderer that allows the deployment of the digital twin of a space with accurate spatial audio effects consistent with the actual acoustic properties of this space.
The first stage consists of preparation of the 3D model and running the SRIR simulations using the state-of-the-art wave-based method for arbitrary pairs of source-receiver positions. This stage provides a set of learning data being used in the second stage - training the SRIR reconstruction model. The training stage aims to learn the model of the acoustic properties of the digitalized space using the Acoustic Volume Rendering approach (AVR). The last stage is the construction of a plugin with a dedicated 3D audio renderer where rendering comprises reconstruction of the early part of the SRIR, estimation of the reverb part, and HOA-based binauralization.
N3DAR allows the building of tailored audio rendering plugins that can be deployed along with visual 3D models of digitalized spaces, where users can freely navigate through the space with 6 degrees of freedom and experience high-fidelity binaural playback in real time.;We provide a detailed description of the challenges and considerations for each of the stages. We also conduct an extensive evaluation of the audio rendering capabilities with both, objective metrics and subjective methods using a dedicated evaluation platform.
Author (s): Janusz kiewicz, Lukasz; Cenda, Piotr; Pensko, Maria; Wasilewski, Jakub; Wozniak, Tomasz
Affiliation:
SoftServe; SoftServe; SoftServe; SoftServe; SoftServe
(See document for exact affiliation information.)
AES Convention: 158
Paper Number:335
Publication Date:
2025-05-12
Import into BibTeX
Permalink: https://aes2.org/publications/elibrary-page/?id=22886
(1422KB)
Click to purchase paper as a non-member or login as an AES member. If your company or school subscribes to the E-Library then switch to the institutional version. If you are not an AES member Join the AES. If you need to check your member status, login to the Member Portal.
Janusz kiewicz, Lukasz; Cenda, Piotr; Pensko, Maria; Wasilewski, Jakub; Wozniak, Tomasz; 2025; Neural 3D Audio Renderer for acoustic digital twin creation [PDF]; SoftServe; SoftServe; SoftServe; SoftServe; SoftServe; Paper 335; Available from: https://aes2.org/publications/elibrary-page/?id=22886
Janusz kiewicz, Lukasz; Cenda, Piotr; Pensko, Maria; Wasilewski, Jakub; Wozniak, Tomasz; Neural 3D Audio Renderer for acoustic digital twin creation [PDF]; SoftServe; SoftServe; SoftServe; SoftServe; SoftServe; Paper 335; 2025 Available: https://aes2.org/publications/elibrary-page/?id=22886
@article{janusz2025neural,
author={janusz kiewicz lukasz and cenda piotr and pensko maria and wasilewski jakub and wozniak tomasz},
journal={journal of the audio engineering society},
title={neural 3d audio renderer for acoustic digital twin creation},
year={2025},
number={335},
month={may},}