Proceedings of the 11th Convention of the
European Acoustics Association
Forum Acusticum / EuroNoise 2025


Málaga, Spain
June 23 - 26, 2025





Session: A15.10 Modeling and estimation of room impulse responses with machine learning
Date: Tuesday 24 June 2025
Time: 09:40
Title: DiffusionRIR: Room Impulse Response Interpolation using Diffusion Models
Author(s): Sagi Della Torre
Mirco Pezzoli
Fabio Antonacci
Sharon Gannot
Pages: 4079-4086
DOI: https://www.doi.org/10.61782/fa.2025.0301
PDF: https://dael.euracoustics.org/confs/fa2025/data/articles/000301.pdf
Conference proceedings
Abstract

Room Impulse Responses (RIRs) characterize acoustic environments and are crucial in multiple audio signal processing tasks. High-quality RIR estimates drive applications such as virtual microphones, sound source localization, augmented reality, and data augmentation. However, obtaining RIR measurements with high spatial resolution is resource-intensive, making it impractical for large spaces or when dense sampling is required. This research addresses the challenge of estimating RIRs at unmeasured locations within a room using Denoising Diffusion Probabilistic Models (DDPM). Our method leverages the analogy between RIR matrices and image inpainting, transforming RIR data into a format suitable for diffusion-based reconstruction. Using simulated RIR data based on the image method, we demonstrate our approach’s effectiveness on microphone arrays of different curvatures, from linear to semi-circular. Our method successfully reconstructs missing RIRs, even in large gaps between microphones. Under these conditions, it achieves accurate reconstruction, significantly outperforming baseline Spline Cubic Interpolation (SCI) in terms of Normalized Mean Square Error (NMSE) and Cosine Distance (CD) between actual and interpolated RIRs. This research highlights the potential of using generative models for effective RIR interpolation, paving the way for generating additional data from limited real-world measurements.