EscFOA: Enhancing Spatial Learning for Visually Impaired Learners via Generative Spatial Audio in 360-Degree Educational Environments
This paper proposes EscFOA, a framework that uses geometry-aware spatial audio to enhance immersive educational environments for visually impaired learners, outperforming conventional audio methods.
The paper introduces EscFOA, a novel framework that generates geometry-consistent spatial audio to enhance immersive educational environments for visually impaired learners.
Keywords
Before reading this…
Applications
- →Education
- →Accessibility
To understand this paper, make sure you know these concepts first:
- Understanding of spatial audio and immersive educational environments.find papers →
- Familiarity with 3D Gaussian Splatting and conditional diffusion models.find papers →
Abstract
More Like ThisImmersive 360-degree educational environments often lack accessible spatial structure, limiting visually impaired learners' ability to orient, explore, and construct mental representations. This paper proposes EscFOA, a geometry-aware spatial audio generation framework designed as an \emph{acoustic scaffolding} to support spatial cognition. By integrating 3D Gaussian Splatting (3DGS) with conditional diffusion models, EscFOA reconstructs scene geometry from 360-degree videos to synthesize high-fidelity spatial audio consistent with the environmental structure. Explicitly targeting learning outcomes like independent spatial orientation and reduced cognitive load, EscFOA significantly outperforms conventional monaural and stereo audio in supporting spatial learning behaviors among blindfolded sighted participants (simulating visually impaired learners). These findings demonstrate that geometry-consistent generative audio can effectively enable inclusive access to complex spatial learning materials.