20 results for “spectrograms”
CS papers onlyHybrid search: Keyword + semantic, ranked by combined score.ⓘ
Want pure semantic search? Try claim verification →
The paper introduces CaReCoS, a benchmark for multimodal reasoning over medical acoustic spectrograms, and evaluates the performance of vision and omni models, finding a maximum accuracy of 51.2%.
A spectrogram-based framework is proposed for event detection, localization, and classification in power system waveforms using short-time Fourier transform.
The paper proposes a dual-encoder architecture that fuses processed acoustic waveforms and spectrograms using a differentiable Choquet integral to improve underwater acoustic classification while main…
This paper introduces SSTMark, a training-free speech watermarking framework that encodes watermark information into the semantic content of generated speech.
This paper investigates neural activity during five auditory conditions using EEG recordings from a 5-year-old participant, revealing condition-specific modulation of neural oscillatory activity and d…
The paper introduces Reflector, an interactive audio workstation that adapts pitch-class retrieval as compositions evolve, using a learned embedding space based on a hand-designed oracle.
This paper formulates rag identification in Rabindra Sangeet as a supervised classification problem using symbolic music-sheet notations and constructs a rag-labelled dataset. It introduces a weighted…
The paper proposes the Morlet Spectral Transformer (MST), a novel architecture that effectively decodes cross-subject emotion from EEG by designing specialized spectral and spatial representations, ou…
This paper proposes a data-driven method for automatic blind audio equalization using a deep neural network and semantic embeddings.
This paper presents ADAC, a compiler that converts trained differentiable audio models into efficient FAUST code for real-time audio effects.
This paper identifies scattering network architectures that maximize separation capacity for data with low intrinsic dimension by characterizing and bounding the separation capacity of general feature…
The paper introduces Noise-Aware BEATs (NABEATs), a noise-aware audio self-supervised learning framework that estimates clean BEATs representations from noisy audio signals using an auxiliary referenc…
This paper introduces PolSeT, a Polish psychoacoustic and Music Information Retrieval dataset with 1901 descriptors and 18 instrument sound ratings.
This paper proposes a new definition of source in source tracing as a compositional tuple of Architecture, Training Data, and other training factors, and introduces a framework using Structured Orthon…
This paper shows that sub-dictionaries sampled from a coherent dictionary using a coherence rejective Poisson sampling model are well-conditioned with high probability, as long as their expected size…