ArXivCSExplorer
☆☆Bookmarks🏆RSSHow to UseFAQ
Built with and by Teycir Ben Soltane•
How to Use•FAQ•GitHub•arXiv.org•
Share:

20 results for “spectrograms”

CS papers only

Hybrid search: Keyword + semantic, ranked by combined score.ⓘ

Want pure semantic search? Try claim verification →

eess.ASEmpiricalRecentJul 3, 2026

CaReCoS: A Spectrogram based Visual Benchmark for Cardiac, Respiratory and Cough Sounds

Harshit Rajgarhia, Shuubham Ojha, Akhil Pothanapalli, Rachuri Lokesh +3 more

The paper introduces CaReCoS, a benchmark for multimodal reasoning over medical acoustic spectrograms, and evaluates the performance of vision and omni models, finding a maximum accuracy of 51.2%.

View →
cs.SDeess.SYEmpiricalRecentJul 23, 2026

Spectrogram-Based Joint Detection, Localization, and Classification of Events in Continuously Recorded IBR Waveforms

Shivanshu Tripathi, Maziar Raissi, Hamed Mohsenian-Rad

A spectrogram-based framework is proposed for event detection, localization, and classification in power system waveforms using short-time Fourier transform.

View →
cs.SDcs.LGRecentJun 1, 2026

Parameter-efficient Dual-encoder Architecture with Differentiable Choquet Integral Fusion for Underwater Acoustic Classification

Amirmohammad Mohammadi, Joshua Peeples, Alexandra Van Dine

The paper proposes a dual-encoder architecture that fuses processed acoustic waveforms and spectrograms using a differentiable Choquet integral to improve underwater acoustic classification while main…

View →
cs.SDEmpiricalRecentJul 20, 2026

SSTMark: Robust Training-Free Semantic-Level Speech Watermarking

Kuan-Lin Chu, Jun-Cheng Chen, Chun-Shien Lu

This paper introduces SSTMark, a training-free speech watermarking framework that encodes watermark information into the semantic content of generated speech.

View →
q-bio.NCeess.SPEmpiricalRecentJun 23, 2026

EEG Interpretation Across Chant Listening: A Single-Subject Pilot Investigation Using Spectral and Functional Connectivity Analysis

Prerna Singh, Aishwarya Ghosh, Neelam Sinha, Deepti Navaratna

This paper investigates neural activity during five auditory conditions using EEG recordings from a 5-year-old participant, revealing condition-specific modulation of neural oscillatory activity and d…

View →
cs.SDcs.IRcs.LGEmpiricalRecentJul 24, 2026

Reflector: Arrangement-Aware Harmonic Retrieval for Sample-Based Composition

Austin Rockman

The paper introduces Reflector, an interactive audio workstation that adapts pitch-class retrieval as compositions evolve, using a learned embedding space based on a hand-designed oracle.

View →
cs.SDEmpiricalRecentJul 8, 2026

Rag Classification of Tagore Songs using Symbolic Music Notation and Novel Weighted Distance Measures

Chandan Misra, Swarup Chattopadhyay

This paper formulates rag identification in Rabindra Sangeet as a supervised classification problem using symbolic music-sheet notations and constructs a rag-labelled dataset. It introduces a weighted…

View →
cs.LGcs.AIRecentMay 30, 2026

Dive into Waves: Morlet Spectral Transformer for Cross-Subject Emotion Decoding from EEG

Jiaxin Qing, Lexin Li

The paper proposes the Morlet Spectral Transformer (MST), a novel architecture that effectively decodes cross-subject emotion from EEG by designing specialized spectral and spatial representations, ou…

View →
cs.SDeess.ASEmpiricalRecentJul 26, 2026

Automatic Audio Equalization with Semantic Embeddings

Eloi Moliner, Vesa Välimäki, Konstantinos Drossos, Matti S. Hämäläinen

This paper proposes a data-driven method for automatic blind audio equalization using a deep neural network and semantic embeddings.

View →
eess.AScs.PLcs.SDEmpiricalRecentJun 19, 2026

Compiling Differentiable Audio Graphs to Real-Time DSP

Facundo Franchino, Sebastian J. Schlecht

This paper presents ADAC, a compiler that converts trained differentiable audio models into efficient FAUST code for real-time audio effects.

View →
stat.MLcs.ITcs.LGTheoreticalRecentJul 7, 2026

Separation Capacity of Scattering Networks on Low-Dimensional Datasets

Konstantin Häberle, Helmut Bölcskei

This paper identifies scattering network architectures that maximize separation capacity for data with low intrinsic dimension by characterizing and bounding the separation capacity of general feature…

View →
eess.ASEmpiricalRecentJul 18, 2026

NABEATs: Noise-Aware Audio Representation Learning

Takuya Fujimura, Yoshiki Masuyama, Gordon Wichern, Christoph Boeddeker +2 more

The paper introduces Noise-Aware BEATs (NABEATs), a noise-aware audio self-supervised learning framework that estimates clean BEATs representations from noisy audio signals using an auxiliary referenc…

View →
cs.SDeess.ASDatasetRecentJun 18, 2026

PolSeT: Polish Semantics of Timbre Dataset

Jan Jasiński

This paper introduces PolSeT, a Polish psychoacoustic and Music Information Retrieval dataset with 1901 descriptors and 18 instrument sound ratings.

View →
eess.AScs.LGEmpiricalRecentJul 3, 2026

Open-Set Source Tracing as Compositional Factors via Structured Prototypes

Santiago Rubio, Antonio Almudévar, Antonio Miguel, Eduardo Lleida +1 more

This paper proposes a new definition of source in source tracing as a compositional tuple of Architecture, Training Data, and other training factors, and introduces a framework using Structured Orthon…

View →
math.PRcs.ITTheoreticalRecentJun 23, 2026

Conditioning of incoherent sub-dictionaries sampled from a coherent dictionary

Karin Schnass

This paper shows that sub-dictionaries sampled from a coherent dictionary using a coherence rejective Poisson sampling model are well-conditioned with high probability, as long as their expected size…

View →