ArXivCSExplorer
☆☆Bookmarks🏆RSSHow to UseFAQ
Built with and by Teycir Ben Soltane•
How to Use•FAQ•GitHub•arXiv.org•
Share:

20 results for “keystroke sounds”

CS papers only

Hybrid search: Keyword + semantic, ranked by combined score.ⓘ

Want pure semantic search? Try claim verification →

cs.CRcs.SDEmpiricalRecentJul 24, 2026

Transforming Keystroke Noise to Text: Self-Supervised Acoustic Eavesdropping Attacks on Keyboards

Atsunori Okada, Akira Ito, Rei Ueno, Yuichi Hayashi +1 more

Researchers present a self-supervised method for reconstructing typed text from keystroke sounds without labeled data, achieving high accuracy in various scenarios.

View →
cs.CRcs.SDRecentMay 5, 2026

DECKER: Domain-invariant Embedding for Cross-Keyboard Extraction and Recognition

Bikrant Bikram Pratap Maurya, Nitin Choudhury, Daksh Agarwal, Arun Balaji Buduru

The paper introduces DECKER, a domain-invariant framework that significantly improves cross-keyboard keystroke inference by normalizing device variations and leveraging linguistic context, demonstrati…

View →
cs.CRRecentApr 17, 2026

QUACK! Making the (Rubber) Ducky Talk: A Systematic Study of Keystroke Dynamics for HID Injection Detection

Alessandro Lotto, Francesco Marchiori, Mauro Conti

This paper introduces a systematic, privacy-preserving method using keystroke dynamics to robustly distinguish between human typing and automated HID injection attacks, independent of user identity.

View →
eess.AScs.LGEmpiricalRecentJul 3, 2026

An Intervention-Based Framework for Shortcut Diagnosis in Spoofing Countermeasures

Santiago Rubio, Pilar Bello, Dayana Ribas, Antonio Miguel +2 more

The paper proposes a diagnostic framework using controlled acoustic perturbations to identify shortcut dependencies in deepfake audio detection models, revealing non-speech intervals as a dominant sho…

View →
eess.ASEmpiricalRecentJul 18, 2026

NABEATs: Noise-Aware Audio Representation Learning

Takuya Fujimura, Yoshiki Masuyama, Gordon Wichern, Christoph Boeddeker +2 more

The paper introduces Noise-Aware BEATs (NABEATs), a noise-aware audio self-supervised learning framework that estimates clean BEATs representations from noisy audio signals using an auxiliary referenc…

View →
eess.AScs.AIcs.SDDatasetRecentJul 18, 2026

RealDESED: A Real-World Domestic Sound Event Detection Benchmark

Florian Schmid, Paul Primus, Alexander Fichtinger, Tara Jadidi +2 more

This paper introduces RealDESED, a new benchmark for domestic sound event detection with 5,710 recordings, precise annotations, and multi-annotator labeling.

View →
cs.CRRecentApr 22, 2026

VRSafe: A Secure Virtual Keyboard to Mitigate Keystroke Inference in Virtual Reality

Yijun Yuan, Na Du, Adam J. Lee, Balaji Palanisamy

The paper introduces VRSafe, a novel virtual QWERTY keyboard designed to significantly mitigate keystroke inference attacks in virtual reality by introducing false positive keystrokes and incorporatin…

View →
eess.ASEmpiricalRecentJul 3, 2026

CaReCoS: A Spectrogram based Visual Benchmark for Cardiac, Respiratory and Cough Sounds

Harshit Rajgarhia, Shuubham Ojha, Akhil Pothanapalli, Rachuri Lokesh +3 more

The paper introduces CaReCoS, a benchmark for multimodal reasoning over medical acoustic spectrograms, and evaluates the performance of vision and omni models, finding a maximum accuracy of 51.2%.

View →
cs.SDEmpiricalRecentJul 20, 2026

SSTMark: Robust Training-Free Semantic-Level Speech Watermarking

Kuan-Lin Chu, Jun-Cheng Chen, Chun-Shien Lu

This paper introduces SSTMark, a training-free speech watermarking framework that encodes watermark information into the semantic content of generated speech.

View →
cs.CLeess.ASEmpiricalRecentJul 17, 2026

Controlling Implicit Shortcut Reliance in L2 Spoken English Auto-markers

Shilin Gao, Mark J. F. Gales, Kate M. Knill

This paper proposes a new training criterion to reduce a classifier's reliance on shortcuts in language proficiency assessment systems, improving their correlation with human references.

View →
eess.AScs.CVEmpiricalRecentJul 16, 2026

WanSong v1.0 Technical Report

Binghui Chen, Pandeng Li, Yu Liu, Jingren Zhou

This paper introduces WanSong, a diffusion-based model for long-form, commercial-grade song generation that directly generates high-fidelity, multilingual songs up to 5 minutes and outputs dual stems,…

View →
cs.SDcs.AIcs.CREmpiricalRecentJul 18, 2026

Do Speech Tokens Leak Voiceprints? Speaker Inversion Attacks Against End-to-End Speech Language Models

Ye Lu, Yihan Yan, Zhaoyang Zhang, Zhitao Ou +3 more

This paper introduces Audio BERT (AuB) and SpInv, methods for recovering embeddings from speech tokens and performing speaker inversion attacks using only three seconds of frontend output.

View →
cs.CRcs.AIcs.SDRecentApr 16, 2026

Hijacking Large Audio-Language Models via Context-Agnostic and Imperceptible Auditory Prompt Injection

Meng Chen, Kun Wang, Li Lu, Jiaheng Zhang +1 more

The paper introduces AudioHijack, a framework that successfully demonstrates context-agnostic and imperceptible auditory prompt injection attacks, showing that commercial Large Audio-Language Models c…

View →
cs.CRRecentApr 4, 2026

Perceptual Gaps: ASCII Art and Overlapping Audio as CAPTCHA

Choon-Hou Rafael Chong

The paper proposes two novel CAPTCHA types—ASCII art and overlapping audio—and demonstrates that current frontier LLMs struggle significantly to solve them, suggesting they are highly effective anti-b…

View →
cs.CRcs.LGcs.SDRecentMar 18, 2026

STEP: Detecting Audio Backdoor Attacks via Stability-based Trigger Exposure Profiling

Kun Wang, Meng Chen, Junhao Wang, Yuli Wu +5 more

STEP introduces a novel, black-box, retraining-free detector that profiles audio samples using dual perturbation branches to detect backdoor attacks by exploiting the characteristic instability of hid…

View →