Detailed Information

Cited 0 time in webofscience Cited 0 time in scopus
Metadata Downloads

A Momentum-Based Framework with Contrastive Data Generation for Robust Sound Source Localization

Full metadata record
DC Field Value Language
dc.contributor.authorKim, Hyun-Soo-
dc.contributor.authorYang, Da-Hee-
dc.contributor.authorChang, Joon-Hyuk-
dc.date.accessioned2026-07-20T01:30:13Z-
dc.date.available2026-07-20T01:30:13Z-
dc.date.issued2026-04-
dc.identifier.issn2997-6928-
dc.identifier.issn2997-6995-
dc.identifier.urihttps://scholarworks.bwise.kr/hanyang/handle/2021.sw.hanyang/219344-
dc.description.abstractWe propose MoCo-SSL, a momentum-based contrastive learning framework for multi-channel sound source localization (SSL) that enhances azimuth-aware representation learning. While prior SSL studies have used contrastive learning to handle varied acoustic conditions, we emphasize hard negatives-pairs with distinct azimuths recorded in the same room-for learning fine-grained spatial cues. A curriculum-based strategy gradually increases the proportion of such samples to raise task difficulty. The momentum contrast design employs a key encoder that maintains stable embeddings during curriculum transitions and receives audio with less noise and reverberation to produce clearer azimuth cues, thereby guiding the query encoder toward robust representations. Experiments show that MoCo-SSL consistently surpasses baselines, demonstrating the value of structured and noise-resilient representation learning in challenging SSL scenarios.-
dc.format.extent7-
dc.language영어-
dc.language.isoENG-
dc.publisherInstitute of Electrical and Electronics Engineers Inc.-
dc.titleA Momentum-Based Framework with Contrastive Data Generation for Robust Sound Source Localization-
dc.typeArticle-
dc.identifier.doi10.1109/ASRU65441.2025.11434773-
dc.identifier.scopusid2-s2.0-105036504456-
dc.identifier.bibliographicCitationASRU 2025 - 2025 IEEE Automatic Speech Recognition and Understanding Workshop, pp 1 - 7-
dc.citation.titleASRU 2025 - 2025 IEEE Automatic Speech Recognition and Understanding Workshop-
dc.citation.startPage1-
dc.citation.endPage7-
dc.type.docTypeConference paper-
dc.description.isOpenAccessN-
dc.description.journalRegisteredClassscopus-
dc.subject.keywordPlusAcoustic generators-
dc.subject.keywordPlusAcoustic noise-
dc.subject.keywordPlusAcoustic noise measurement-
dc.subject.keywordPlusArchitectural acoustics-
dc.subject.keywordPlusAudio acoustics-
dc.subject.keywordPlusContrastive Learning-
dc.subject.keywordPlusCurricula-
dc.subject.keywordAuthorcontrastive learning-
dc.subject.keywordAuthorcurriculum learning-
dc.subject.keywordAuthorexponentially moving average-
dc.subject.keywordAuthormomentum-
dc.subject.keywordAuthorsound source localization-
dc.identifier.urlhttps://ieeexplore.ieee.org/document/11434773-
Files in This Item
Go to Link
Appears in
Collections
서울 공과대학 > 서울 융합전자공학부 > 1. Journal Articles

qrcode

Items in ScholarWorks are protected by copyright, with all rights reserved, unless otherwise indicated.

Related Researcher

Researcher Chang, Joon-Hyuk photo

Chang, Joon-Hyuk
COLLEGE OF ENGINEERING (SCHOOL OF ELECTRONIC ENGINEERING)
Read more

Altmetrics

Total Views & Downloads

BROWSE