Cited 0 time in
A Momentum-Based Framework with Contrastive Data Generation for Robust Sound Source Localization
| DC Field | Value | Language |
|---|---|---|
| dc.contributor.author | Kim, Hyun-Soo | - |
| dc.contributor.author | Yang, Da-Hee | - |
| dc.contributor.author | Chang, Joon-Hyuk | - |
| dc.date.accessioned | 2026-07-20T01:30:13Z | - |
| dc.date.available | 2026-07-20T01:30:13Z | - |
| dc.date.issued | 2026-04 | - |
| dc.identifier.issn | 2997-6928 | - |
| dc.identifier.issn | 2997-6995 | - |
| dc.identifier.uri | https://scholarworks.bwise.kr/hanyang/handle/2021.sw.hanyang/219344 | - |
| dc.description.abstract | We propose MoCo-SSL, a momentum-based contrastive learning framework for multi-channel sound source localization (SSL) that enhances azimuth-aware representation learning. While prior SSL studies have used contrastive learning to handle varied acoustic conditions, we emphasize hard negatives-pairs with distinct azimuths recorded in the same room-for learning fine-grained spatial cues. A curriculum-based strategy gradually increases the proportion of such samples to raise task difficulty. The momentum contrast design employs a key encoder that maintains stable embeddings during curriculum transitions and receives audio with less noise and reverberation to produce clearer azimuth cues, thereby guiding the query encoder toward robust representations. Experiments show that MoCo-SSL consistently surpasses baselines, demonstrating the value of structured and noise-resilient representation learning in challenging SSL scenarios. | - |
| dc.format.extent | 7 | - |
| dc.language | 영어 | - |
| dc.language.iso | ENG | - |
| dc.publisher | Institute of Electrical and Electronics Engineers Inc. | - |
| dc.title | A Momentum-Based Framework with Contrastive Data Generation for Robust Sound Source Localization | - |
| dc.type | Article | - |
| dc.identifier.doi | 10.1109/ASRU65441.2025.11434773 | - |
| dc.identifier.scopusid | 2-s2.0-105036504456 | - |
| dc.identifier.bibliographicCitation | ASRU 2025 - 2025 IEEE Automatic Speech Recognition and Understanding Workshop, pp 1 - 7 | - |
| dc.citation.title | ASRU 2025 - 2025 IEEE Automatic Speech Recognition and Understanding Workshop | - |
| dc.citation.startPage | 1 | - |
| dc.citation.endPage | 7 | - |
| dc.type.docType | Conference paper | - |
| dc.description.isOpenAccess | N | - |
| dc.description.journalRegisteredClass | scopus | - |
| dc.subject.keywordPlus | Acoustic generators | - |
| dc.subject.keywordPlus | Acoustic noise | - |
| dc.subject.keywordPlus | Acoustic noise measurement | - |
| dc.subject.keywordPlus | Architectural acoustics | - |
| dc.subject.keywordPlus | Audio acoustics | - |
| dc.subject.keywordPlus | Contrastive Learning | - |
| dc.subject.keywordPlus | Curricula | - |
| dc.subject.keywordAuthor | contrastive learning | - |
| dc.subject.keywordAuthor | curriculum learning | - |
| dc.subject.keywordAuthor | exponentially moving average | - |
| dc.subject.keywordAuthor | momentum | - |
| dc.subject.keywordAuthor | sound source localization | - |
| dc.identifier.url | https://ieeexplore.ieee.org/document/11434773 | - |
Items in ScholarWorks are protected by copyright, with all rights reserved, unless otherwise indicated.
222, Wangsimni-ro, Seongdong-gu, Seoul, 04763, Korea+82-2-2220-1366
COPYRIGHT © 2024 HANYANG UNIVERSITY.
Certain data included herein are derived from the © Web of Science of Clarivate Analytics. All rights reserved.
You may not copy or re-distribute this material in whole or in part without the prior written consent of Clarivate Analytics.
