Clustering-based Image-Text Graph Matching for Domain Generalization

Park, Nokyung; Chae, Daewon; Shim, Jeongyong; Kim, Sangpil; Kim, Eun-Sol; Kim, Jinkyu

doi:10.1007/978-3-031-78192-6_26

Detailed Information

Cited 0 time in webofscience

Cited 0 time in scopus

Metadata Downloads

Clustering-based Image-Text Graph Matching for Domain Generalization

Full metadata record

DC Field	Value	Language
dc.contributor.author	Park, Nokyung	-
dc.contributor.author	Chae, Daewon	-
dc.contributor.author	Shim, Jeongyong	-
dc.contributor.author	Kim, Sangpil	-
dc.contributor.author	Kim, Eun-Sol	-
dc.contributor.author	Kim, Jinkyu	-
dc.date.accessioned	2025-01-02T09:02:02Z	-
dc.date.available	2025-01-02T09:02:02Z	-
dc.date.issued	2024-12	-
dc.identifier.issn	0302-9743	-
dc.identifier.issn	1611-3349	-
dc.identifier.uri	https://scholarworks.bwise.kr/hanyang/handle/2021.sw.hanyang/204235	-
dc.description.abstract	Learning domain-invariant visual representations is important to train a model that can generalize well to unseen target task domains. Recent works demonstrate that text descriptions contain high-level class-discriminative information and such auxiliary semantic cues can be used as effective pivot embedding for domain generalization problems. However, they use pivot embedding in a global manner (i.e., aligning an image embedding with sentence-level text embedding), which does not fully utilize the semantic cues of given text description. In this work, we advocate for the use of local alignment between image regions and corresponding textual descriptions to get domain-invariant features. To this end, we first represent image and text inputs as graphs. We then cluster nodes within these graphs and match the graph-based image node features to the nodes of textual graphs. This matching process is conducted both globally and locally, tightly aligning visual and textual semantic sub-structures. We experiment with large-scale public datasets, such as CUB-DG and DomainBed, and our model achieves matched or better state-of-the-art performance on these datasets. The code is available at: https://github.com/noparkee/Graph-Clustering-based-DG.	-
dc.format.extent	17	-
dc.language	영어	-
dc.language.iso	ENG	-
dc.publisher	Springer Verlag	-
dc.title	Clustering-based Image-Text Graph Matching for Domain Generalization	-
dc.type	Article	-
dc.publisher.location	미국	-
dc.identifier.doi	10.1007/978-3-031-78192-6_26	-
dc.identifier.scopusid	2-s2.0-85212509146	-
dc.identifier.wosid	001565035500026	-
dc.identifier.bibliographicCitation	Lecture Notes in Computer Science, v.15310, pp 390 - 406	-
dc.citation.title	Lecture Notes in Computer Science	-
dc.citation.volume	15310	-
dc.citation.startPage	390	-
dc.citation.endPage	406	-
dc.type.docType	Proceedings Paper	-
dc.description.isOpenAccess	N	-
dc.description.journalRegisteredClass	scopus	-
dc.relation.journalResearchArea	Computer Science	-
dc.relation.journalWebOfScienceCategory	Computer Science, Artificial Intelligence	-
dc.relation.journalWebOfScienceCategory	Computer Science, Interdisciplinary Applications	-
dc.relation.journalWebOfScienceCategory	Computer Science, Theory & Methods	-
dc.subject.keywordPlus	Contrastive Learning	-
dc.subject.keywordPlus	Graph embeddings	-
dc.subject.keywordPlus	Image matching	-
dc.subject.keywordAuthor	Domain Generalization	-
dc.subject.keywordAuthor	Multimodal Learning	-
dc.identifier.url	https://link.springer.com/chapter/10.1007/978-3-031-78192-6_26	-

Files in This Item: Go to Link

Appears in Collections: 서울 공과대학 > 서울 컴퓨터소프트웨어학부 > 1. Journal Articles

Show simple item record

qrcode

Related Researcher

Researcher Kim, Eun Sol photo

Kim, Eun Sol: COLLEGE OF ENGINEERING (SCHOOL OF COMPUTER SCIENCE)

Read more

Altmetrics

Total Views & Downloads

RSS_1.0 RSS_2.0 ATOM_1.0

222, Wangsimni-ro, Seongdong-gu, Seoul, 04763, Korea+82-2-2220-1366

Certain data included herein are derived from the © Web of Science of Clarivate Analytics. All rights reserved.
You may not copy or re-distribute this material in whole or in part without the prior written consent of Clarivate Analytics.

Detailed Information

Related Researcher

Altmetrics

Total Views & Downloads

BROWSE