Self-Training Boosted Multi-Factor Matching Network for Composed Image Retrieval

The composed image retrieval (CIR) task aims to retrieve the desired target image for a given multimodal query, i.e., a reference image with its corresponding modification text. The key limitations encountered by existing efforts are two aspects: 1) ignoring the multiple query-target matching factor...

Ausführliche Beschreibung

Bibliographische Detailangaben
Veröffentlicht in:IEEE transactions on pattern analysis and machine intelligence. - 1979. - 46(2024), 5 vom: 04. Apr., Seite 3665-3678
1. Verfasser: Wen, Haokun (VerfasserIn)
Weitere Verfasser: Song, Xuemeng, Yin, Jianhua, Wu, Jianlong, Guan, Weili, Nie, Liqiang
Format: Online-Aufsatz
Sprache:English
Veröffentlicht: 2024
Zugriff auf das übergeordnete Werk:IEEE transactions on pattern analysis and machine intelligence
Schlagworte:Journal Article