West Ashley Library

9 a.m. - 7 p.m.

Phone: (843) 766-6635

Wando Mount Pleasant Library

9 a.m. – 8 p.m.

Phone: (843) 805-6888

Email Us

Village Library

9 a.m. – 6 p.m.

Phone: (843) 884-9741

St. Paul's/Hollywood Library

9 a.m. – 8 p.m.

Phone: (843) 889-3300

Otranto Road Library

9 a.m. – 8 p.m.

Phone: (843) 572-4094

Mt. Pleasant Library

9 a.m. – 8 p.m.

Phone: (843) 849-6161

McClellanville Library

9 a.m. - 6 p.m.

Phone: (843) 887-3699

Keith Summey North Charleston Library

9 a.m. – 8 p.m.

Phone: (843) 744-2489

John's Island Library

9 a.m. – 8 p.m.

Phone: (843) 559-1945

Hurd/St. Andrews Library

9 a.m. – 8 p.m.

Phone: (843) 766-2546

Folly Beach Library

Closed

Phone: (843) 588-2001

Edisto Island Library

9 a.m. - 6 p.m.

Phone: (843) 869-2355

Dorchester Road Library

9 a.m. – 8 p.m.

Phone: (843) 552-6466

John L. Dart Library

9 a.m. – 7 p.m.

Phone: (843) 722-7550

Baxter-Patrick James Island

9 a.m. – 8 p.m.

Phone: (843) 795-6679

Main Library

9 a.m. – 6 p.m.

Phone: (843) 805-6930

Bees Ferry West Ashley Library

9 a.m. – 8 p.m.

Phone: (843) 805-6892

Edgar Allan Poe/Sullivan's Island Library

Closed for renovations

Phone: (843) 883-3914

Mobile Library

9 a.m. - 5 p.m.

Phone: (843) 805-6909

Email Us

Item request has been placed!

Item request cannot be made.

Processing Request

Adaptive multimodal prompt for human-object interaction with local feature enhanced transformer.

Item request has been placed!

Item request cannot be made.

Processing Request

Read More Add to Saved list

Author(s): Xue, Kejun; Gao, Yongbin; Fang, Zhijun; Jiang, Xiaoyan; Yu, Wenjun; Chen, Mingxuan; Wu, Chenmou
Source:
Applied Intelligence; Dec2024, Vol. 54 Issue 23, p12492-12504, 13p
Subject Terms:
TRANSFORMER models; COMPUTER vision; FEATURE extraction; LEARNING strategies; DATA distribution

Additional Information
- Abstract:
  Human-object interaction (HOI) detection is an important computer vision task for recognizing the interaction between humans and surrounding objects in an image or video. The HOI datasets have a serious long-tailed data distribution problem because it is challenging to have a dataset that contains all potential interactions. Many HOI detectors have addressed this issue by utilizing visual-language models. However, due to the calculation mechanism of the Transformer, the visual-language model is not good at extracting the local features of input samples. Therefore, we propose a novel local feature enhanced Transformer to motivate encoders to extract multi-modal features that contain more information. Moreover, it is worth noting that the application of prompt learning in HOI detection is still in preliminary stages. Consequently, we propose a multi-modal adaptive prompt module, which uses an adaptive learning strategy to facilitate the interaction of language and visual prompts. In the HICO-DET and SWIG-HOI datasets, the proposed model achieves full interaction with 24.21% mAP and 14.29% mAP, respectively. Our code is available at https://github.com/small-code-cat/AMP-HOI. [ABSTRACT FROM AUTHOR]
- Abstract:
  Copyright of Applied Intelligence is the property of Springer Nature and its content may not be copied or emailed to multiple sites or posted to a listserv without the copyright holder's express written permission. However, users may print, download, or email articles for individual use. This abstract may be abridged. No warranty is given about the accuracy of the copy. Users should refer to the original published version of the material for the full abstract. (Copyright applies to all Abstracts.)

Comments

No Comments.

menu

Adaptive multimodal prompt for human-object interaction with local feature enhanced transformer.

Contact CCPL

Patron Login

menu

Adaptive multimodal prompt for human-object interaction with local feature enhanced transformer.

Engage with CCPL

Contact CCPL