Skip to menu Skip to content Skip to footer

2026

Journal Article

Beyond standard views: enhancing quarantine true fruit fly identification with a multi-angle deep learning approach (Diptera, Tephritidae)

Li, Zitao, Jiang, Fan, Wu, Zhuojie, Yu, Xin, Yang, Ding, Cheng, Daifeng and Li, Xuankun (2026). Beyond standard views: enhancing quarantine true fruit fly identification with a multi-angle deep learning approach (Diptera, Tephritidae). Royal Society Open Science, 13 (8) rsos261109. doi: 10.1098/rsos.261109

Beyond standard views: enhancing quarantine true fruit fly identification with a multi-angle deep learning approach (Diptera, Tephritidae)

2026

Journal Article

Thinking outside the frame: viewpoint-aware spatial reasoning in large multimodal models

Ke, Yan, Shiri, Fatemeh, Zhang, Hu and Yu, Xin (2026). Thinking outside the frame: viewpoint-aware spatial reasoning in large multimodal models. Visual Intelligence, 4 (1) 21, 1. doi: 10.1007/s44267-026-00126-0

Thinking outside the frame: viewpoint-aware spatial reasoning in large multimodal models

2026

Journal Article

TBCNet: Twin-branch collaborative network for hyperspectral anomaly detection

Zhao, Dong, You, Mingtao, Xiang, Pei, Hu, Jianling, Asano, Yuta, Yu, Xin, Hsu, Chih-Chung, Zhou, Huixin and Ren, Jinchang (2026). TBCNet: Twin-branch collaborative network for hyperspectral anomaly detection. Pattern Recognition, 180 114647, 114647-180. doi: 10.1016/j.patcog.2026.114647

TBCNet: Twin-branch collaborative network for hyperspectral anomaly detection

2026

Journal Article

Tephritid26: A standardized, multi-angle image dataset of quarantine-significant true fruit flies for deep learning-based identification

Li, Zitao, Wang, Xingkai, Wu, Zhuojie, Yong, Mengyuan, De Meyer, Marc, Bourtzis, Kostas, Wei, Bingbing, Zheng, Tianyu, Zeng, Qian, Mo, Jin, Liu, Ruosi, Susanto, Agus, Liu, Weiqi, Guo, Wenchao, Ding, Xinhua, Huang, Xiaolei, Yang, Ding, Cheng, Daifeng, Yu, Xin, Jiang, Fan and Li, Xuankun (2026). Tephritid26: A standardized, multi-angle image dataset of quarantine-significant true fruit flies for deep learning-based identification. Scientific Data. doi: 10.1038/s41597-026-07713-2

Tephritid26: A standardized, multi-angle image dataset of quarantine-significant true fruit flies for deep learning-based identification

2026

Journal Article

UAV-based multispectral object tracking with positive-negative prompt mining network

Teng, Xiang, Zhao, Dong, Xiang, Pei, Yu, Xin, Li, Zhuanfeng, Hsu, Chih-Chung, Yang, Tianfang, Zhou, Huixin and Ren, Jinchang (2026). UAV-based multispectral object tracking with positive-negative prompt mining network. IEEE Transactions on Geoscience and Remote Sensing, 64 5515418, 1-18. doi: 10.1109/tgrs.2026.3698119

UAV-based multispectral object tracking with positive-negative prompt mining network

2026

Journal Article

Safe and reliable diffusion models via subspace projection

Chen, Huiqiang, Zhu, Tianqing, Wang, Linlin, Yu, Xin, Gao, Longxiang and Zhou, Wanlei (2026). Safe and reliable diffusion models via subspace projection. IEEE Transactions on Dependable and Secure Computing, 23 (4), 9272-9285. doi: 10.1109/TDSC.2026.3692493

Safe and reliable diffusion models via subspace projection

2026

Journal Article

DFBSNet: dual frequency-domain branch fusion and selection network for hyperspectral anomaly detection

Yao, Yiming, Wang, Qing, Zhao, Dong, You, Mingtao, Xiang, Pei, Asano, Yuta, Yu, Xin, Wang, Chao, Zhou, Huixin and Ren, Jinchang (2026). DFBSNet: dual frequency-domain branch fusion and selection network for hyperspectral anomaly detection. Pattern Recognition, 180 113967, 1-13. doi: 10.1016/j.patcog.2026.113967

DFBSNet: dual frequency-domain branch fusion and selection network for hyperspectral anomaly detection

2026

Journal Article

Compression-oriented video super-resolution

Wang, Shuyun, Liu, Yanbin, Lu, Ming, Wu, Zhuojie, Tian, Senmao, Guo, Yandong and Yu, Xin (2026). Compression-oriented video super-resolution. IEEE Transactions on Image Processing, 35, 4040-4050. doi: 10.1109/tip.2026.3682128

Compression-oriented video super-resolution

2026

Journal Article

Cluster-aware prompt ensemble learning for few-shot vision-language model adaptation

Chen, Zhi, Yu, Xin, Tao, Xiaohui, Li, Yan and Huang, Zi (2026). Cluster-aware prompt ensemble learning for few-shot vision-language model adaptation. Pattern Recognition, 172 (C) 112596. doi: 10.1016/j.patcog.2025.112596

Cluster-aware prompt ensemble learning for few-shot vision-language model adaptation

2026

Journal Article

Mobile Auslan: A multimodal dialogue-centered sign language learning system

Sheng, Hongwei, Shen, Xin, Du, Heming and Yu, Xin (2026). Mobile Auslan: A multimodal dialogue-centered sign language learning system. Computer Vision and Image Understanding, 265 104646, 1-18. doi: 10.1016/j.cviu.2026.104646

Mobile Auslan: A multimodal dialogue-centered sign language learning system

2026

Journal Article

Distributed zero-shot learning for visual recognition

Chen, Zhi, Luo, Yadan, Huang, Zi, Li, Jingjing, Wang, Sen and Yu, Xin (2026). Distributed zero-shot learning for visual recognition. IEEE Transactions on Multimedia, 28, 7451-7463. doi: 10.1109/TMM.2026.3673561

Distributed zero-shot learning for visual recognition

2026

Journal Article

Preface

Liu, Miaomiao, Yu, Xin, Xu, Chang and Song, Yiliao (2026). Preface. Lecture Notes in Computer Science, 16370 LNAI, v-vi.

Preface

2025

Journal Article

Hyperspectral video object tracking with cross-modal spectral complementary and memory prompt network

Jiang, Wenhao, Zhao, Dong, Wang, Chen, Yu, Xin, Arun, Pattathal V., Asano, Yuta, Xiang, Pei and Zhou, Huixin (2025). Hyperspectral video object tracking with cross-modal spectral complementary and memory prompt network. Knowledge-Based Systems, 330 (Part B) 114595, 1-16. doi: 10.1016/j.knosys.2025.114595

Hyperspectral video object tracking with cross-modal spectral complementary and memory prompt network

2025

Journal Article

Analytical survey of learning with low-resource data: from analysis to investigation

Cao, Xiaofeng, Xu, Mingwei, Yu, Xin, Yao, Jiangchao, Ye, Wei, Huang, Shengjun, Zhang, Minling, Tsang, Ivor, Ong, Yew-Soon, Kwok, James T. and Shen, Heng Tao (2025). Analytical survey of learning with low-resource data: from analysis to investigation. ACM Computing Surveys, 58 (6) 3773075, 1-47. doi: 10.1145/3773075

Analytical survey of learning with low-resource data: from analysis to investigation

2025

Journal Article

ICE: interactive 3D game character facial editing via dialogue

Wu, Haoqian, Zhao, Minda, Hu, Zhipeng, Fan, Changjie, Li, Lincheng, Chen, Weijie, Zhao, Rui and Yu, Xin (2025). ICE: interactive 3D game character facial editing via dialogue. IEEE Transactions on Multimedia, 27, 3210-4223. doi: 10.1109/tmm.2025.3557611

ICE: interactive 3D game character facial editing via dialogue

2025

Journal Article

DreamCar: leveraging car-specific prior for in-the-wild 3D car reconstruction

Du, Xiaobiao, Sun, Haiyang, Lu, Ming, Zhu, Tianqing and Yu, Xin (2025). DreamCar: leveraging car-specific prior for in-the-wild 3D car reconstruction. IEEE Robotics and Automation Letters, 10 (2), 1840-1847. doi: 10.1109/lra.2024.3523231

DreamCar: leveraging car-specific prior for in-the-wild 3D car reconstruction

2025

Journal Article

TalkCLIP: talking head generation with text-guided expressive speaking styles

Ma, Yifeng, Wang, Suzhen, Ding, Yu, Ma, Bowen, Lv, Tangjie, Fan, Changjie, Hu, Zhipeng, Deng, Zhidong and Yu, Xin (2025). TalkCLIP: talking head generation with text-guided expressive speaking styles. IEEE Transactions on Multimedia, 27, 6335-6346. doi: 10.1109/tmm.2025.3581808

TalkCLIP: talking head generation with text-guided expressive speaking styles

2024

Journal Article

M3 A: A multimodal misinformation dataset for media authenticity analysis

Xu, Qingzheng, Chen, Huiqiang, Du, Heming, Zhang, Hu, Łukasik, Szymon, Zhu, Tianqing and Yu, Xin (2024). M3 A: A multimodal misinformation dataset for media authenticity analysis. Computer Vision and Image Understanding, 249 104205, 104205. doi: 10.1016/j.cviu.2024.104205

M3 A: A multimodal misinformation dataset for media authenticity analysis

2024

Journal Article

Ethics-aware face recognition aided by synthetic face images

Du, Xiaobiao, Yu, Xin, Liu, Jinhui, Dai, Beifen and Xu, Feng (2024). Ethics-aware face recognition aided by synthetic face images. Neurocomputing, 600 128129, 128129. doi: 10.1016/j.neucom.2024.128129

Ethics-aware face recognition aided by synthetic face images

2024

Journal Article

Proactive image manipulation detection via deep semi-fragile watermark

Zhao, Yuan, Liu, Bo, Zhu, Tianqing, Ding, Ming, Yu, Xin and Zhou, Wanlei (2024). Proactive image manipulation detection via deep semi-fragile watermark. Neurocomputing, 585 127593. doi: 10.1016/j.neucom.2024.127593

Proactive image manipulation detection via deep semi-fragile watermark