Xin Yu

Email:: xin.yu@uq.edu.au

Background

My name is Xin Yu, a Senior Lecturer at the University of Queensland. I am an Australian Research Council Discovery Early Career Researcher Award 2023-2025 (DECRA) recipient and an awardee of the prestigious Google Research Scholar Program in 2021. I am also a Google Visiting Faculty. Previously, I was a research fellow at the Australian National University (ANU). I received my PhD degree from the Australian National Unversity under the supervision of Prof. Richard Hartley, Prof. Fatih Porikli and Dr. Basura Fernando. I also received a PhD degree from Tsinghua University supervised by Prof. Li Zhang. I am interested in Computer Vision and Machine Learning topics.

My research topics includes various computer vision and machine learning tasks, especially in efficient low-level image processing, image retrieval and localization, action recognition, 3D pose estimation, visual navigation and sign language recognition and translation.

Availability

Dr Xin Yu is:: Available for supervision

Research impacts

One of my research papers has been awarded "Best Paper Honorable Mention" award in the premium computer vision conference WACV 2020, and one paper has been nominated for the Best Paper Award in CVPR 2020.

I was awarded the Outstanding Reviewer Award in ECCV 2020, CVPR 2021 and ICCV 2021. CVPR, ICCV and ECCV are internationally world-leading computer vision and machine learning conferences. My research interests include deep learning techniques, image processing, and computer vision tasks. I am a program committee member of top-tier computer vision and machine learning conferences, such as CVPR, ICCV, ECCV, ICML, ICLR and NeurIPS, and a reviewer of prestigious journals, such as TPAMI, IJCV and TIP.

I am happy to supervise self-motivated PhD and MPhil students. If you are an undergraduate student and willing to conduct your honour project, please drop me an email.

Search Professor Xin Yu’s works on UQ eSpace

159 works between 2011 and 2025

All (159) Journal Article (58) Book Chapter (2) Conference Publication (99)

2025

Journal Article

Multi-contrast computed tomography atlas of healthy pancreas with dense displacement sampling registration

Zhou, Yinchi, Lee, Ho Hin, Tang, Yucheng, Yu, Xin, Yang, Qi, Kim, Michael E., Remedios, Lucas W., Bao, Shunxing, Spraggins, Jeffrey M., Huo, Yuankai and Landman, Bennett A. (2025). Multi-contrast computed tomography atlas of healthy pancreas with dense displacement sampling registration. Journal of Medical Imaging, 12 (02). doi: 10.1117/1.jmi.12.2.024006

Multi-contrast computed tomography atlas of healthy pancreas with dense displacement sampling registration

2025

Conference Publication

FlashVTG: Feature Layering and Adaptive Score Handling Network for Video Temporal Grounding

Cao, Zhuo, Zhang, Bingqing, Du, Heming, Yu, Xin, Li, Xue and Wang, Sen (2025). FlashVTG: Feature Layering and Adaptive Score Handling Network for Video Temporal Grounding. IEEE. doi: 10.1109/wacv61041.2025.00894

FlashVTG: Feature Layering and Adaptive Score Handling Network for Video Temporal Grounding

2025

Conference Publication

TokenBinder: Text-Video Retrieval with One-to-Many Alignment Paradigm

Zhang, Bingqing, Cao, Zhuo, Du, Heming, Yu, Xin, Li, Xue, Liu, Jiajun and Wang, Sen (2025). TokenBinder: Text-Video Retrieval with One-to-Many Alignment Paradigm. IEEE. doi: 10.1109/wacv61041.2025.00485

TokenBinder: Text-Video Retrieval with One-to-Many Alignment Paradigm

2025

Journal Article

DreamCar: leveraging car-specific prior for in-the-wild 3D car reconstruction

Du, Xiaobiao, Sun, Haiyang, Lu, Ming, Zhu, Tianqing and Yu, Xin (2025). DreamCar: leveraging car-specific prior for in-the-wild 3D car reconstruction. IEEE Robotics and Automation Letters, 10 (2), 1840-1847. doi: 10.1109/lra.2024.3523231

DreamCar: leveraging car-specific prior for in-the-wild 3D car reconstruction

2025

Conference Publication

Vision-based abnormal action dataset for recognising body motion disorders

Ying, Jiaying, Shen, Xin and Yu, Xin (2025). Vision-based abnormal action dataset for recognising body motion disorders. 37th Australasian Joint Conference on Artificial Intelligence, AI 2024, Melbourne, VIC, Australia, 25 - 29 November 2024. Singapore, Singapore: Springer Nature Singapore. doi: 10.1007/978-981-96-0351-0_33

Vision-based abnormal action dataset for recognising body motion disorders

2025

Journal Article

ICE: Interactive 3D Game Character Facial Editing via Dialogue

Wu, Haoqian, Zhao, Minda, Hu, Zhipeng, Fan, Changjie, Li, Lincheng, Chen, Weijie, Zhao, Rui and Yu, Xin (2025). ICE: Interactive 3D Game Character Facial Editing via Dialogue. IEEE Transactions on Multimedia, PP (99), 1-14. doi: 10.1109/tmm.2025.3557611

ICE: Interactive 3D Game Character Facial Editing via Dialogue

2025

Conference Publication

Transferable attacks for semantic segmentation

He, Mengqi, Zhang, Jing and Yu, Xin (2025). Transferable attacks for semantic segmentation. 35th Australasian Database Conference, Gold Coast, QLD, Australia, 16-18 December 2024. Heidelberg, Germany: Springer. doi: 10.1007/978-981-96-1242-0_28

Transferable attacks for semantic segmentation

2025

Book Chapter

Who is Being Impersonated? Deepfake Audio Detection and Impersonated Identification via Extraction of Id-Specific Features

Guo, Tianchen, Du, Heming, Huo, Huan, Liu, Bo and Yu, Xin (2025). Who is Being Impersonated? Deepfake Audio Detection and Impersonated Identification via Extraction of Id-Specific Features. Lecture Notes in Computer Science. (pp. 301-320) Singapore: Springer Nature Singapore. doi: 10.1007/978-981-96-1548-3_21

Who is Being Impersonated? Deepfake Audio Detection and Impersonated Identification via Extraction of Id-Specific Features

2024

Journal Article

"Understanding robustness lottery": a geometric visual comparative analysis of neural network pruning approaches

Li, Zhimin, Liu, Shusen, Yu, Xin, Bhavya, Kailkhura, Cao, Jie, Diffenderfer, James Daniel, Bremer, Peer-Timo and Pascucci, Valerio (2024). "Understanding robustness lottery": a geometric visual comparative analysis of neural network pruning approaches. IEEE transactions on visualization and computer graphics, 31 (9), 1-16. doi: 10.1109/TVCG.2024.3514996

"Understanding robustness lottery": a geometric visual comparative analysis of neural network pruning approaches

2024

Conference Publication

CPT-VR: Improving Surface Rendering via Closest Point Transform with View-Reflection Appearance

Hu, Zhipeng, Zhang, Yongqiang, Liu, Chen, Li, Lincheng, Peng, Sida, Zhou, Xiaowei, Fan, Changjie and Yu, Xin (2024). CPT-VR: Improving Surface Rendering via Closest Point Transform with View-Reflection Appearance. 18th European Conference on Computer Vision, ECCV 2024, Milan, Italy, 29 September –4 October 2024. Cham, Switzerland: Springer. doi: 10.1007/978-3-031-73464-9_14

CPT-VR: Improving Surface Rendering via Closest Point Transform with View-Reflection Appearance

2024

Conference Publication

Snap and diagnose: an advanced multimodal retrieval system for identifying plant diseases in the wild

Wei, Tianqi, Chen, Zhi and Yu, Xin (2024). Snap and diagnose: an advanced multimodal retrieval system for identifying plant diseases in the wild. MMASIA ’24, Auckland, New Zealand, 3-6 December 2024. New York, United States: ACM. doi: 10.1145/3696409.3700293

Snap and diagnose: an advanced multimodal retrieval system for identifying plant diseases in the wild

2024

Conference Publication

FreeAvatar: robust 3D facial animation transfer by learning an expression foundation model

Qiu, Feng, Zhang, Wei, Liu, Chen, An, Rudong, Li, Lincheng, Ding, Yu, Fan, Changjie, Hu, Zhipeng and Yu, Xin (2024). FreeAvatar: robust 3D facial animation transfer by learning an expression foundation model. SA '24: SIGGRAPH Asia 2024, Tokyo, Japan, 3-6 December 2024. New York, NY, United States: ACM. doi: 10.1145/3680528.3687669

FreeAvatar: robust 3D facial animation transfer by learning an expression foundation model

2024

Journal Article

M3 A: A multimodal misinformation dataset for media authenticity analysis

Xu, Qingzheng, Chen, Huiqiang, Du, Heming, Zhang, Hu, Łukasik, Szymon, Zhu, Tianqing and Yu, Xin (2024). M3 A: A multimodal misinformation dataset for media authenticity analysis. Computer Vision and Image Understanding, 249 104205. doi: 10.1016/j.cviu.2024.104205

M3 A: A multimodal misinformation dataset for media authenticity analysis

2024

Book Chapter

OpenSight: A Simple Open-Vocabulary Framework for LiDAR-Based Object Detection

Zhang, Hu, Xu, Jianhua, Tang, Tao, Sun, Haiyang, Yu, Xin, Huang, Zi and Yu, Kaicheng (2024). OpenSight: A Simple Open-Vocabulary Framework for LiDAR-Based Object Detection. Lecture Notes in Computer Science. (pp. 1-19) Cham: Springer Nature Switzerland. doi: 10.1007/978-3-031-72907-2_1

OpenSight: A Simple Open-Vocabulary Framework for LiDAR-Based Object Detection

2024

Conference Publication

Benchmarking in-the-wild multimodal disease recognition and a versatile baseline

Wei, Tianqi, Chen, Zhi, Huang, Zi and Yu, Xin (2024). Benchmarking in-the-wild multimodal disease recognition and a versatile baseline. MM '24: The 32nd ACM International Conference on Multimedia, Melbourne, VIC, Australia, 28 October-1 November 2024. New York, United States: Association for Computing Machinery. doi: 10.1145/3664647.3680599

Benchmarking in-the-wild multimodal disease recognition and a versatile baseline

2024

Journal Article

Ethics-aware face recognition aided by synthetic face images

Du, Xiaobiao, Yu, Xin, Liu, Jinhui, Dai, Beifen and Xu, Feng (2024). Ethics-aware face recognition aided by synthetic face images. Neurocomputing, 600 128129, 128129. doi: 10.1016/j.neucom.2024.128129

Ethics-aware face recognition aided by synthetic face images

2024

Conference Publication

High-frequency trans-spinal magnetic stimulation for chronic neuropathic pain treatment

Marturano, Francesca, Gomez-Cid, Lidia, Straney, Don, Chen, Iris Yin-Ching, Albrecht, Alice Marie Cécile, Yu, Xin, Ay, Ilknur and Bonmassar, Giorgio (2024). High-frequency trans-spinal magnetic stimulation for chronic neuropathic pain treatment. 2024 46th Annual International Conference of the IEEE Engineering in Medicine and Biology Society (EMBC), Orlando, FL, United States, 15-19 July 2024. Piscataway, NJ, United States: Institute of Electrical and Electronics Engineers. doi: 10.1109/embc53108.2024.10781608

High-frequency trans-spinal magnetic stimulation for chronic neuropathic pain treatment

2024

Conference Publication

Recent update on the Tsinghua tabletop Kibble balance

Li, S., Ma, Y., Ma, K., Liu, W., Li, N., Liu, X., Peng, L., Zhao, W., Huang, S. and Yu, X. (2024). Recent update on the Tsinghua tabletop Kibble balance. Conference on Precision Electromagnetic Measurements (CPEM) / Joint NCSL-International Annual Workshop and Symposium (NCSLI), Denver, CO United States, 8-12 July 2024. Piscataway, NJ United States: Institute of Electrical and Electronics Engineers. doi: 10.1109/cpem61406.2024.10645985

Recent update on the Tsinghua tabletop Kibble balance

2024

Journal Article

Restoring consciousness with pharmacologic therapy: mechanisms, targets, and future directions

Barra, Megan E., Solt, Ken, Yu, Xin and Edlow, Brian L. (2024). Restoring consciousness with pharmacologic therapy: mechanisms, targets, and future directions. Neurotherapeutics, 21 (4) e00374, 1-10. doi: 10.1016/j.neurot.2024.e00374

Restoring consciousness with pharmacologic therapy: mechanisms, targets, and future directions

2024

Conference Publication

Language-guided multi-modal emotional mimicry intensity estimation

Qiu, Feng, Zhang, Wei, Liu, Chen, Li, Lincheng, Du, Heming, Guo, Tianchen and Yu, Xin (2024). Language-guided multi-modal emotional mimicry intensity estimation. 2024 IEEE/CVF Conference on Computer Vision and Pattern Recognition Workshops (CVPRW), Seattle, WA, United States, 17-18 June 2024. Piscataway, NJ, United States: Institute of Electrical and Electronics Engineers. doi: 10.1109/cvprw63382.2024.00477

Language-guided multi-modal emotional mimicry intensity estimation

Current funding

2025 - 2026

Creation of an interactive online seaweed production map to support policy-making for the Indonesian seaweed industry

KONEKSI Environment and Climate Change Extension Support

Open grant
2024 - 2027

AI-Empowered and Video-Based uplift of Paralympic classification systems (AQIRP project administered by Follow Me AI)

Follow Me AI Pty Ltd

Open grant
2023 - 2028

Breaking the Communication Barrier for the Australian Deaf Community: Vision Based Australian Sign Language Translation and Production

Google Asia Pacific Pte Ltd

Open grant
2023 - 2026

Advancing Human Perception: Countering Evolving Malicious Fake Visual Data

ARC Discovery Early Career Researcher Award

Open grant
2023 - 2027

Analytics for the Australian Grains Industry (AAGI)

Grains Research & Development Corporation

Open grant

Past funding

2024

Breaking the Communication Barrier for the Australian Deaf Community: Vision Based Australian Sign Language Translation and Production

Google Inc

Open grant
2023 - 2024

Developing applications of satellite imagery for modelling environmental and social impacts of climate change on seaweed farming in Indonesia (KONEKSI Grant administered by Griffith University)

Griffith University

Open grant
2023 - 2025

Two-way Auslan: Automatic Machine Translation of Australian Sign Language (ARC Discovery Project administered by ANU)

The Australian National University

Open grant

Availability

Dr Xin Yu is:: Available for supervision

Before you email them, read our advice on how to contact a supervisor.

Supervision history

Current supervision

Doctor Philosophy

Two way Auslan Translation

Principal Advisor

Other advisors: Associate Professor Mahsa Baktashmotlagh
Doctor Philosophy

Towards Efficient Pest Detection in Agriculture

Principal Advisor

Other advisors: Associate Professor Sen Wang
Doctor Philosophy

Multimodal foundation model design and analysis

Principal Advisor

Other advisors: Dr Miao Xu
Doctor Philosophy

Pose Estimation for Human with Disabilities

Principal Advisor

Other advisors: Professor Brian Lovell
Doctor Philosophy

The prediction, diagnosis, and severity estimation models for plant disease

Principal Advisor

Other advisors: Associate Professor Sen Wang
Doctor Philosophy

Two way Auslan Translation

Principal Advisor

Other advisors: Professor Helen Huang
Doctor Philosophy

Automatic Retinal Health Monitoring through Multi-modal Medical Imaging

Principal Advisor

Other advisors: Associate Professor Mahsa Baktashmotlagh
Doctor Philosophy

Compressed Video Restoration

Principal Advisor

Other advisors: Dr Miao Xu
Doctor Philosophy

Understanding Human Intention and Performance

Principal Advisor

Other advisors: Associate Professor Sen Wang
Doctor Philosophy

Combating evolving deceptive fake visual information through deepfake detection

Principal Advisor

Other advisors: Associate Professor Sen Wang
Doctor Philosophy

Understanding Human Intention and Performance

Principal Advisor

Other advisors: Dr Miao Xu
Doctor Philosophy

Human Posture Recognition Applied to Physical Activity

Principal Advisor

Other advisors: Professor Sean Tweedy
Doctor Philosophy

Combating evolving deceptive fake visual information through deepfake detection

Principal Advisor

Other advisors: Dr Miao Xu
Doctor Philosophy

Remote Sensing Analysis in computer vision

Associate Advisor

Other advisors: Professor Helen Huang
Doctor Philosophy

Enhancing Robustness and Generalizability in Computational Models

Associate Advisor

Other advisors: Associate Professor Mahsa Baktashmotlagh
Doctor Philosophy

Data driven approaches for smart farming

Associate Advisor

Other advisors: Professor Helen Huang

Enquiries

For media enquiries about Dr Xin Yu's areas of expertise, story ideas and help finding experts, contact our Media team:

communications@uq.edu.au

External profiles

Personal links

Update my profile

Xin Yu

Overview

Background

Availability

Research impacts

Works

Multi-contrast computed tomography atlas of healthy pancreas with dense displacement sampling registration

FlashVTG: Feature Layering and Adaptive Score Handling Network for Video Temporal Grounding

TokenBinder: Text-Video Retrieval with One-to-Many Alignment Paradigm

DreamCar: leveraging car-specific prior for in-the-wild 3D car reconstruction

Vision-based abnormal action dataset for&nbsp;recognising body motion disorders

ICE: Interactive 3D Game Character Facial Editing via Dialogue

Transferable attacks for&nbsp;semantic segmentation

Who is Being Impersonated? Deepfake Audio Detection and Impersonated Identification via Extraction of Id-Specific Features

"Understanding robustness lottery": a geometric visual comparative analysis of neural network pruning approaches

CPT-VR: Improving Surface Rendering via Closest Point Transform with View-Reflection Appearance

Snap and diagnose: an advanced multimodal retrieval system for identifying plant diseases in the wild

FreeAvatar: robust 3D facial animation transfer by learning an expression foundation model

M3 A: A multimodal misinformation dataset for media authenticity analysis

OpenSight: A Simple Open-Vocabulary Framework for LiDAR-Based Object Detection

Benchmarking in-the-wild multimodal disease recognition and a versatile baseline

Ethics-aware face recognition aided by synthetic face images

High-frequency trans-spinal magnetic stimulation for chronic neuropathic pain treatment

Recent update on the Tsinghua tabletop Kibble balance

Restoring consciousness with pharmacologic therapy: mechanisms, targets, and future directions

Language-guided multi-modal emotional mimicry intensity estimation

Funding

Current funding

Past funding

Supervision

Availability

Supervision history

Current supervision

Two way Auslan Translation

Towards Efficient Pest Detection in Agriculture

Multimodal foundation model design and analysis

Pose Estimation for Human with Disabilities

The prediction, diagnosis, and severity estimation models for plant disease

Two way Auslan Translation

Automatic Retinal Health Monitoring through Multi-modal Medical Imaging

Compressed Video Restoration

Understanding Human Intention and Performance

Combating evolving deceptive fake visual information through deepfake detection

Understanding Human Intention and Performance

Human Posture Recognition Applied to Physical Activity

Combating evolving deceptive fake visual information through deepfake detection

Remote Sensing Analysis in computer vision

Enhancing Robustness and Generalizability in Computational Models

Data driven approaches for smart farming

Media

Enquiries

Vision-based abnormal action dataset for recognising body motion disorders

Transferable attacks for semantic segmentation