Please refer to the Google Scholar profile for more up-to-date papers.
|
Yifan Ye, Jiaqi Ma, Jun Cen, Zhihe Lu† † Corresponding author RAL, 2026 A training-free token compression framework that accelerates VLA inference with token expanding and merging. |
|
Yifan Ye, Jun Cen, Jing Chen†, Zhihe Lu† † Corresponding author RAL, 2026 A framework that progressively improves a few-shot model through simulator interactions. |
|
Rui Zhu, Liang Bai, Yanming Guo, Yirun Ruan, Tianyuan Yu†, Zhihe Lu† † Corresponding author CVPR, 2026 A very early attempt to conduct test-time prompt learning in federated domain. |
|
Aiming Zhang, Tianyuan Yu, Liang Bai, Jun Tang, Yanming Guo, Yirun Ruan, Yun Zhou, Zhihe Lu† † Corresponding author TIP, 2025 code A method that leverages domain information and language guidance to adapt VLMs at test time. |
|
Zhihe Lu*, Jiawang Bai*, Xin Li, Zeyu Xiao, Xinchao Wang * Equal Contribution TIP, 2025 code A two-stage pipeline to alleivate the difficulty of VLMs adaptation at test time. |
|
Zhihe Lu, Jiawang Bai, Xin Li, Zeyu Xiao, Xinchao Wang ICML, 2024 code The first investigation of ensemble learning for VLMs. |
|
Xin Li*, Dongze Lian*, Zhihe Lu*, Jiawang Bai, Zhibo Chen, Xinchao Wang * Equal Contribution NeurIPS, 2023 code Introduce knowledge graph for tuning vision and language models. |
|
Keji He, Chenyang Si, Zhihe Lu, Yan Huang, Liang Wang, Xinchao Wang NeurIPS, 2023 code A work to explore the significance of high-frequency information for enhanced Vision-and-Language Navigation. |
|
Zhihe Lu, Da Li, Yi-Zhe Song, Tao Xiang, Timothy M. Hospedales TIP, 2023 An uncertainty-aware solution for SFDASS. |
|
Zhihe Lu, Sen He, Da Li, Yi-Zhe Song, Tao Xiang TIP, 2023 Investigating the feature-prediction covariance based Transformer for calibrating the biases in GFSS. |
|
Tao Yu*, Zhihe Lu*, Xin Jin, Zhibo Chen, Xinchao Wang * Equal Contribution CVPR, 2023 code A simple yet effective tuning method for vision and language models. |
|
Zhihe Lu, Sen He, Xiatian Zhu, Li Zhang, Yi-Zhe Song, Tao Xiang ICCV, 2021 code A novel training pipeline for few-shot segmentation with classifier weight transformer. |
|
Zhihe Lu, Yongxin Yang, Xiatian Zhu, Cong Liu, Yi-Zhe Song, Tao Xiang CVPR, 2020 code A novel way to use infinite number of classifiers without extra parameters to identify misaligned regions. |
|
Zhihe Lu, Tanhao Hu, Lingxiao Song, Zhaoxiang Zhang, Ran He ACM MM, 2018 A Couple-Agent Face Parsing based Generative Adversarial Network (CAFP-GAN) that unites the knowledge of facial semantic regions and controllable expression signals. |
|
Lingxiao Song, Zhihe Lu, Ran He, Zhenan Sun, Tieniu Tan ACM MM, 2018 A Geometry-Guided Generative Adversarial Network (G2-GAN) was proposed for photorealistic and identity-preserving facial expression synthesis. |
|
Zhihe Lu, Zhihang Li, Jie Cao, Ran He, Zhenan Sun ACPR, 2017 A very early survey for face image synthesis. |