


default search action
Yunzhi Zhuge
Person information
Refine list

refinements active!
zoomed in on ?? of ?? records
view refined list in
Journal Articles
- 2026
[j22]Jiazuo Yu, Yunzhi Zhuge, Lu Zhang
, Zichen Huang, Wei Zhou, Dong Wang, Huchuan Lu, You He, Long Chen:
One Aligned LLM to Serve Them All: A Transfer Recipe for Training VLMs without Visual-Language Re-Alignment. Int. J. Comput. Vis. 134(7): 331 (2026)
[j21]Hengrun Zhao
, Yifan Wang, Yunzhi Zhuge
, Lijun Wang, Huchuan Lu:
Aggregating global-scale pixel-wise forgery cues within a graph. Neural Networks 204: 109272 (2026)
[j20]Kecheng Zhang, Zongxin Yang
, Mingfei Han
, Yunzhi Zhuge
, Haihong Hao
, Changlin Li
, Zhihui Li
, Xiaojun Chang
:
SELongVLM: Empowering Long Video Language Models With Self-Corrective Clip Selection. IEEE Trans. Pattern Anal. Mach. Intell. 48(7): 8694-8709 (2026)
[j19]Xiaojian Shen, Dahu Shi, Jianrong Zhang, Hai Li, Hongwei Zhao, Dawei Zhang, Yunzhi Zhuge, Zhiliang Wu
, Guanghui Yue, Wei Zhou:
BeatDance: Generating beat-consistent 3D dance with hierarchical spatial-temporal modeling. Pattern Recognit. 180: 114344 (2026)
[j18]Yunzhi Zhuge, Mengyuan Zhu, Yizhuang Peng, Lu Zhang, Jin Zhan, Pingping Zhang, Huchuan Lu:
Bridging CLIP and CLAP for open-vocabulary audio-visual segmentation with semantic coherence. Pattern Recognit. 180: 114586 (2026)
[j17]Xinzhuo Yu, Yunzhi Zhuge
, Sitong Gong
, Lu Zhang
, Pingping Zhang, Huchuan Lu
:
Parameter-Aware Mamba Model for Multitask Dense Prediction. IEEE Trans. Cybern. 56(5): 2650-2662 (2026)
[j16]Yunzhi Zhuge
, Sitong Gong
, Lu Zhang
, Qi Xu
, Wenda Zhao
, Jin Zhan, Huchuan Lu
:
Context-Infused Trajectories: Enhancing Context and Frame Consistency in Reasoning Video Object Segmentation. IEEE Trans. Image Process. 35: 5239-5252 (2026)
[j15]Yunzhi Zhuge
, Xinzhuo Yu, Lu Zhang
, Xu Jia
, Jin Zhan
, Huchuan Lu
:
Exploiting Cross-Task Synergy via Frequency-Driven Hierarchical Learning for Multi-Task Dense Prediction. IEEE Trans. Image Process. 35: 6198-6210 (2026)
[j14]Wenbo Zhang
, Yunzhi Zhuge
, Lu Zhang
, Ping Hu
, Dong Wang
, Huchuan Lu
:
DiMuS: Disentangled Multi-Signal Learning for Weakly Supervised Point-Based 3D Object Detection. IEEE Trans. Image Process. 35: 6815-6829 (2026)
[j13]Yongqi Shan, Lu Zhang
, Jiazuo Yu
, Yunzhi Zhuge
, Huchuan Lu
:
3D-SceneQ: Empowering 3D LLM With Query-Guided Adaptive Pruning and Multi-Modal Feature Enhancement. IEEE Trans. Multim. 28: 5972-5983 (2026)
[j12]Yongqi Shan
, Yunzhi Zhuge
, Huchuan Lu
:
HDVS: semi-supervised semantic segmentation via heterogeneous dual-branch voting supervision. Vis. Intell. 4(1) (2026)- 2025
[j11]Jiazuo Yu
, Zichen Huang
, Yunzhi Zhuge
, Lu Zhang
, Ping Hu
, Dong Wang
, Huchuan Lu
, You He
:
MoE-Adapters++: Toward More Efficient Continual Learning of Vision-Language Models Via Dynamic Mixture-of-Experts Adapters. IEEE Trans. Pattern Anal. Mach. Intell. 47(12): 11912-11928 (2025)
[j10]Zhenyu Chen
, Jiawen Zhu
, Lu Zhang
, Ping Hu
, Yunzhi Zhuge, Huchuan Lu
, You He
:
RVMamba: Selective Text-Vision Mamba for Referring Video Object Segmentation. IEEE Signal Process. Lett. 32: 4349-4353 (2025)
[j9]Haomiao Xiong, Yunzhi Zhuge
, Jiawen Zhu
, Lu Zhang
, Huchuan Lu
:
3UR-LLM: An End-to-End Multimodal Large Language Model for 3D Scene Understanding. IEEE Trans. Multim. 27: 2899-2911 (2025)
[j8]Sitong Gong
, Yunzhi Zhuge
, Lu Zhang
, Yifan Wang
, Pingping Zhang, Lijun Wang
, Huchuan Lu
:
AVS-Mamba: Exploring Temporal and Multi-Modal Mamba for Audio-Visual Segmentation. IEEE Trans. Multim. 27: 5413-5425 (2025)
[j7]Sitong Gong
, Yunzhi Zhuge
, Lu Zhang
, Pingping Zhang, Huchuan Lu
:
Complementary and Contrastive Learning for Audio-Visual Segmentation. IEEE Trans. Multim. 27: 7407-7418 (2025)
[j6]Qinghe Wang
, Xu Jia
, Xiaomin Li
, Taiqing Li
, Liqian Ma
, Yunzhi Zhuge
, Huchuan Lu
:
StableIdentity: Inserting Anybody Into Anywhere at First Sight. IEEE Trans. Multim. 27: 9342-9353 (2025)
[j5]Yunzhi Zhuge
, Hongyu Gu, Lu Zhang
, Jinqing Qi
, Huchuan Lu
:
Learning Motion and Temporal Cues for Unsupervised Video Object Segmentation. IEEE Trans. Neural Networks Learn. Syst. 36(5): 9084-9097 (2025)
[j4]Yue Zhang
, Chao Wang
, Fei Fang
, Yun-Zhi Zhuge
, Hehe Fan
, Xiaojun Chang
, Cheng Deng
, Yi Yang:
SAMControl: Controlling Pose and Object for Image Editing with Soft Attention Mask. ACM Trans. Multim. Comput. Commun. Appl. 21(11): 320:1-320:28 (2025)- 2024
[j3]Yue Wang
, Lu Zhang
, Pingping Zhang, Yunzhi Zhuge
, Junfeng Wu, Hong Yu, Huchuan Lu
:
Learning Local-Global Representation for Scribble-Based RGB-D Salient Object Detection via Transformer. IEEE Trans. Circuits Syst. Video Technol. 34(11): 11592-11604 (2024)- 2018
[j2]Yun-Zhi Zhuge
, Gang Yang, Pingping Zhang, Huchuan Lu
:
Boundary-Guided Feature Aggregation Network for Salient Object Detection. IEEE Signal Process. Lett. 25(12): 1800-1804 (2018)- 2015
[j1]Yun-Zhi Zhuge, Huchuan Lu:
Robust Video Text Detection with Morphological Filtering Enhanced MSER. J. Comput. Sci. Technol. 30(2): 353-363 (2015)
Conference and Workshop Papers
- 2026
[c21]Long Chen, Wei Miao
, Xin Gao, Yunzhi Zhuge, Hongming Xu, Yaxin Li, Qi Xu:
Spatial-Frequency Spiking Neural Network for Underwater Object Detection. AAAI 2026: 20217-20225
[c20]Baolu Li
, Yiming Zhang
, Qinghe Wang
, Liqian Ma
, Xiaoyu Shi
, Xintao Wang
, Pengfei Wan
, Zhenfei Yin
, Yunzhi Zhuge
, Huchuan Lu
, Xu Jia
:
VFXMaster: Unlocking Dynamic Visual Effect Generation via In-Context Learning. SIGGRAPH (Conference Paper Track) 2026: 130:1-130:9- 2025
[c19]Chengyang Ye, Yunzhi Zhuge, Pingping Zhang:
Towards Open-Vocabulary Remote Sensing Image Semantic Segmentation. AAAI 2025: 9436-9444
[c18]Wenbo Zhang
, Lu Zhang, Ping Hu, Liqian Ma, Yunzhi Zhuge, Huchuan Lu:
Bootstraping Clustering of Gaussians for View-consistent 3D Scene Understanding. AAAI 2025: 10166-10175
[c17]Sitong Gong, Yunzhi Zhuge, Lu Zhang, Zongxin Yang, Pingping Zhang, Huchuan Lu:
The Devil is in Temporal Token: High Quality Video Reasoning Segmentation. CVPR 2025: 29183-29192
[c16]Haomiao Xiong, Zongxin Yang, Jiazuo Yu, Yunzhi Zhuge, Lu Zhang, Jiawen Zhu, Huchuan Lu:
Streaming Video Understanding and Multi-round Interaction with Memory-enhanced Knowledge. ICLR 2025
[c15]Mengyuan Zhu, Yunzhi Zhuge, Sitong Gong, Lu Zhang, Huchuan Lu:
FDAVS: Exploring Frequency-Driven Modality Enhancement in Audio-Visual Segmentation. ICME 2025: 1-6
[c14]Yue Zhu
, Haiwen Diao
, Shang Gao
, Jiazuo Yu
, Jiawen Zhu
, Yunzhi Zhuge
, Shuai Hao
, Xu Jia
, Lu Zhang
, Ying Zhang
, Huchuan Lu
:
Regularizing Subspace Redundancy of Low-Rank Adaptation. ACM Multimedia 2025: 1666-1675
[c13]Lu Zhang, Jiazuo Yu, Haomiao Xiong, Ping Hu, Yunzhi Zhuge, Huchuan Lu, You He:
FineRS: Fine-grained Reasoning and Segmentation of Small Objects with Reinforcement Learning. NeurIPS 2025- 2024
[c12]Songsong Yu, Yifan Wang, Yunzhi Zhuge, Lijun Wang, Huchuan Lu:
DME: Unveiling the Bias for Better Generalized Monocular Depth Estimation. AAAI 2024: 6817-6825
[c11]Bocen Li, Yunzhi Zhuge, Shan Jiang, Lijun Wang, Yifan Wang, Huchuan Lu:
3D Prompt Learning for RGB-D Tracking. ACCV (2) 2024: 394-411
[c10]Jiazuo Yu, Yunzhi Zhuge, Lu Zhang, Ping Hu, Dong Wang, Huchuan Lu, You He:
Boosting Continual Learning of Vision-Language Models via Mixture-of-Experts Adapters. CVPR 2024: 23219-23230
[c9]Haiwen Diao, Bo Wan, Xu Jia, Yunzhi Zhuge, Ying Zhang, Huchuan Lu, Long Chen:
SHERL: Synthesizing High Accuracy and Efficient Memory for Resource-Limited Transfer Learning. ECCV (44) 2024: 75-95
[c8]Jiazuo Yu, Haomiao Xiong, Lu Zhang, Haiwen Diao, Yunzhi Zhuge, Lanqing Hong, Dong Wang, Huchuan Lu, You He, Long Chen:
LLMs Can Evolve Continually on Modality for X-Modal Reasoning. NeurIPS 2024- 2023
[c7]Kaining Ying, Qing Zhong
, Weian Mao, Zhenhua Wang, Hao Chen, Lin Yuanbo Wu, Yifan Liu, Chengxiang Fan
, Yunzhi Zhuge, Chunhua Shen:
CTVIS: Consistent Training for Online Video Instance Segmentation. ICCV 2023: 899-908
[c6]Hongyu Gu, Yunzhi Zhuge
, Lu Zhang, Jinqing Qi, Huchuan Lu:
Few-shot Semantic Segmentation by Exploiting Dynamic and Regional Contexts. ICME 2023: 834-839- 2022
[c5]Yun-Zhi Zhuge
, Xu Jia
:
Multi-granularity Transformer for Image Super-Resolution. ACCV (3) 2022: 138-154- 2021
[c4]Yunzhi Zhuge
, Chunhua Shen:
Deep Reasoning Network for Few-shot Semantic Segmentation. ACM Multimedia 2021: 5344-5352- 2019
[c3]Yun-Zhi Zhuge, Yu Zeng, Huchuan Lu:
Deep Embedding Features for Salient Object Detection. AAAI 2019: 9340-9347
[c2]Yu Zeng, Yun-Zhi Zhuge, Huchuan Lu, Lihe Zhang, Mingyang Qian, Yizhou Yu:
Multi-Source Weak Supervision for Saliency Detection. CVPR 2019: 6074-6083
[c1]Yu Zeng, Yun-Zhi Zhuge, Huchuan Lu, Lihe Zhang:
Joint Learning of Saliency Detection and Weakly Supervised Semantic Segmentation. ICCV 2019: 7222-7232
Informal and Other Publications
- 2026
[i29]Xiyan Feng, Wenbo Zhang, Lu Zhang, Yunzhi Zhuge, Huchuan Lu, You He:
Towards Cross-Platform Generalization: Domain Adaptive 3D Detection with Augmentation and Pseudo-Labeling. CoRR abs/2601.08174 (2026)
[i28]Qing'an Liu, Juntong Feng, Yuhao Wang, Xinzhe Han, Yujie Cheng, Yue Zhu, Haiwen Diao, Yunzhi Zhuge, Huchuan Lu:
VISTA-Bench: Do Vision-Language Models Really Understand Visualized Text as Well as Pure Text? CoRR abs/2602.04802 (2026)
[i27]Kecheng Zhang, Zongxin Yang, Mingfei Han, Haihong Hao, Yunzhi Zhuge, Changlin Li, Junhan Zhao, Zhihui Li, Xiaojun Chang:
Progressive Online Video Understanding with Evidence-Aligned Timing and Transparent Decisions. CoRR abs/2604.18459 (2026)
[i26]Yuhao Wang, Mu Qiao, Haiwen Diao, Yunzhi Zhuge, Pingping Zhang, Xindong Zhang, Lei Zhang, Huchuan Lu:
ERA: Entropy-Guided Visual Token Pruning with Rectified Attention for Efficient MLLMs. CoRR abs/2606.31982 (2026)- 2025
[i25]Yunzhi Zhuge, Hongyu Gu, Lu Zhang, Jinqing Qi, Huchuan Lu:
Learning Motion and Temporal Cues for Unsupervised Video Object Segmentation. CoRR abs/2501.07806 (2025)
[i24]Sitong Gong, Yunzhi Zhuge, Lu Zhang, Yifan Wang, Pingping Zhang, Lijun Wang, Huchuan Lu:
AVS-Mamba: Exploring Temporal and Multi-modal Mamba for Audio-Visual Segmentation. CoRR abs/2501.07810 (2025)
[i23]Haomiao Xiong, Yunzhi Zhuge, Jiawen Zhu, Lu Zhang, Huchuan Lu:
3UR-LLM: An End-to-End Multimodal Large Language Model for 3D Scene Understanding. CoRR abs/2501.07819 (2025)
[i22]Sitong Gong, Yunzhi Zhuge, Lu Zhang, Zongxin Yang, Pingping Zhang, Huchuan Lu:
The Devil is in Temporal Token: High Quality Video Reasoning Segmentation. CoRR abs/2501.08549 (2025)
[i21]Haomiao Xiong, Zongxin Yang, Jiazuo Yu, Yunzhi Zhuge, Lu Zhang, Jiawen Zhu, Huchuan Lu:
Streaming Video Understanding and Multi-round Interaction with Memory-enhanced Knowledge. CoRR abs/2501.13468 (2025)
[i20]Hengrun Zhao, Yunzhi Zhuge, Yifan Wang, Lijun Wang, Huchuan Lu, Yu Zeng:
Learning Universal Features for Generalizable Image Forgery Localization. CoRR abs/2504.07462 (2025)
[i19]Yue Zhu, Haiwen Diao, Shang Gao, Jiazuo Yu, Jiawen Zhu, Yunzhi Zhuge, Shuai Hao, Xu Jia, Lu Zhang, Ying Zhang, Huchuan Lu:
Regularizing Subspace Redundancy of Low-Rank Adaptation. CoRR abs/2507.20745 (2025)
[i18]Sitong Gong, Lu Zhang, Yunzhi Zhuge, Xu Jia, Pingping Zhang, Huchuan Lu:
Reinforcing Video Reasoning Segmentation to Think Before It Segments. CoRR abs/2508.11538 (2025)
[i17]Zirui Zheng, Takashi Isobe, Tong Shen, Xu Jia, Jianbin Zhao, Xiaomin Li, Mengmeng Ge, Baolu Li
, Qinghe Wang, Dong Li, Dong Zhou, Yunzhi Zhuge, Huchuan Lu, Emad Barsoum:
Layout-Conditioned Autoregressive Text-to-Image Generation via Structured Masking. CoRR abs/2509.12046 (2025)
[i16]Sitong Gong, Yunzhi Zhuge, Lu Zhang, Pingping Zhang, Huchuan Lu:
Complementary and Contrastive Learning for Audio-Visual Segmentation. CoRR abs/2510.10051 (2025)
[i15]Lu Zhang, Jiazuo Yu, Haomiao Xiong, Ping Hu, Yunzhi Zhuge, Huchuan Lu, You He:
FineRS: Fine-grained Reasoning and Segmentation of Small Objects with Reinforcement Learning. CoRR abs/2510.21311 (2025)
[i14]Baolu Li, Yiming Zhang, Qinghe Wang, Liqian Ma, Xiaoyu Shi, Xintao Wang, Pengfei Wan, Zhenfei Yin, Yunzhi Zhuge, Huchuan Lu, Xu Jia:
VFXMaster: Unlocking Dynamic Visual Effect Generation via In-Context Learning. CoRR abs/2510.25772 (2025)
[i13]Xinzhuo Yu, Yunzhi Zhuge, Sitong Gong, Lu Zhang, Pingping Zhang, Huchuan Lu:
Parameter Aware Mamba Model for Multi-task Dense Prediction. CoRR abs/2511.14503 (2025)- 2024
[i12]Qinghe Wang, Xu Jia, Xiaomin Li, Taiqing Li, Liqian Ma, Yunzhi Zhuge, Huchuan Lu:
StableIdentity: Inserting Anybody into Anywhere at First Sight. CoRR abs/2401.15975 (2024)
[i11]Jiazuo Yu, Yunzhi Zhuge, Lu Zhang, Ping Hu, Dong Wang, Huchuan Lu, You He:
Boosting Continual Learning of Vision-Language Models via Mixture-of-Experts Adapters. CoRR abs/2403.11549 (2024)
[i10]Haiwen Diao, Bo Wan, Xu Jia, Yunzhi Zhuge, Ying Zhang, Huchuan Lu, Long Chen:
SHERL: Synthesizing High Accuracy and Efficient Memory for Resource-Limited Transfer Learning. CoRR abs/2407.07523 (2024)
[i9]Jiazuo Yu, Haomiao Xiong, Lu Zhang, Haiwen Diao, Yunzhi Zhuge, Lanqing Hong, Dong Wang, Huchuan Lu, You He, Long Chen:
LLMs Can Evolve Continually on Modality for X-Modal Reasoning. CoRR abs/2410.20178 (2024)
[i8]Yicheng Yang, Pengxiang Li, Lu Zhang, Liqian Ma, Ping Hu, Siyu Du, Yunzhi Zhuge, Xu Jia, Huchuan Lu:
DreamMix: Decoupling Object Attributes for Enhanced Editability in Customized Image Inpainting. CoRR abs/2411.17223 (2024)
[i7]Wenbo Zhang, Lu Zhang, Ping Hu, Liqian Ma, Yunzhi Zhuge, Huchuan Lu:
Bootstraping Clustering of Gaussians for View-consistent 3D Scene Understanding. CoRR abs/2411.19551 (2024)
[i6]Chengyang Ye, Yunzhi Zhuge, Pingping Zhang:
Towards Open-Vocabulary Remote Sensing Image Semantic Segmentation. CoRR abs/2412.19492 (2024)- 2023
[i5]Kaining Ying, Qing Zhong, Weian Mao, Zhenhua Wang, Hao Chen, Lin Yuanbo Wu, Yifan Liu, Chengxiang Fan, Yunzhi Zhuge, Chunhua Shen:
CTVIS: Consistent Training for Online Video Instance Segmentation. CoRR abs/2307.12616 (2023)
[i4]Pengxiang Li, Zhili Liu, Kai Chen, Lanqing Hong, Yunzhi Zhuge, Dit-Yan Yeung, Huchuan Lu, Xu Jia:
TrackDiffusion: Multi-object Tracking Data Generation via Diffusion Models. CoRR abs/2312.00651 (2023)- 2019
[i3]Yu Zeng, Yun-Zhi Zhuge, Huchuan Lu, Lihe Zhang, Mingyang Qian, Yizhou Yu:
Multi-source weak supervision for saliency detection. CoRR abs/1904.00566 (2019)
[i2]Yu Zeng, Yun-Zhi Zhuge, Huchuan Lu, Lihe Zhang:
Joint Learning of Saliency Detection and Weakly Supervised Semantic Segmentation. CoRR abs/1909.04161 (2019)- 2018
[i1]Yun-Zhi Zhuge, Pingping Zhang, Huchuan Lu:
Boundary-guided Feature Aggregation Network for Salient Object Detection. CoRR abs/1809.10821 (2018)
Coauthor Index

manage site settings
To protect your privacy, all features that rely on external API calls from your browser are turned off by default. You need to opt-in for them to become active. All settings here will be stored as cookies with your web browser. For more information see our F.A.Q.
Unpaywalled article links
Add open access links from
to the list of external document links (if available).
Privacy notice: By enabling the option above, your browser will contact the API of unpaywall.org to load hyperlinks to open access articles. Although we do not have any reason to believe that your call will be tracked, we do not have any control over how the remote server uses your data. So please proceed with care and consider checking the Unpaywall privacy policy.
Archived links via Wayback Machine
For web page which are no longer available, try to retrieve content from the
of the Internet Archive (if available).
Privacy notice: By enabling the option above, your browser will contact the API of archive.org to check for archived content of web pages that are no longer available. Although we do not have any reason to believe that your call will be tracked, we do not have any control over how the remote server uses your data. So please proceed with care and consider checking the Internet Archive privacy policy.
Reference lists
Add a list of references from
,
, and
to record detail pages.
load references from crossref.org and opencitations.net
Privacy notice: By enabling the option above, your browser will contact the APIs of crossref.org, opencitations.net, and semanticscholar.org to load article reference information. Although we do not have any reason to believe that your call will be tracked, we do not have any control over how the remote server uses your data. So please proceed with care and consider checking the Crossref privacy policy and the OpenCitations privacy policy, as well as the AI2 Privacy Policy covering Semantic Scholar.
Citation data
Add a list of citing articles from
and
to record detail pages.
load citations from opencitations.net
Privacy notice: By enabling the option above, your browser will contact the API of opencitations.net and semanticscholar.org to load citation information. Although we do not have any reason to believe that your call will be tracked, we do not have any control over how the remote server uses your data. So please proceed with care and consider checking the OpenCitations privacy policy as well as the AI2 Privacy Policy covering Semantic Scholar.
OpenAlex data
Load additional information about publications from
.
Privacy notice: By enabling the option above, your browser will contact the API of openalex.org to load additional information. Although we do not have any reason to believe that your call will be tracked, we do not have any control over how the remote server uses your data. So please proceed with care and consider checking the information given by OpenAlex.
last updated on 2026-08-25 22:21 CEST by the dblp team
all metadata released as open data under CC0 1.0 license
see also: Terms of Use | Privacy Policy | Imprint


Google
Google Scholar
Semantic Scholar
Internet Archive Scholar
CiteSeerX
ORCID






