


default search action
Yidi Li 0001
Person information
- affiliation: Taiyuan University of Technology, College of Computer Science and Technology, Taiyuan, China
- affiliation: Peking University, Shenzhen Graduate School, Key Laboratory of Machine Perception, Shenzhen, China
Other persons with the same name
- Yidi Li — disambiguation page
Refine list

refinements active!
zoomed in on ?? of ?? records
view refined list in
2020 – today
- 2026
[j18]Yihan Li
, Yidi Li
, Zhenhuan Xu
, Hao Guo, Mengyuan Liu
, Weiwei Wan
:
AVCLNet: Multimodal Multispeaker Tracking Network Using Audio-Visual Contrastive Learning. CAAI Trans. Intell. Technol. 11(1): 238-255 (2026)
[j17]Yixin Guo
, Zhenxue Chen
, Xuewen Rong
, Chengyun Liu, Lili Song, Yidi Li:
ASNet: An adaptive scene-aware network for RGB-thermal urban scene semantic segmentation. J. Vis. Commun. Image Represent. 118: 104822 (2026)
[j16]Yidi Li
, Yihan Li, Zhenhuan Xu, Weiwei Wan, Hong Liu
:
Global-local distillation network-based audio-visual speaker tracking with incomplete modalities. Pattern Recognit. 179: 113916 (2026)
[j15]Bin Ren, Eduard Zamfir, Zongwei Wu, Yawei Li, Yidi Li, Danda Pani Paudel, Radu Timofte, Ming-Hsuan Yang, Luc Van Gool, Nicu Sebe:
Any Image Restoration via Efficient Spatial-Frequency Degradation Adaptation. Trans. Mach. Learn. Res. 2026 (2026)
[c14]Chaoyi Guo, Jianyu Zhou
, Zijie Wang, Mingzhe Liu, Yun Zhao, Yidi Li:
Enhancing Document Layout Analysis Through Frequency Decomposition Convolution and Multi-scale Adaptive Linear Attention. ICIC (10) 2026: 530-541
[c13]Zijie Wang, Chaoyi Guo, Zichu Zhang, Muhan Guo, Yidi Li:
XS-YOLO: A Lightweight NMS-Free Framework for Small Object Detection in Remote Sensing Imagery. ICIC (18) 2026: 575-585
[i10]Wei Yu, Runjia Qian, Yumeng Li, Liquan Wang, Songheng Yin, Sri Siddarth Chakaravarthy P, Dennis Anthony, Yang Ye, Yidi Li, Weiwei Wan, Animesh Garg:
MosaicMem: Hybrid Spatial Memory for Controllable Video World Models. CoRR abs/2603.17117 (2026)
[i9]Qian Feng, Pengfei Li, Rongshan Gao, Jiale Xu, Rui Gong, Yidi Li:
EMPD: An Event-based Multimodal Physiological Dataset for Remote Pulse Wave Detection. CoRR abs/2603.26699 (2026)
[i8]Qian Feng, Hao Guo, Yan Niu, Zhenhuan Xu, Yidi Li:
Fusion-E2Pulse: A Multimodal Event-RGB Fusion Network for Non-contact Pulse Wave Reconstruction. CoRR abs/2606.15597 (2026)- 2025
[j14]Jie Xiang
, Ang Zhao
, Xia Li
, Xubin Wu, Yanqing Dong, Yan Niu, Xin Wen
, Yidi Li:
Enhancing Brain MRI Super-Resolution Through Multi-Slice Aware Matching and Fusion. CAAI Trans. Intell. Technol. 10(5): 1411-1421 (2025)
[j13]Xubin Wu, Yan Niu, Xia Li
, Jie Xiang
, Yidi Li:
A Prior Causality-Guided Multi-View Diffusion Network for Brain Disorder Classification. CAAI Trans. Intell. Technol. 10(6): 1731-1744 (2025)
[j12]Hangbei Cheng
, Xueyu Liu
, Jun Zhang
, Xiaorong Dong
, Xuetao Ma
, Yansong Zhang
, Hao Meng
, Xing Chen
, Guanghui Yue
, Yidi Li
, Yongfei Wu
:
GLMKD: Joint global and local mutual knowledge distillation for weakly supervised lesion segmentation in histopathology images. Expert Syst. Appl. 279: 127425 (2025)
[j11]Yidi Li
, Jiahao Wen
, Rui Gong, Bin Ren
, Wenhao Li
, Chen Cheng
, Hong Liu
, Nicu Sebe
:
PVAFN: Point-Voxel Attention Fusion Network with Multi-Pooling Enhancing for 3D Object Detection. Expert Syst. Appl. 281: 127608 (2025)
[j10]Yixin Guo
, Zhenxue Chen
, Xuewen Rong, Chengyun Liu, Lili Song, Yidi Li
:
3CNet: Cross-modal cooperative correction network for RGB-T semantic segmentation. Image Vis. Comput. 161: 105638 (2025)
[j9]Zihao Mi, Jianan Zhang, Xueyu Liu, Guanghui Yue, Junhong Yue, Mingqiang Wei, Yidi Li
, Yongfei Wu
:
Multi-instance curriculum learning for histopathology image classification with bias reduction. Medical Image Anal. 105: 103647 (2025)
[j8]Ying Zhu
, Hong Liu
, Guoliang Hua
, Hao Tang
, Yidi Li
, Weibo Huang:
Dual Attention Guidance Network for Self-Supervised Monocular Depth Estimation. IEEE Trans. Circuits Syst. Video Technol. 35(12): 12023-12037 (2025)
[j7]Yidi Li
, Hong Liu
, Bing Yang
:
STNet: Deep Audio-Visual Fusion Network for Robust Speaker Tracking. IEEE Trans. Multim. 27: 1835-1847 (2025)
[c12]Yidi Li, Wenkai Zhao, Zeyu Wang, Zhenhuan Xu, Bin Ren, Nicu Sebe
:
Multi-Stage Multimodal Distillation for Audio-Visual Speaker Tracking. ICASSP 2025: 1-5
[c11]Yidi Li, Kairan Zhang, Chenxu Yang, Chongwei Yan, Rongshan Gao, Mingliang Dou, Bin Ren:
Vision-Guided Acoustic Localization with Decoupled Inference for Moving Speakers. ICIC (19) 2025: 15-27
[c10]Yihong Wu, Jinqiao Wei, Xionghui Zhao, Yidi Li, Shaoyi Du, Bin Ren, Nicu Sebe
:
DSGC-Net: A Dual-Stream Graph Convolutional Network for Crowd Counting via Feature Correlation Mining. PRCV (17) 2025: 400-413
[i7]Bin Ren, Eduard Zamfir, Zongwei Wu, Yawei Li
, Yidi Li, Danda Pani Paudel, Radu Timofte, Ming-Hsuan Yang, Luc Van Gool, Nicu Sebe:
Any Image Restoration via Efficient Spatial-Frequency Degradation Adaptation. CoRR abs/2504.14249 (2025)
[i6]Yihong Wu, Jinqiao Wei, Xionghui Zhao, Yidi Li, Shaoyi Du, Bin Ren, Nicu Sebe:
DSGC-Net: A Dual-Stream Graph Convolutional Network for Crowd Counting via Feature Correlation Mining. CoRR abs/2509.02261 (2025)- 2024
[j6]Yidi Li
, Jiale Ren
, Yawei Wang, Guoquan Wang, Xia Li
, Hong Liu:
Audio-visual keyword transformer for unconstrained sentence-level keyword spotting. CAAI Trans. Intell. Technol. 9(1): 142-152 (2024)
[j5]Wanruo Zhang
, Hong Liu, Jianbing Wu
, Yidi Li
:
MVSSC: Meta-reinforcement learning based visual indoor navigation using multi-view semantic spatial context. Pattern Recognit. Lett. 177: 75-81 (2024)
[j4]Tao Wang
, Mengyuan Liu
, Hong Liu
, Wenhao Li
, Miaoju Ban
, Tianyu Guo
, Yidi Li
:
Feature Completion Transformer for Occluded Person Re-Identification. IEEE Trans. Multim. 26: 8529-8542 (2024)
[c9]Zihao Mi, Xueyu Liu, Jianan Zhang, Guangze Shi
, Yidi Li, Yongfei Wu
:
Multi-instance Curriculum Learning for Histopathology Image Classifications with Hard Negative Mining and Positive Augmentation. BIBM 2024: 2311-2316
[c8]Ruijia Fan, Hong Liu, Yidi Li, Peini Guo, Guoquan Wang, Ti Wang:
AttA-NET: Attention Aggregation Network for Audio-Visual Emotion Recognition. ICASSP 2024: 8030-8034
[c7]Zhenhuan Xu, Yongfei Wu
, Liming Zhang, Yidi Li:
Adaptive Fourier Decomposition Based Signal Extraction on Weak Electromagnetic Field. ICASSP 2024: 9446-9450
[i5]Yidi Li, Jiahao Wen, Bin Ren, Wenhao Li, Zhenhuan Xu, Hao Guo, Hong Liu, Nicu Sebe:
PVAFN: Point-Voxel Attention Fusion Network with Multi-Pooling Enhancing for 3D Object Detection. CoRR abs/2408.14600 (2024)
[i4]Yidi Li, Hong Liu, Bing Yang:
STNet: Deep Audio-Visual Fusion Network for Robust Speaker Tracking. CoRR abs/2410.05964 (2024)- 2023
[j3]Yidi Li
, Guoquan Wang, Zhan Chen, Hao Tang
, Hong Liu:
On-device audio-visual multi-person wake word spotting. CAAI Trans. Intell. Technol. 8(4): 1578-1589 (2023)
[j2]Jian Zhang
, Ge Yang
, Runwei Ding
, Yidi Li
:
Cascade RDN: Towards Accurate Localization in Industrial Visual Anomaly Detection With Structural Anomaly Generation. IEEE Robotics Autom. Lett. 8(9): 5560-5567 (2023)
[c6]Xingyue Shi, Hong Liu, Wei Shi, Zihui Zhou, Yidi Li:
Boosting Person Re-Identification with Viewpoint Contrastive Learning and Adversarial Training. ICASSP 2023: 1-5
[c5]Guoquan Wang, Hong Liu, Tianyu Guo
, Jingwen Guo, Ti Wang, Yidi Li:
Self-Supervised 3D Skeleton Representation Learning with Active Sampling and Adaptive Relabeling for Action Recognition. ICIP 2023: 56-60
[i3]Tao Wang, Hong Liu, Wenhao Li, Miaoju Ban, Tuanyu Guo, Yidi Li:
Feature Completion Transformer for Occluded Person Re-identification. CoRR abs/2303.01656 (2023)
[i2]Tianyu Guo, Mengyuan Liu, Hong Liu, Wenhao Li, Jingwen Guo, Tao Wang, Yidi Li:
Joint Adversarial and Collaborative Learning for Self-Supervised Action Recognition. CoRR abs/2307.07791 (2023)- 2022
[c4]Yidi Li
, Hong Liu, Hao Tang
:
Multi-Modal Perception Attention Network with Self-Supervised Learning for Audio-Visual Speaker Tracking. AAAI 2022: 1456-1463
[c3]Peini Guo
, Zhengyan Chen
, Yidi Li
, Hong Liu
:
Audio-Visual Fusion Network Based on Conformer for Multimodal Emotion Recognition. CICAI (2) 2022: 315-326- 2021
[i1]Yidi Li, Hong Liu, Hao Tang:
Multi-Modal Perception Attention Network with Self-Supervised Learning for Audio-Visual Speaker Tracking. CoRR abs/2112.07423 (2021)- 2020
[j1]Yidi Li
, Hong Liu, Bing Yang, Runwei Ding, Yang Chen:
Deep Metric Learning-Assisted 3D Audio-Visual Speaker Tracking via Two-Layer Particle Filter. Complex. 2020: 3764309:1-3764309:8 (2020)
[c2]Hong Liu, Yongheng Sun, Yidi Li, Bing Yang:
3D Audio-Visual Speaker Tracking with A Novel Particle Filter. ICPR 2020: 7343-7348
2010 – 2019
Coauthor Index

manage site settings
To protect your privacy, all features that rely on external API calls from your browser are turned off by default. You need to opt-in for them to become active. All settings here will be stored as cookies with your web browser. For more information see our F.A.Q.
Unpaywalled article links
Add open access links from
to the list of external document links (if available).
Privacy notice: By enabling the option above, your browser will contact the API of unpaywall.org to load hyperlinks to open access articles. Although we do not have any reason to believe that your call will be tracked, we do not have any control over how the remote server uses your data. So please proceed with care and consider checking the Unpaywall privacy policy.
Archived links via Wayback Machine
For web page which are no longer available, try to retrieve content from the
of the Internet Archive (if available).
Privacy notice: By enabling the option above, your browser will contact the API of archive.org to check for archived content of web pages that are no longer available. Although we do not have any reason to believe that your call will be tracked, we do not have any control over how the remote server uses your data. So please proceed with care and consider checking the Internet Archive privacy policy.
Reference lists
Add a list of references from
,
, and
to record detail pages.
load references from crossref.org and opencitations.net
Privacy notice: By enabling the option above, your browser will contact the APIs of crossref.org, opencitations.net, and semanticscholar.org to load article reference information. Although we do not have any reason to believe that your call will be tracked, we do not have any control over how the remote server uses your data. So please proceed with care and consider checking the Crossref privacy policy and the OpenCitations privacy policy, as well as the AI2 Privacy Policy covering Semantic Scholar.
Citation data
Add a list of citing articles from
and
to record detail pages.
load citations from opencitations.net
Privacy notice: By enabling the option above, your browser will contact the API of opencitations.net and semanticscholar.org to load citation information. Although we do not have any reason to believe that your call will be tracked, we do not have any control over how the remote server uses your data. So please proceed with care and consider checking the OpenCitations privacy policy as well as the AI2 Privacy Policy covering Semantic Scholar.
OpenAlex data
Load additional information about publications from
.
Privacy notice: By enabling the option above, your browser will contact the API of openalex.org to load additional information. Although we do not have any reason to believe that your call will be tracked, we do not have any control over how the remote server uses your data. So please proceed with care and consider checking the information given by OpenAlex.
last updated on 2026-09-02 23:52 CEST by the dblp team
all metadata released as open data under CC0 1.0 license
see also: Terms of Use | Privacy Policy | Imprint


Google
Google Scholar
Semantic Scholar
Internet Archive Scholar
CiteSeerX
ORCID






