I am currently a researcher at Ant Group, working on multimodal large language models. Previously, I was a postdoctoral researcher at The Hong Kong Polytechnic University (PolyU). I received my Ph.D. from Peking University in 2022 under the supervision of Prof. Shiliang Zhang. My current research interests include MLLM agents, multimodal reasoning, and related areas.
2026.09, two papers about latent reasoning and reliability-aware RL in MLLM are accepted by EMNLP2026.
2026.06, one paper about counting in MLLM is accepted by ECCV2026.
2025.12, one journal paper about zero-shot keypoint localization in MLLM is accepted by Pattern Recognition.
2025.10, We have released the MLLM techinical report, Ming-Flash-Omni, a unified architecture for multimodal perception and generation.
2025.07, one paper is accepted by NeruIPS.
2024.02, one journal paper about pose estimation is accepted by Chinese Journal of Computer.
2022.06, I graduated from PKU in 2022.
2022.01, one journal paper about pose estimation in polar coordinates is accepted by TIP.
2020.12, I won the National Scholarship.
2020.06, I won the President Scholarship.
2020.02, one journal paper is accepted by TIP.
2019.12, I won the Qingyun Shi Best Paper Award.
2019.07, one journal paper is accepted by TPAMI.
2019.07, one paper is accepted by ICCV2019.
2018.11, one paper is accepted by AAAI2019 (oral).
2018.06, one demo paper is accepted by ICMR2018.
2017.08, I won the 2nd Prize in 4th China Graduate Contest on Smart City Technology and Creative Design. Thank Prof. Zhang and all my teammates.
2017.07, one paper is accepted by ICCV2017.
Senior Program Committee of AAAI.
Editor of IEIE SPC.
Reviewer of ICCV, CVPR, ECCV, NeurIPS, ICML, AAAI, ACL, EMNLP, etc.
Reviewer of T-PAMI, T-IP, IJCV, T-CSVT, TOMM, Neurocomputing, IET computer vision, etc.
(First / Corresponding Author)
Learning When to Think in Latent Visual Space: Reasoning Mode Adaptive Reinforcement Learning.
Jade Chein, Jinpeng Ou, Leo Wei, Qingpei Guo, Jianing Li
EMNLP 2026. [Paper]
To Answer or Not to Answer? Reliability-Oriented Reinforcement Learning for Visual Language Models.
Jade Chien, Julian Wu, Leo Wei, Dylan Chen, Qingpei Guo, Jianing Li
EMNLP 2026. [Paper]
HER-Count: Learning Hyper-Exemplar Representation for Generalized Zero-Shot Object Counting.
Jianing Li, Xiaobin Liu, Ruihan Xu
ECCV 2026. [Paper]
Generalizable Large Language Model Based Human Keypoint Localization.
Jianing Li, Xiaobin Liu, Chanho Eom, Shuang Yang, Jianzhong He, Hantao Yao, Jing Yuan
Pattern Recognition, 2025. [Paper]
BMW: Bidirectionally Memory bank reWriting for Unsupervised Person Re-Identification.
Xiaobin Liu, Jianing Li, Baiwei Guo, WenbinZhu, Jing Yuan
NeurIPS 2025. [Paper]
Deep Learning-Based 2D Human Pose Estimation: Current Status and Future Prospects.
Jianing Li, Dongkai Wang, Shiliang Zhang
Chinese Journal of Computer 2024. [Paper]
PolarPose: Single-stage Multi-person Pose Estimation in Polar Coordinates
Jianing Li, Yaowei Wang, Shiliang Zhang
TIP 2022. [Paper]
Joint Visual and Temporal Consistency for Unsupervised Domain Adaptive Person Re-Identification
Jianing Li, Shiliang Zhang
ECCV 2020. [Paper]
Multi-Scale Temporal Cues Learning for Video Person Re-Identification
Jianing Li, Shiliang Zhang, Tiejun Huang
TIP 2020. [Paper]
Pose-Guided Representation Learning for Person Re-Identification
Jianing Li, Shiliang Zhang, Qi Tian, Meng Wang, Wen Gao
PAMI 2019. [Paper]
Global-Local Temporal Representations For Video Person Re-Identification
Jianing Li, Jingdong Wang, Qi Tian, Wen Gao, Shiliang Zhang
ICCV 2019. [Paper]
Multi-Scale 3D Convolution Network for Video Based Person Re-Identification
Jianing Li, Shiliang Zhang, Tiejun Huang
AAAI 2019 (oral).
[Paper]
VP-ReID: Vehicle and Person Re-Identification System
Longhui Wei, Xiaobin Liu, Jianing Li, Shiliang Zhang
ICMR Demo, 2018.
[Paper]
Pose-driven Deep Convolutional Model for Person Re-identification
Chi Su*, Jianing Li*, Shiliang Zhang, Junliang Xing, Wen Gao, Qi Tian (*equal contribution)
ICCV 2017.
[Paper]