Kai (Kevin) Xu徐凯

Professor at the Institute of AI for Industries (IAII), Chinese Academy of Sciences · PI of PAI Lab · Co-lead of iGRAPE Lab @ NUDT

Hi, I am Kai Xu, a Professor at the Institute of AI for Industries, Chinese Academy of Sciences. I lead the PAI Lab @ IAII and co-lead the iGRAPE Lab @ NUDT. Our work sits at the intersection of computer graphics, embodied intelligence, and physical & spatial intelligence — currently focused on world models, neural simulators, and robotic manipulation & navigation.

Before joining IAII I was a professor at the National University of Defense Technology (NUDT), where I earned my Ph.D. in 2011, and an Adjunct Professor of Simon Fraser University. From 2008 to 2010 I was a visiting Ph.D. student in the GrUVi Lab at SFU with Richard (Hao) Zhang. I did postdoctoral research at VCC@SIAT from 2012 to 2014 with Baoquan Chen, and visited the Vision and Robotics Group at Princeton in 2017–2018, working with Thomas Funkhouser and Szymon Rusinkiewicz.

I serve as an Associate Editor of ACM Transactions on Graphics and IEEE TVCG, and as Executive Area Editor of Computational Visual Media. The full paper list is on the publications page.

We are recruiting. PostDocs, Ph.D./Master's students, and research engineers with a CS/EE background are welcome to join PAI Lab and iGRAPE Lab. If you are interested in world models, neural simulation, or robot learning, email me your CV.
Selected Projects
SAFE-EQA semantic-aware embodied question answering
SAFE-EQA: Semantic-Aware Efficient Exploration for Embodied Question Answering
Kai Xu et al.
ECCV 2026
“Guide embodied agents toward question-relevant regions by combining semantic maps, efficient exploration, and VLM reasoning.”
HoloTetSphere unified tetrahedral mesh reconstruction
HoloTetSphere: Unified TetSphere Mesh Reconstruction for Physical Simulations
YaQiao Dai, Renjiao Yi, Zhirui Gao, Wei Chen, Kai Xu, Chenyang Zhu
ECCV 2026
“Topology-adaptive reconstruction of unified, single-connected tetrahedral meshes for reliable downstream physical simulation.”
RoboBPP benchmark
RoboBPP: Benchmarking Robotic Online Bin Packing with Physics-based Simulation
Zhoufeng Wang, Hang Zhao, Juzhan Xu, Shishun Zhang, Zeyu Xiong, Ruizhen Hu, Chenyang Zhu, Zecui Zeng, Kai Xu
Submitted 2026
“A physics-based benchmark that finally asks whether a packing plan is physically feasible.”
World models survey
From Specialist to Generalist: A Comprehensive Survey on World Models
Kai Xu, Hang Zhao, Ruizhen Hu, Yuhang Huang, Ziqiao Zhou, Wancheng Feng, Li Yi, Sida Peng, et al.
Computational Visual Media 2026
“Specialist and generalist world models must coexist — precision and generalization pull in opposite directions.”
PIN-WM
PIN-WM: Learning Physics-INformed World Models for Non-Prehensile Manipulation
Wenxuan Li, Hang Zhao, Zhiyuan Yu, Yu Du, Qin Zou, Ruizhen Hu, Kai Xu
RSS 2025
“Identify mass, friction and restitution from vision alone — then learn a policy that transfers.”
LaDi-WM
LaDi-WM: A Latent Diffusion-Based World Model for Predictive Manipulation
Yuhang Huang, Jiazhao Zhang, Shilong Zou, Xinwang Liu, Ruizhen Hu, Kai Xu
CoRL 2025
“Predicting the evolution of latent space is easier to learn — and generalizes better — than predicting pixels.”
Pin-pression gripper
Designing Pin-pression Gripper and Learning its Dexterous Grasping with Online In-hand Adjustment
Hewen Xiao, Xiuping Liu, Hang Zhao, Jian Liu, Kai Xu
SIGGRAPH 2025 · ACM TOG
“A gripper inspired by pin-pression toys: each finger reshapes itself around the object.”
RemixFusion
RemixFusion: Residual-based Mixed Representation for Large-scale Online RGB-D Reconstruction
Yuqing Lan, Chenyang Zhu, Shuaifeng Zhi, Jiazhao Zhang, Zhoufeng Wang, Renjiao Yi, Yijie Wang, Kai Xu
SIGGRAPH Asia 2025 · ACM TOG
“High-quality, large-scale online RGB-D reconstruction from a residual-based mixed representation.”
CogNav
CogNav: Cognitive Process Modeling for Object Goal Navigation with LLMs
Yihan Cao, Jiazhao Zhang, Zhinan Yu, Shuzhen Liu, Zheng Qin, Qin Zou, Bo Du, Kai Xu
ICCV 2025
“Object search is a cognitive process, not just a perceptual one — so we model it with an LLM.”
Packing configuration trees
Deliberate Planning of 3D Bin Packing on Packing Configuration Trees
Hang Zhao, Juzhan Xu, Kexiong Yu, Ruizhen Hu, Chenyang Zhu, Bo Du, Kai Xu
IJRR 2025
“A full-fledged tree description of the state and action space of online 3D bin packing.”
LLM-enhanced rearrangement
LLM-enhanced Scene Graph Learning for Household Rearrangement
Wenhao Li, Zhiyuan Yu, Qijin She, Zhinan Yu, Yuqing Lan, Chenyang Zhu, Ruizhen Hu, Kai Xu
SIGGRAPH Asia 2024 · extended in TOG 2026
“Mining object functionality and user preference directly from the scene itself.”
MIPS-Fusion
MIPS-Fusion: Multi-Implicit-Submaps for Scalable and Robust Online Neural RGB-D Reconstruction
Yijie Tang, Jiazhao Zhang, Zhinan Yu, He Wang, Kai Xu
SIGGRAPH Asia 2023 · ACM TOG
“Divide-and-conquer neural SLAM: many small implicit submaps instead of one big field.”
IBS-Grasp
Learning High-DOF Reaching-and-Grasping via Dynamic Representation of Gripper-Object Interaction
Qijin She, Ruizhen Hu, Juzhan Xu, Min Liu, Kai Xu, Hui Huang
SIGGRAPH 2022 · ACM TOG
“Interaction Bisector Surface as a compact, gripper-agnostic state for dexterous grasping.”
ROSEFusion
ROSEFusion: Random Optimization for Online Dense Reconstruction under Fast Camera Motion
Jiazhao Zhang, Chenyang Zhu, Lintao Zheng, Kai Xu
SIGGRAPH 2021 · ACM TOG
“Real-time RGB-D reconstruction that survives violently fast camera motion.”
SymmetryNet
SymmetryNet: Learning to Predict Reflectional and Rotational Symmetries of 3D Shapes from Single-View RGB-D Images
Yifei Shi, Junwen Huang, Hongjia Zhang, Xin Xu, Szymon Rusinkiewicz, Kai Xu
SIGGRAPH Asia 2020 · ACM TOG
“End-to-end learnable symmetry prediction from a single RGB-D view.”
Multi-robot scene reconstruction
Multi-Robot Collaborative Dense Scene Reconstruction
Siyan Dong, Kai Xu, Qiang Zhou, Andrea Tagliasacchi, Shiqing Xin, Matthias Nießner, Baoquan Chen
SIGGRAPH 2019 · ACM TOG
“A fleet of robots that plans, divides and conquers an unknown indoor scene together.”
SCORES
SCORES: Shape Composition with Recursive Substructure Priors
Chenyang Zhu, Kai Xu, Siddhartha Chaudhuri, Renjiao Yi, Hao Zhang
SIGGRAPH Asia 2018 · ACM TOG
“Composing a coherent shape out of incompatible parts, recursively.”
Fine-grained component labeling
Learning to Group and Label Fine-Grained Shape Components
Xiaogang Wang, Bin Zhou, Haiyue Fang, Xiaowu Chen, Qinping Zhao, Kai Xu
SIGGRAPH Asia 2018 · ACM TOG
“Stock 3D models already encode the artist's decomposition — don't throw those components away.”
Shape style from projected lines
Semi-Supervised Co-Analysis of 3D Shape Styles from Projected Lines
Fenggen Yu, Yan Zhang, Kai Xu, Ali Mahdavi-Amiri, Hao Zhang
SIGGRAPH 2018 · Graphics Replicability Stamp
“Localizing style patches on 3D shapes with only weak supervision.”
Object-aware autoscanning
Object-Aware Guidance for Autonomous Scene Reconstruction
Ligang Liu, Xi Xia, Han Sun, Hui Huang, Kai Xu
SIGGRAPH 2018 · ACM TOG
“Interleaving next-best-object with next-best-view, in a single navigational pass.”
Tensor field navigation
Autonomous Reconstruction of Unknown Indoor Scenes Guided by Time-varying Tensor Fields
Kai Xu, Lintao Zheng, Zihao Yan, Guohang Yan, Eugene Zhang, Matthias Nießner, Oliver Deussen, Daniel Cohen-Or, Hui Huang
SIGGRAPH Asia 2017 · ACM TOG
“Tensor fields as a global planner for a robot exploring an unknown room.”
GRASS
GRASS: Generative Recursive Autoencoders for Shape Structures
Jun Li, Kai Xu, Siddhartha Chaudhuri, Ersin Yumer, Hao Zhang, Leonidas Guibas
SIGGRAPH 2017 · Press Release Feature
“Learning a generative model of 3D shape structure, not just geometry.”
Deformation-driven shape correspondence
Deformation-Driven Shape Correspondence
Chenyang Zhu, Renjiao Yi, Wallace Lira, Ibraheem Alhashim, Kai Xu, Hao Zhang
SIGGRAPH 2017 · ACM TOG
“Structure-aware correspondence found by deforming one shape into the other.”
3D attention-driven depth acquisition
3D Attention-Driven Depth Acquisition for Object Identification
Kai Xu, Yifei Shi, Lintao Zheng, Junyu Zhang, Min Liu, Hui Huang, Hao Zhang, Daniel Cohen-Or, Baoquan Chen
SIGGRAPH Asia 2016 · ACM TOG
“Where should the depth camera look next? A 3D recurrent attention model decides.”
Autoscanning
Autoscanning for Coupled Scene Reconstruction and Proactive Object Analysis
Kai Xu, Hui Huang, Yifei Shi, Hao Li, Pinxin Long, Jianong Caichen, Wei Sun, Baoquan Chen
SIGGRAPH Asia 2015 · ACM TOG
“A robot that reconstructs a scene and pokes at its objects to understand them.”
Topology-varying shape correspondence
Deformation-Driven Topology-Varying 3D Shape Correspondence
SIGGRAPH Asia 2015 · ACM TOG
“Corresponding two shapes whose part topologies do not even match.”
Contextual focal points
Organizing Heterogeneous Scene Collections through Contextual Focal Points
Kai Xu, Rui Ma, Hao Zhang, Chenyang Zhu, Ariel Shamir, Daniel Cohen-Or, Hui Huang
SIGGRAPH 2014 · ACM TOG
“Finding the recurring ‘focal points’ that organize a messy collection of 3D scenes.”
Structural blending
Topology-Varying 3D Shape Creation via Structural Blending
SIGGRAPH 2014 · ACM TOG
“Blending two shapes through their structures, letting topology change along the way.”
Symmetry maximization
Layered Analysis of Irregular Facades via Symmetry Maximization
SIGGRAPH 2013 · ACM TOG
“Peeling an irregular building facade into layers by maximizing symmetry.”
Co-hierarchical analysis
Co-Hierarchical Analysis of Shape Structures
SIGGRAPH 2013 · ACM TOG
“One consistent hierarchy shared across a whole set of structurally different shapes.”
Fit and Diverse
Fit and Diverse: Set Evolution for Inspiring 3D Shape Galleries
Kai Xu, Hao Zhang, Daniel Cohen-Or, Baoquan Chen
SIGGRAPH 2012 · ACM TOG
“Evolving a whole set of shapes at once, balancing fitness against diversity.”
Multi-scale partial intrinsic symmetry
Multi-Scale Partial Intrinsic Symmetry Detection
SIGGRAPH Asia 2012 · ACM TOG
“Symmetry is scale-dependent — so detect it across scales, and only partially.”
Photo-inspired modeling
Photo-Inspired Model-Driven 3D Object Modeling
Kai Xu, Hanlin Zheng, Hao Zhang, Daniel Cohen-Or, Ligang Liu, Yueshan Xiong
SIGGRAPH 2011 · ACM TOG
“Turning a single photograph into a 3D model by deforming a structured template.”
Style-content separation
Style-Content Separation by Anisotropic Part Scales
SIGGRAPH Asia 2010 · ACM TOG
“Style lives in the anisotropic scales of parts; content lives in what the parts are.”
Partial intrinsic reflectional symmetry
Partial Intrinsic Reflectional Symmetry of 3D Shapes
Kai Xu, Hao Zhang, Andrea Tagliasacchi, Ligang Liu, Guo Li, Min Meng, Yueshan Xiong
SIGGRAPH Asia 2009 · ACM TOG
“A voting scheme that turns partial symmetries into an intrinsic reflectional symmetry axis transform.”
Feature-aligned shape texturing
Feature-Aligned Shape Texturing
Kai Xu, Daniel Cohen-Or, Tao Ju, Ligang Liu, Hao Zhang, Shizhe Zhou, Yueshan Xiong
SIGGRAPH Asia 2009 · ACM TOG
“Texture that follows the shape's own feature lines.”
The complete publication list, including papers not shown here, is on the publications page.
News
See all publications →
Professional Activities
Journal Editorial Boards
Chairing & Committees
Awards & Honors
Courses & Tutorials
Current and Past Affiliations
IAII, Chinese Academy of Sciences NUDT Simon Fraser University Princeton University VCC @ SIAT