(* stands for equal contribution)
Preprints
|
Tele-Catch: Adaptive Teleoperation for Dexterous Dynamic 3D Object Catching
W. Zhao, J. Dong, R. Zhang, K. Li, Q. Zhao, K. Huang
arXiv, 2026 ,Under Review
[ arXiv] [Code]
|
|
Towards Robotic Dexterous Hand Intelligence: A Survey
W. Zhao*, T. Liang*, X. Guo*, Rui Zhang, I. King, K. Huang
arXiv, 2026 ,Under Review
[ arXiv] [Project]
|
|
InfiniHand: Streaming World-Space Hand Motion Estimation from Egocentric Video
K. Ren, K. Song, W. Zhao, Y. Wang, Y. Liu, B. Dai, H. Guo, C. Shen, M. Yu, T. Lu, J. Dong
arXiv, 2026 ,Under Review
[ arXiv] [ Project]
|
|
InternW0-Δ: A World Action Model Bridging Predictive Dynamics and Actions
Physical Intelligence Team (42 Authors), Shanghai AI Laboratory
arXiv, 2026 ,Under Review
[ Technical Report] [ Project]
|
Conferences
|
SynthVerse: A Large-Scale Diverse Synthetic Dataset for Point Tracking
W. Zhao*, H. Xu*, X. Miao, Q. Zhao, R. Zhang, K. Huang, N. Gao, P. Cao, M. Sun, M. Yu, T. Lu, L. Xu, J. Dong, J. Pang
ACM SIGGRAPH Annual Conference (SIGGRAPH), 2026,CCF-A, Core-A*
[ Paper] [ Project] [ Code]
|
|
TrajVG: 3D Trajectory-Coupled Visual Geometry Learning
X. Miao, W. Zhao, T. Lu, L. Xu, M. Yu, Y. Long, J. Pang, J. Dong
ACM SIGGRAPH Annual Conference (SIGGRAPH), 2026,CCF-A, Core-A*
[ Paper] [ Project] [ Code]
|
|
Geo-DPO: Aligning Semantic Intent with Geometry for 3D Affordance Segmentation
Z. Yang, X. Wang, X. Qiu W. Zhao, Q. Zhang, J. Xiao
European Conference on Computer Vision (ECCV), 2026,CCF-B, Core-A*
[ Paper] [Code]
|
|
BFANet: Revisiting 3D Semantic Segmentation with Boundary Feature Analysis
W. Zhao, R. Zhang, Q. Wang, G. Cheng, K. Huang
Conference on Computer Vision and Pattern Recognition (CVPR), 2025,CCF-A, Core-A*
[ Paper] [ Code]
|
|
PO3AD: Predicting Point Offsets toward Better 3D Point Cloud Anomaly Detection
J. Ye*, W. Zhao*, X. Yang, G. Cheng, K. Huang
Conference on Computer Vision and Pattern Recognition (CVPR), 2025,CCF-A, Core-A*
[ Paper] [ Code]
|
|
Towards Training-Free Open-World Classification with 3D Generative Models
X. Xia*, W. Zhao*, Y. Yan, G. Yang, R. Zhang, K.Huang, X. Yang
ACM International Conference on Multimedia (ACM MM), 2025,CCF-A, Core-A*
[ Paper] [ Code]
|
|
Divide and Conquer: 3D Point Cloud Instance Segmentation With Point-Wise Binarization
W. Zhao, Y. Yan, C. Yang, J. Ye, X. Yang, K. Huang
International Conference on Computer Vision (ICCV), 2023,CCF-A, Core-A*
[ Paper] [ Code] [ Video]
|
Journals
|
Unlock Pose Diversity: Accurate and Efficient Implicit Keypoint-based Spatiotemporal Diffusion for Audio-driven Talking Portrait
C. Yang, K. Yao, Y. Yan, C. Jiang, W. Zhao, J. Sun, etc.
International Journal of Computer Vision (IJCV) , 2026,CCF-A,JCR-Q1
[ Paper] [ Code] [ Demo]
|
|
Open-Pose 3D Zero-Shot Learning: Benchmark and Challenges
W. Zhao*, G. Yang*, R. Zhang, C. Jiang, C. Yang, Y. Yan, A. Hussain, K. Huang
Neural Networks (NN), 2025,CCF-B,JCR-Q1
[ Paper] [ Code]
|
|
Revisiting 3D point cloud analysis with Markov process
C. Jiang, W. Ma, K. Huang, Q. Wang, X. Yang,W. Zhao, J. Wu, X. Wang, etc.
Pattern Recognition (PR), 2025,CCF-B,JCR-Q1
[ Paper] [ Code]
|