韩锴
Assistant Professor
Visual AI Lab
School of Computing and Data Science
The University of Hong Kong
Room 220, Run Run Shaw Building
Pokfulam Road, Hong Kong
kaihanx at hku.hk
I am an Assistant Professor at The University of Hong Kong, where I direct the Visual AI Lab. My research interests lie in computer vision, machine learning, and artificial intelligence. My current research focuses on spatial intelligence, open-world learning, foundation models, generative AI, and their relevant fields. The overarching goal of my research is to achieve principled and comprehensive visual understanding, close the intelligence gap between machines and humans, and build reliable AI systems for open-world use. Previously, I was a Visiting Faculty Researcher at Google Research, an Assistant Professor in the Department of Computer Science at University of Bristol, and a Postdoctoral Researcher in VGG at the University of Oxford, working with Prof. Andrew Zisserman and Prof. Andrea Vedaldi. I received my Ph.D. from the Department of Computer Science at The University of Hong Kong, advised by Prof. Kenneth K.Y. Wong. During my Ph.D., I also worked with Prof. Jean Ponce and Prof. Minsu Cho at the WILLOW team of Inria and École Normale Supérieure (ENS).
Interested? Send me your resume and transcripts.
| July 2026 | Will serve as Associate Editor for IEEE Robotics and Automation Letters (RA-L). |
| June 2026 | Will serve as Action Editor for Transactions on Machine Learning Research (TMLR). |
| June 2026 | Four papers (Spatial Panoptic Captioning, JoVA, Physical Simulation, Category Discovery) are accepted to ECCV 2026. |
| Apr 2026 | Four papers (PartCo, Geometric Reciprocity, iVGR, ZeroBench) are accepted to ICML 2026. |
| Apr 2026 | ZeroBench is covered by Meta's Muse Spark. |
| Apr 2026 | Two papers (Infinite Ladder; CodeBind to Findings) are accepted to ACL 2026. |
| Mar 2026 | One paper (AI in Oral Health Surveillance) is accepted to Journal of Dental Research (JDR). |
| Feb 2026 | Three papers (Sculpt4D; Speed3R and Scene-Level Heterogeneous Physics to Findings) are accepted to CVPR 2026. |
| Dec 2025 | Invited to serve as Area Chair for ECCV 2026. |
| Nov 2025 | One paper (Deepfake Detection with Graph Neural Network) is accepted to KDD 2026. |
| Nov 2025 | Two papers (LooC, OnlineVPO) are accepted to WACV 2026. |
| Nov 2025 | One paper on semantic correspondence is accepted to TPAMI. |
| Sept 2025 | Seven papers (Panoptic Captioning, 3DRS, Fin3R, Wukong, VaMP, SEAL, GSPN-2) are accepted to NeurIPS 2025. |
| Sept 2025 | Invited talks at University of Cambridge and University of Birmingham in the UK. |
| Aug 2025 | Invited to serve as: Area Chair for CVPR 2026, Area Chair for ICLR 2026. |
| July 2025 | ZeroBench is covered by Google's Gemini 2.5 Report. |
| June 2025 | Two papers (Inpaint4Drag, GRAB) are accepted to ICCV 2025. |
| June 2025 | Talks @ CVPR 2025 in Nashville: Lightning talk at CVPR Area Chair Workshop; Keynote talks at CVPR Workshops on Fine-Grained Visual Categorization, Visual Anomaly and Novelty Detection, and Domain Generalization. |
| June 2025 | Invited to serve as Area Chair for AAAI 2026. |
| May 2025 | Two papers (GAMEBot, PruneVid) are accepted to ACL 2025. |
| Mar 2025 | Splat4D is accepted to SIGGRAPH 2025. |
| Feb 2025 | Six papers (ICE, HypCD, Mr. DETR, v-CLR, GSPN, PASS) are accepted to CVPR 2025. |
| Jan 2025 | Five papers (HiLo, DebGCD, BiGR, Needle Threading, AvatarGO) are accepted to ICLR 2025. |
| Sept 2024 | SciFIBench is accepted to NeurIPS 2024. |
| Sept 2024 | Invited to serve as: Area Chair for ICLR 2025, Area Chair for CVPR 2025. |
| Aug 2024 | Our Dissecting OOD and OSR paper is accepted to IJCV. |
| Jul 2024 | Three papers (RegionDrag, PromptCCD, and ConceptExpress) are accepted to ECCV 2024. |
| Feb 2024 | Three papers (IBD-SLAM, DreamAvatar, and SD4Match) are accepted to CVPR 2024. |
| Feb 2024 | CiPR is accepted to TMLR 2024. |
| Jan 2024 | Two papers (SPTNet and FROSTER) are accepted to ICLR 2024. |
| Oct 2023 | Invited to serve as an Area Chair for ECCV 2024. |
| Sept 2023 | One paper on text-guided 3D head avatar generation and editing is accepted to NeurIPS 2023. |
| Aug 2023 | One paper on visual correspondence is accepted to TPAMI. |
| July 2023 | Two papers (on generalized category discovery/open-vocabulary semantic segmentation) are accepted to ICCV 2023. |
| June 2023 | Invited to serve as an Area Chair for CVPR 2024. |
| Mar 2023 | OOD-CV workshop @ ICCV 2023. Welcome participants from all over! |
| Feb 2023 | Two papers (on compositional zero-shot learning/3D human digitization) are accepted to CVPR 2023. |
| Jul 2022 | One paper on novel category discovery without forgetting is accepted to ECCV 2022. |
| Jun 2022 | Best Paper Runner-Up Award at CVPR 2022 Workshop on Continual Learning in Computer Vision. |
| Mar 2022 | Three papers (about generalized category discovery/3D human digitization/instance segmentation) are accepted to CVPR 2022. |
| Jan 2022 | One paper about open-set recognition is accepted to ICLR 2022. |
| Oct 2021 | One paper about visual correspondence is accepted to BMVC 2021. |
| Sept 2021 | One paper about novel category discovery is accepted to NeurIPS 2021. |
| Sept 2021 | Recognized as an Outstanding Reviewer for ICCV 2021, in the top 5% of experienced reviewers. |
| July 2021 | One paper about single- and multi-modal novel category discovery is accepted to ICCV 2021. |
| June 2021 | Our AutoNovel paper is accepted to TPAMI. |
| June 2021 | Recognized as an Outstanding Reviewer for CVPR 2021. |
| May 2021 | One paper about dynamic convolution for semantic scene completion is accepted to TPAMI. |
| Mar 2021 | One paper about long-tailed recognition is accepted to CVPR 2021. |
| Dec 2020 | One paper about mirror surface reconstruction is accepted to IEEE TIP. |
| Sept 2020 | One paper about dense correspondence is accepted to NeurIPS 2020. |
| Jun 2020 | One paper about deep photometric stereo is accepted to TPAMI. |
| June 2020 | Recognized as an Outstanding Reviewer for CVPR 2020. |
| Feb 2020 | Two papers (about semantic correspondence / 3D semantic scene completion) are accepted to CVPR 2020. |
| Dec 2019 | One paper about novel category discovery is accepted to ICLR 2020. |
| Jul 2019 | One paper about novel category discovery is accepted to ICCV 2019. |
| Jul 2019 | One paper about transparent object matting is accepted to IJCV. |
| Mar 2019 | Two papers (about unsupervised object discovery and matching / uncalibrated photometric stereo) are accepted to CVPR 2019. |
![]() |
Beyond Single Expert: Harmonizing Diverse Visual Priors in MLLMs for Spatial Understanding Xiao Lin, Xiaohu Huang, Kai Han arXiv preprint arXiv:2607.15054, 2026. BibTeX PDF arXiv Project
|
![]() |
Surgical Post-Training: Proximal On-Policy Distillation for Reasoning with Knowledge Retention Wenye Lin, Kai Han arXiv preprint arXiv:2603.01683, 2026. BibTeX PDF arXiv
|
![]() |
JoVA: Unified Multimodal Learning for Joint Video-Audio Generation Xiaohu Huang, Hao Zhou, Qiangpeng Yang, Shilei Wen, Kai Han arXiv preprint arXiv:2512.13677, 2026. BibTeX PDF arXiv Project Code
|
![]() |
Generalized Category Discovery under Domain Shifts: From Vision to Vision-Language Models Hongjun Wang, Po Hu, Kai Han arXiv preprint arXiv:2605.00906, 2026. BibTeX PDF arXiv Project
|
![]() |
Effective Prompt Pool Learning for Continual Category Discovery Fernando Julio Cendra, Xinghui Li, Kai Han arXiv preprint arXiv:2407.19001, 2026. BibTeX PDF arXiv
|
![]() |
MHPO: Modulated Hazard-aware Policy Optimization for Stable Reinforcement Learning Hongjun Wang, Wei Liu, Weibo Gu, Xing Sun, Kai Han arXiv preprint arXiv:2603.16929, 2026. BibTeX PDF arXiv
|
![]() |
Category Discovery: An Open-World Perspective Zhenqi He, Yuanpei Liu, Kai Han arXiv preprint arXiv:2509.22542, 2026 BibTeX PDF arXiv Project
|
![]() |
PartCo: Part-Level Correspondence Priors Enhance Category Discovery Fernando Julio Cendra, Kai Han International Conference on Machine Learning (ICML), 2026. BibTeX PDF arXiv Project
|
![]() |
Geometric Reciprocity: Unlocking Self-Supervision for Stereoscopic Video Generation Jingyi Lu, Kai Han International Conference on Machine Learning (ICML), 2026. BibTeX PDF arXiv Project
|
![]() |
iVGR: Internalizing Visually Grounded Reasoning for MLLMs with Reinforcement Learning Chang-Bin Zhang, Yujie Zhong, Qiang Zhang, Kai Han International Conference on Machine Learning (ICML), 2026. BibTeX PDF arXiv Project Code
|
![]() |
ZeroBench: An Impossible Visual Benchmark for Contemporary Large Multimodal Models Jonathan Roberts, Mohammad Reza Taesiri, Ansh Sharma, Akash Gupta, et al., Kai Han†, Samuel Albanie† International Conference on Machine Learning (ICML), 2026. BibTeX PDF arXiv Project Code
|
![]() |
How Long Is a Piece of String? A Brief Empirical Analysis of Tokenizers Jonathan Roberts, Kai Han, Samuel Albanie ICML 2026 Workshop on Combining Theory and Benchmarks (CTB). BibTeX PDF arXiv
|
![]() |
Sculpt4D: Generating 4D Shapes via Sparse-Attention Diffusion Transformers Minghao Yin, Wenbo Hu, Jiale Xu, Ying Shan, Kai Han IEEE Conference on Computer Vision and Pattern Recognition (CVPR), 2026. BibTeX PDF arXiv Project Code
|
![]() |
Speed3R: Sparse Feed-forward 3D Reconstruction Models Weining Ren, Xiao Tan, Kai Han IEEE Conference on Computer Vision and Pattern Recognition (CVPR) Findings, 2026. BibTeX PDF arXiv Project Code
|
![]() |
Scene-Level Heterogeneous Physics Simulation with 3D Gaussian Splats Xiaoyang Liu, Shangzhe Wu, Kai Han IEEE Conference on Computer Vision and Pattern Recognition (CVPR) Findings, 2026. BibTeX PDF arXiv Project Code
|
![]() |
When Deepfake Detection Meets Graph Neural Network: a Unified and Lightweight Learning Framework Haoyu Liu, Chaoyu Gong, Mengke He, Jiate Li, Kai Han, Siqiang Luo ACM SIGKDD Conference on Knowledge Discovery and Data Mining (KDD), 2026. BibTeX PDF arXiv
|
![]() |
Ascending the Infinite Ladder: Benchmarking Spatial Deformation Reasoning in Vision-Language Models Jiahuan Zhang, Shunwen Bai, Tianheng Wang, Kaiwen Guo, Zijia Song, Hanqing Wu, Guozheng Rao, Kai Han, Kaicheng Yu Annual Meeting of the Association for Computational Linguistics (ACL), 2026. BibTeX PDF arXiv
|
![]() |
CodeBind: Decoupled Representation Learning for Multimodal Alignment with Unified Compositional Codebook Zeyu Chen, Jie Li, Kai Han Annual Meeting of the Association for Computational Linguistics (ACL) Findings, 2026. BibTeX PDF arXiv Project Code
|
![]() |
AI in Oral Health Surveillance: Critical Review Zeyu Chen, Pei Liu, Kai Han, Peixi Liao, Yanqi Yang, May C.M. Wong, Cynthia K.Y. Yiu, Edward C.M. Lo Journal of Dental Research (JDR), 2026. BibTeX DOI
|
![]() |
LooC: Effective Low-Dimensional Codebook for Compositional Vector Quantization Jie Li, Kwan-Yee K. Wong, Kai Han IEEE/CVF Winter Conference on Applications of Computer Vision (WACV), 2026. BibTeX PDF arXiv Project Code
|
![]() |
Align Video Diffusion Model with Online Video-Centric Preference Optimization Jiacheng Zhang, Jie Wu, Weifeng Chen, Yatai Ji, Weilin Huang, Xuefeng Xiao, Kai Han IEEE/CVF Winter Conference on Applications of Computer Vision (WACV), 2026. BibTeX PDF arXiv Project
|
![]() |
Semantic Correspondence: Unified Benchmarking and a Strong Baseline Kaiyan Zhang, Xinghui Li, Jingyi Lu, Kai Han IEEE Transactions on Pattern Analysis and Machine Intelligence (TPAMI), 2025 BibTeX PDF arXiv Project Code
|
![]() |
Panoptic Captioning: An Equivalence Bridge for Image and Text Kun-Yu Lin, Hongjun Wang, Weining Ren, Kai Han Conference on Neural Information Processing Systems (NeurIPS), 2025 BibTeX PDF arXiv Project Code
|
![]() |
3DRS: MLLMs Need 3D-Aware Representation Supervision for Scene Understanding Xiaohu Huang, Jingjing Wu, Qunyi Xie, Kai Han Conference on Neural Information Processing Systems (NeurIPS), 2025 BibTeX PDF arXiv Project Code
|
![]() |
Wukong's 72 Transformations: High-fidelity 3D Morphing via Flow Models Minghao Yin, Yukang Cao, Kai Han Conference on Neural Information Processing Systems (NeurIPS), 2025 BibTeX PDF arXiv Project
|
![]() |
VaMP: Variational Multi-Modal Prompt Learning for Vision-Language Models Silin Cheng, Kai Han Conference on Neural Information Processing Systems (NeurIPS), 2025 BibTeX PDF arXiv Project
|
![]() |
Fin3R: Fine-Tuning Feed-Foward 3D Reconstruction Models via Monocular Knowledge Distillation Weining Ren, Hongjun Wang, Xiao Tan, Kai Han Conference on Neural Information Processing Systems (NeurIPS), 2025 BibTeX PDF arXiv Project Code
|
![]() |
SEAL: Semantic-Aware Hierarchical Learning for Generalized Category Discovery Zhenqi He*, Yuanpei Liu*, Kai Han Conference on Neural Information Processing Systems (NeurIPS), 2025 BibTeX PDF arXiv Project Code
|
![]() |
GSPN-2: Efficient Parallel Sequence Modeling Hongjun Wang, Yitong Jiang, Collin McCarthy, David Wehr, Hanrong Ye, Xinhao Li, Ka Chun Cheung, Wonmin Byeon, Jinwei Gu, Ke Chen, Kai Han, Hongxu Yin, Pavlo Molchanov, Jan Kautz, Sifei Liu Conference on Neural Information Processing Systems (NeurIPS), 2025 BibTeX PDF arXiv
|
![]() |
ELIP: Enhanced Visual-Language Foundation Models for Image Retrieval Guanqi Zhan*, Yuanpei Liu*, Kai Han, Weidi Xie, Andrew Zisserman IEEE International Conference on Content-Based Multimedia Indexing (CBMI), 2025. BibTeX PDF arXiv Project Code
|
![]() |
Inpaint4Drag: Repurposing Inpainting Models for Drag-Based Image Editing via Bidirectional Warping Jingyi Lu, Kai Han International Conference on Computer Vision (ICCV), 2025. BibTeX PDF arXiv Project Code Demo
|
![]() |
GRAB: A Challenging GRaph Analysis Benchmark for Large Multimodal Models Jonathan Roberts, Kai Han, Samuel Albanie International Conference on Computer Vision (ICCV), 2025. BibTeX PDF arXiv Project Data Code
|
![]() |
GAMEBoT: Transparent Assessment of LLM Reasoning in Games Wenye Lin, Jonathan Roberts, Yunhan Yang, Samuel Albanie, Zongqing Lu, Kai Han Annual Meeting of the Association for Computational Linguistics (ACL), 2025. BibTeX PDF arXiv Project Code
|
![]() |
PruneVid: Visual Token Pruning for Efficient Video Large Language Models Xiaohu Huang, Hao Zhou, Kai Han Annual Meeting of the Association for Computational Linguistics (ACL) Findings, 2025. BibTeX PDF arXiv Project Code
|
![]() |
Splat4D: Diffusion-Enhanced 4D Gaussian Splatting for Temporally and Spatially Consistent Content Creation Minghao Yin, Yukang Cao, Songyou Peng, Kai Han SIGGRAPH, 2025. BibTeX PDF arXiv DOI Project
|
![]() |
ICE: Intrinsic Concept Extraction from a Single Image via Diffusion Models Fernando Julio Cendra, Kai Han IEEE Conference on Computer Vision and Pattern Recognition (CVPR), 2025. Highlight presentation · 2.8% of submissions
BibTeX PDF arXiv Project Code
|
![]() |
Hyperbolic Category Discovery Yuanpei Liu*, Zhenqi He*, Kai Han IEEE Conference on Computer Vision and Pattern Recognition (CVPR), 2025. BibTeX PDF arXiv Project Code
|
![]() |
Mr. DETR: Instructive Multi-Route Training for Detection Transformers Chang-Bin Zhang, Yujie Zhong, Kai Han IEEE Conference on Computer Vision and Pattern Recognition (CVPR), 2025. BibTeX PDF arXiv Project Code
|
![]() |
v-CLR: View-Consistent Learning for Open-World Instance Segmentation Chang-Bin Zhang, Jinhong Ni, Yujie Zhong, Kai Han IEEE Conference on Computer Vision and Pattern Recognition (CVPR), 2025. Highlight presentation · 2.8% of submissions
BibTeX PDF arXiv Project Code
|
![]() |
Parallel Sequence Modeling via Generalized Spatial Propagation Network Hongjun Wang, Wonmin Byeon, Jiarui Xu, Jinwei Gu, Ka Chun Cheung, Xiaolong Wang, Kai Han, Jan Kautz, Sifei Liu IEEE Conference on Computer Vision and Pattern Recognition (CVPR), 2025. BibTeX PDF arXiv Project Code
|
![]() |
Detecting Open World Objects via Partial Attribute Assignment Muli Yang, Gabriel James Goenawan, Huaiyuan Qin, Kai Han, Xi Peng, Yanhua Yang, Hongyuan Zhu IEEE Conference on Computer Vision and Pattern Recognition (CVPR), 2025. BibTeX PDF Code
|
![]() |
DebGCD: Debiased Learning with Distribution Guidance for Generalized Category Discovery Yuanpei Liu, Kai Han International Conference on Learning Representations (ICLR), 2025. BibTeX PDF arXiv Project Code
|
![]() |
HiLo: A Learning Framework for Generalized Category Discovery Robust to Domain Shifts Hongjun Wang, Sagar Vaze, Kai Han International Conference on Learning Representations (ICLR), 2025. BibTeX PDF arXiv Project Code
|
![]() |
Needle Threading: Can LLMs Follow Threads Through Near-Million-Scale Haystacks? Jonathan Roberts, Kai Han, Samuel Albanie International Conference on Learning Representations (ICLR), 2025. BibTeX PDF arXiv Project Code Dataset
|
![]() |
BiGR: Harnessing Binary Latent Codes for Image Generation and Improved Visual Representation Capabilities Shaozhe Hao, Xuantong Liu, Xianbiao Qi, Shihao Zhao, Bojia Zi, Rong Xiao, Kai Han, Kwan-Yee K. Wong International Conference on Learning Representations (ICLR), 2025. BibTeX PDF arXiv Project Code
|
|
AvatarGO: Zero-shot 4D Human-Object Interaction Generation and Animation Yukang Cao, Liang Pan, Kai Han, Kwan-Yee K. Wong, Ziwei Liu International Conference on Learning Representations (ICLR), 2025. BibTeX PDF arXiv Project Code
|
|
![]() |
VipDiff: Towards Coherent and Diverse Video Inpainting via Training-free Denoising Diffusion Models Chaohao Xie, Kai Han, Kwan-Yee K. Wong IEEE/CVF Winter Conference on Applications of Computer Vision (WACV), 2025. BibTeX PDF arXiv Project Code
|
![]() |
CusConcept: Customized Visual Concept Decomposition with Diffusion Models Zhi Xu, Shaozhe Hao, Kai Han IEEE/CVF Winter Conference on Applications of Computer Vision (WACV), 2025. BibTeX PDF arXiv Code
|
![]() |
SciFIBench: Benchmarking Large Multimodal Models for Scientific Figure Interpretation Jonathan Roberts, Kai Han, Neil Houlsby, Samuel Albanie Conference on Neural Information Processing Systems (NeurIPS), 2024. BibTeX PDF arXiv Data & Code
|
![]() |
Dissecting Out-of-Distribution Detection and Open-Set Recognition: A Critical Analysis of Methods and Benchmarks Hongjun Wang, Sagar Vaze, Kai Han International Journal of Computer Vision (IJCV), 2024. BibTeX PDF arXiv Project Code
|
![]() |
RegionDrag: Fast Region-Based Image Editing with Diffusion Models Jingyi Lu, Xinghui Li, Kai Han European Conference on Computer Vision (ECCV), 2024. BibTeX PDF arXiv Project Code
|
![]() |
PromptCCD: Learning Gaussian Mixture Prompt Pool for Continual Category Discovery Fernando Julio Cendra, Bingchen Zhao, Kai Han European Conference on Computer Vision (ECCV), 2024. BibTeX PDF arXiv Project Code
|
![]() |
ConceptExpress: Harnessing Diffusion Models for Single-image Unsupervised Concept Extraction Shaozhe Hao, Kai Han, Zhengyao Lv, Shihao Zhao, Kwan-Yee K. Wong European Conference on Computer Vision (ECCV), 2024. Oral presentation · 2.3% of submissions
BibTeX PDF arXiv Project Code
|
![]() |
IBD-SLAM: Learning Image-Based Depth Fusion for Generalizable SLAM Minghao Yin, Shangzhe Wu, Kai Han IEEE Conference on Computer Vision and Pattern Recognition (CVPR), 2024. BibTeX PDF Project
|
|
DreamAvatar: Text-and-Shape Guided 3D Human Avatar Generation via Diffusion Models Yukang Cao*, Yan-Pei Cao*, Kai Han, Ying Shan, Kwan-Yee K. Wong IEEE Conference on Computer Vision and Pattern Recognition (CVPR), 2024. BibTeX PDF arXiv Project Code
|
|
![]() |
SD4Match: Learning to Prompt Stable Diffusion Model for Semantic Matching Xinghui Li, Jingyi Lu, Kai Han, Victor Prisacariu IEEE Conference on Computer Vision and Pattern Recognition (CVPR), 2024. BibTeX PDF arXiv Project Code
|
![]() |
What’s in a Name? Beyond Class Indices for Image Recognition Kai Han*, Xiaohu Huang*, Yandong Li*, Sagar Vaze*, Jie Li, Xuhui Jia CVPR Workshop on Computer Vision in the Wild, 2024. BibTeX PDF arXiv Code
|
![]() |
Charting New Territories: Exploring the Geographic and Geospatial Capabilities of Multimodal LLMs Jonathan Roberts, Timo Lüddecke, Rehan Sheikh, Kai Han, Samuel Albanie CVPR Workshop on EarthVision, 2024. BibTeX PDF arXiv Dataset
|
![]() |
CiPR: An Efficient Framework with Cross-instance Positive Relations for Generalized Category Discovery Shaozhe Hao, Kai Han, Kwan-Yee K. Wong Transactions on Machine Learning Research (TMLR), 2024. BibTeX PDF arXiv Code
|
![]() |
SPTNet: An Efficient Alternative Framework for Generalized Category Discovery with Spatial Prompt Tuning Hongjun Wang, Sagar Vaze, Kai Han International Conference on Learning Representations (ICLR), 2024. BibTeX PDF arXiv Project Code
|
![]() |
FROSTER: Frozen CLIP is a Strong Teacher for Open-Vocabulary Action Recognition Xiaohu Huang, Hao Zhou, Kun Yao, Kai Han International Conference on Learning Representations (ICLR), 2024. BibTeX PDF arXiv Project Code
|
![]() ![]() |
HeadSculpt: Crafting 3D Head Avatars with Text Xiao Han*, Yukang Cao*, Kai Han, Xiatian Zhu, Jiankang Deng, Yi-Zhe Song, Tao Xiang, Kwan-Yee K. Wong Conference on Neural Information Processing Systems (NeurIPS), 2023. BibTeX PDF arXiv Project Code
|
![]() |
GPT4GEO: How a Language Model Sees the World's Geography Jonathan Roberts, Timo Lüddecke, Sowmen Das, Kai Han, Samuel Albanie NeurIPS Workshop on Foundation Models for Decision Making, 2023. BibTeX PDF arXiv Code
|
![]() |
DualRC: A Dual-Resolution Learning Framework with Neighbourhood Consensus for Visual Correspondences Xinghui Li, Kai Han, Shuda Li, Victor Prisacariu IEEE Transactions on Pattern Analysis and Machine Intelligence (TPAMI), 2023. BibTeX DOI Project Code
|
![]() |
Learning Semi-supervised Gaussian Mixture Models for Generalized Category Discovery Bingchen Zhao, Xin Wen, Kai Han International Conference on Computer Vision (ICCV), 2023. BibTeX PDF arXiv Code
|
![]() |
Open-Vocabulary Semantic Segmentation with Decoupled One-Pass Network Cong Han*, Yujie Zhong*, Dengjie Li, Kai Han, Lin Ma International Conference on Computer Vision (ICCV), 2023. BibTeX PDF arXiv Code
|
![]() |
SATIN: A Multi-Task Metadataset for Classifying Satellite Imagery using Vision-Language Models Jonathan Roberts, Kai Han, Samuel Albanie ICCV TNGCV Workshop, 2023. BibTeX PDF arXiv Project
|
![]() |
Learning Attention as Disentangler for Compositional Zero-shot Learning Shaozhe Hao, Kai Han, Kwan-Yee K. Wong IEEE Conference on Computer Vision and Pattern Recognition (CVPR), 2023. BibTeX PDF arXiv Project Code
|
![]() |
SeSDF: Self-evolved Signed Distance Field for Implicit 3D Clothed Human Reconstruction Yukang Cao, Kai Han, Kwan-Yee K. Wong IEEE Conference on Computer Vision and Pattern Recognition (CVPR), 2023. BibTeX PDF arXiv Project Code
|
![]() |
ViCo: Detail-Preserving Visual Condition for Personalized Text-to-Image Generation Shaozhe Hao, Kai Han, Shihao Zhao, Kwan-Yee K. Wong arXiv preprint, 2023. BibTeX PDF arXiv Code
|
![]() |
SimSC: A Simple Framework for Semantic Correspondence with Temperature Learning Xinghui Li, Kai Han, Xingchen Wan, Victor Adrian Prisacariu arXiv preprint, 2023. BibTeX PDF arXiv
|
![]() |
Novel Class Discovery without Forgetting K J Joseph, Sujoy Paul, Gaurav Aggarwal, Soma Biswas, Piyush Rai, Kai Han, Vineeth N Balasubramanian European Conference on Computer Vision (ECCV), 2022. BibTeX PDF arXiv
|
![]() |
Generalized Category Discovery Sagar Vaze, Kai Han, Andrea Vedaldi, Andrew Zisserman IEEE Conference on Computer Vision and Pattern Recognition (CVPR), 2022. BibTeX PDF arXiv Project Code
|
![]() |
JIFF: Jointly-aligned Implicit Face Function for High Quality Single View Clothed Human Reconstruction Yukang Cao, Guanying Chen, Kai Han, Wenqi Yang, Kwan-Yee K. Wong IEEE Conference on Computer Vision and Pattern Recognition (CVPR), 2022. Oral presentation · 4.2% of submissions
BibTeX PDF arXiv
Project Code
|
![]() |
SharpContour: A Contour-based Boundary Refinement Approach for Efficient and Accurate Instance Segmentation Chenming Zhu, Xuanye Zhang, Yanran Li, Liangdong Qiu, Kai Han, Xiaoguang Han IEEE Conference on Computer Vision and Pattern Recognition (CVPR), 2022. BibTeX PDF arXiv Project
|
![]() |
Spacing Loss for Discovering Novel Categories K J Joseph, Sujoy Paul, Gaurav Aggarwal, Soma Biswas, Piyush Rai, Kai Han, Vineeth N Balasubramanian CVPR Workshop on Continual Learning in Computer Vision, 2022. Best Paper Runner-Up Award BibTeX PDF arXiv
|
![]() |
Open-Set Recognition: A Good Closed-Set Classifier is All You Need? Sagar Vaze, Kai Han, Andrea Vedaldi, Andrew Zisserman International Conference on Learning Representations (ICLR), 2022. Oral presentation · 1.6% of submissions
BibTeX PDF arXiv OpenReview
Project Code
|
![]() |
Novel Visual Category Discovery with Dual Ranking Statistics and Mutual Knowledge Distillation Bingchen Zhao, Kai Han Conference on Neural Information Processing Systems (NeurIPS), 2021. BibTeX PDF arXiv OpenReview Supp Project Code
|
![]() |
Joint Representation Learning and Novel Category Discovery on Single- and Multi-modal Data Xuhui Jia, Kai Han, Yukun Zhu, Bradley Green International Conference on Computer Vision (ICCV), 2021. BibTeX PDF arXiv
|
![]() |
𝕏Resolution Correspondence Networks Georgi Tinchev, Shuda Li, Kai Han, David Mitchell, Rigas Kouskouridas British Machine Vision Conference (BMVC), 2021. BibTeX PDF arXiv Project Code
|
![]() |
LSD-C: Linearly Separable Deep Clusters Sylvestre-Alvise Rebuffi*, Sebastien Ehrhardt*, Kai Han*, Andrea Vedaldi, Andrew Zisserman ICCV Workshop on Visual Inductive Priors for Data-Efficient Deep Learning, 2021. (* indicates equal contribution.) BibTeX PDF arXiv Code
|
![]() |
Contrastive Learning based Hybrid Networks for Long-Tailed Image Classification Peng Wang, Kai Han, Xiu-Shen Wei, Lei Zhang, Lei Wang IEEE Conference on Computer Vision and Pattern Recognition (CVPR), 2021. BibTeX PDF arXiv
|
![]() |
AutoNovel: Automatically Discovering and Learning Novel Visual Categories Kai Han, Sylvestre-Alvise Rebuffi, Sebastien Ehrhardt, Andrea Vedaldi, Andrew Zisserman IEEE Transactions on Pattern Analysis and Machine Intelligence (TPAMI), 2021. BibTeX PDF arXiv DOI Project Code
|
![]() |
Anisotropic Convolutional Neural Networks for RGB-D based Semantic Scene Completion Jie Li, Peng Wang, Kai Han, Yu Liu IEEE Transactions on Pattern Analysis and Machine Intelligence (TPAMI), 2021. BibTeX PDF arXiv DOI Project Code
|
![]() |
Fixed Viewpoint Mirror Surface Reconstruction under an Uncalibrated Camera Kai Han, Miaomiao Liu, Dirk Schnieders, Kwan-Yee K. Wong IEEE Transactions on Image Processing (TIP), 2021. BibTeX PDF arXiv DOI Supp Project Code
|
![]() |
Dual-Resolution Correspondence Networks Xinghui Li, Kai Han, Shuda Li, Victor Prisacariu Conference on Neural Information Processing Systems (NeurIPS), 2020. BibTeX PDF arXiv Supp Project Code
|
![]() |
Automatically Discovering and Learning New Visual Categories with Ranking Statistics Kai Han*, Sylvestre-Alvise Rebuffi*, Sebastien Ehrhardt*, Andrea Vedaldi, Andrew Zisserman International Conference on Learning Representations (ICLR), 2020. (* indicates equal contribution.) BibTeX PDF arXiv OpenReview Video Project Code
|
![]() |
Correspondence Networks with Adaptive Neighbourhood Consensus Shuda Li*, Kai Han*, Theo W. Costain, Henry Howard-Jenkins, Victor Prisacariu IEEE Conference on Computer Vision and Pattern Recognition (CVPR), 2020. (* indicates equal contribution.) BibTeX PDF arXiv Project Code
|
![]() |
Anisotropic Convolutional Networks for 3D Semantic Scene Completion Jie Li, Kai Han, Peng Wang, Yu Liu, Xia Yuan IEEE Conference on Computer Vision and Pattern Recognition (CVPR), 2020. BibTeX PDF arXiv Project Code
|
![]() |
Semi-Supervised Learning with Scarce Annotations Sylvestre-Alvise Rebuffi*, Sebastien Ehrhardt*, Kai Han*, Andrea Vedaldi, Andrew Zisserman CVPR Deep Vision Workshop, 2020. (* indicates equal contribution.) BibTeX PDF arXiv Project Code
|
![]() |
Deep Photometric Stereo for Non-Lambertian Surfaces Guanying Chen, Kai Han, Boxin Shi, Yasuyuki Matsushita, Kwan-Yee K. Wong IEEE Transactions on Pattern Analysis and Machine Intelligence (TPAMI), 2020. BibTeX PDF arXiv DOI Supp Project Code
|
![]() |
Learning to Discover Novel Visual Categories via Deep Transfer Clustering Kai Han, Andrea Vedaldi, Andrew Zisserman International Conference on Computer Vision (ICCV), 2019. BibTeX PDF arXiv Supp Project Code
|
![]() |
Unsupervised Image Matching and Object Discovery as Optimization Huy V. Vo, Francis Bach, Minsu Cho, Kai Han, Yann LeCun, Patrick Pérez, Jean Ponce IEEE Conference on Computer Vision and Pattern Recognition (CVPR), 2019. BibTeX PDF arXiv Code
|
![]() |
Self-calibrating Deep Photometric Stereo Networks Guanying Chen, Kai Han, Boxin Shi, Yasuyuki Matsushita, Kwan-Yee K. Wong IEEE Conference on Computer Vision and Pattern Recognition (CVPR), 2019. Oral presentation · 5.6% of submissions
BibTeX PDF arXiv Video Poster Project Code
|
![]() |
Learning Transparent Object Matting Guanying Chen*, Kai Han*, Kwan-Yee K. Wong International Journal of Computer Vision (IJCV), 2019. (* indicates equal contribution.) BibTeX PDF arXiv DOI Project Code
|
![]() |
PS-FCN: A Flexible Learning Framework for Photometric Stereo Guanying Chen, Kai Han, Kwan-Yee K. Wong European Conference on Computer Vision (ECCV), 2018. BibTeX PDF arXiv Video Poster Project Code
|
![]() |
TOM-Net: Learning Transparent Object Matting from a Single Image Guanying Chen*, Kai Han*, Kwan-Yee K. Wong IEEE Conference on Computer Vision and Pattern Recognition (CVPR), 2018. (* indicates equal contribution.) Spotlight presentation · 6.7% of submissions
BibTeX PDF arXiv Video Poster Project Code
|
![]() |
Dense Reconstruction of Transparent Objects by Altering Incident Light Paths Through Refraction Kai Han, Kwan-Yee K. Wong, Miaomiao Liu International Journal of Computer Vision (IJCV), 2018. BibTeX PDF arXiv DOI
|
![]() |
SCNet: Learning Semantic Correspondence Kai Han, Rafael S. Rezende, Bumsub Ham, Kwan-Yee K. Wong, Minsu Cho, Cordelia Schmid, Jean Ponce International Conference on Computer Vision (ICCV), 2017. BibTeX PDF arXiv HAL Poster Project Code
|
![]() |
Single View 3D Reconstruction under an Uncalibrated Camera and an Unknown Mirror Sphere Kai Han, Kwan-Yee K. Wong, Xiao Tan International Conference on 3D Vision (3DV), 2016. BibTeX PDF Poster
|
![]() |
Mirror Surface Reconstruction Under an Uncalibrated Camera Kai Han, Kwan-Yee K. Wong, Dirk Schnieders, Miaomiao Liu IEEE Conference on Computer Vision and Pattern Recognition (CVPR), 2016. BibTeX PDF Project Code
|
![]() |
A Fixed Viewpoint Approach for Dense Reconstruction of Transparent Objects Kai Han, Kwan-Yee K. Wong, Miaomiao Liu IEEE Conference on Computer Vision and Pattern Recognition (CVPR), 2015. BibTeX PDF
|
![]() |
Single view reconstruction of transparent, mirror and diffuse surfaces Kai Han The University of Hong Kong, Pokfulam, Hong Kong, 2018. BibTeX HKU Theses Online
|
CVPR (2018–2023) · ICCV (2019, 2021, 2023) · ECCV (2020, 2022) · NeurIPS (2022–2024) · ICLR (2022–2024) · ICML (2023–2026) · ACL (2026) · SIGGRAPH Asia (2024–2025) · ICRA (2023) · AAAI (2020–2022) · ACCV (2020, 2022)
TPAMI · IJCV · TOG · JMLR · TIP · etc.