Huan Wang

My name is ๐˜๐˜ถ๐˜ข๐˜ฏ ๐˜ž๐˜ข๐˜ฏ๐˜จ (Chinese: ็Ž‹ๆฌข). I work for (myself at) Westlake University as a tenure-track Assistant Professor, leading the ENCODE Lab.

I earned my Ph.D. degree (2024) at Northeastern University (Boston, USA), supervised by Prof. Yun (Raymond) Fu. Before that, I obtained my M.S. and B.E. degrees from Zhejiang University (Hangzhou, China), advised by Prof. Haoji Hu. I also visited VLLab at UC Merced, luckily falling under the guidance of Prof. Ming-Hsuan Yang. I was fortunate to intern at Google/Snap/MERL/Alibaba, working with many fantastic industrial researchers. I am a Senior Member of the IEEE and serve as an Area Chair for AAAI/ICLR/CVPR/NeurIPS 2026, AAAI/ICLR 2027.

My research interests revolve around Efficient AI (at the implementation level of AI; ๐˜™๐˜˜: how to build effective and efficient intelligent systems) and AI Evaluation (at the theoretical and modeling level of AI; ๐˜™๐˜˜: what is intelligence, how to measure and model it), spanning many tasks in multimodal AI and generative AI, particularly in the field of Computer Vision. The long-term goal of my research is to understand (human) intelligence and develop associated techniques for human well-being. Here is my research statement.

On top of my research in the technical sense, I am also interested in a few philosophical and sociological problems concerning being and postmodernism, with a focus on the critical theory of technology.

profile photo

To create something beautiful.

Openings / Collaborations
  • We hire PhD students, research assistants, and visiting students. We do not have postdoc positions now unfortunately.
  • I am actively looking for PhD (Fall 2027) and visiting students (targetting CVPR'27, ICML'27) for research-oriented projects.
  • How to apply? Please check this admission.pdf (this doc contains all the guidelines you need to apply).
Recent News
Filter news
  • 2026/08 [ECCV'26] I will attend ECCV 2026 with my students to present three papers. I will also host a tutorial on efficient MLLM inference and speak at a workshop on efficient visual generation. Both events are on Sep 8th. See you in Malmรถ!
  • 2026/08 [Services] I will serve as an Area Chair (AC) for ICLR 2027.
  • 2026/08 [TPAMI'26] One paper "Random Sparse Networks Training with Sharpness-Aware Regularization" is accepted by Transactions on Pattern Analysis and Machine Intelligence (TPAMI, IF=20.4). Congrats to Yue and Mingyuan!
  • 2026/06 [TIP'26] One paper about ego-centric 3D scene generation is accepted by Transactions on Image Processing (TIP). Congrats to Zhenyu!
  • 2026/06 [Services] I will serve as an Area Chair (AC) for AAAI 2027.
  • 2025/06 [Award-to-Students] ๐ŸŽ‰Congrats to Junhan Zhu on receiving the SenseTime Scholarship (ๅ•†ๆฑคๅฅ–ๅญฆ้‡‘). Junhan, an undergrad research intern in my group, is the very first undergrad stuent of Westlake University who published a top-tier AI conference paper (OBS-Diff) as the sole first author. I was fortunate to advise him in the work.
  • 2026/06 [ECCV'26] Three papers accepted by ECCV 2026, about efficient and reliable AI (topics: speculative decoding, hallucination mitigation, and mobile VAR). Congrats to my students and postdoc (Kejia, Zefang, and Ying)!
  • 2026/06 [CVPR'26] ENCODE Lab is attending CVPR 2026 in Denver. See you in altitude!
  • 2026/05 [Recent works] Can LLMs write efficient kernels for mobile devices? Check it in our MobileKernelBench!
  • 2026/05 [ICML'26] 4 papers accepted by ICML 2026! Congrats to my students and collaborators! See you in Seoul! ๐ŸŽ‰
  • 2026/02 [CVPR'26] 6 papers accepted by CVPR 2026! Congrats to my students and collaborators! ๐ŸŽ‰ - LinkedIn post
  • 2026/01 [ICLR'26] 4 papers accepted by ICLR 2026! Congrats to my students and collaborators! ๐ŸŽ‰ - LinkedIn post
  • 2026/01 [TMLR'26, Survey] Our systematic review of the long-context MLLM token compression methods is accepted by TMLR! [arxiv] [paper-repo]
  • 2026/01 [CPAL'26] Two papers on efficient LLMs/MLLMs via sparsity and low-rank decomposition are accepted by CPAL'26. Congrats to Mingluo (Fall'26 incoming PhD student) and Haolei (visiting student)! Codes will be released soon.
  • 2025/11 [Funds] Received a research fund from CAAI-Ant Group (CAAI-่š‚่š็ง‘็ ”ๅŸบ้‡‘). Thank you CAAI and Ant Group!
  • 2025/09 [Funds] Received a research fund from a Hangzhou Municipal Talent Program (ๆญๅทžๅธ‚้’ๅนดไบบๆ‰้กน็›ฎ). Thanks!
  • 2025/09 [NeurIPS'25] 4 papers accepted by NeurIPS 2025 in the field of efficient and reliable AI. Congrats to my students and collaborators from CWRU! ๐ŸŽ‰ The three papers from our group:
    • HoliTom: As a top-performing video LLM token compression method, HoliTom can maintain 99.1% performance while reducing the FLOPs to only 6.9%. And, it's training-free! [arxiv] [code] [webpage]
    • Poison as Cure: Adversarial visual noise is always malicious to our models like "poison"? No, we find it can also be a cure to mitigate the hallucination problem of VLMs. [arxiv] [code] [webpage]
    • FreqExit: FreqExit is a dynamic inference framework for visual autoregressive (VAR) models via early exit equiped with a novel proposed frequency-aware guidance. [openreview] [code] [webpage]
  • 2025/09 [Services] I will serve as an Area Chair for ICLR 2026 and CVPR 2026.
  • 2025/07 [Preprint] We are excited to present the first systematic review of multimodal long-context token compression methods. [arxiv] [code]
  • 2025/07 [MM'25] A paper about efficient video diffusion model via network pruning is accepted by MM'25. Congrats again to Yiming! Code will be released soon.
  • 2025/06 [ICCV'25] A paper about efficient robot manipulation is accepted by ICCV'25. Congrats to Yiming! Code will be released soon.
  • 2025/06 [Award-to-Students] ๐ŸŽ‰Congrats to my PhD student Keda Tao on receiving the "2025 Westlake University Xinrui Award (่ฅฟๆน–ๅคงๅญฆๅšๅฃซ็ ”็ฉถ็”Ÿๆ–ฐ้”ๅฅ–)" (only 2 recipients in AI among all the 2025 Fall PhD students in School of Engineering).
  • 2025/06 [Services] I will serve as an Area Chair for AAAI 2026.
  • 2025/02 [CVPR'25] DyCoke is accepted by CVPR'25! Congrats to my PhD student Keda! DyCoke is a training-free, plug-and-play token compression method for fast video LLMs: 1.5x wall-clock inference speedup and 1.4x memory reduction with no performance drop. [arxiv][code]
  • 2025/02 [Preprint] Can diffusion models blend visual concepts that are semantically very unsimilar (e.g., an orange and a teddy bear)? Yes, we introduce FreeBlend, a new method to blend arbitrary concepts. [arxiv] [code] [webpage]
  • 2025/01 [ICLR'25] One paper about distilling large foundation models with low cost "Compressing Vision Foundation Models at ImageNet-level Costs" is accepted by ICLR'25. Thanks to the lead author Yitian!
  • 2024/12 [Preprint] We present empirical evidence to show that oracle pruning, the "ground-truth" pruning paradigm that has been followed for around 35 years in the pruning community, does not hold in practice. [arxiv][webpage]
  • 2024/09 [NeurIPS'24] We introduce a training framework Scala to learn slimmable ViTs. Using Scala, a ViT model is trained once but can inference at different widths, up to the need of devices with different resources. The project is led by Yitian. Congrats!
  • 2024/07 [MM'24] We present the first real-time on-device video SCI (Snapshot Compressive Imaging) framework via dedicated network design and a distillation-based training strategy. Congrats to Miao!
  • 2024/07 [ECCV'24] One paper about efficient video SCI (Snapshot Compressive Imaging) via network quantization is accepted by ECCV'24 as an oral. Congrats to Miao! [code]
  • 2024/06 [New Start] Join the beautiful Westlake University as a tenure-track Assistant Professor.
  • 2024/04 [Graduation] ๐ŸŽˆPhD studentโ†’PhD. So many thanks to my Ph.D. committee (Prof. Yun Raymond Fu, Prof. Octavia Camps, Prof. Zhiqiang Tao) and co-authors! I am grateful to my family and friends for their support and encouragement during the PhD journey.
Contact
  • Work: ๐˜ธ๐˜ข๐˜ฏ๐˜จ๐˜ฉ๐˜ถ๐˜ข๐˜ฏ@๐˜ธ๐˜ฆ๐˜ด๐˜ต๐˜ญ๐˜ข๐˜ฌ๐˜ฆ.๐˜ฆ๐˜ฅ๐˜ถ.๐˜ค๐˜ฏ
  • Review: ๐˜ฉ๐˜ถ๐˜ข๐˜ฏ.๐˜ธ๐˜ข๐˜ฏ๐˜จ.๐˜ค๐˜ฐ๐˜ฐ๐˜ญ@๐˜จ๐˜ฎ๐˜ข๐˜ช๐˜ญ.๐˜ค๐˜ฐ๐˜ฎ

(This webpage template is stolen from Jon Barron, and it chooses not go gentle into that good night.)