|
Huan Wang
My name is 𝘏𝘶𝘢𝘯 𝘞𝘢𝘯𝘨 (Chinese: 王欢). I work for (myself at) Westlake University as a tenure-track
Assistant Professor, leading the ENCODE
Lab.
I earned my Ph.D. degree (2024) at Northeastern University (Boston, USA), supervised by Prof. Yun (Raymond) Fu. Before that, I obtained my M.S. and
B.E. degrees from Zhejiang University (Hangzhou, China), advised by Prof. Haoji Hu. I also visited VLLab at UC Merced, luckily falling under the guidance of
Prof. Ming-Hsuan Yang. I was fortunate to intern
at
Google/Snap/MERL/Alibaba, working with many fantastic industrial researchers. I am a Senior Member
of the IEEE and serve as an Area
Chair for AAAI/ICLR/CVPR/NeurIPS 2026, AAAI/ICLR 2027.
Google Scholar ·
OpenReview · GitHub · LinkedIn · X · Rednote (小红书)
|
To create something beautiful.
|
|
|
Research Overview
My research spans two intertwined directions -- Efficient AI and AI Evaluation
-- bridging the implementation, algorithmic, and theoretical levels of intelligent systems (an
illustration about their relationship is shown in the figure below). For
the data and task domains, I focus on the vision-centered multimodal AI concerning its
understanding,
generation, and self-improvement.
My publications cover the following topics:
-
Model compression for various vision and language models:
Collaborative Distillation [CVPR'20], GReg [ICLR'21], SRP [ICLR'22], KD+DA
[NeurIPS'23], GASSL
[TPAMI'23] for CNNs; R2L
[ECCV'22] for NeRF; SnapFusion
[NeurIPS'23] for diffusion models; ROSE [CPAL'26]
for LLMs/MLLMs; SparseSSM [ICML'26]
for Mambas; etc.
-
Token and KV cache compression for LLMs/MLLMs:
DyCoke [CVPR'25], HoliTom [NeurIPS'25], StreamingTom [CVPR'26], EarlyTom [CVPR'26], RLKV [ICML'26], etc.
-
Dynamic computing for visual autoregressive models:
FreqExit [NeurIPS'25], LISA [ECCV'26], Prism-MoE [ICML'26], etc.
-
Mobile and edge AI:
MobileR2L [CVPR'23], SnapFusion [NeurIPS'23], MobileSCI [MM'24], LightDP [ICCV'25], EVAR [ECCV'26], etc.
-
Kernel generation for AI4AI:
ConCuR, MobileKernelBench, etc.
Here is my research statement. On top of my
research in the
technical sense, I am also leisurely interested in a few philosophical and
sociological problems concerning being and
postmodernism, esp. the critical theory of technology, with a simple idea in my mind: What we
spent so much time on at least should bring us a better life.
|
Openings / Collaborations
- We hire PhD students, research assistants, and visiting students. We do not have postdoc
positions now unfortunately.
- I am actively looking for PhD (Fall 2027)
and visiting students (targetting CVPR'27, ICML'27) for research-oriented projects.
- How to apply? Please check this admission.pdf
(this doc contains all the guidelines you need to apply).
- Sep 18, 2026: Our lab is recruiting new PhD and visiting students for a project `AI4AI:
Interactive World
Model`
along
with LIN's Lab, see
xhs
post.
|
|
Recent News
-
2026/09 [Services] I will serve as an Area Chair (AC) for
CVPR
2027.
-
2026/08 [ECCV'26] I will attend ECCV 2026 with my
students to present three papers. I
will also host a tutorial on
efficient MLLM
inference and speak at a workshop on efficient
visual generation. Both events are on Sep 8th. See you
in Malmö!
-
2026/08 [Services] I will serve as an Area Chair (AC) for
ICLR
2027.
-
2026/08 [TPAMI'26] One paper "Random Sparse Networks Training
with Sharpness-Aware Regularization" is accepted by Transactions on
Pattern Analysis and Machine Intelligence (TPAMI, IF=20.4). Congrats to Yue and Mingyuan!
-
2026/06 [TIP'26] One paper about ego-centric 3D scene
generation is accepted by Transactions on
Image Processing (TIP). Congrats to Zhenyu!
-
2026/06 [Services] I will serve as an Area Chair (AC) for
AAAI
2027.
-
2025/06 [Award-to-Students] 🎉Congrats to Junhan Zhu on receiving the SenseTime Scholarship
(商汤奖学金). Junhan, an undergrad research intern in my group, is the very first undergrad stuent of
Westlake University who published a top-tier AI conference paper (OBS-Diff) as the sole first author.
I was fortunate to advise him in the work.
-
2026/06 [ECCV'26] Three papers accepted by ECCV 2026,
about efficient and reliable AI (topics: speculative decoding, hallucination mitigation, and
mobile VAR).
Congrats to my students and postdoc (Kejia, Zefang, and Ying)!
-
2026/06 [CVPR'26] ENCODE Lab is attending CVPR 2026 in Denver. See you in altitude!
-
2026/05 [Recent works] Can LLMs write
efficient kernels for mobile devices? Check it in our MobileKernelBench!
-
2026/05 [ICML'26] 4 papers accepted by ICML 2026!
Congrats to my students and collaborators! See you in Seoul! 🎉
-
2026/02 [CVPR'26] 6 papers accepted by CVPR 2026!
Congrats to my students and collaborators! 🎉 - LinkedIn
post
-
2026/01 [ICLR'26] 4 papers accepted by ICLR 2026!
Congrats to my students and collaborators! 🎉 - LinkedIn
post
-
2026/01 [TMLR'26, Survey] Our systematic
review of the long-context MLLM token compression methods is accepted by TMLR! [arxiv] [paper-repo]
-
2026/01 [CPAL'26] Two papers on efficient LLMs/MLLMs via
sparsity and low-rank decomposition are accepted by CPAL'26.
Congrats to Mingluo (Fall'26 incoming PhD
student) and
Haolei (visiting student)! Codes will be released
soon.
-
2025/11 [Funds] Received a research fund from CAAI-Ant
Group (CAAI-蚂蚁科研基金).
Thank you CAAI and Ant Group!
-
2025/09 [Funds] Received a research fund from a Hangzhou
Municipal Talent Program (杭州市青年人才项目). Thanks!
-
2025/09 [NeurIPS'25] 4 papers accepted by NeurIPS 2025
in the field of efficient and reliable AI.
Congrats to my students and collaborators from CWRU! 🎉 The three papers from our
group:
-
HoliTom: As a top-performing video LLM
token compression method, HoliTom can maintain 99.1% performance while reducing the FLOPs to
only 6.9%. And, it's training-free! [arxiv]
[code] [webpage]
-
Poison as Cure:
Adversarial visual noise is always malicious to our models like "poison"? No, we find it can
also be a cure to mitigate the
hallucination problem of VLMs. [arxiv] [code] [webpage]
- FreqExit: FreqExit is a dynamic inference framework for visual autoregressive (VAR)
models via early exit equiped with a novel proposed frequency-aware guidance. [openreview] [code] [webpage]
-
2025/09 [Services] I will serve as an Area Chair for ICLR
2026 and CVPR 2026.
-
2025/07 [Preprint] We are excited to present the first
systematic
review of multimodal long-context token compression methods. [arxiv] [code]
-
2025/07 [MM'25] A paper about
efficient video diffusion
model via network pruning is accepted by MM'25. Congrats again to Yiming! Code will be released soon.
-
2025/06 [ICCV'25] A paper about efficient robot
manipulation is accepted by ICCV'25. Congrats to Yiming! Code will be released soon.
-
2025/06 [Award-to-Students] 🎉Congrats to my PhD
student Keda Tao on receiving the "2025 Westlake
University
Xinrui Award (西湖大学博士研究生新锐奖)" (only 2 recipients in AI among all the 2025 Fall PhD students
in School of Engineering).
-
2025/06 [Services] I will serve as an Area Chair for AAAI
2026.
-
2025/02 [CVPR'25] DyCoke is accepted by CVPR'25!
Congrats to my PhD student Keda!
DyCoke is a training-free, plug-and-play token compression method for fast video LLMs: 1.5x
wall-clock inference speedup and 1.4x memory reduction with no performance drop. [arxiv][code]
-
2025/02 [Preprint] Can diffusion models blend visual
concepts that are semantically very unsimilar (e.g., an orange and a
teddy bear)? Yes, we introduce FreeBlend, a new method to blend
arbitrary concepts. [arxiv] [code] [webpage]
-
2025/01 [ICLR'25] One paper about distilling large
foundation models with low cost "Compressing Vision Foundation Models at ImageNet-level
Costs" is accepted by ICLR'25.
Thanks to the lead author Yitian!
-
2024/12 [Preprint] We present empirical evidence
to show that oracle pruning, the "ground-truth" pruning paradigm that has been followed for
around 35 years in the pruning community, does not hold in practice.
[arxiv][webpage]
-
2024/09 [NeurIPS'24] We introduce a training framework
Scala to learn slimmable ViTs. Using Scala, a ViT model is trained once but can inference
at different widths, up to the need of devices with different resources. The project is led by
Yitian. Congrats!
-
2024/07 [MM'24] We present the first real-time
on-device video SCI (Snapshot Compressive Imaging) framework via dedicated network design
and a distillation-based training strategy. Congrats to Miao!
-
2024/07 [ECCV'24] One paper about efficient video SCI
(Snapshot Compressive Imaging) via network quantization is accepted by ECCV'24 as an
oral. Congrats to Miao!
[code]
-
2024/06 [New Start] Join the beautiful Westlake
University as a tenure-track Assistant Professor.
-
2024/04 [Graduation] 🎈PhD student→PhD. So many
thanks to my Ph.D. committee (Prof. Yun Raymond
Fu, Prof. Octavia
Camps, Prof.
Zhiqiang Tao) and co-authors! I am grateful to my family and
friends for their support and encouragement during the PhD journey.
|
|
Contact
- Work: 𝘸𝘢𝘯𝘨𝘩𝘶𝘢𝘯@𝘸𝘦𝘴𝘵𝘭𝘢𝘬𝘦.𝘦𝘥𝘶.𝘤𝘯
- Review: 𝘩𝘶𝘢𝘯.𝘸𝘢𝘯𝘨.𝘤𝘰𝘰𝘭@𝘨𝘮𝘢𝘪𝘭.𝘤𝘰𝘮
|
(This webpage template is stolen from Jon
Barron, and it chooses not go gentle into that good night.)
|
|