|
I am a second-year Ph.D. student in Computer Science at College of Computer Science and Technology, Zhejiang University, advised by Prof.
Bohan Zhuang. My research interests lie
in the fields of efficient large 3D model/world model.
I am currently a research intern at ByteDance Seed.
Email  /  Google Scholar  /  Github /  X /  LinkedIn /  CV |
|
|
|
Weijie Wang*, Haoyu Zhao*, Yifan Yang, Feng Chen, Zeyu Zhang, Yefei He, Zicheng Duan, Donny Y. Chen, Yuqing Yang, Bohan Zhuang† Latent Spatial Memory stores persistent 3D scene content directly as latent tokens for efficient, spatially consistent video world models. |
|
Weijie Wang*, Zimu Li*, Jinchuan Shi, Zeyu Zhang, Botao Ye, Marc Pollefeys, Donny Y. Chen, Bohan Zhuang† TriSplat predicts simulation-ready triangle primitives for feed-forward sparse-view 3D scene reconstruction. |
|
Weijie Wang*, Xiaoxuan He*, Youping Gu*, Yifan Yang†, Zeyu Zhang, Yefei He, Yanbo Ding, Xirui Hu, Donny Y. Chen, Zhiyuan He, Yuqing Yang†, Bohan Zhuang† ICML 2026
World-R1 aligns text-to-video generation with 3D constraints through reinforcement learning, improving geometric consistency while preserving visual quality and motion diversity. |
|
Weijie Wang*, Qihang Cao*, Sensen Gao*, Donny Y. Chen, Haofei Xu, Wenjing Bian, Songyou Peng, Tat-Jen Cham, Chuanxia Zheng, Andreas Geiger, Jianfei Cai, Jia-Wang Bian†, Bohan Zhuang† We present a problem-driven survey of feed-forward 3D scene modeling, covering representative methods, datasets, and downstream applications. |
|
Haoyu Zhao*, Akide Liu*, Zeyu Zhang*, Weijie Wang*, Feng Chen, Ruihan Zhu, Gholamreza Haffari, Bohan Zhuang† ACL 2026 Findings
We propose Chain-of-View (CoV) prompting, a training-free test-time reasoning framework that transforms VLMs into active viewpoint reasoners for spatial reasoning in 3D environments. |
|
Weijie Wang*, Jiagang Zhu*, Zeyu Zhang, Xiaofeng Wang, Zheng Zhu†, Guosheng Zhao, Chaojun Ni, Haoxiao Wang, Guan Huang, Xinze Chen, Yukun Zhou, Wenkang Qin, Duochao Shi, Haoyun Li, Yicheng Xiao, Donny Y. Chen, Jiwen Lu ICME 2026 Oral (Top 3%)
We present DriveGen3D, a novel framework for generating high-quality and highly controllable dynamic 3D driving scenes that addresses critical limitations in existing methodologies. |
|
Weijie Wang*, Yeqing Chen*, Zeyu Zhang, Hengyu Liu, Haoxiao Wang, Zhiyuan Feng, Wenkang Qin, Jia-Wang Bian, Zheng Zhu†, Donny Y. Chen, Bohan Zhuang† ECCV 2026
VolSplat improves multi-view consistency and geometric accuracy for feed-forward 3DGS with voxel-aligned prediction. |
|
Duochao Shi*, Weijie Wang*, Donny Y. Chen, Zeyu Zhang, Jia-Wang Bian, Bohan Zhuang† 3DV 2026
We introduce PM-Loss, a novel regularization loss based on a learned point map for feed-forward 3DGS, leading to smoother 3D geometry and better rendering. |
|
Weijie Wang†, Donny Y. Chen†, Zeyu Zhang, Duochao Shi, Akide Liu, Bohan Zhuang NeurIPS 2025
ZPressor is an architecture-agnostic module that compresses multi-view inputs for scalable feed-forward 3DGS. |
|
Chaojun Ni*, Xiaofeng Wang*, Zheng Zhu*†, Weijie Wang*, Haoyun Li, Guosheng Zhao, Jie Li, Wenkang Qin, Guan Huang, Wenjun Mei† ICCV 2025
We introduce WonderTurbo, the first real-time interactive 3D scene generation framework capable of generating novel perspectives of 3D scenes within 0.72 seconds. |
|
Xiaoxuan He*, Siming Fu*, Zeyue Xue*, Weijie Wang, Ruizhe He, Yuming Li, Dacheng Yin, Shuai Dong, Haoyang Huang, Hongfa Wang, Nan DUAN, Bohan Zhuang† ICML 2026
We introduce Flash-GRPO, a one-step policy optimization framework that improves video diffusion alignment efficiency with iso-temporal grouping and temporal gradient rectification. |
|
Haoxiao Wang*, Antao Xiang*, Haiyang Sun*, Peilin Sun, Changhao Pan, Yifu Chen, Minjie Hong, Weijie Wang, Shuang Chen, Yue Chen, Zhou Zhao ECCV 2026
We introduce DiGSeg, a diffusion-based generalist segmentation framework that repurposes pretrained diffusion models for text-conditioned semantic and open-vocabulary segmentation. |
|
Zhiyuan Feng*, Zhaolu Kang*, Qijie Wang, Zhiying Du, Jiongrui Yan, Shubin Shi, Chengbo Yuan, Huizhi Liang, Yu Deng, Qixiu Li, Rushuai Yang, Arctanx An, Leqi Zheng, Weijie Wang, Shawn Chen, Sicheng Xu, Yaobo Liang, Jiaolong Yang†, Baining Guo ICLR 2026
We introduce MV-RoboBench, a benchmark for evaluating multi-view spatial reasoning in robotic manipulation, revealing large gaps between state-of-the-art VLMs and human performance. |
|
Shuang Chen*, Yue Guo*, Zhaochen Su, Yafu Li, Yulun Wu, Jiacheng Chen, Jiayu Chen, Weijie Wang, Xiaoye Qu†, Yu Cheng† ICLR 2026
We introduce ReVisual-R1, achieving a new state-of-the-art among open-source 7B MLLMs on challenging benchmarks. |
|
Haoxiao Wang*, Kaichen Zhou*, Binrui Gu, Zhiyuan Feng, Weijie Wang, Peilin Sun, Yicheng Xiao, Jianhua Zhang, Hao Dong† ICRA 2025
We propose a single-view RGB-D-based depth completion framework, TransDiff, that leverages the Denoising Diffusion Probabilistic Models(DDPM) to achieve material-agnostic object grasping in desktop. |
|
Yunzhi Yan, Haotong Lin, Chenxu Zhou, Weijie Wang, Haiyang Sun, Kun Zhan, Xianpeng Lang, Xiaowei Zhou, Sida Peng† ECCV 2024
This paper aims to tackle the problem of modeling dynamic urban street scenes from monocular videos. We introduce Street Gaussians, a new explicit scene representation that tackles some major limitations. |
|
ByteDance
2026.01 - Present Research Intern at ByteDance Seed Advisors: Jianfeng Zhang and Qianyi Wu |
|
![]() |
Zhejiang University
2025.09 - Present Ph.D. Student in Computer Science and Technology ZIP Lab, State Key Lab of CAD&CG, College of Computer Science and Technology Advisor: Prof. Bohan Zhuang |
|
Microsoft
2025.07 - 2026.01 Research Intern at Microsoft Research Asia Advisors: Yuqing Yang, Yifan Yang and Zhiyuan He |
|
![]() |
Zhejiang University
2021.09 - 2025.06 Undergraduate B.E. with Honors in Software Engineering from College of Computer Science and Technology Advanced Class of Engineering Education, Chu Kochen Honors College Research Intern at ZJU3DV, advised by Prof. Xiaowei Zhou and Prof. Sida Peng Research Intern at HICAI-ZJU, advised by Prof. Keyan Ding |
|
|
|
|
|
|
|
This template is a modification to Jon Barron's website. |
