I am a Research Scientist at Adobe Research, where I lead multiple research-to-production projects in generative image and video editing. My work spans unified editing frameworks (EditVerse, ICLR'26 Oral), generative video propagation (GenProp, Adobe MAX Sneak 2025), and physically-grounded shadow understanding (MetaShadow, shipped in Lightroom & Photoshop). I am equally driven by fundamental questions โ how should we train visual generative models from scratch, and can editing and generation be unified into a single paradigm? These threads are converging on my broader vision: building world-aware visual foundation models that inherently understand the physical constraints of our world. I received my Ph.D. from The Chinese University of Hong Kong, advised by Prof. Chi-Wing Fu. When I'm not pushing pixels through neural networks, I'm exposing them on film โ a slower way of seeing that keeps me close to the beauty I'm trying to teach machines to understand.
Publications
Click a paper to see details
Selected Works
Chimera: Designing and Chinchilla-Scaling Hybrid Visual Diffusion Transformers
Summary & resources
A hybrid visual diffusion backbone that replaces most full attention with linear attention, paired with Chinchilla-style scaling laws that make its compute-optimal training recipe predictable.
paperFull details & citationChongjian Ge*ยงโ , Hanwen Jiang*ยง, Tianyu Wang*ยง, Jiuxiang Guยง, Yiran Xuยง, Ziwen Chenยง, Shaoteng Liu, Jing Shi, Yicong Hong, Zefan Cai, Hailin Jin, Hao Tanโ
EditVerse: Unifying Image and Video Editing and Generation with In-Context Learning
Summary & resources
A unified framework for image and video generation and editing that leverages self-attention for in-context learning and cross-modal knowledge transfer.
paperprojectFull details & citationXuan Ju, Tianyu Wang, Yuqian Zhou, He Zhang, Qing Liu, Nanxuan Zhao, Zhifei Zhang, Yijun Li, Yuanhao Cai, Shaoteng Liu, Daniil Pakhomov, Zhe Lin, Soo Ye Kimโ , Qiang Xuโ
Self-Evaluation Unlocks Any-Step Text-to-Image Generation
Summary & resources
A from-scratch training approach for text-to-image generation that supports any-step inference using self-evaluation mechanism.
paperprojectFull details & citationXin Yu, Xiaojuan Qiโ , Zhengqi Li, Kai Zhang, Richard Zhang, Zhe Lin, Eli Shechtman, Tianyu Wangโ , Yotam Nitzanโ
LightMover: Towards Precise and Efficient Control for Light Movement
Summary & resources
A framework for precise and efficient control of light movement in images, enabling physically plausible relighting and shadow manipulation.
Full details & citationGengze Zhou, Tianyu Wang, Soo Ye Kim, Zhixin Shu, Xin Yu, Yannick Hold-Geoffroy, Sumit Chaturvedi, Qi Wu, Zhe Lin, Scott Cohen
MetaShadow: Object-Centered Shadow Detection, Removal, and Synthesis
Summary & resources
A three-in-one framework for object-centered shadow detection, removal, and controllable synthesis in natural images.
paperpreprintFull details & citationTianyu Wang, Jianming Zhang, Haitian Zheng, Zhihong Ding, Scott Cohen, Zhe Lin, Wei Xiong, Chi-Wing Fu, Luis Figueroaโ , Soo Ye Kimโ
Generative Video Propagation
Summary & resources
A unified framework that propagates edits from the first frame to all following frames using video generation models.
paperprojectpreprintFull details & citationShaoteng Liu, Tianyu Wang, Jui-Hsien Wang, Qing Liu, Zhifei Zhang, Joon-Young Lee, Yijun Li, Bei Yu, Zhe Lin, Soo Ye Kimโ , Jiaya Jiaโ
ObjectMover: Generative Object Movement with Video Prior
Summary & resources
A generative model for moving objects in images with accurate pose adjustment, lighting re-harmonization, and shadow/reflection synthesis.
paperprojectvideoFull details & citationXin Yu, Tianyu Wang, Soo Ye Kim, Paul Guerrero, Xi Chen, Qing Liu, Zhe Lin, Xiaojuan Qiโ
More Publications (16)
RetouchIQ: MLLM Agents for Instruction-Based Image Retouching with Generalist Reward
Summary & resources
An MLLM-based agent framework for instruction-driven image retouching with a generalist reward model.
Full details & citationQiucheng Wu, Jing Shi, Simon Jenni, Kushal Kafle, Tianyu Wang, Shiyu Chang, Handong Zhao
OmniVCus: Feedforward Subject-driven Video Customization with Multimodal Control Conditions
Summary & resources
A diffusion Transformer framework for multi-subject video customization with multimodal control conditions, featuring a data construction pipeline for training without labels.
paperprojectcodeFull details & citationYuanhao Cai, He Zhang, Xi Chen, Jinbo Xing, Yiwei Hu, Yuqian Zhou, Kai Zhang, Zhifei Zhang, Soo Ye Kim, Tianyu Wang, Yulun Zhang, Xiaokang Yang, Zhe Lin, Alan Yuille
Unveiling Deep Shadows: A Survey and Benchmark on Image and Video Shadow Detection, Removal, and Generation in the Deep Learning Era
Summary & resources
A comprehensive survey and benchmark covering the past decade of shadow detection, removal, and generation research in deep learning.
papercodeFull details & citationXiaowei Hu, Zhenghao Xing, Tianyu Wang, Chi-Wing Fu, Pheng-Ann Heng
Video Instance Shadow Detection Under the Sun and Sky
Summary & resources
A semi-supervised framework for video instance shadow detection using contrastive learning and cycle consistency.
paperFull details & citationZhenghao Xing, Tianyu Wang, Xiaowei Hu, Haoran Wu, Chi-Wing Fu, Pheng-Ann Heng
Learning Weather-General and Weather-Specific Features for Image Restoration Under Multiple Adverse Weather Conditions
Summary & resources
A unified framework for image restoration that learns both weather-general and weather-specific features to handle multiple adverse conditions.
paperFull details & citationYi Zhu, Tianyu Wang, Xueyang Fuโฏ, Xin Yang, Xuejin Guo, Jiabin Dai, Yu Qiao, Xiaowei Huโฏ
H2ONet: Hand-Occlusion-and-Orientation-aware Network for Real-time 3D Hand Mesh Reconstruction
Summary & resources
A network for real-time 3D hand mesh reconstruction that handles hand occlusion and orientation challenges.
papervideocodeFull details & citationHao Xu, Tianyu Wang, Xiao Tang, Chi-Wing Fu
SILT: Shadow-aware Iterative Label Tuning for Learning to Detect Shadows from Noisy Labels
Summary & resources
A method for learning shadow detection from noisy labels using iterative label tuning with shadow-aware mechanisms.
paperFull details & citationHan Yang*, Tianyu Wang*, Xiaowei Hu, Chi-Wing Fu
Instance Shadow Detection with a Single-Stage Detector
Summary & resources
Journal extension of CVPR 2021 work with improved single-stage detector for instance shadow detection and shadow-object association.
papercodeFull details & citationTianyu Wang, Xiaowei Hu, Pheng-Ann Heng, Chi-Wing Fu
Sparse2Dense: Learning to Densify 3D Features for 3D Object Detection
Summary & resources
A framework to boost 3D detection by learning to densify point clouds in latent space, improving detection of small and distant objects.
papercodeFull details & citationTianyu Wang, Xiaowei Hu, Zhengzhe Liu, Chi-Wing Fu
Single-Stage Instance Shadow Detection With Bidirectional Relation Learning
Summary & resources
A single-stage approach for detecting instance-level shadows using bidirectional relation learning between objects and their shadows.
papervideocodeFull details & citationTianyu Wang, Xiaowei Huโ , Chi-Wing Fu, Pheng-Ann Heng
Revisiting Shadow Detection: A New Benchmark Dataset for Complex World
Summary & resources
A comprehensive benchmark dataset for shadow detection in complex real-world scenarios.
paperdatasetcodeFull details & citationXiaowei Hu, Tianyu Wang, Chi-Wing Fu, Yitong Jiang, Qiong Wang, Pheng-Ann Heng
Instance Shadow Detection
Summary & resources
First work to tackle instance-level shadow detection, associating each shadow with its corresponding object instance.
paperdatasetcodeFull details & citationTianyu Wang, Xiaowei Hu, Qiong Wang, Pheng-Ann Heng, Chi-Wing Fu
Single-Image Real-Time Rain Removal Based on Depth-Guided Non-Local Features
Summary & resources
Real-time rain removal using depth-guided non-local features for robust image restoration.
paperdatasetcodeFull details & citationXiaowei Hu, Lei Zhu, Tianyu Wang, Chi-Wing Fu, Pheng-Ann Heng
SAC-Net: Spatial Attenuation Context for Salient Object Detection
Summary & resources
Spatial attention mechanism for salient object detection using context attenuation.
paperresultscodeFull details & citationXiaowei Hu, Chi-Wing Fu, Lei Zhu, Tianyu Wang, Pheng-Ann Heng
Spatial Attentive Single-Image Deraining with a High Quality Real Rain Dataset
Summary & resources
Spatial attention network for single-image deraining with a new high-quality real-world rain dataset.
paperdatasetdataset_baiducodeFull details & citationTianyu Wang, Xin Yang, Ke Xu, Shaozhe Chen, Qiang Zhang, Rynson W.H. Lauโ
Towards Accurate Alignment in Real-time 3D Hand-Mesh Reconstruction
Summary & resources
Real-time 3D hand mesh reconstruction with improved alignment accuracy.
papercodeFull details & citationXiao Tang, Tianyu Wang, Chi-Wing Fu
Film Photography
01โฒ
02โฒ
03โฒ
04โฒ
05โฒ
06โฒ
07โฒ
08โฒ
09โฒ
10โฒ
11โฒ
12โฒ
13โฒ
14โฒ
15โฒ
16โฒ
17โฒ
18โฒ
19โฒ
20โฒ
21โฒ
22โฒ
23โฒ
24โฒ
25โฒ
26โฒ
27โฒ
28โฒ
29โฒ
30โฒ
31โฒ
32โฒ
33โฒ
34โฒ
35โฒ
36โฒ
37โฒ
38โฒ
39โฒ
40โฒ
41โฒ
42โฒ
43โฒ
44โฒ
45โฒ
46โฒ
47โฒ
48โฒ
49โฒ
50โฒ
51โฒ
52โฒ
53โฒ
54โฒ
55โฒ
56โฒ
57โฒ
58โฒ
59โฒ
60โฒ
61โฒ
62โฒ
63โฒ
64โฒ
65โฒ
66โฒ
67โฒ
68โฒ
69โฒ
70โฒ
71โฒ
72โฒ
73โฒ
74โฒ
75โฒ
76โฒ
77โฒ
78โฒ
79โฒ
80โฒ
81โฒ
82โฒ
83โฒ
84โฒ
85โฒ
86โฒ
87โฒ
88โฒ
89โฒ
90โฒ
91โฒ
92โฒ
93โฒ
94โฒ
95โฒ
96โฒ
97โฒ
98โฒ
99โฒ
100โฒ
101โฒ
102โฒ
103โฒ
104โฒ
105โฒ
106โฒ
107โฒ
108โฒ
109โฒ
110โฒ
111โฒ
112โฒ
113โฒ
114โฒ
115โฒ
116โฒ
117โฒ
118โฒ
119โฒ
120โฒ
121โฒ
122โฒ
123โฒ
124โฒ
125โฒ
126โฒ
127โฒ
128โฒ
129โฒ
130โฒ
131โฒ
132โฒ
133โฒ
134โฒ
135โฒ
136โฒ
137โฒ
138โฒ
139โฒ
140โฒ
141โฒ
142โฒ
143โฒ
144โฒ
145โฒ
146โฒ
147โฒ
148โฒ
149โฒ
150โฒ
151โฒ
152โฒ
153โฒ
154โฒ
155โฒ
156โฒ
157โฒ
158โฒ
159โฒ
160โฒ
161โฒ
162โฒ
163โฒ
164โฒ
165โฒ
166โฒ
167โฒ
168โฒ
169โฒ
170โฒ
171โฒ
172โฒ
173โฒ
174โฒ
175โฒ
176โฒ
177โฒ
178โฒ
179โฒ
180โฒ
181โฒ
182โฒ
183โฒ
184โฒ
185โฒ
186โฒ
187โฒ





