Scholar
Dongdong Yu
Google Scholar ID: B2RmjSYAAAAJ
AISphere Tech.
Computer Vision
Human Pose Estimation
Video Object Segmentation
Scene Parsing
Radiomics
Follow
Homepage
↗
Google Scholar
↗
Citations & Impact
All-time
Citations
5,185
H-index
21
i10-index
31
Publications
20
Co-authors
15
list available
Publications
6 items
PixWorld: Unifying 3D Scene Generation and Reconstruction in Pixel Space
2026
Cited
0
OneWorld: Taming Scene Generation with 3D Unified Representation Autoencoder
2026
Cited
0
Taming Camera-Controlled Video Generation with Verifiable Geometry Reward
2025
Cited
0
MMGen: Unified Multi-modal Image Generation and Understanding in One Go
2025
Cited
0
LaVin-DiT: Large Vision Diffusion Transformer
arXiv.org · 2024
Cited
1
TrackGo: A Flexible and Efficient Method for Controllable Video Generation
arXiv.org · 2024
Cited
4
Co-authors
9 total
Di Dong
Institute of Automation, Chinese Academy of Sciences
Kai Su
ByteDance
Mu Zhou
Stanford University; Rutgers University
Olivier Gevaert
Stanford University
Zhenchao Jin
USTC > HKU
Mengjie Fang
Beihang University
Shuo Wang
Beihang University; institute of automation, chinese academy of sciences
Xin Geng
School of Computer Science and Engineering, Southeast University