Research Intern
2025 · San Jose, CASony Corporation of America.
I am a 2nd-year Ph.D. student at Texas A&M University, supervised by Prof. Wenping Wang and Prof. Xin Li. People call me John.
I am passionate about bringing intelligence to 3D and eventually the real world, through Robot Foundational Policies, Physics-based Motion, and 3D Real-to-Sim.
Previously, I obtained my B.Eng. and M.Eng. degrees from Shanghai Jiao Tong University, where I was fortunate to be supervised by Prof. Li Song.
A humanoid policy steered by future goal images: a predictor turns subtask prompts into goal images, and a planner drives whole-body actions toward them across multi-task routines.
We learn dense rewards from action-free videos by modeling task-relevant latent dynamics, so the reward tracks task progress rather than visual distractors such as lighting.
PDT is a novel diffusion-based framework for transforming dense point distribution into a target distribution that is semantically meaningful.
In this paper, we introduced a novel text-to-avatar generation method that separately generates the human body and the clothes and allows high-quality animation on the generated avatar.
We proposed a novel diffusion-based pipeline that generates complete 360 panoramas using one or more unregistered NFoV images captured from arbitrary angles.
SPGen leverages Spherical Projection (SP) to generate high-quality 3D shapes with diffusion models.
Sony Corporation of America.
miHoYo (HoYoverse).
Intel Asia Pacific Development.