ylzhu@nullcs.stanford.edu
I am a Machine Learning Engineer at Apple, where I work on multimodal AI for human perception and the creation of realistic digital humans.
I received my Ph.D. degree from the Computer Science Department, Stanford University, advised by Ron Fedkiw at the Stanford AI Lab and Graphics Lab. Previously, I worked at Epic Games on facial performance capture and avatar creation.
My current research centers on multimodal AI for digital humans: building models that perceive, simulate, and animate people with realism and physical grounding. Bringing together generative models, computer vision, and computational physics, I extend the same principles to embodied AI (robotics), building agents that reason about and operate in the physical world. Broadly, I aim to democratize the creation of agents that understand and act in the world, turning expensive, expert-only pipelines into tools anyone can use.
|
|
Leveraging Deepfakes to Close the Domain Gap between Real and Synthetic Images in Facial Capture Pipelines
arXiv /
Paper
|
|
|
Learning Topological Motion Primitives for Knot Planning
|
|
|
Self-Supervised Learning of State Estimation for Manipulating Deformable Linear Objects
|
|
|
A Pixel-Based Framework for Data-Driven Clothing
|
|
|
Three Dimensional Reconstruction of Botanical Trees with Simulatable Geometry
|
|
|
Scientific visualization of genealogical data
SIGGRAPH 2020 Art Papers -
Best Art Paper Award
Paper / Project Video / Laser Engraving / Rendering / Animation Media Coverage: Business Wire |
|
|
Naughty AlphaGo
Game of Computer Go transformed into an Emotional Tangible Playground
|