Ziqi Ma
Computer Vision @ California Institute of Technology.
Hello! I am Ziqi, a PhD student at Caltech. My goal is to let AI understand and act in the physical world. Currently I do this via “world modeling’’ – letting AI learn spatial and physical priors by generating and modeling the world. I am advised by Georgia Gkioxari and Yisong Yue. I interned at World Labs and Meta SuperIntelligence Labs. I am also a Microsoft and UChicago alumni.
News
| Oct 05, 2026 | LoGo is released! Reward design actually matters a great deal in post-training long-horizon video gen for consistency! Learned a ton working on this during my internship at World Labs! |
|---|---|
| Sep 24, 2026 | DynaTokens got accepted to NeurIPS. See you in Atlanta! |
| Sep 01, 2026 | Atlas is released! Best camera controlled video gen, plus everything you get with an omni AR model! |
| Jun 17, 2026 | Out of Sight, Out of Mind? Evaluating State Evolution in Video World Models is accepted to ECCV! |
| Jun 06, 2026 | SAM 3D got CVPR 2026 Best Paper Award Honorable Mention! |
| May 21, 2026 | I passed my candidacy, and am honored to receive the Simoudis Discovery Prize! |