arxiv:2605.04647
Pengxiang Li
pengxiang
AI & ML interests
Video generation, Image editing, AD
Recent Activity
upvoted a paper 30 days ago
GST-Bench: Can VLMs Develop Global Spatial Awareness from Video? upvoted a paper about 1 month ago
Towards Physics of Multimodal Pretraining: Knowledge Flow, Modality Synergy, Early Unification, and Recipes