Spatial AI for the Physical World
3D/4D World Modeling · Efficient Vision-Language Models · Embodied Intelligence
CVSP Lab at Pusan National University develops spatial multimodal AI for understanding, generating, and interacting with real-world environments.
Our research focuses on:
•
3D/4D World Modeling — Gaussian splatting, dynamic scene reconstruction, and immersive media
•
Efficient Vision-Language Models — token pruning, model compression, and reliable multimodal reasoning
•
Embodied Spatial Intelligence — real-world perception and decision-making for robots and media


.gif&blockId=396abcc9-0e34-80ee-a30c-cb7b55c3131b)









.png&blockId=39fabcc9-0e34-80ea-a477-c4e8f77b0d15&width=1024)

















