I build practical multimodal AI systems across visual editing, scene understanding, video understanding, and audio-visual learning.
Hello! I am a master's student in the 3D Vision & Robotics Lab at UNIST, advised by Prof. Kyungdon Joo.
I was a visiting student in the CARTE, MIE, at the University of Toronto through the AI Convergence Program, supported by the IITP, Korean Government.
My research has focused on 3D scene understanding, image manipulation using generative models, video understanding, and visual-language modeling. Currently, I am expanding my research scope to audio-visual modeling, where I am developing and evaluating multimodal learning frameworks.