I am an integrated M.S.–Ph.D. student in the DAVIAN Lab at KAIST AI, advised by Professor Jaegul Choo. My research centers on Vision-Language-Action (VLA) models for robot manipulation.
One question runs through my work: how to ground a model in the spatial structure of a scene. I spent several years localizing sound sources in visual scenes and understanding video, learning to align perception across modalities in space and time. That thread now runs into robotics — predicting the 3D trajectories a robot should follow, and giving policies a robot-centric view of the scene they act in. VLA is where that grounding becomes action.
Alongside this, I work on LLM safety: content moderation for specialized domains and defense for LLM agents.
Research Interests
2026
2025
2024
Integrated M.S.–Ph.D. in Artificial Intelligence
KAIST AI · DAVIAN Lab
Advisor: Prof. Jaegul Choo
Research focus: Vision-Language-Action, multimodal, LLM safety
B.S. in Computer Science & Engineering
Kyung Hee University · Visual AI Lab
Advisor: Prof. Jung Uk Kim
Research focus: Multi-modal learning, Video Understanding, Sound Source Localization
AI Engineer
LETSUR (AI Startup)
Development of AI-based services utilizing LLMs and RAG.
AI Research Intern
ETRI (Electronics and Telecommunications Research Institute)
Research on Video Moment Retrieval and Highlight Detection.
Undergraduate Researcher
Kyung Hee University · Visual AI Lab
Supervisor: Jung Uk Kim.