Multimodal AI · Research & Engineering
I work on vision-language models, video understanding, and AI systems —
from research experiments and evaluation to deployment.
- Multimodal & Vision-Language Learning
- Video / Temporal Understanding
- Action-centric & Embodied Intelligence
- Reliable evaluation and applied AI systems
Currently exploring how multimodal models can better understand actions, time, and interaction.



