Skip to main content


Research Overview
#

# Visual Understanding & Recognition
We develop AI algorithms for image and video understanding across diverse visual recognition tasks. Our research covers image classification, object detection, segmentation, video understanding, and multimodal visual perception for intelligent vision systems.
Category #1
Keyword: Image Understanding, Video Understanding, Object Detection, Segmentation, Visual Recognition

# Reliable & Generalizable Visual AI
We develop vision AI that can reliably operate beyond its training environments. Our research investigates out-of-distribution (OOD) detection, domain generalization, and leverages vision foundation models, Vision-Language Models (VLMs), Large Multimodal Models (LMMs), and Mixture-of-Experts (MoE) to improve robustness and generalization.
Category #2
Keyword: Out-of-distribution (OOD) Detection, Domain Generalization, VLM/LMM, Mixture-of-Experts (MoE)

# Efficient Learning & Embeeded Vision AI
We investigate efficient learning and adaptation techniques for vision and multimodal AI. Our research includes prompt learning, parameter-efficient fine-tuning (PEFT), domain-specific model adaptation, efficient foundation model utilization, and on-device deployment for practical AI systems.
Category #3
Keyword: Prompt Learning, PEFT, Domain Adaptation, On-Device AI, Embedded Vision

# Intelligent Vision Systems & Applications
We apply advanced computer vision and multimodal AI to solve real-world problems. Our research includes intelligent visual perception, event and anomaly detection, human-centered AI, and practical vision applications across diverse domains.
Category #4
Keyword: Event Detection, Anomaly Detection, Intelligent Vision Systems, AI Applications

Research Projects
#

Ongoing Projects
  • (MSIT/NRF) Modular Memory-based Augmentation and Unlearning for LMM
  • (ETRI) Large Multimodal Model Technology for On-Device AI
  • (CNU) Synergistic Learning of Foundation Models for Visual Understanding
  • (MSIT/IITP) Adaptive Vision-Language Understanding under Domain Shift
  • (MOE/NRF) Intelligent Defense Unmanned Systems Research Institute - Agentic AI
Completed Projects
  • (MSIT/IITP) Active Reasoning Cooperative Multimodal Agent Technology
  • (AICOSS) Leveraging Public Data to Enhance the Quality of AI Education
  • (CNU) Bias Assessment and Quantification in Audio-Visual Generative AI
  • (MSIT/IITP) AI Tech. for Self-improving Competency-aware Learning Capabilities
  • (MSIT/IITP) High-Performance Visual Big Data Discovery Platform (DeepView)
  • (MSIT/IITP) Previsional Intelligence based on Long-term Visual Memory Network
  • (MSIT/IITP) Presymptomatic Alzheimer Prevention PoC based on Self-adaptive and Progressive Machine Learning Framework