Publications
- All
- International
- Domestic
Different Changes Require Different Reasoning: Change-Type-Specialized Experts for Robust Change Captioning
European Conference on Computer Vision (ECCV), 2026
PersonaDrive: Controllable Trajectory Prediction with Multi-Dimensional Driving Personas
European Conference on Computer Vision (ECCV), 2026
Online Versatile Incremental Learning: Towards Class and Domain-Agnostic Adaptation at Any Time
European Conference on Computer Vision (ECCV), 2026
Adverse Weather Removal via Dynamic Enhancement Diffusion with Weather-Adaptive Prompting
IEEE Transactions on Image Processing (TIP, IF: 13.7), 2026
Learning from Oblivion: Predicting Knowledge-Overflowed Weights via Retrodiction of Forgetting
IEEE/CVF Conference on Computer Vision and Pattern Recognition (CVPR), 2026
From Adaptation to Generalization: Adaptive Visual Prompting for Medical Image Segmentation
IEEE/CVF Conference on Computer Vision and Pattern Recognition (CVPR) Findings, 2026
Rationale-Guided Learning for Multimodal Emotion Recognition
IEEE International Conference on Acoustics, Speech and Signal Processing (ICASSP), 2026
See, Rank, and Filter: Important Word-Aware Clip Filtering via Scene Understanding for Moment Retrieval and Highlight Detection
AAAI Conference on Artificial Intelligence (AAAI), 2026
Task Prototype-Based Knowledge Retrieval for Multi-Task Learning from Partially Annotated Data
AAAI Conference on Artificial Intelligence (AAAI), 2026
Do We Need Perfect Data? Leveraging Noise for Domain Generalized Segmentation
AAAI Conference on Artificial Intelligence (AAAI), 2026
Unsupervised Domain Adaptation for Medical Image Segmentation Using Adaptogen-Perturbation
Medical Image Analysis (IF: 11.8), 2026
Object-aware Sound Source Localization via Audio-Visual Scene Understanding
IEEE/CVF Conference on Computer Vision and Pattern Recognition (CVPR), 2025
SSMPD: Semi-Supervised Learning for Multispectral Pedestrian Detection
IEEE Transactions on Multimedia (TMM, IF: 9.7), 2025
Enabling Visual Object Detection with Object Sounds via Visual Modality Recalling Memory
IEEE Transactions on Neural Networks and Learning Systems (TNNLS, IF: 8.9), 2025
Spatial Mask-based Adaptive Robust Training for Video Object Segmentation with Noisy Labels
IEEE Transactions on Circuits and Systems for Video Technology (TCSVT, IF: 11.1), 2025
Unified Link Prediction Modeling for Enhanced Knowledge Graph Completion Task
Expert Systems With Applications (IF: 7.5), 2025
Moment Retrieval and Highlight Detection Framework via Semantic Alignment of Phrases
Korean Institute of Broadcast and Media Engineers (KIBME), 2025
Grounded VideoQA via Optimal Transport-based Fine-Grained
Korean Institute of Broadcast and Media Engineers (KIBME), 2025
Vision Language Model Based Approach for Change Captioning
Korea Institute of Military Science and Technology (KIMST), 2025
Video Moment Retrieval and Highlight Detection via Effective Fusion of Captions Generated by Vision-Language Models
Journal of Broadcast Engineering / KIBME, 2025
The Necessity of Training Strategies for Monocular 3D Object Detection under Adverse Weather Conditions
Korean Institute of Broadcast and Media Engineers (KIBME), 2025
A Semi-Supervised Learning Framework for Rain-Robust Multispectral Pedestrian Detection with Limited Labels
Korean Institute of Broadcast and Media Engineers (KIBME), 2025
Performance Evaluation and Analysis of Visual Information Removal in Dense Video Captioning
Korean Institute of Broadcast and Media Engineers (KIBME), 2025
Siamese Network-Based Knowledge Distillation Method for Domain-Robust Visual Question Answering
Korean Institute of Broadcast and Media Engineers (KIBME), 2025
Moment Retrieval and Highlight Detection via Keyword Extraction from Natural Language
Korean Institute of Broadcast and Media Engineers (KIBME), 2025
A Real-time Free-Viewpoint Videos Streaming Model Robust to Fast-Moving Objects
Korean Institute of Broadcast and Media Engineers (KIBME), 2025
Hierarchical Semantic Prompt Design for Robust Open-Vocabulary Object Detection
Journal of KIISE, 2025
Learning Trimodal Relation for Audio-Visual Question Answering with Missing Modality
European Conference on Computer Vision (ECCV), 2024
MonoWAD: Weather-Adaptive Diffusion Model for Robust Monocular 3D Object Detection
European Conference on Computer Vision (ECCV), 2024
Towards Model-Agnostic Dataset Condensation by Heterogeneous Models
European Conference on Computer Vision (ECCV), 2024
Learning to Visually Localize Sound Sources from Mixtures without Prior Source Knowledge
IEEE/CVF Conference on Computer Vision and Pattern Recognition (CVPR), 2024
Enhancing Audio-Visual Question Answering with Missing Modality via Trans-Modal Associative Learning
IEEE International Conference on Acoustics, Speech and Signal Processing (ICASSP), 2024
Robust Airway Generation Labeling with Airway Segmentation for Reliable Airway Assessment
IEEE Access, 2024
Video Moment Retrieval and Highlight Detection Using Captions Generated by Multimodal Large Language Models
Korea Software Congress (KSC), 2024
Study For Automatic Pedestrian Data Augmented Learning Technique For Robust Multispectral Pedestrian Detection
Korea Computer Congress (KCC), 2024
Robust Video Moment Retrieval and Highlight Detection with Using Large Vision-Language Model
Korea Computer Congress (KCC), 2024
Robust 3D Object Detection Using 2D-Based Object Detection Information and Symmetry Knowledge
Journal of Broadcast Engineering, 2024
Efficient Sampling Method for Repetitive Data for Label Data Selection In Multispectral Pedestrian Detection
Korean Institute of Broadcast and Media Engineers (KIBME), 2024
Effective Change Captioning via Feature Map Restoration in Noisy Environments
Korean Institute of Broadcast and Media Engineers (KIBME), 2024
Handling Missing Modality Using Multimodal Relational Knowledge
Korea Institute of Military Science and Technology (KIMST), 2024
Audio-Visual Spatial Integration and Recursive Attention for Robust Sound Source Localization
ACM International Conference on Multimedia (ACM MM), 2023
Online Class Incremental Learning on Stochastic Blurry Task Boundary via Mask and Visual Prompt Tuning
IEEE/CVF International Conference on Computer Vision (ICCV), 2023
Stereoscopic Vision Recalling Memory for Monocular 3D Object Detection
IEEE Transactions on Image Processing (TIP, IF: 13.7), 2023
Robust Multispectral Pedestrian Detection Via Spectral Position-Free Feature Mapping
IEEE International Conference on Image Processing (ICIP), 2023
Towards Robust Audio-Based Vehicle Detection via Importance-Aware Audio-Visual Learning
IEEE International Conference on Acoustics, Speech and Signal Processing (ICASSP), 2023
Similarity Relation Preserving Cross-Modal Learning for Multispectral Pedestrian Detection Against Adversarial Attacks
IEEE International Conference on Acoustics, Speech and Signal Processing (ICASSP), 2023
Multimodal Pedestrian Detection Using Attention and Multi Modal Guide
Korea Software Congress (KSC), 2023
Single-Modal Pedestrian Detection Leveraging Multimodal Knowledge for Blackout Situations
Journal of KIISE, 2023
Effective 3D Object Detection using 2D Object Detection information
Korean Institute of Broadcast and Media Engineers (KIBME), 2023
Enhancing Multi-Modal Audio-Visual QA Model Performance via Single Modal Feature Maps with Added Noise Input
Korean Institute of Broadcast and Media Engineers (KIBME), 2023
Automatic Notification of Dangerous Situations in Blind Spots via Image-based Pedestrian Detection using Deep Learning
Korea Software Congress (KSC), 2022