Skip to main content

송환준 교수

Hwanjun Song

KAIST 산업및시스템공학과 · 컴퓨터과학

연구실 소개

송환준 교수의 연구실은 딥러닝의 실용성과 안정성을 높이기 위한 고도화된 학습 기법을 중심으로 연구를 진행하고 있습니다. 특히 레이블 노이즈에 강한 신경망 학습, 대량 데이터 환경에서의 효율적 클러스터링 및 객체 검출 기술 개발에 초점을 맞추고 있으며, 실제 응용에서 흔히 발생하는 데이터 품질 문제 해결에 기여하고자 합니다. 연구는 이론적 분석과 실증적 검증을 병행하여, 정확도와 효율성을 동시에 확보하는 알고리즘 설계를 목표로 합니다.

노이즈 레이블 학습병렬 클러스터링비교적 효율적인 객체 검출데이터 증강딥러닝 안정성

연구 현황

논문 수
99
총 인용 수
2,087
최근 5년 논문
72
주요 분야
컴퓨터과학

연구 성과 추이

표시된 성과는 수집된 데이터 기준으로 산출되며, 일부 차이가 있을 수 있습니다.

5개년 연도별 논문 게재 수
72총합
2022
2023
2024
2025
2026
5개년 연도별 피인용 수
1,473총합
20222023202420252026

주요 논문

15
1
논문|인용수 1,079·2022
Learning From Noisy Labels With Deep Neural Networks: A Survey
Hwanjun Song, Minseok Kim, Dongmin Park, Yooju Shin, Jae-Gil Lee
SJR Q1IEEE Transactions on Neural Networks and Learning Systems

Deep learning has achieved remarkable success in numerous domains with help from large amounts of big data. However, the quality of data labels is a concern because of the lack of high-quality labels in many real-world scenarios. As noisy labels severely degrade the generalization performance of deep neural networks, learning from noisy labels (robust training) is becoming an important task in modern deep learning applications. In this survey, we first describe the problem of learning with label

Artificial IntelligenceComputer Science
2
논문|인용수 196·2019
SELFIE: Refurbishing Unclean Samples for Robust Deep Learning
Hwanjun Song, Minseok Kim, Jae-Gil Lee
International Conference on Machine Learning
Artificial IntelligenceComputer Science
3
preprint|인용수 91·2020
Learning from Noisy Labels with Deep Neural Networks: A Survey
Hwanjun Song, Minseok Kim, Dongmin Park, Yooju Shin, Jae-Gil Lee
arXiv (Cornell University)OA

Deep learning has achieved remarkable success in numerous domains with help from large amounts of big data. However, the quality of data labels is a concern because of the lack of high-quality labels in many real-world scenarios. As noisy labels severely degrade the generalization performance of deep neural networks, learning from noisy labels (robust training) is becoming an important task in modern deep learning applications. In this survey, we first describe the problem of learning with label

Artificial IntelligenceComputer Science
4
논문|인용수 59·2018
RP-DBSCAN
Hwanjun Song, Jae-Gil Lee

In most parallel DBSCAN algorithms, neighboring points are assigned to the same data partition for parallel processing to facilitate calculation of the density of the neighbors. This data partitioning scheme causes a few critical problems including load imbalance between data partitions, especially in a skewed data set. To remedy these problems, we propose a cell-based data partitioning scheme, pseudo random partitioning , that randomly distributes small cells rather than the points themselves.

Artificial IntelligenceComputer Science
5
preprint|인용수 46·2021
ViDT: An Efficient and Effective Fully Transformer-based Object Detector
Hwanjun Song, Sun Deqing, Sanghyuk Chun, Varun Jampani, Dongyoon Han, Byeongho Heo, Wonjae Kim, Ming–Hsuan Yang
arXiv (Cornell University)OA

Transformers are transforming the landscape of computer vision, especially for recognition tasks. Detection transformers are the first fully end-to-end learning systems for object detection, while vision transformers are the first fully transformer-based architecture for image classification. In this paper, we integrate Vision and Detection Transformers (ViDT) to build an effective and efficient object detector. ViDT introduces a reconfigured attention module to extend the recent Swin Transforme

Computer Vision and Pattern RecognitionComputer Science
6
논문|인용수 41·2017
PAMAE
Hwanjun Song, Jae-Gil Lee, Wook-Shin Han

The k-medoids algorithm is one of the best-known clustering algorithms. Despite this, however, it is not as widely used for big data analytics as the k-means algorithm, mainly because of its high computational complexity. Many studies have attempted to solve the efficiency problem of the k-medoids algorithm, but all such studies have improved efficiency at the expense of accuracy. In this paper, we propose a novel parallel k-medoids algorithm, which we call PAMAE, that achieves both high accurac

Signal ProcessingComputer Science
7
논문|인용수 25·2019
Prestopping: How Does Early Stopping Help Generalization Against Label Noise?
Hwanjun Song, Minseok Kim, Dongmin Park, Jae-Gil Lee
arXiv (Cornell University)OA
Signal ProcessingComputer Science
8
논문|인용수 22·2020
Ada-boundary: accelerating DNN training via adaptive boundary batch selection
Hwanjun Song, Sundong Kim, Minseok Kim, Jae-Gil Lee
SJR Q1Machine LearningOA
Computer Vision and Pattern RecognitionComputer Science
9
논문|인용수 20·2024
Toward Robustness in Multi-Label Classification: A Data Augmentation Strategy against Imbalance and Noise
Hwanjun Song, Minseok Kim, Jae-Gil Lee
Proceedings of the AAAI Conference on Artificial IntelligenceOA

Multi-label classification poses challenges due to imbalanced and noisy labels in training data. In this paper, we propose a unified data augmentation method, named BalanceMix, to address these challenges. Our approach includes two samplers for imbalanced labels, generating minority-augmented instances with high diversity. It also refines multi-labels at the label-wise granularity, categorizing noisy labels as clean, re-labeled, or ambiguous for robust optimization. Extensive experiments on thre

Artificial IntelligenceComputer Science
10
논문|인용수 20·2024
FineSurE: Fine-grained Summarization Evaluation using LLMs
Hwanjun Song, Hang Su, Igor Shalyminov, Jason Cai, Saab Mansour
OA

Automated evaluation is crucial for streamlining text summarization benchmarking and model development, given the costly and timeconsuming nature of human evaluation.Traditional methods like ROUGE do not correlate well with human judgment, while recently proposed LLM-based metrics provide only summary-level assessment using Likertscale scores.This limits deeper model analysis, e.g., we can only assign one hallucination score at the summary level, while at the sentence level, we can count sentenc

Artificial IntelligenceComputer Science
11
preprint|인용수 19·2019
How does Early Stopping Help Generalization against Label Noise?
Hwanjun Song, Minseok Kim, Dongmin Park, Jae-Gil Lee
arXiv (Cornell University)OA

Noisy labels are very common in real-world training data, which lead to poor generalization on test data because of overfitting to the noisy labels. In this paper, we claim that such overfitting can be avoided by "early stopping" training a deep neural network before the noisy labels are severely memorized. Then, we resume training the early stopped network using a "maximal safe set," which maintains a collection of almost certainly true-labeled samples at each epoch since the early stop point.

Artificial IntelligenceComputer Science
12
논문|인용수 18·2024
Prompt-guided DETR with RoI-pruned masked attention for open-vocabulary object detection
Hwanjun Song, Jihwan Bang
SJR Q1Pattern Recognition
Computer Vision and Pattern RecognitionComputer Science
13
preprint|인용수 11·2022
An Extendable, Efficient and Effective Transformer-based Object Detector
Hwanjun Song, Deqing Sun, Sanghyuk Chun, Varun Jampani, Dongyoon Han, Byeongho Heo, Wonjae Kim, Ming–Hsuan Yang
arXiv (Cornell University)OA

Transformers have been widely used in numerous vision problems especially for visual recognition and detection. Detection transformers are the first fully end-to-end learning systems for object detection, while vision transformers are the first fully transformer-based architecture for image classification. In this paper, we integrate Vision and Detection Transformers (ViDT) to construct an effective and efficient object detector. ViDT introduces a reconfigured attention module to extend the rece

Computer Vision and Pattern RecognitionComputer Science
14
논문|인용수 8·2020
Carpe Diem, Seize the Samples Uncertain "at the Moment" for Adaptive Batch Selection
Hwanjun Song, Minseok Kim, Sundong Kim, Jae-Gil Lee

The accuracy of deep neural networks is significantly affected by how well mini-batches are constructed during the training step. In this paper, we propose a novel adaptive batch selection algorithm called Recency Bias that exploits the uncertain samples predicted inconsistently in recent iterations. The historical label predictions of each training sample are used to evaluate its predictive uncertainty within a sliding window. Then, the sampling probability for the next mini-batch is assigned t

Artificial IntelligenceComputer Science
15
preprint|인용수 3·2023
Prompt-Guided Transformers for End-to-End Open-Vocabulary Object Detection
Hwanjun Song, Jihwan Bang
arXiv (Cornell University)OA

Prompt-OVD is an efficient and effective framework for open-vocabulary object detection that utilizes class embeddings from CLIP as prompts, guiding the Transformer decoder to detect objects in both base and novel classes. Additionally, our novel RoI-based masked attention and RoI pruning techniques help leverage the zero-shot classification ability of the Vision Transformer-based CLIP, resulting in improved detection performance at minimal computational cost. Our experiments on the OV-COCO and

Computer Vision and Pattern RecognitionComputer Science

대표 연구 분야

Artificial IntelligenceComputer Vision and Pattern RecognitionInformation SystemsSignal ProcessingManagement Science and Operations ResearchEconomics and Econometrics

송환준 교수의 연구를 Nubint에서 더 깊이 살펴보세요

이 연구실의 논문을 앱에서 열어 AI와 함께 읽고, 핵심을 요약하고, 내 글에 인용하세요.