Skip to main content

임성훈 교수

Sung-Hoon Lim

KAIST 문화기술대학원 · 컴퓨터과학

연구실 소개

임성훈 교수의 연구실은 주로 단일 카메라 기반 깊이 추정, 다중 시점 스테레오 복원, 그리고 자율주행 차량의 동적 객체 인식에 초점을 맞춘다. 특히, 자기지도 학습 기반 깊이 추정, 비정상적인 조건(텍스처가 부족한 영역, 반사성 표면 등)에서도 안정적인 복원을 위한 기하학적·광학적 제약을 통합한 신경망 설계에 힘쓰고 있으며, LiDAR 범위 영상 기반의 움직임 객체 분할 및 도메인 적응 기반의 비지도 학습 기법도 함께 개발하고 있다. 연구는 실세계 적용에 초점을 맞춰, 캘리브레이션 불필요, 실시간 처리, 다양한 환경에서의 일반화 능력 향상을 목표로 한다.

단일 카메라 깊이 추정다중 시점 스테레오자율주행 동적 객체 인식비지도 도메인 적응자기지도 학습

연구 현황

논문 수
82
총 인용 수
1,078
최근 5년 논문
49
주요 분야
컴퓨터과학

연구 성과 추이

표시된 성과는 수집된 데이터 기준으로 산출되며, 일부 차이가 있을 수 있습니다.

5개년 연도별 논문 게재 수
49총합
2022
2023
2024
2025
2026
5개년 연도별 피인용 수
305총합
20222023202420252026

주요 논문

15
1
논문|인용수 114·2023
Deep Digging into the Generalization of Self-Supervised Monocular Depth Estimation
Jin Woo Bae, Sungho Moon, Sunghoon Im
Proceedings of the AAAI Conference on Artificial IntelligenceOA

Self-supervised monocular depth estimation has been widely studied recently. Most of the work has focused on improving performance on benchmark datasets, such as KITTI, but has offered a few experiments on generalization performance. In this paper, we investigate the backbone networks (e.g., CNNs, Transformers, and CNN-Transformer hybrid models) toward the generalization of monocular depth estimation. We first evaluate state-of-the-art models on diverse public datasets, which have never been see

Media TechnologyEngineering
2
논문|인용수 97·2016
High-Quality Depth from Uncalibrated Small Motion Clip
Hyowon Ha, Sunghoon Im, Jaesik Park, Hae‐Gon Jeon, In So Kweon

We propose a novel approach that generates a highquality depth map from a set of images captured with a small viewpoint variation, namely small motion clip. As opposed to prior methods that recover scene geometry and camera motions using pre-calibrated cameras, we introduce a self-calibrating bundle adjustment tailored for small motion. This allows our dense stereo algorithm to produce a high-quality depth map for the user without the need for camera calibration. In the dense matching, the distr

Computer Vision and Pattern RecognitionComputer Science
3
논문|인용수 95·2021
Learning Monocular Depth in Dynamic Scenes via Instance-Aware Projection Consistency
Seokju Lee, Sunghoon Im, Stephen Lin, In So Kweon

We present an end-to-end joint training framework that explicitly models 6-DoF motion of multiple dynamic objects, ego-motion, and depth in a monocular camera setup without supervision. Our technical contributions are three-fold. First, we highlight the fundamental difference between inverse and forward projection while modeling the individual motion of each rigid object, and propose a geometrically correct projection pipeline using a neural forward projection module. Second, we design a unified

Computer Vision and Pattern RecognitionComputer Science
4
논문|인용수 83·2019
DPSNet: End-to-end Deep Plane Sweep Stereo
Sunghoon Im, Hae‐Gon Jeon, Stephen Lin, In So Kweon
arXiv (Cornell University)OA

Multiview stereo aims to reconstruct scene depth from images acquired by a camera under arbitrary motion. Recent methods address this problem through deep learning, which can utilize semantic cues to deal with challenges such as textureless and reflective regions. In this paper, we present a convolutional neural network called DPSNet (Deep Plane Sweep Network) whose design is inspired by best practices of traditional geometry-based approaches for dense depth reconstruction. Rather than directly

Computer Vision and Pattern RecognitionComputer Science
5
book chapter|인용수 78·2016
All-Around Depth from Small Motion with a Spherical Panoramic Camera
Sunghoon Im, Hyowon Ha, François Rameau, Hae‐Gon Jeon, Gyeongmin Choe, In So Kweon
SJR Q2Lecture notes in computer science
Computer Vision and Pattern RecognitionComputer Science
6
논문|인용수 61·2022
RVMOS: Range-View Moving Object Segmentation Leveraged by Semantic and Motion Features
Jae-Yeul Kim, Jungwan Woo, Sunghoon Im
SJR Q1IEEE Robotics and Automation Letters

Detecting traffic participants is an essential and age-old problem in autonomous driving. Recently, the recognition of moving objects has emerged as a major issue in this field for safe driving. In this paper, we present RVMOS, a LiDAR Range-View-based Moving Object Segmentation framework that segments moving objects given a sequence of range-view images. In contrast to the conventional method, our network incorporates both motion and semantic features, each of which encodes the motion of object

Computer Vision and Pattern RecognitionComputer Science
7
논문|인용수 58·2021
DRANet: Disentangling Representation and Adaptation Networks for Unsupervised Cross-Domain Adaptation
Seunghun Lee, Sunghyun Cho, Sunghoon Im

In this paper, we present DRANet, a network architecture that disentangles image representations and transfers the visual attributes in a latent space for unsupervised cross-domain adaptation. Unlike the existing domain adaptation methods that learn associated features sharing a domain, DRANet preserves the distinctiveness of each domain’s characteristics. Our model encodes individual representations of content (scene structure) and style (artistic appearance) from both source and target images.

Artificial IntelligenceComputer Science
8
논문|인용수 47·2015
High Quality Structure from Small Motion for Rolling Shutter Cameras
Sunghoon Im, Hyowon Ha, Gyeongmin Choe, Hae‐Gon Jeon, Kyungdon Joo, In So Kweon

We present a practical 3D reconstruction method to obtain a high-quality dense depth map from narrow-baseline image sequences captured by commercial digital cameras, such as DSLRs or mobile phones. Depth estimation from small motion has gained interest as a means of various photographic editing, but important limitations present themselves in the form of depth uncertainty due to a narrow baseline and rolling shutter. To address these problems, we introduce a novel 3D reconstruction method from n

Computer Vision and Pattern RecognitionComputer Science
9
논문|인용수 44·2019
Ring Difference Filter for Fast and Noise Robust Depth From Focus
Hae‐Gon Jeon, Jaeheung Surh, Sunghoon Im, In So Kweon
SJR Q1IEEE Transactions on Image Processing

Depth from focus (DfF) is a method of estimating the depth of a scene by using information acquired through changes in the focus of a camera. Within the DfF framework of, the focus measure (FM) forms the foundation which determines the accuracy of the output. With the results from the FM, the role of a DfF pipeline is to determine and recalculate unreliable measurements while enhancing those that are reliable. In this paper, we propose a new FM, which we call the "ring difference filter" (RDF),

Media TechnologyEngineering
10
논문|인용수 43·2018
RANUS: RGB and NIR Urban Scene Dataset for Deep Scene Parsing
Gyeongmin Choe, Seong‐heum Kim, Sunghoon Im, Joon‐Young Lee, Srinivasa G. Narasimhan, In So Kweon
SJR Q1IEEE Robotics and Automation Letters

In this letter, we present a data-driven method for scene parsing of road scenes to utilize single-channel near-infrared (NIR) images. To overcome the lack of data problem in non-RGB spectrum, we define a new color space and decompose the task of deep scene parsing into two subtasks with two separate CNN architectures for chromaticity channels and semantic masks. For chromaticity estimation, we build a spatially-aligned RGB-NIR image database (40k urban scenes) to infer color information from RG

Computer Vision and Pattern RecognitionComputer Science
11
논문|인용수 42·2019
Depth Completion with Deep Geometry and Context Guidance
Byeong-Uk Lee, Hae‐Gon Jeon, Sunghoon Im, In So Kweon

In this paper, we present an end-to-end convolutional neural network (CNN) for depth completion. Our network consists of a geometry network and a context network. The geometry network, a single encoder-decoder network, learns to optimize a multi-task loss to generate an initial propagated depth map and a surface normal. The complementary outputs allow it to correctly propagate initial sparse depth points in slanted surfaces. The context network extracts a local and a global feature of an image t

Computer Vision and Pattern RecognitionComputer Science
12
논문|인용수 42·2016
Stereo Matching with Color and Monochrome Cameras in Low-Light Conditions
Hae‐Gon Jeon, Joon‐Young Lee, Sunghoon Im, Hyowon Ha, In So Kweon

Consumer devices with stereo cameras have become popular because of their low-cost depth sensing capability. However, those systems usually suffer from low imaging quality and inaccurate depth acquisition under low-light conditions. To address the problem, we present a new stereo matching method with a color and monochrome camera pair. We focus on the fundamental trade-off that monochrome cameras have much better light-efficiency than color-filtered cameras. Our key ideas involve compensating fo

Computer Vision and Pattern RecognitionComputer Science
13
논문|인용수 34·2017
Noise Robust Depth from Focus Using a Ring Difference Filter
Jaeheung Surh, Hae‐Gon Jeon, Yunwon Park, Sunghoon Im, Hyowon Ha, In So Kweon

Depth from focus (DfF) is a method of estimating depth of a scene by using the information acquired through the change of the focus of a camera. Within the framework of DfF, the focus measure (FM) forms the foundation on which the accuracy of the output is determined. With the result from the FM, the role of a DfF pipeline is to determine and recalculate unreliable measurements while enhancing those that are reliable. In this paper, we propose a new FM that more accurately and robustly measures

Media TechnologyEngineering
14
논문|인용수 30·2018
Accurate 3D Reconstruction from Small Motion Clip for Rolling Shutter Cameras
Sunghoon Im, Hyowon Ha, Gyeongmin Choe, Hae‐Gon Jeon, Kyungdon Joo, In So Kweon
SJR Q1IEEE Transactions on Pattern Analysis and Machine Intelligence

Structure from small motion has become an important topic in 3D computer vision as a method for estimating depth, since capturing the input is so user-friendly. However, major limitations exist with respect to the form of depth uncertainty, due to the narrow baseline and the rolling shutter effect. In this paper, we present a dense 3D reconstruction method from small motion clips using commercial hand-held cameras, which typically cause the undesired rolling shutter artifact. To address these pr

Computer Vision and Pattern RecognitionComputer Science
15
논문|인용수 8·2019
Deep Depth from Uncalibrated Small Motion Clip
Sunghoon Im, Hyowon Ha, Hae‐Gon Jeon, Stephen Lin, In So Kweon
SJR Q1IEEE Transactions on Pattern Analysis and Machine Intelligence

We propose a novel approach to infer a high-quality depth map from a set of images with small viewpoint variations. In general, techniques for depth estimation from small motion consist of camera pose estimation and dense reconstruction. In contrast to prior approaches that recover scene geometry and camera motions using pre-calibrated cameras, we introduce in this paper a self-calibrating bundle adjustment method tailored for small motion which enables computation of camera poses without the ne

Computer Vision and Pattern RecognitionComputer Science

대표 연구 분야

Computer Vision and Pattern RecognitionArtificial IntelligenceMedia TechnologyAerospace EngineeringElectrical and Electronic EngineeringCognitive Neuroscience

임성훈 교수의 연구를 Nubint에서 더 깊이 살펴보세요

이 연구실의 논문을 앱에서 열어 AI와 함께 읽고, 핵심을 요약하고, 내 글에 인용하세요.