Skip to main content

Sung-Hoon Lim

Korea Advanced Institute of Science and Technology · 情報科学

研究室紹介

Professor Sung-Hoon Lim's research lab specializes in computer vision and deep learning for autonomous systems, with a strong focus on monocular and multiview depth estimation, self-supervised learning, and 3D scene understanding. The lab develops geometrically principled neural networks that model motion, depth, and object dynamics in complex environments, particularly for autonomous driving applications. Key research directions include self-calibrating stereo reconstruction, disentangled representation learning for domain adaptation, and joint estimation of ego-motion and object motion from monocular video. The lab emphasizes robustness and generalization across diverse real-world scenarios through innovative architectural designs and self-supervisory signal engineering.

monocular depth estimationself-supervised learning3D scene reconstructionmotion estimationdomain adaptation

Research Overview

Papers
82
Total Citations
1,078
Papers (5y)
49
Primary Field
情報科学

Research Output Trend

Figures are computed from collected data and may differ slightly.

Publications per year (5y)
49total
2022
2023
2024
2025
2026
Citations per year (5y)
305total
20222023202420252026

Selected Papers

15
1
Article|114 citations·2023
Deep Digging into the Generalization of Self-Supervised Monocular Depth Estimation
Jin Woo Bae, Sungho Moon, Sunghoon Im
Proceedings of the AAAI Conference on Artificial IntelligenceOA

Self-supervised monocular depth estimation has been widely studied recently. Most of the work has focused on improving performance on benchmark datasets, such as KITTI, but has offered a few experiments on generalization performance. In this paper, we investigate the backbone networks (e.g., CNNs, Transformers, and CNN-Transformer hybrid models) toward the generalization of monocular depth estimation. We first evaluate state-of-the-art models on diverse public datasets, which have never been see

Media TechnologyEngineering
2
Article|97 citations·2016
High-Quality Depth from Uncalibrated Small Motion Clip
Hyowon Ha, Sunghoon Im, Jaesik Park, Hae‐Gon Jeon, In So Kweon

We propose a novel approach that generates a highquality depth map from a set of images captured with a small viewpoint variation, namely small motion clip. As opposed to prior methods that recover scene geometry and camera motions using pre-calibrated cameras, we introduce a self-calibrating bundle adjustment tailored for small motion. This allows our dense stereo algorithm to produce a high-quality depth map for the user without the need for camera calibration. In the dense matching, the distr

Computer Vision and Pattern RecognitionComputer Science
3
Article|95 citations·2021
Learning Monocular Depth in Dynamic Scenes via Instance-Aware Projection Consistency
Seokju Lee, Sunghoon Im, Stephen Lin, In So Kweon

We present an end-to-end joint training framework that explicitly models 6-DoF motion of multiple dynamic objects, ego-motion, and depth in a monocular camera setup without supervision. Our technical contributions are three-fold. First, we highlight the fundamental difference between inverse and forward projection while modeling the individual motion of each rigid object, and propose a geometrically correct projection pipeline using a neural forward projection module. Second, we design a unified

Computer Vision and Pattern RecognitionComputer Science
4
Article|83 citations·2019
DPSNet: End-to-end Deep Plane Sweep Stereo
Sunghoon Im, Hae‐Gon Jeon, Stephen Lin, In So Kweon
arXiv (Cornell University)OA

Multiview stereo aims to reconstruct scene depth from images acquired by a camera under arbitrary motion. Recent methods address this problem through deep learning, which can utilize semantic cues to deal with challenges such as textureless and reflective regions. In this paper, we present a convolutional neural network called DPSNet (Deep Plane Sweep Network) whose design is inspired by best practices of traditional geometry-based approaches for dense depth reconstruction. Rather than directly

Computer Vision and Pattern RecognitionComputer Science
5
Book Chapter|78 citations·2016
All-Around Depth from Small Motion with a Spherical Panoramic Camera
Sunghoon Im, Hyowon Ha, François Rameau, Hae‐Gon Jeon, Gyeongmin Choe, In So Kweon
SJR Q2Lecture notes in computer science
Computer Vision and Pattern RecognitionComputer Science
6
Article|61 citations·2022
RVMOS: Range-View Moving Object Segmentation Leveraged by Semantic and Motion Features
Jae-Yeul Kim, Jungwan Woo, Sunghoon Im
SJR Q1IEEE Robotics and Automation Letters

Detecting traffic participants is an essential and age-old problem in autonomous driving. Recently, the recognition of moving objects has emerged as a major issue in this field for safe driving. In this paper, we present RVMOS, a LiDAR Range-View-based Moving Object Segmentation framework that segments moving objects given a sequence of range-view images. In contrast to the conventional method, our network incorporates both motion and semantic features, each of which encodes the motion of object

Computer Vision and Pattern RecognitionComputer Science
7
Article|58 citations·2021
DRANet: Disentangling Representation and Adaptation Networks for Unsupervised Cross-Domain Adaptation
Seunghun Lee, Sunghyun Cho, Sunghoon Im

In this paper, we present DRANet, a network architecture that disentangles image representations and transfers the visual attributes in a latent space for unsupervised cross-domain adaptation. Unlike the existing domain adaptation methods that learn associated features sharing a domain, DRANet preserves the distinctiveness of each domain’s characteristics. Our model encodes individual representations of content (scene structure) and style (artistic appearance) from both source and target images.

Artificial IntelligenceComputer Science
8
Article|47 citations·2015
High Quality Structure from Small Motion for Rolling Shutter Cameras
Sunghoon Im, Hyowon Ha, Gyeongmin Choe, Hae‐Gon Jeon, Kyungdon Joo, In So Kweon

We present a practical 3D reconstruction method to obtain a high-quality dense depth map from narrow-baseline image sequences captured by commercial digital cameras, such as DSLRs or mobile phones. Depth estimation from small motion has gained interest as a means of various photographic editing, but important limitations present themselves in the form of depth uncertainty due to a narrow baseline and rolling shutter. To address these problems, we introduce a novel 3D reconstruction method from n

Computer Vision and Pattern RecognitionComputer Science
9
Article|44 citations·2019
Ring Difference Filter for Fast and Noise Robust Depth From Focus
Hae‐Gon Jeon, Jaeheung Surh, Sunghoon Im, In So Kweon
SJR Q1IEEE Transactions on Image Processing

Depth from focus (DfF) is a method of estimating the depth of a scene by using information acquired through changes in the focus of a camera. Within the DfF framework of, the focus measure (FM) forms the foundation which determines the accuracy of the output. With the results from the FM, the role of a DfF pipeline is to determine and recalculate unreliable measurements while enhancing those that are reliable. In this paper, we propose a new FM, which we call the "ring difference filter" (RDF),

Media TechnologyEngineering
10
Article|43 citations·2018
RANUS: RGB and NIR Urban Scene Dataset for Deep Scene Parsing
Gyeongmin Choe, Seong‐heum Kim, Sunghoon Im, Joon‐Young Lee, Srinivasa G. Narasimhan, In So Kweon
SJR Q1IEEE Robotics and Automation Letters

In this letter, we present a data-driven method for scene parsing of road scenes to utilize single-channel near-infrared (NIR) images. To overcome the lack of data problem in non-RGB spectrum, we define a new color space and decompose the task of deep scene parsing into two subtasks with two separate CNN architectures for chromaticity channels and semantic masks. For chromaticity estimation, we build a spatially-aligned RGB-NIR image database (40k urban scenes) to infer color information from RG

Computer Vision and Pattern RecognitionComputer Science
11
Article|42 citations·2019
Depth Completion with Deep Geometry and Context Guidance
Byeong-Uk Lee, Hae‐Gon Jeon, Sunghoon Im, In So Kweon

In this paper, we present an end-to-end convolutional neural network (CNN) for depth completion. Our network consists of a geometry network and a context network. The geometry network, a single encoder-decoder network, learns to optimize a multi-task loss to generate an initial propagated depth map and a surface normal. The complementary outputs allow it to correctly propagate initial sparse depth points in slanted surfaces. The context network extracts a local and a global feature of an image t

Computer Vision and Pattern RecognitionComputer Science
12
Article|42 citations·2016
Stereo Matching with Color and Monochrome Cameras in Low-Light Conditions
Hae‐Gon Jeon, Joon‐Young Lee, Sunghoon Im, Hyowon Ha, In So Kweon

Consumer devices with stereo cameras have become popular because of their low-cost depth sensing capability. However, those systems usually suffer from low imaging quality and inaccurate depth acquisition under low-light conditions. To address the problem, we present a new stereo matching method with a color and monochrome camera pair. We focus on the fundamental trade-off that monochrome cameras have much better light-efficiency than color-filtered cameras. Our key ideas involve compensating fo

Computer Vision and Pattern RecognitionComputer Science
13
Article|34 citations·2017
Noise Robust Depth from Focus Using a Ring Difference Filter
Jaeheung Surh, Hae‐Gon Jeon, Yunwon Park, Sunghoon Im, Hyowon Ha, In So Kweon

Depth from focus (DfF) is a method of estimating depth of a scene by using the information acquired through the change of the focus of a camera. Within the framework of DfF, the focus measure (FM) forms the foundation on which the accuracy of the output is determined. With the result from the FM, the role of a DfF pipeline is to determine and recalculate unreliable measurements while enhancing those that are reliable. In this paper, we propose a new FM that more accurately and robustly measures

Media TechnologyEngineering
14
Article|30 citations·2018
Accurate 3D Reconstruction from Small Motion Clip for Rolling Shutter Cameras
Sunghoon Im, Hyowon Ha, Gyeongmin Choe, Hae‐Gon Jeon, Kyungdon Joo, In So Kweon
SJR Q1IEEE Transactions on Pattern Analysis and Machine Intelligence

Structure from small motion has become an important topic in 3D computer vision as a method for estimating depth, since capturing the input is so user-friendly. However, major limitations exist with respect to the form of depth uncertainty, due to the narrow baseline and the rolling shutter effect. In this paper, we present a dense 3D reconstruction method from small motion clips using commercial hand-held cameras, which typically cause the undesired rolling shutter artifact. To address these pr

Computer Vision and Pattern RecognitionComputer Science
15
Article|8 citations·2019
Deep Depth from Uncalibrated Small Motion Clip
Sunghoon Im, Hyowon Ha, Hae‐Gon Jeon, Stephen Lin, In So Kweon
SJR Q1IEEE Transactions on Pattern Analysis and Machine Intelligence

We propose a novel approach to infer a high-quality depth map from a set of images with small viewpoint variations. In general, techniques for depth estimation from small motion consist of camera pose estimation and dense reconstruction. In contrast to prior approaches that recover scene geometry and camera motions using pre-calibrated cameras, we introduce in this paper a self-calibrating bundle adjustment method tailored for small motion which enables computation of camera poses without the ne

Computer Vision and Pattern RecognitionComputer Science

Research Areas

Computer Vision and Pattern RecognitionArtificial IntelligenceMedia TechnologyAerospace EngineeringElectrical and Electronic EngineeringCognitive Neuroscience

Sung-Hoon Limの研究をNubintでさらに深く

この研究室の論文をアプリで開き、AIと共に読み、要約し、引用しましょう。