Skip to main content

Gunhee Kim

Seoul National University · 情報科学

研究室紹介

Professor Gunhee Kim's research lab specializes in computer vision and multimedia analysis, with a focus on unsupervised and weakly supervised learning for large-scale image and video understanding. The lab develops innovative methods for cosegmentation, storyline graph reconstruction, and region-of-interest detection, emphasizing scalability, structural modeling of visual narratives, and real-world applicability in web-scale image collections. Key research directions include joint image and video summarization, structural event modeling, and intelligent visual data mining for applications such as photo recommendation and robot navigation.

cosegmentationstoryline graphsunsupervised learninglarge-scale visual analysisregion detection

Research Overview

Papers
222
Total Citations
6,087
Papers (5y)
73
Primary Field
情報科学

Research Output Trend

Figures are computed from collected data and may differ slightly.

Publications per year (5y)
73total
2022
2023
2024
2025
2026
Citations per year (5y)
351total
20222023202420252026

Selected Papers

15
1
Article|276 citations·2011
Distributed cosegmentation via submodular optimization on anisotropic diffusion
Gunhee Kim, Eric P. Xing, Li Fei-Fei, Takeo Kanade
OA

The saliency of regions or objects in an image can be significantly boosted if they recur in multiple images. Leveraging this idea, cosegmentation jointly segments common regions from multiple images. In this paper, we propose CoSand, a distributed cosegmentation approach for a highly variable large-scale image collection. The segmentation task is modeled by temperature maximization on anisotropic heat diffusion, of which the temperature maximization with finite K heat sources corresponds to a K

Computer Vision and Pattern RecognitionComputer Science
2
Article|186 citations·2014
Joint Summarization of Large-Scale Collections of Web Images and Videos for Storyline Reconstruction
Gunhee Kim, Leonid Sigal, Eric P. Xing
OA

In this paper, we address the problem of jointly summarizing large sets of Flickr images and YouTube videos. Starting from the intuition that the characteristics of the two media types are different yet complementary, we develop a fast and easily-parallelizable approach for creating not only high-quality video summaries but also novel structural summaries of online images as storyline graphs. The storyline graphs can illustrate various events or activities associated with the topic in a form of

Computer Vision and Pattern RecognitionComputer Science
3
Article|140 citations·2012
On multiple foreground cosegmentation
Gunhee Kim, Eric P. Xing
OA

In this paper, we address a challenging image segmentation problem called multiple foreground cosegmentation (MFC), which concerns a realistic scenario in general Webuser photo sets where a finite number of K foregrounds of interest repeatedly occur cross the entire photo set, but only an unknown subset of them is presented in each image. This contrasts the classical cosegmentation problem dealt with by most existing algorithms, which assume a much simpler but less realistic setting where the sa

Computer Vision and Pattern RecognitionComputer Science
4
Article|100 citations·2009
Unsupervised Detection of Regions of Interest Using Iterative Link Analysis
Gunhee Kim, Antonio Torralba
DSpace@MIT (Massachusetts Institute of Technology)OA

This paper proposes a fast and scalable alternating optimization technique to de-tect regions of interest (ROIs) in cluttered Web images without labels. The pro-posed approach discovers highly probable regions of object instances by itera-tively repeating the following two functions: (1) choose the exemplar set (i.e. a small number of highly ranked reference ROIs) across the dataset and (2) refine the ROIs of each image with respect to the exemplar set. These two subproblems are formulated as ra

Computer Vision and Pattern RecognitionComputer Science
5
Article|78 citations·2014
Reconstructing Storyline Graphs for Image Recommendation from Web Community Photos
Gunhee Kim, Eric P. Xing
OA

In this paper, we investigate an approach for reconstructing storyline graphs from large-scale collections of Internet images, and optionally other side information such as friendship graphs. The storyline graphs can be an effective summary that visualizes various branching narrative structure of events or activities recurring across the input photo sets of a topic class. In order to explore further the usefulness of the storyline graphs, we leverage them to perform the image sequential predicti

Computer Vision and Pattern RecognitionComputer Science
6
Article|52 citations·2005
The autonomous tour-guide robot Jinny
Gunhee Kim, Woojin Chung, Kyung-Rock Kim, Munsang Kim, Sang-Mok Han, Richard H. Shinn

This paper explains a new tour-guide robot Jinny. The Jinny is developed by focusing on human robot interaction and autonomous navigation. In order to achieve reliable and safe navigation performance, an integrated navigation strategy is established based on the analysis of a robot's states and the decision making process of robot behaviors. According to the condition of environments, the robot can select its motion algorithm among four types of navigation strategy. Also, we emphasized the manag

Computer Vision and Pattern RecognitionComputer Science
7
Article|46 citations·2019
Video Question Answering with Spatio-Temporal Reasoning
Yunseok Jang, Yale Song, Chris Dongjoo Kim, Youngjae Yu, Youngjin Kim, Gunhee Kim
SJR Q1International Journal of Computer Vision
Computer Vision and Pattern RecognitionComputer Science
8
Article|41 citations·2013
Jointly Aligning and Segmenting Multiple Web Photo Streams for the Inference of Collective Photo Storylines
Gunhee Kim, Eric P. Xing
OA

With an explosion of popularity of online photo sharing, we can trivially collect a huge number of photo streams for any interesting topics such as scuba diving as an outdoor recreational activity class. Obviously, the retrieved photo streams are neither aligned nor calibrated since they are taken in different temporal, spatial, and personal perspectives. However, at the same time, they are likely to share common storylines that consist of sequences of events and activities frequently recurred w

Computer Vision and Pattern RecognitionComputer Science
9
Article|34 citations·2015
Ranking and retrieval of image sequences from multiple paragraph queries
Gunhee Kim, Seungwhan Moon, Leonid Sigal

We propose a method to rank and retrieve image sequences from a natural language text query, consisting of multiple sentences or paragraphs. One of the method's key applications is to visualize visitors' text-only reviews on TRIPADVISOR or YELP, by automatically retrieving the most illustrative image sequences. While most previous work has dealt with the relations between a natural language sentence and an image or a video, our work extends to the relations between paragraphs and image sequences

Computer Vision and Pattern RecognitionComputer Science
10
Article|30 citations·2015
Joint photo stream and blog post summarization and exploration
Gunhee Kim, Seungwhan Moon, Leonid Sigal

We propose an approach that utilizes large collections of photo streams and blog posts, two of the most prevalent sources of data on the Web, for joint story-based summarization and exploration. Blogs consist of sequences of images and associated text; they portray events and experiences with concise sentences and representative images. We leverage blogs to help achieve story-based semantic summarization of collections of photo streams. In the opposite direction, blog posts can be enhanced with

Computer Vision and Pattern RecognitionComputer Science
11
Article|29 citations·2007
Navigation Behavior Selection Using Generalized Stochastic Petri Nets for a Service Robot
Gunhee Kim, Woojin Chung
IEEE Transactions on Systems Man and Cybernetics Part C (Applications and Reviews)

<para xmlns:mml="http://www.w3.org/1998/Math/MathML" xmlns:xlink="http://www.w3.org/1999/xlink"> Appropriate design and control of behaviors of mobile robots are important for their successful autonomous navigation in a real dynamic environment. This paper proposes a formal selection framework of multiple navigation behaviors for a service robot. In the presented approach, modeling, analysis, and performance evaluation are carried out based on generalized stochastic Petri nets (GSPNs). By adopti

Computer Vision and Pattern RecognitionComputer Science
12
Article|23 citations·2008
Segmentation of Salient Regions in Outdoor Scenes Using Imagery and 3-D Data
Gunhee Kim, Daniel Huber, Martial Hebert
OA

This paper describes a segmentation method for extracting salient regions in outdoor scenes using both 3-D laser scans and imagery information. Our approach is a bottom- up attentive process without any high-level priors, models, or learning. As a mid-level vision task, it is not only robust against noise and outliers but it also provides valuable information for other high-level tasks in the form of optimal segments and their ranked saliency. In this paper, we propose a new saliency definition

Computer Vision and Pattern RecognitionComputer Science
13
Article|20 citations·2013
Time-sensitive web image ranking and retrieval via dynamic multi-task regression
Gunhee Kim, Eric P. Xing

In this paper, we investigate a time-sensitive image retrieval problem, in which given a query keyword, a query time point, and optionally user information, we retrieve the most relevant and temporally suitable images from the database. Inspired by recently emerging interests on query dynamics in information retrieval research, our time-sensitive image retrieval algorithm can infer users' implicit search intent better and provide more engaging and diverse search results according to temporal tre

Computer Vision and Pattern RecognitionComputer Science
14
Preprint|17 citations·2020
Augmenting Data for Sarcasm Detection with Unlabeled Conversation Context
Hankyol Lee, Youngjae Yu, Gunhee Kim
OA

We present a novel data augmentation technique, CRA (Contextual Response Augmentation), which utilizes conversational context to generate meaningful samples for training. We also mitigate the issues regarding unbalanced context lengths by changing the inputoutput format of the model such that it can deal with varying context lengths effectively. Specifically, our proposed model, trained with the proposed data augmentation technique, participated in the sarcasm detection task of FigLang2020, have

Artificial IntelligenceComputer Science
15
Article|14 citations·2008
Unsupervised modeling and recognition of object categories with combination of visual contents and geometric similarity links
Gunhee Kim, Christos Faloutsos, Martial Hebert
OA

This paper proposes a probabilistic approach for unsupervised modeling and recognition of object categories which combines two types of complementary visual evidence, visual contents and inter-connected links between the images. By doing so, our approach not only increases modeling and recognition performance but also provides possible solutions to several problems including modeling of geometric information, computational complexity, and the inherent ambiguity of visual words. Our approach can

Computer Vision and Pattern RecognitionComputer Science

Research Areas

Computer Vision and Pattern RecognitionArtificial IntelligenceSignal ProcessingInformation SystemsComputer Networks and CommunicationsPublic Health, Environmental and Occupational Health

Gunhee Kimの研究をNubintでさらに深く

この研究室の論文をアプリで開き、AIと共に読み、要約し、引用しましょう。