김동진 교수
Dongjin Kim
한양대학교 데이터사이언스학부 · 컴퓨터과학
연구실 소개
김동진 교수의 연구실은 시각적 관계 기반 이미지 이해와 다중_caption 생성 기술을 핵심으로 삼고 있습니다. 특히 '관계적 캡션 생성'(relational captioning) 기반의 다중 정보적 이미지 설명 생성 기술을 개발하여, 객체 간의 상호관계를 정확하고 풍부하게 표현하는 데 초점을 맞추고 있습니다. 또한, 다중 작업 학습을 활용한 인간 행동 이해 및 3D 물리 시뮬레이션 기반 충돌 탐지 기술 등 응용 분야로의 확장도 진행 중입니다. 연구는 자연어 생성, 컴퓨터 비전, 인공지능의 융합을 통해 실질적인 지능형 시스템 구현에 기여하고자 합니다.
연구 현황
연구 성과 추이
표시된 성과는 수집된 데이터 기준으로 산출되며, 일부 차이가 있을 수 있습니다.
주요 논문
15Our goal in this work is to train an image captioning model that generates more dense and informative captions. We introduce "relational captioning," a novel image captioning task which aims to generate multiple captions with respect to relational information between objects in an image. Relational captioning is a framework that is advantageous in both diversity and amount of information, leading to image understanding based on relationships. Part-of-speech (POS, i.e. subject-object-predicate ca
This paper presents an event-driven approach that efficiently detects collisions among multiple ballistic spheres moving in the 3D space. Adopting a hierarchical uniform space subdivision scheme, we are able to trace the trajectories of spheres and their time-varying spatial distribution. We identify three types of events to detect the sequence of all collisions during our simulation: collision, entering, and leaving. The first type of event is due to actual collisions, and the other two types o
We introduce dense relational captioning, a novel image captioning task which aims to generate multiple captions with respect to relational information between objects in a visual scene. Relational captioning provides explicit descriptions for each relationship between object combinations. This framework is advantageous in both diversity and amount of information, leading to a comprehensive image understanding based on relationships, e.g., relational proposal generation. For relational understan
Human behavior understanding is arguably one of the most important mid-level components in artificial intelligence. In order to efficiently make use of data, multi-task learning has been studied in diverse computer vision tasks including human behavior understanding. However, multitask learning relies on task specific datasets and constructing such datasets can be cumbersome. It requires huge amounts of data, labeling efforts, statistical consideration etc. In this paper, we leverage existing si
ADVERTISEMENT RETURN TO ISSUEPREVCommunication to the...Communication to the EditorNEXTSynthesis of a New Class of Processable Electroluminescent Poly(cyanoterephthalylidene) Derivative with a Tertiary Amine LinkageDong-Jin Kim, Sung-Hyun Kim, Taehyoung Zyung, Jang-Joo Kim, Iwhan Cho, and Sam Kwon ChoiView Author Information Department of Advanced Materials Engineering, Korea Advanced Institute of Science and Technology, P.O. Box 201, Cheongryang, Seoul, Korea, Department of Chemistry, Korea Adv
We designed and fabricated a high performance spring-type piezoelectric energy harvester that selectively collects current from the inner part of a spring shell.
A common problem in the task of human-object interaction (HOI) detection is that numerous HOI classes have only a small number of labeled examples, resulting in training sets with a long-tailed distribution. The lack of positive labels can lead to low classification accuracy for these classes. Towards addressing this issue, we observe that there exist natural correlations and anti-correlations among human-object interactions. In this paper, we model the correlations as action co-occurrence matri
A common problem in human-object interaction (HOI) detection task is that numerous HOI classes have only a small number of labeled examples, resulting in training sets with a long-tailed distribution. The lack of positive labels can lead to low classification accuracy for these classes. Towards addressing this issue, we observe that there exist natural correlations and anti-correlations among human-object interactions. In this paper, we model the correlations as action co-occurrence matrices and
Abstract A new type of processable electroluminescent polymer containing tertiary amine linkage was prepared by typical Wittig reaction between bis(4-formylphenyl) n-butylamine and diphosphonium salt. The resulting polymer was highly soluble in common organic solvents so that it could be spun-cast onto glass plate coated with ITO electrode to give highly transparent homogeneous thin film. The molecular weight of the polymer determined by gel permeation chromatography using polystyrene standards
The software industry has been increasingly aware of the need to employ copyright protection techniques against software piracy or illegal distribution. Recently, as the value of software as intellectual property grows, so does the rate of global software piracy. Software piracy is a real threat to software industry. Moreover, Pirated software is more vulnerable to attacks since such software is less likely to be supported with critical security patches and updates. In this paper, we propose a b
A software birthmark is unique, as certain native characteristics of a program, hence can be used to measure the similarity between programs. In general, a static software birthmark does not need program execution, but is more vulnerable to attacks by semantic-preserving transformations. A dynamic software birthmark is applicable to packed executables, but cannot cover all the possible program paths. In this paper, we propose a novel effective technique to measure the similarity of Microsoft Win
This paper addresses the problem in Web page ranking of effectively combining link and content information with efficiency high enough to be applicable to real-world search engines. Unlike previous surfer models, our approach is based on the viewpoint of a Web page author. Based on this viewpoint, we formulate the concept of contribution score, which indicates the amount to which a term in each page is utilized by other pages. To improve efficiency without loss of effectiveness, we exploit the e
대표 연구 분야
김동진 교수의 연구를 Nubint에서 더 깊이 살펴보세요
이 연구실의 논문을 앱에서 열어 AI와 함께 읽고, 핵심을 요약하고, 내 글에 인용하세요.