Skip to main content

남호성 교수

Nan Hoesung

고려대학교 영어영문학과 · 심리학

연구실 소개

남호성 교수의 연구실은 말의 운동 제어 모델과 음성 제어의 신경생물학적 기반을 탐구하며, 특히 말동작을 '자세 제어'로 해석하는 Task Dynamic 모델을 기반으로 자연어 음성의 비정형적 구조를 정량적으로 분석하는 데 전문성을 가진다. 음성의 비음성적 요소인 ‘게슈처 스코어’를 추출하고, 이를 기반으로 자연어 음성 데이터에 대한 정교한 주제도 분석 및 자동 음성 인식 성능 향상 기법을 개발한다. 최근에는 대규모 언어 모델의 도메인 특화 미세조정과 언어 습득 이론을 접목한 L2 어휘 학습 연구까지 확장하고 있다.

게슈처 스코어Task Dynamic 모델자연어 음성 분석자동 음성 인식L2 어휘 습득

연구 현황

논문 수
138
총 인용 수
1,908
최근 5년 논문
27
주요 분야
심리학

연구 성과 추이

표시된 성과는 수집된 데이터 기준으로 산출되며, 일부 차이가 있을 수 있습니다.

5개년 연도별 논문 게재 수
27총합
2021
2023
2024
2025
2026
5개년 연도별 피인용 수
23총합
20212023202420252026

주요 논문

15
1
book chapter|인용수 175·2009
Self-organization of syllable structure: a coupled oscillator model
Hosung Nam, Louis Goldstein, Elliot Saltzman
Experimental and Cognitive PsychologyPsychology
2
논문|인용수 88·2004
TADA: An enhanced, portable Task Dynamics model in M A T L A B
Hosung Nam, Louis Goldstein, Elliot Saltzman, Dani Byrd
SJR Q1The Journal of the Acoustical Society of America

A portable computational system called TADA was developed for the Task Dynamic model of speech motor control [Saltzman and Munhall, Ecol. Psychol. 1, 333–382 (1989)]. The model maps from a set of linguistic gestures, specified as activation functions with corresponding constriction goal parameters, to time functions for a set of model articulators. The original Task Dynamic code was ported to the (relatively) platform-independent MATLAB environment and includes a MATLAB version of the Haskins ar

Experimental and Cognitive PsychologyPsychology
3
논문|인용수 59·2016
MEASURES OF IMPLICIT KNOWLEDGE REVISITED
Jeong-eun Kim, Hosung Nam
SJR Q1Studies in Second Language Acquisition

Timed grammaticality judgment tests (TGJT) and oral elicited imitation tests (OEIT) are considered reliable and valid measures of implicit linguistic knowledge, but studies consistently observe better performances on the TGJT than the OEIT due to the different types of processing they require: comprehension for the TGJT and production for the OEIT. This study examines whether degree of access to implicit knowledge is a function of processing type. Results from a series of factor analyses suggest

Language and LinguisticsArts and Humanities
4
논문|인용수 50·2017
Phonetic drift in Spanish-English bilinguals: Experiment and a self-organizing model
Stephen J Tobin, Hosung Nam, Carol A. Fowler
SJR Q1Journal of PhoneticsOA
Experimental and Cognitive PsychologyPsychology
5
논문|인용수 33·2012
A procedure for estimating gestural scores from speech acoustics
Hosung Nam, Vikramjit Mitra, Mark Tiede, Mark Hasegawa‐Johnson, Carol Espy-Wilson, Elliot Saltzman, Louis Goldstein
SJR Q1The Journal of the Acoustical Society of AmericaOA

Speech can be represented as a constellation of constricting vocal tract actions called gestures, whose temporal patterning with respect to one another is expressed in a gestural score. Current speech datasets do not come with gestural annotation and no formal gestural annotation procedure exists at present. This paper describes an iterative analysis-by-synthesis landmark-based time-warping architecture to perform gestural annotation of natural speech. For a given utterance, the Haskins Laborato

Signal ProcessingComputer Science
6
논문|인용수 30·2013
Computational simulation of CV combination preferences in babbling
Hosung Nam, Louis M. Goldstein, Sara Giulivi, Andrea G. Levitt, D. H. Whalen
SJR Q1Journal of Phonetics
Experimental and Cognitive PsychologyPsychology
7
논문|인용수 12·2019
How do textual features of L2 argumentative essays differ across proficiency levels? A multidimensional cross-sectional study
Jeong-eun Kim, Hosung Nam
SJR Q1Reading and Writing
Literature and Literary TheoryArts and Humanities
8
논문|인용수 9·2017
The pedagogical relevance of processing instruction in second language idiom acquisition
Jeong-eun Kim, Hosung Nam
SJR Q1IRAL - International Review of Applied Linguistics in Language Teaching

Abstract This study explores the relevance and effectiveness of processing instruction in second language (L2) idiom learning by examining (1) whether structured input (SI) is more effective than non-SI, comprehension-based activities and (2) whether explicit information (EI) in addition to SI can facilitate L2 idiom learning. One hundred adult L2 English speakers were randomly assigned to one of six conditions: four groups who participated in SI activities in one of four EI conditions (i. e., n

Developmental and Educational PsychologyPsychology
9
논문|인용수 7·2023
Exploring the feasibility of fine-tuning large-scale speech recognition models for domain-specific applications: A case study on Whisper model and KsponSpeech dataset
Jungwon Chang, Hosung Nam
Phonetics and Speech SciencesOA

This study investigates the fine-tuning of large-scale Automatic Speech Recognition (ASR) models, specifically OpenAI’s Whisper model, for domain-specific applications using the KsponSpeech dataset. The primary research questions address the effectiveness of targeted lexical item emphasis during fine-tuning, its impact on domain-specific performance, and whether the fine-tuned model can maintain generalization capabilities across different languages and environments. Experiments were conducted u

Artificial IntelligenceComputer Science
10
논문|인용수 7·2010
A procedure for estimating gestural scores from natural speech
Hosung Nam, Vikramjit Mitra, Mark Tiede, Elliot Saltzman, Louis Goldstein, Carol Espy-Wilson, Mark Hasegawa‐Johnson

Speech can be represented as a constellation of constricting events, gestures, which are defined at distinct vocal tract sites, in the form of a gestural score. Gestures and their output trajectories, tract variables, which are available only in synthetic speech, have recently been shown to improve automatic speech recognition (ASR) performance. In this paper we propose an iterative analysis-by-synthesis landmark based time-warping architecture to obtain gestural scores for natural speech. Given

Signal ProcessingComputer Science
11
논문|인용수 6·2013
Hearing tongue loops: Perceptual sensitivity to acoustic signatures of articulatory dynamics
Hosung Nam, Christine Mooshammer, Khalil Iskarous, D. H. Whalen
SJR Q1The Journal of the Acoustical Society of AmericaOA

Previous work has shown that velar stops are produced with a forward movement during closure, forming a forward (anterior) loop for a VCV sequence, when the preceding vowels are back or mid. Are listeners aware of this aspect of articulatory dynamics? The current study used articulatory synthesis to examine how such kinematic patterns are reflected in the acoustics, and whether those acoustic patterns elicit different goodness ratings. In Experiment I, the size and direction of loops was modulat

Experimental and Cognitive PsychologyPsychology
12
논문|인용수 5·2020
Identification of English vowels by non-native listeners: Effects of listeners’ experience of the target dialect and talkers’ language background
Shinsook Lee, Jaekoo Kang, Hosung Nam
SJR Q1Second language Research

This study investigates how second language (L2) listeners’ perception is affected by two factors: the listeners’ experience with the target dialect – North American English (NAE) vs. Standard Southern British English (SSBE) – and talkers’ language background: native vs. non-native talkers; i.e. interlanguage speech intelligibility benefit (ISIB) talker effects. Two groups of native-Korean-speaking listeners with different target English dialects – L1-Korean listeners of English as a second lang

Experimental and Cognitive PsychologyPsychology
13
논문|인용수 4·2021
Hyperparameter experiments on end-to-end automatic speech recognition*
Hyungwon Yang, Hosung Nam
Phonetics and Speech SciencesOA

End-to-end (E2E) automatic speech recognition (ASR) has achieved promising performance gains with the introduced self-attention network, Transformer. However, due to training time and the number of hyperparameters, finding the optimal hyperparameter set is computationally expensive. This paper investigates the impact of hyperparameters in the Transformer network to answer two questions: which hyperparameter plays a critical role in the task performance and training speed. The Transformer network

Artificial IntelligenceComputer Science
14
논문|인용수 3·2024
Automatic speech recognition (ASR) for the diagnosis of pronunciation of speech sound disorders in Korean children
Taekyung Ahn, Yeonjung Hong, Younggon Im, Do Hyung Kim, Dayoung Kang, Joo Won Jeong, Jae Won Kim, Min Jung Kim, Ah‐Ra Cho, Hosung Nam, Dae‐Hyun Jang
SJR Q1Clinical Linguistics & Phonetics

This study presents a model of automatic speech recognition (ASR) that is designed to diagnose pronunciation issues in children with speech sound disorders (SSDs) to replace manual transcriptions in clinical procedures. Because ASR models trained for general purposes mainly predict input speech into standard spelling words, well-known high-performance ASR models are not suitable for evaluating pronunciation in children with SSDs. We fine-tuned the wav2vec2.0 XLS-R model to recognise words as the

Artificial IntelligenceComputer Science
15
논문|인용수 3·2010
A procedure for estimating gestural scores from articulatory data.
Hosung Nam, Vikramjit Mitra, Mark Tiede, Elliot Saltzman, Louis Goldstein, Carol Espy-Wilson, Mark Hasegawa‐Johnson
SJR Q1The Journal of the Acoustical Society of America

Speech can be represented as a set of discrete vocal tract constriction gestures (gestural score) defined at functionally distinct speech organs [tract variables (TVs)]. Using such gestures as sub-word units in an ASR system, variation in speech arising from coarticulation and reduction can be addressed. Since there is a lack of test corpora annotated with gestural scores, we develop a semi-automatic procedure for estimating and annotating gestural scores from natural speech databases using the

Experimental and Cognitive PsychologyPsychology

대표 연구 분야

Experimental and Cognitive PsychologyArtificial IntelligenceSignal ProcessingDevelopmental and Educational PsychologyLanguage and LinguisticsMechanical Engineering

남호성 교수의 연구를 Nubint에서 더 깊이 살펴보세요

이 연구실의 논문을 앱에서 열어 AI와 함께 읽고, 핵심을 요약하고, 내 글에 인용하세요.