RISS 학술연구정보서비스

검색

인기 검색어

    다국어 입력

    http://chineseinput.net/에서 pinyin(병음)방식으로 중국어를 변환할 수 있습니다.

    변환된 중국어를 복사하여 사용하시면 됩니다.

    예시)
    • 中文 을 입력하시려면 zhongwen을 입력하시고 space를누르시면됩니다.
    • 北京 을 입력하시려면 beijing을 입력하시고 space를 누르시면 됩니다.
    닫기

    한국어의 확장된 어휘 단위 연구 : 분포와 의미를 중심으로 = Distribution and Semantic Functions on Korean Extended Lexical Units

    한글로보기

    https://www.riss.kr/link?id=T13847602

    • 0

      상세조회
    • 0

      다운로드
    서지정보 열기
    • 내보내기
    • 내책장담기
    • 공유하기
    • 오류접수

    부가정보

    다국어 초록 (Multilingual Abstract) kakao i 다국어 번역

    The objective of this study is twofold. First, it aims to investigate the extent to which a Korean speaker makes use of Extended Lexical Units (ELU), that is, what its distribution is, by taking a corpus-driven approach. Second, it examines the semantic functions of the identified Korean ELU in their context. Analysing the properties of ELU with respect to distribution and semantic functions does not only imply to consider the properties of ELU as they appear in the common Korean language use, but it also means to examine the various aspects of these ELU by defining as parameters a given community and the formal textual structures of novels, newspapers, and academic papers.
    ELU is the generic term for the semantic function units (that is, beyond individual words) which are assigned enunciative meanings in language use, such as word sequences, contextual meanings, and grammatical categories. In this study, are considered as core ELU the word sequences that both appear in at least five texts and occur at least ten times in a million words.
    Chapter 2 examines the significance of corpus linguistics as theory which constitutes the theoretical and methodological background of this study. Furthermore, as ELU are integral parts of language use, they are recognized as crucial semantic units in corpus linguistics as well as essential elements that can relate to traditional linguistics.
    Chapter 3 discusses the methodological issues encountered in this study. The corpora analysed here are the Sejong spoken corpus and the Sejong written corpus both with part-of-speech annotations, as well as the Kyungpook National University corpus of academic papers. Unlike Indo-European languages in which orthographically spaced units generally correspond to words, the orthographically spaced units of the Korean language are called ecel which are semantic units combined with functional units such as case-markers or word-endings and present all the typological features of agglutinative languages. Thus, in order to detect the components of the Korean ELU it has been necessary to use part-of-speech annotated corpora. In addition, to obtain a richer and refined list, as regards case-markers and word-endings, variant forms (allomorphs, allophones, and contracted forms) have been merged with the basic forms.
    Chapter 4 looks into the various aspects of the ELU distribution. The analysis of the ELU occurring in the Sejong spoken and written corpora sheds light on the distributive aspects of the ELU in the common language use of native Korean speakers. If we consider N-grams of size three to five, they are not only more varied in types but they also are more frequently used in spoken rather than written genres. The rate of three word n-gram use in both spoken and written languages reaches respectively 55% and 48% of the language use to the maximum and 36% and 33% after deducting the duplicates. Besides, in order to examine the distributive aspects of the ELU that appear in the language use of a given community, the distribution of n-grams in the three registers of novels, newspapers, and academic papers has been investigated. Compared to common language use, these n-grams display a generally more varied typology as well as a higher usage frequency.
    Chapter 5 analyzes the meanings which the ELU uncover in context according to the metafunctional system of language that consists of the conceptual function, the textual function, and the interpersonal function, as described by Halliday & Matthiessen (2014). In both spoken and written languages, the conceptual function appears to be more prominent in short n-grams, while the interpersonal function is more prevalent in longer n-grams.
    As to examine the semantic features of ELU depending on the register, this study focuses on key n-grams by applying the statistical concept of keyness. In the case of novels, the units that ensure the textual cohesiveness correspond to negative key n-grams, whereas in newspapers time reference and quotation expressions reveal positive key n-grams. As for academic articles, expressions suggesting the intention or attitude of the writer correspond to positive key n-grams which show a high level of keyness.
    Chapter 6 investigates a number of case examples selected among the Korean ELU identified by taking a corpus-driven approach so as to discuss the general contribution of this study to theoretical linguistics as well as applied linguistic.
    By defining as parameters the speaker as main agent of language use, the register as the context to which the speaker belongs, and a text’s structural context, this study has established a list of the ELU which most frequently occur in the Korean language use and investigated their distribution as well as their semantic functions. The list established by this study is a list of semantic units which are acquired by the native Korean speaker through language use. For all speakers who acquired the Korean language or are still at a learning stage, the meaning of these ELU includes a great deal of semantic units that are hard to understand by combining the individual meanings of their constituents. Therefore, the results yielded here are believed to be a great help for both native speakers who acquired Korean and learners of Korean as a second language, and to put forward new perspectives in lexicography which until now has only focused on rarely used phrases.
    번역하기

    The objective of this study is twofold. First, it aims to investigate the extent to which a Korean speaker makes use of Extended Lexical Units (ELU), that is, what its distribution is, by taking a corpus-driven approach. Second, it examines the semant...

    The objective of this study is twofold. First, it aims to investigate the extent to which a Korean speaker makes use of Extended Lexical Units (ELU), that is, what its distribution is, by taking a corpus-driven approach. Second, it examines the semantic functions of the identified Korean ELU in their context. Analysing the properties of ELU with respect to distribution and semantic functions does not only imply to consider the properties of ELU as they appear in the common Korean language use, but it also means to examine the various aspects of these ELU by defining as parameters a given community and the formal textual structures of novels, newspapers, and academic papers.
    ELU is the generic term for the semantic function units (that is, beyond individual words) which are assigned enunciative meanings in language use, such as word sequences, contextual meanings, and grammatical categories. In this study, are considered as core ELU the word sequences that both appear in at least five texts and occur at least ten times in a million words.
    Chapter 2 examines the significance of corpus linguistics as theory which constitutes the theoretical and methodological background of this study. Furthermore, as ELU are integral parts of language use, they are recognized as crucial semantic units in corpus linguistics as well as essential elements that can relate to traditional linguistics.
    Chapter 3 discusses the methodological issues encountered in this study. The corpora analysed here are the Sejong spoken corpus and the Sejong written corpus both with part-of-speech annotations, as well as the Kyungpook National University corpus of academic papers. Unlike Indo-European languages in which orthographically spaced units generally correspond to words, the orthographically spaced units of the Korean language are called ecel which are semantic units combined with functional units such as case-markers or word-endings and present all the typological features of agglutinative languages. Thus, in order to detect the components of the Korean ELU it has been necessary to use part-of-speech annotated corpora. In addition, to obtain a richer and refined list, as regards case-markers and word-endings, variant forms (allomorphs, allophones, and contracted forms) have been merged with the basic forms.
    Chapter 4 looks into the various aspects of the ELU distribution. The analysis of the ELU occurring in the Sejong spoken and written corpora sheds light on the distributive aspects of the ELU in the common language use of native Korean speakers. If we consider N-grams of size three to five, they are not only more varied in types but they also are more frequently used in spoken rather than written genres. The rate of three word n-gram use in both spoken and written languages reaches respectively 55% and 48% of the language use to the maximum and 36% and 33% after deducting the duplicates. Besides, in order to examine the distributive aspects of the ELU that appear in the language use of a given community, the distribution of n-grams in the three registers of novels, newspapers, and academic papers has been investigated. Compared to common language use, these n-grams display a generally more varied typology as well as a higher usage frequency.
    Chapter 5 analyzes the meanings which the ELU uncover in context according to the metafunctional system of language that consists of the conceptual function, the textual function, and the interpersonal function, as described by Halliday & Matthiessen (2014). In both spoken and written languages, the conceptual function appears to be more prominent in short n-grams, while the interpersonal function is more prevalent in longer n-grams.
    As to examine the semantic features of ELU depending on the register, this study focuses on key n-grams by applying the statistical concept of keyness. In the case of novels, the units that ensure the textual cohesiveness correspond to negative key n-grams, whereas in newspapers time reference and quotation expressions reveal positive key n-grams. As for academic articles, expressions suggesting the intention or attitude of the writer correspond to positive key n-grams which show a high level of keyness.
    Chapter 6 investigates a number of case examples selected among the Korean ELU identified by taking a corpus-driven approach so as to discuss the general contribution of this study to theoretical linguistics as well as applied linguistic.
    By defining as parameters the speaker as main agent of language use, the register as the context to which the speaker belongs, and a text’s structural context, this study has established a list of the ELU which most frequently occur in the Korean language use and investigated their distribution as well as their semantic functions. The list established by this study is a list of semantic units which are acquired by the native Korean speaker through language use. For all speakers who acquired the Korean language or are still at a learning stage, the meaning of these ELU includes a great deal of semantic units that are hard to understand by combining the individual meanings of their constituents. Therefore, the results yielded here are believed to be a great help for both native speakers who acquired Korean and learners of Korean as a second language, and to put forward new perspectives in lexicography which until now has only focused on rarely used phrases.

    더보기

    목차 (Table of Contents)

    • 1. 서론 1
    • 1.1. 연구 목적 및 의의 1
    • 1.2. 연구 대상 5
    • 1.3. 선행 연구 검토 10
    • 1.4. 논의의 구성 14
    • 1. 서론 1
    • 1.1. 연구 목적 및 의의 1
    • 1.2. 연구 대상 5
    • 1.3. 선행 연구 검토 10
    • 1.4. 논의의 구성 14
    • 2. 이론적 배경 16
    • 2.1. 이론으로서의 말뭉치 언어학 16
    • 2.2. 말뭉치 언어학 내에서의 확장된 어휘 단위 20
    • 3. 확장된 어휘 단위 연구의 방법론 26
    • 3.1. 말뭉치의 구성 26
    • 3.2. 확장된 어휘 단위의 구성 요소에 대한 한정 28
    • 3.2.1. 확장된 어휘 단위 분석의 기본 단위 28
    • 3.2.2. 분석 단위의 ‘기본형’과 ‘변이형’ 33
    • 3.3. 확장된 어휘 단위의 분포 기준 37
    • 3.4. 확장된 어휘 단위 분석을 위한 도구의 설계 42
    • 4. 확장된 어휘 단위의 분포적 양상 45
    • 4.1. 구어 및 문어 장르별 n-gram의 분포적 양상 45
    • 4.1.1. 장르별 n-gram의 편재성 46
    • 4.1.2. 장르별 n-gram의 형태?통사적 특성 51
    • 4.2. 사용역별 n-gram의 분포적 양상 68
    • 4.2.1. 대상 사용역의 개요 68
    • 4.2.2. 사용역별 n-gram의 편재성 72
    • 5. 확장된 어휘 단위의 의미 기능적 특성 82
    • 5.1. 확장된 어휘 단위의 의미 기능적 분류 체계 82
    • 5.2. 구어 및 문어 장르에서 나타나는 n-gram의 의미 기능적 양상 87
    • 5.3. 사용역별 핵심 n-gram의 의미 기능적 양상 94
    • 5.3.1. 사용역별 핵심 n-gram의 추출 94
    • 5.3.2. 사용역별 핵심 n-gram의 의미 기능 97
    • 5.3.3. 학술 논문의 텍스트 구조에 따른 핵심 n-gram의 의미 기능 116
    • 6. 확장된 어휘 단위의 사례 연구 121
    • 6.1. 단어의 연접 범주와 의미 기능: 것/거 * 121
    • 6.2. 사용역별 의미 기능: ‘*와 같은’ 125
    • 6.3. 맥락적 의미 기능: ‘-다고/라고 할 수 있-’ 128
    • 6.4. 사례 분석에서 나타나는 확장된 어휘 단위 연구의 의의 132
    • 7. 결론 135
    • 참고문헌 143
    • 영문초록 151
    • 부록 1: 조사, 어미 변이형 통합 목록 154
    • 부록 2: 구어 n-gram 목록(3gram, 5gram) 157
    • 부록 3: 문어 n-gram 목록(3gram, 5gram) 171
    • 부록 4: 소설 n-gram 빈도 및 핵심도 비교(3gram, 5gram) 184
    • 부록 5: 신문 n-gram 빈도 및 핵심도 비교(3gram, 5gram) 192
    • 부록 6: 학술 논문 n-gram 빈도 및 핵심도 비교(3gram, 5gram) 200
    더보기

    참고문헌 (Reference)

    1. 연어연구, 김진해, 서울: 한국문화사, , 2000

    2. 우리말본, 최 현배, 서울: 정음문화사, , 1937

    3. 인지의미론, 임지룡, 탑, 서울: 탑출판사, , 1997

    4. 한글공동체, 이상규, 박문사, 서울: 박문사, , 2014

    5. 국어 어휘론, 김종택, 서울: 탑출판사, , 1993

    6. 국어 의미론, 임지룡, 서울: 탑출판사, , 1992

    7. 논항과 부가어, 유현경, 우리말글 연구 1, 175-196, , 1994

    8. 시제, 상, 양태, 박진호, 국어학 60, 289-322, , 2011

    9. 국어문법론강의, 이익섭, 채완, 서울: 학연사, , 2000

    10. 표준국어문법론, 남기심, 탑, 서울: 탑출판사, , 1993

    1. 연어연구, 김진해, 서울: 한국문화사, , 2000

    2. 우리말본, 최 현배, 서울: 정음문화사, , 1937

    3. 인지의미론, 임지룡, 탑, 서울: 탑출판사, , 1997

    4. 한글공동체, 이상규, 박문사, 서울: 박문사, , 2014

    5. 국어 어휘론, 김종택, 서울: 탑출판사, , 1993

    6. 국어 의미론, 임지룡, 서울: 탑출판사, , 1992

    7. 논항과 부가어, 유현경, 우리말글 연구 1, 175-196, , 1994

    8. 시제, 상, 양태, 박진호, 국어학 60, 289-322, , 2011

    9. 국어문법론강의, 이익섭, 채완, 서울: 학연사, , 2000

    10. 표준국어문법론, 남기심, 탑, 서울: 탑출판사, , 1993

    11. 한국어 연어 연구, 임근석, 월인, 서울: 월인, , 2010

    12. 『우리말 문법론』, 고영근, 집문당, 파주: 집문당, , 2008

    13. 의미, 텍스트, 교육, 신명선, 서울: 한국문화사, , 2008

    14. 통사론과 통사 단위, 임동훈, 어학연구 제31권 제1호, 87-138, , 1995

    15. 한국어 지시사 연구, 민경모, 연세대학교 대학원, 연세대학교 박사학위 논문, , 2008

    16. 국어의 관용 표현 연구, 문금현, 서울: 태학사, , 1999

    17. 인문 지식․정보의 미래, 이상규, 서울: 박문사, , 2015

    18. 한국어 동사구문의 연구, 홍재성, 서울: 탑출판사, , 1987

    19. 한국어의 구어와 말뭉치, 서상규, 한국어 교육 , 24-3, 71-107, , 2013

    20. 국어 어휘부와 단어 형성, 송원용, 서울: 태학사, , 2005

    21. 한국어 구어 말뭉치 연구, 서상규, 한국문화사, 서울: 한국문화사, , 2013

    22. ?담화와 문법 그리고 의미?, 정희자, 한국문화사, 서울: 한국문화사, , 2009

    23. 어휘의미론의 흐름과 특성, 임지룡, 한말연구 31, 195-227, , 2012

    24. 어미 '-다고'의 의미와 용법, 유현경, 배달말 31, 99-122, , 2002

    25. 어휘부 등재소와 복합 구성, 문병열, 국어학 69, 135-166, , 2014

    26. 핵심어 분석의 절차와 쟁점, 이수진, 남길임, 텍스트언어학 32, 98-212, , 2012

    27. 한국어 텍스트의 구성소 분석, 황미향, 한국텍스트언어학회, 텍스트언어학 6, 187-207, , 1999

    28. 한국어의 정형화된 표현 연구, 최준, 송현주, 남길임, 담화·인지 언어학회, 담화와 인지 제17권 2 호, 163-190, , 2010

    29. 경험 동사의 의미적 운율 연구, 최준, 한국사전학 제18호, 209-226, , 2011

    30. 띄어쓰기에 관한 몇 가지 문제, 이선웅, 국어국문학 134, 123-153, , 2003

    31. 「시간부사의 문장의미 구성」, 임채훈, 한국어의미학회, 한국어 의미학 12, 155-170, , 2003

    32. 국어의 어휘부 사전에 대한 연구, 시정곤, 한국현대언어학회, 언어연구 17-1, 163-184, , 2001

    33. 유추에 의한 복합명사 형성 연구, 채현식, 서울: 태학사, , 2003

    34. 국어 동기화의 인지언어학적 탐색, 송현주, 한국문화사, 서울: 한국문화사, , 2015

    35. 「국어의 문장 제시어에 대하여」, 李善雄, 한국어문교육연구회, 어문연구 33-1, 59-84, , 2005

    36. 국어의 동사연결 구성에 대한 연구, 강현화, 서울: 한국문화사, , 1998

    37. 어휘의 공기 경향성과 의미적 운율, 남길임, 한글 제298호, 135-164, , 2012

    38. 연어를 이용한 어휘 교육 방안 연구, 한송화, 한국어 교육 15권 3 호, 295-318, , 2004

    39. 「국어 상용어구의 의미구성 연구」, 정수진, 담화·인지 언어학회, 담화와 인지 제9권 2호, 215-238, , 2002

    40. ‘아니다’의 사용패턴과 부정의 의미, 남길임, 한국어 의미학 33, 41-65, , 2010

    41. 「국어 단어의 형태․통사론적 연구」, 최형용, 서울대 박사학위 논문, , 2002

    42. [체언+용언]꼴의 연어 구성에 대한 연구, 강현화, 사전 편찬학 연구 8, 191-224, , 1998

    43. 한국어의 동사와 문법요소의 결합 양상, 박진호, 서울대학교 박사학위 논문, , 2003

    44. 말뭉치와 언어연구: 외국의 사례와 경향, 최재웅, 한국어학 63, 71-102, , 2014

    45. 학술 텍스트에 나타난 핵심 구문의 추출, 최준, 남길임, 어문론총 60호, 65-92, , 2014

    46. 한국어 관용 표현의 정보화와 전산 처리, 이동혁, 서울: 역락, , 2007

    47. 한국어 빈도 사전 편찬을 위한 기초 연구, 안의정, 한국사전학 20, 234-258, , 2012

    48. 「‘나름, 때문, 마련’의 문법적 기능」, 송창선, 문학과언어연구회, 문학과 언어 16-1, 59-80, , 1995

    49. 「자연언어처리를 위한 관용표현 연구」, 김한샘, 한국어의미학회, 한국어 의미학 13, 43-67, , 2003

    50. 국어에 내재한 도상성의 양상과 의미 특성, 임지룡, 한글 266, 169-205, , 2004

    51. 한국어 연어의 개념과 그 통사·의미적 성격, 임홍빈, 국어학 39, 279-311, , 2002

    52. 한국어 연어 정보의 분석 , 응용에 관한 연구, 최호철, 강범모, 흥종선, 한국어학회, 한국어학 11, 73-158, , 2000

    53. 어휘의 텍스트 형성 기능과 어휘 지도의 방향, 황미향, 언어과학회, 언어과학연구 31, 297-318, , 2004

    54. 통계적 방법을 이용한 문법적 연어 후보 추출, 임근석, 한국어학 45권, 305-333, , 2009

    55. ‘-고 있-’과 ‘-어 있-’의 기능과 의미 연구, 송창선, 언어과학연구 62, 179-204, , 2012

    56. 국어 어휘범주의 기본층위 탐색 및 의미특성 연구, 임지룡, 담화와 인지 18-1, 153-182, , 2011

    57. 국어사전에서의 구어 어휘 선정과 기술 방안 연구, 안의정, 서울: 한국문화사, , 2009

    58. 「‘단어결합’과 ‘단어어울림’에 대한 고찰」, 서상규, 연세대학교 국학연구원, 동방학지 98, 419-468, , 1997

    59. 「통계에 기반한 한국어 연어 결합 측정의 평가」, 이두행, 연세대학교 석사학위 논문, , 2010

    60. 「한국어 어순 변이 경향과 그 요인에 대한 연구」, 신서인, 국어학 50, 213-240, , 2007

    61. 확장된 어휘 단위에 대한 연구 동향과 한국어의 기술, 남길임, 한국사전학 18, 73-98, , 2011

    62. 「‘이/가’, ‘을/를’의 비전형적인 분포와 기능」, 신서인, 국어학 69, 69-103, , 2014

    63. 「학문 목적 한국어 교육용 말뭉치 설계 방안 연구」, 민경모, 한국언어문화교육학회, 언어와 문화 6-1, 137-156, , 2010

    64. 국어 학술텍스트에 드러난 헤지(Hedge) 표현에 대한 연구, 신명선, 배달말 38, 151-180, , 2006

    65. 「언어 유형론적 관점에서 본 한국어의 연속동사구문」, 이선웅, 한국현대언어학회, 언어연구 27-1. 165-182, , 2011

    66. 「핵심어 분석을 통한 영어 성별 어휘 사용 양상 연구」, 전지은, 고려대학교 박 사학위논문, , 2010

    67. ‘이론으로서의 말뭉치언어학’에 대한연구 현황과 쟁점, 남길임, 한국어 의 미학 46, 163-187, , 2014

    68. 「한국어 문화문법(ethno-grammar)의 설정 가능성에 대하여」, 임채훈, 국제한국어교육학회, 한국어 교육 22-4, 109-129, , 2011

    69. 「한국어 정형표현 연구-대규모 말뭉치 분석을 중심으로-」, 장석배, 연세대학교 박사학위 논문, , 2014

    70. 둥지 밖의 언어(지혜의 심장, 우리말 사전 지식의 진화를 위하여), 이상규, 생각의나무, 서울: 생각의나무, , 2008

    71. 21세기 세종계획 전자사전구축분과 연어사전의 정보구조와기술내용, 임홍빈, 임근석, 한국사전학 4호, 99-130, , 2004

    72. “-ㄹ 예정이다”류 구문 연구-말뭉치 용례의 통사 정보 분석을 중심으 로-, 남길임, 한국어학회, 한국어학 22, 69-94, , 2004

    73. 부표제어의 범위와 유형: 속담․관용표현․연어․패턴․상투표현․자유표현의 기술, 남길임, 한국사전학 9, 143-161, , 2007

    74. 「말뭉치 분석에 기반을 둔 낱말 빈도의 조사와 그 응용-‘연세 말뭉치’ 를 중심으로」, 서상규, 한글학회, 한글 242, 225-270, , 1998

    75. 「쓰기 영역의 실천적 양상 연구-고등학교 ‘작문’과 대학 ‘교양작문’을 중심으로-」, 김선정, 국어교과교육학회, 국어교과교육연구 8, 107-130, , 2004

    76. 한국어 정형화된 표현의 분석 단위에 대한 연구: 형태 기반 분석과 어절 기반 분석의 비교를 중심으로, 남길임, 담화와 인지 제20권 1호, 113-136, , 2013

    더보기

    분석정보

    View

    상세정보조회

    0

    Usage

    원문다운로드

    0

    대출신청

    0

    복사신청

    0

    EDDS신청

    0

    동일 주제 내 활용도 TOP

    더보기

    주제

    연도별 연구동향

    연도별 활용동향

    연관논문

    연구자 네트워크맵

    공동연구자 (7)

    유사연구자 (20) 활용도상위20명

    이 자료와 함께 이용한 RISS 자료

    나만을 위한 추천자료

    해외이동버튼