이 연구는 독일어-한국어 병렬코퍼스(German-Korean parallel corpus)의 구축을 최초로 시도하고 이를 활용할 수 있는 방법론과 모델을 개발함으로써 이것이 독일어 교육 및 통ㆍ번역, 사전 편찬, 독...

http://chineseinput.net/에서 pinyin(병음)방식으로 중국어를 변환할 수 있습니다.
변환된 중국어를 복사하여 사용하시면 됩니다.
https://www.riss.kr/link?id=G3752809
-
2006년
Korean
한국연구재단(NRF)
0
상세조회0
다운로드이 연구는 독일어-한국어 병렬코퍼스(German-Korean parallel corpus)의 구축을 최초로 시도하고 이를 활용할 수 있는 방법론과 모델을 개발함으로써 이것이 독일어 교육 및 통ㆍ번역, 사전 편찬, 독...
이 연구는 독일어-한국어 병렬코퍼스(German-Korean parallel corpus)의 구축을 최초로 시도하고 이를 활용할 수 있는 방법론과 모델을 개발함으로써 이것이 독일어 교육 및 통ㆍ번역, 사전 편찬, 독일어-한국어 비교 연구, 자동번역 등 여러 분야에서 유용하게 이용될 수 있는 기반을 마련할 목적에서 수행되었다. 본 연구는 1) 병렬코퍼스 구축을 위한 텍스트 선정 지침 연구, 2) 코퍼스 텍스트의 수집 및 OCR 작업, 3) 독-한 병렬코퍼스 구축을 위한 정렬 작업(목표치: 독일어, 한국어 각각 50만 어절), 4) 병렬코퍼스 검색도구 연구, 5) 병렬코퍼스의 활용방안 연구 등으로 목표가 세분화되어 수행되었다. 각 연구의 수행내용과 연구결과는 다음과 같이 요약할 수 있다. 1) 병렬코퍼스 구축을 위한 텍스트 선정 지침 연구: 채택한 양방향 병렬코퍼스 유형의 성격에 맞게 텍스트 선정 지침을 정하고 테스트 조사 및 수집을 진행하였다. 2) 텍스트를 기계 가독형 자료로 변환하는 작업은 OCR 방식과 직접 입수 방식을 취하였다. 3) 독-한 병렬코퍼스 구축을 위한 정렬 작업: 본 연구에서 채택한 병렬코퍼스 정렬(alignment) 작업은 원문과 번역문을 일일이 대조해 가며 정렬 소프트웨어나 텍스트 에디터의 도움으로 문장 대 문장 대응을 시켜 원문 텍스트와 번역문 텍스트를 한 문서에 통합시키는 방법이다. 이렇게 해서 2004년 12월부터 2005년 11월까지 독일어 약 74만 어절, 한국어 약 59만 어절 규모의 병렬코퍼스를 구축했다. 4) 병렬코퍼스 검색도구 연구: 본 연구에서는 독일어와 한국어를 동시에 잘 보여줄 수 있는 이중언어 콘코던스 프로그램을 조사해 보았으나, 기존의 상용 소프트웨어들의 문제를 확인하고, 유니코드와 정규표현이 지원되고, 검색 기능을 이용할 수 있는 EmEditor를 이중언어 콘코던스 프로그램으로 활용하였다. 5) 병렬코퍼스의 활용방안 연구: 본 연구에서는 표본적으로 코퍼스 기반 번역을 중심으로 병렬코퍼스의 구체적인 활용 가능성을 점검해 보았다. 먼저 본 연구과제에서 구축한 병렬 코퍼스의 일부를 직접 이용하여 특히 번역 방향이 한국어-독일어인 경우에 그 유용성을 확인해 보기 위한 번역 실험을 실시하였는데, 여기서 종래 번역 방식의 문제점과 병렬코퍼스의 유용성을 확인할 수 있었다. 이어서 병렬코퍼스를 통해 한국어-독일어 번역에 도움을 받을 수 있는 구체적인 방안들을 연구하였다.
다국어 초록 (Multilingual Abstract)
This research was aimed to initiate the construction of a German-Korean parallel corpus and develop the methodology for practical applications in the fields of German language teaching, translation and interpreting, dictionary compilation, comparative...
This research was aimed to initiate the construction of a German-Korean parallel corpus and develop the methodology for practical applications in the fields of German language teaching, translation and interpreting, dictionary compilation, comparative study of German-Korean, computer-aided translation, etc.
Parallel corpus is defined as a multi-lingual translation process with the display of one language aligned in parallel with its translation into another language. Under the circumstance that the globalization is being accelerated and political, economical, cultural exchanges are rapidly growing, parallel corpus contributes to the efficiency in information exchange between languages through practical applications in translation, interpreting, computer-aided translation, etc. Parallel corpus plays a significant role in the fields as contrastive analysis of different languages, dictionary compilations, literary stylistics, interpreter and translator training.
Significance of parallel corpus drew attention in overseas where vigorous efforts have been made for the construction and expansion of large-scale parallel corpora and their application. On the contrary, the importance of parallel corpus was realized only by a minority in Korea, the methodology being almost undeveloped. By the end of 1990s, a recognition on the necessity of parallel corpus led to a government funded project, the 21st Century Sejong Project, to start the study on the construction of Korean-English, Korean-Japanese, Korean-Chinese, Korean-Russian, Korean-French parallel corpora but Korean-German was not included in the project. This research is essential at this point as an initial attempt to construct a German-Korean parallel corpus and to study its application.
This research was originally proposed as a 3-year project of international cooperative research program in 2004 but the approval was given for a 1-year project, necessitating an adjustment of the research goal. The research has been carried out with its focus on: 1) study to establish a guideline for text selection, 2) collection of corpus texts and OCR works, 3) text alignment works for German-Korean parallel corpus construction (Goal: 500,000 words each from German and Korean), 4) study for searching tools of parallel corpus, 5) study on the applications of parallel corpus.