최근 대규모 언어모델(LLM, Large Language Model)은 인간 수준의 언어 이해 능력을 기반으로 다양한 응용 분야에서 활용되고 있으나, 대부분 범용적인 대화형 상호작용을 목적으로 설계되어 특정 ...

http://chineseinput.net/에서 pinyin(병음)방식으로 중국어를 변환할 수 있습니다.
변환된 중국어를 복사하여 사용하시면 됩니다.
https://www.riss.kr/link?id=T17354333
창원 : 경남대학교 대학원, 2026
2026
한국어
LLM ; LoRA ; 음성인식 ; 파인튜닝 ; Whisper ; automatic speech recognition ; fine-tuning ; Whisper
경상남도
57 ; 26 cm
지도교수: 이세한
I804:48002-200000945196
0
상세조회0
다운로드최근 대규모 언어모델(LLM, Large Language Model)은 인간 수준의 언어 이해 능력을 기반으로 다양한 응용 분야에서 활용되고 있으나, 대부분 범용적인 대화형 상호작용을 목적으로 설계되어 특정 ...
최근 대규모 언어모델(LLM, Large Language Model)은 인간 수준의 언어 이해 능력을 기반으로 다양한 응용 분야에서 활용되고 있으나, 대부분 범용적인 대화형 상호작용을 목적으로 설계되어 특정 도메인의 전문적 명령 해석이나 실시간 제어 응용에는 한계가 존재한다. 이에 본 연구에서는 사용자의 자연어 음성 명령을 실시간으로 인식하고, 이를 구조화된 제어 명령으로 변환할 수 있는 LLM 기반 음성 제어 시스템을 제안하였다. 본 시스템은 Whisper 음성 인식 모델을 통해 발화된 명령을 문장 형태로 변환하고, LoRA(Low-Rank Adaptation) 기법으로 파인튜닝된 LLM이 해당 문장을 해석하여 제어 가능한 JSON 형태의 명령으로 변환하도록 구성되었다. 또한, 이러한 언어모델 구조를 소형 제어보드 환경(Raspberry Pi, Jetson Orin Nano 등) 에서 온디바이스 형태로 동작할 수 있도록 설계함으로써, 클라우드 연산 의존도를 줄이고 실시간 응답성을 확보하였다. 본 연구의 전체 구성은 다음과 같다. 먼저, 기존 음성 제어 시스템의 한계를 분석하고, 자연어 기반 제어의 필요성을 도출하였다. 다음으로, LLM의 효율적 경량화를 위해 LoRA 기반 파인튜닝 구조를 설계하였으며, 음성 인식–명령 해석–제어 명령 변환으로 이어지는 통합 프로세스를 정의하였다. 마지막으로, 제안된 시스템이 실제 소형 제어보드 상에서 실시간 제어 환경에 적용 가능한 형태로 구현될 수 있음을 확인하였다. 이 연구는 자연어 기반의 음성 명령을 구조화된 제어 명령으로 변환하는 온디바이스 LLM 응용 구조를 제시함으로써, 사용자 맞춤형 제어, 산업용 자동화, 이동 로봇 제어 등 다양한 분야에서 활용될 수 있는 경량 지능형 제어 시스템의 새로운 가능성을 제안한다.
다국어 초록 (Multilingual Abstract)
Although recent large language models (LLMs) have been actively applied to various tasks thanks to their human-level language understanding capability, most models have been developed for general conversational interaction, and thus their ability to i...
Although recent large language models (LLMs) have been actively applied to various tasks thanks to their human-level language understanding capability, most models have been developed for general conversational interaction, and thus their ability to interpret domain-specific commands or to support real-time control remains limited. To address this issue, a system is designed in which a user’s natural-language voice command is recognized in real time and converted into a structured control command. Spoken commands are first transcribed into text by a Whisper automatic speech recognition model, after which an LLM fine-tuned with the Low-Rank Adaptation (LoRA) technique interprets the text and converts it into an executable JSON-formatted control command. The overall architecture is further configured to operate in an on-device manner on small control boards (e.g., Raspberry Pi, Jetson Orin Nano), so that dependence on cloud computation is reduced and response latency is improved. The research proceeds as follows. First, the limitations of conventional voice control systems are analyzed and the necessity of natural-language-based control is derived. Next, a LoRA-based fine-tuning architecture is designed to enable efficient lightweight adaptation of the LLM, and an integrated pipeline covering speech recognition, command interpretation, and control-command generation is defined. Finally, it is empirically verified that the proposed system can be implemented on small control boards and can be operated in a form suitable for real-time control scenarios. By presenting an on-device LLM architecture that converts natural-language voice commands into structured control commands, this study suggests a new possibility for lightweight intelligent control systems that can be applied to user-adaptive control interfaces, industrial automation, and mobile robot control.
목차 (Table of Contents)