제조 공정에서 수집되는 시계열 데이터는 정상 구간이 압도적으로 우세하고 이상 패턴이 극히 드물게 나타나는 불균형 구조를 가진다. 이러한 환경에서는 전체 정확도보다 실제 이상을 놓...

http://chineseinput.net/에서 pinyin(병음)방식으로 중국어를 변환할 수 있습니다.
변환된 중국어를 복사하여 사용하시면 됩니다.
https://www.riss.kr/link?id=T17385655
구미 : 국립금오공과대학교 대학원, 2026
학위논문(석사) -- 국립금오공과대학교 대학원 , 산업공학과 , 2026. 2
2026
한국어
경상북도
; 26 cm
지도교수: 이종환
I804:47006-000000017620
0
상세조회0
다운로드제조 공정에서 수집되는 시계열 데이터는 정상 구간이 압도적으로 우세하고 이상 패턴이 극히 드물게 나타나는 불균형 구조를 가진다. 이러한 환경에서는 전체 정확도보다 실제 이상을 놓...
제조 공정에서 수집되는 시계열 데이터는 정상 구간이 압도적으로 우세하고 이상 패턴이 극히 드물게 나타나는 불균형 구조를 가진다. 이러한 환경에서는 전체 정확도보다 실제 이상을 놓치는 2종 오류를 최소화하는 것이 중요한 성능 목표가 된다. 본 연구는 이러한 문제들을 완화하기 위해 Dueling Deep Q-Network(Dueling DQN)를 활용하여 경보/통과 행동 간 상태 가치 차이를 이상 점수로 정의하는 강화학습 기반 이상탐지 방식을 제안한다. 제안 모델은 Dueling 구조를 통해 상태 가치와 행동 간 상대적 차이를 분리해 추정하고, 2종 오류 패널티를 반영한 보상 함수를 통해 정책이 2종 오류 감소를 우선적으로 학습하도록 구성하였다. 학습된 이상 점수는 검증 단계에서 산정된 임계값과 결합하여 최종 이상 여부를 판단한다. 두 종류의 제조 시계열 데이터를 대상으로 한 실험 결과, 제안 방식은 정상·이상 구간의 분포 중첩을 완화하는 경향을 보여, 재구성오차 기반 이상 점수보다 일관된 판단 기준을 제공하였다. 이를 통해 강화학습 기반 이상 점수 재정의가 제조 시계열 이상탐지에서 기존 재구성 기반 모델의 대안을 제시할 수 있음을 확인하였다.
다국어 초록 (Multilingual Abstract)
Manufacturing time-series data exhibit a highly imbalanced structure in which normal patterns dominate and anomalous events occur extremely rarely. In such environments, minimizing Type II errors (False Negatives), rather than maximizing overall accur...
Manufacturing time-series data exhibit a highly imbalanced structure in which normal patterns dominate and anomalous events occur extremely rarely. In such environments, minimizing Type II errors (False Negatives), rather than maximizing overall accuracy, becomes the critical performance objective. To address this challenge, this study proposes a reinforcement learning–based anomaly detection method that employs a Dueling Deep Q-Network (Dueling DQN) and defines an anomaly score as the Q-value difference between the alert and pass actions. The dueling architecture enables the model to separately estimate the state value and the relative advantage between actions, while the reward function incorporates the cost of False Negatives to guide the policy toward suppressing FN occurrences. The learned anomaly score is combined with a threshold determined in the validation phase to classify anomalies during inference. Experiments on two types of manufacturing time-series data show that the proposed method reduces the overlap between normal and abnormal score distributions, providing a more consistent decision boundary compared to reconstruction-error-based anomaly scores. These results demonstrate that redefining anomaly scores via reinforcement learning can serve as an effective alternative to traditional reconstruction-based approaches in manufacturing time-series anomaly detection.
목차 (Table of Contents)