본 연구는 강화학습을 활용하여 행동 경제학의 핵심 편향인 투자자 처분 효과(Disposition Effect)를 모사하고 그 발현 기제를 분석하는 계산 인지 모델(Computational Cognitive Model)을 제시한다. 기존...

http://chineseinput.net/에서 pinyin(병음)방식으로 중국어를 변환할 수 있습니다.
변환된 중국어를 복사하여 사용하시면 됩니다.
https://www.riss.kr/link?id=T17451776
서울 : 서울대학교 대학원, 2026
학위논문(석사) -- 서울대학교 대학원 , 협동과정인지과학전공 , 2026. 2
2026
한국어
153
서울
111 ; 26 cm
지도교수: 고성룡
I804:11032-000000194470
0
상세조회0
다운로드본 연구는 강화학습을 활용하여 행동 경제학의 핵심 편향인 투자자 처분 효과(Disposition Effect)를 모사하고 그 발현 기제를 분석하는 계산 인지 모델(Computational Cognitive Model)을 제시한다. 기존...
본 연구는 강화학습을 활용하여 행동 경제학의 핵심 편향인 투자자 처분 효과(Disposition Effect)를 모사하고 그 발현 기제를 분석하는 계산 인지 모델(Computational Cognitive Model)을 제시한다. 기존의 금융 강화학습 연구가 샤프 지수와 같은 객관적 성과 지표의 최적화에 치중해온 것과 달리, 본 연구는 전망 이론(Prospect Theory)의 가치 함수를 보상 체계로 채택하여 인간의 비합리적 의사결정 과정을 공학적으로 재현하고 실증적으로 검증하였다.
연구를 위해 강화학습 환경 내에 자산별 손익을 추적하는 '심적 회계 추적기'를 구현하였으며, 에이전트의 행동 공간을 Dirichlet 분포로 제약하여 예산 제약 조건 하에서의 연속적인 포트폴리오 최적화를 가능하게 했다. 모델 구조 측면에서는 MLP와 LSTM의 비교를 통해 처분 효과 발현을 위한 시계열 기억 메커니즘의 필수성을 확인하였으며, ‘참조점 의존성’이 형성되는 인지적 과정을 공학적으로 구현하였다.
실험 결과, 첫째, 자산별로 심적 회계를 분리하는 '좁은 프레임(Narrow Frame)' 환경이 전체 자산 통합 프레임(Global Frame)보다 처분 효과를 더욱 강력하게 유도함을 보였다. 특히 통합 프레임 환경에서는 프레임의 전환이 인지 편향을 완화하는 전략으로 작용할 수 있음을 확인하였다. 둘째, Narrow Frame 하에서 손실 회피 계수(λ)와 처분 효과 간의 '역 U자형' 관계를 발견하고 실증 연구와의 수치적 정합성을 규명하였다. 이는 인간의 실제 λ 값이 '손실 보유(미련)'와 '위험 자산 투매(공포)' 사이에서 미련이 최대화되는 인지적 임계점임을 계산 모델이 수치적으로 재현한 것이다. 셋째, L2 Regularization을 통한 일반화 과정을 거친 후에도 처분 효과가 견고하게 유지됨을 보임으로써, 해당 편향이 모델의 오류가 아닌 가치 구조에서 기인하는 본질적 최적해임을 방증하였다.
본 연구는 인지 과학적 통찰을 금융 AI 모델에 통합함으로써 행동 경제학 이론의 계산적 검증 도구를 제공할 뿐만 아니라, 패닉 셀링(Panic Selling)과 같은 시장의 비합리적 현상을 예측하고 관리하기 위한 새로운 기술적 토대를 마련하였다는 데 의의가 있다.
주요어 : 강화 학습, 처분 효과, 전망 이론, 심적 회계, 좁은 프레이밍, 디리클레 분포
다국어 초록 (Multilingual Abstract)
This study presents a computational cognitive model that leverages reinforcement learning to simulate and analyze the underlying mechanisms of the disposition effect, a core behavioral bias in behavioral economics. Departing from conventional financia...
This study presents a computational cognitive model that leverages reinforcement learning to simulate and analyze the underlying mechanisms of the disposition effect, a core behavioral bias in behavioral economics. Departing from conventional financial reinforcement learning research that has predominantly focused on optimizing objective performance metrics such as the Sharpe ratio, this study adopts the value function from Prospect Theory as its reward framework to computationally reproduce and empirically validate irrational human decision-making processes.
To facilitate this investigation, a Mental Account Tracker was implemented within the reinforcement learning environment to monitor asset-level gains and losses, while the agent's action space was constrained using a Dirichlet distribution to enable continuous portfolio optimization under budget constraints. Architecturally, a comparative analysis between MLP and LSTM networks confirmed the necessity of temporal memory mechanisms for manifesting the disposition effect and computationally implemented the cognitive processes underlying the formation of reference dependence.
The experimental findings revealed three key insights. First, narrow frame environments that segregate mental accounts by individual assets induced the disposition effect more strongly than global frame approaches that integrate all assets. Notably, in global frame environments, frame-switching emerged as a viable strategy for mitigating cognitive biases. Second, an inverted U-shaped relationship was discovered between the loss aversion coefficient (λ) and the disposition effect under narrow framing, and numerical consistency with empirical research was established. This demonstrates that the computational model numerically reproduced the cognitive threshold where human actual λ values maximize attachment (loss retention) between "holding losses (attachment)" and "panic selling of risky assets (fear)." Third, the disposition effect remained robust even after generalization through L2 regularization, substantiating that this bias represents an inherent optimal solution arising from the value structure rather than a modeling artifact.
This research contributes to the field by integrating cognitive scientific insights into financial AI models, thereby providing a computational validation tool for behavioral economics theories. Furthermore, it establishes a novel technological foundation for predicting and managing irrational market phenomena such as panic selling.
Keywords : Reinforcement Learning, Disposition Effect, Prospect Theory, Mental Accounting, Narrow Framing, Dirichlet Distribution
목차 (Table of Contents)