최근 인공지능과 컴퓨터 시스템 또는 로봇이 결합한 지능형 시스템과 사람 사이의 상호작용에 대한 다양한 서비스가 연구되고 있는 추세이다. 이러한 지능형 시스템과 사람 사이의 자연스...

http://chineseinput.net/에서 pinyin(병음)방식으로 중국어를 변환할 수 있습니다.
변환된 중국어를 복사하여 사용하시면 됩니다.
https://www.riss.kr/link?id=T17381131
울산 : 울산대학교 일반대학원, 2026
학위논문(석사) -- 울산대학교 일반대학원 , 전기전자컴퓨터공학과 , 2026. 2
2026
한국어
울산
; 26 cm
지도교수: 조강현
I804:48009-200000956071
0
상세조회0
다운로드최근 인공지능과 컴퓨터 시스템 또는 로봇이 결합한 지능형 시스템과 사람 사이의 상호작용에 대한 다양한 서비스가 연구되고 있는 추세이다. 이러한 지능형 시스템과 사람 사이의 자연스...
최근 인공지능과 컴퓨터 시스템 또는 로봇이 결합한 지능형 시스템과 사람 사이의 상호작용에 대한 다양한 서비스가 연구되고 있는 추세이다. 이러한 지능형 시스템과 사람 사이의 자연스러운 상호작용에 있어 표정 인식은 매우 중요한 연구 주제 중 하나 이다. 그러나 현재 딥러닝 기반의 표정 인식은 일반적인 공공 데이터셋 분류에 비해 성능 향상이 더딘 상황이다. 따라서 본 논문은 사람이 표정 인식을 하는 과정에서 공간 고주파 성분을 활용한다는 인지 심리학의 연구 결과에서 영감을 받아 이러한 기능을 합성곱 신경망에 부여하는 집중 모델을 제시한다. 제안된 모델은 특징 지도 텐서 에서 공간 특징 지도 행렬을 추출한 후 그것의 고주파 성분과의 유사성을 특징 벡터 분포의 관점에서 해석하고 계산한다. 이렇게 계산된 유사성을 집중 점수로 환산 후 특징 지도 텐서에 인가하여 공간 고주파 성분을 집중시키는 집중 모델을 제시한다. 또한 유사성을 계산하는 과정에 대한 설명력을 부여하기 위해 기하학적 해석과 주축 가설을 제시한다. 본 논문은 제안된 집중 모델을 ResNet18과 MobileNetV2에 부착하여 대표적인 표정 분류 데이터셋인 FER2013, CK+, JAFFE를 통해 구조와 성능을 검증한다. 또한 공간 특징 지도 행렬과 고주파 성분을 구성하는 행(또는 열) 방향 특징 벡터 분포 사이의 투영 거리 기반의 유사성을 매 학습 Epoch 마다 계산하여 주축 가설의 타당성을 제시한다.
다국어 초록 (Multilingual Abstract)
Facial expression recognition (FER) is a key component for natural human-system interaction in intelligent systems that combine AI with computer platforms or robots. Despite rapid progress in general image classification on public benchmarks, improvem...
Facial expression recognition (FER) is a key component for natural human-system interaction in intelligent systems that combine AI with computer platforms or robots. Despite rapid progress in general image classification on public benchmarks, improvements in deep learning-based FER remain modest. Motivated by findings in cognitive psychology that humans exploit spatial high-frequency cues for expression perception, this paper introduces an attention model that explicitly equips convolutional neural networks with the ability to emphasize such information. Proposed model extracts a spatial feature map matrix from input feature map tensors and its high-frequency component. It computes their simiarity with the perspective of feature vector distributions. The computed similarity is converted into an attention score and applied to the feature map tensor, yielding an attention model that concentrates on spatial high-frequency component. This paper further presents a geometric interpretation and principal axis hypothesis to provide explanatory grounding for the similarity computation. The proposed attention model is attached to ResNet18 and MobileNetV2 and evaluated on representative FER datasets (FER2013, CK+, and JAFFE) to validate both architecture and performance. In addition, this research work quantifies a projection distance-based similarity between the row- (or column) wise feature vector distributions of spatial feature map matrix and its high-frequency component at every training epoch, providing empirical support for the principal axis hypothesis.
목차 (Table of Contents)