학원 원격 탐사 (Remote Sensing, RS) 객체 탐지는 실제 센싱 환경에서 지속적인 어려움을 겪고 있으며, 그 주요 오류 원인은 네트워크 용량 자체보다는 센서 물리 특성과 취득 기하에 의해 지배된...

http://chineseinput.net/에서 pinyin(병음)방식으로 중국어를 변환할 수 있습니다.
변환된 중국어를 복사하여 사용하시면 됩니다.
https://www.riss.kr/link?id=T17355630
서울 : 숭실대학교 대학원, 2025
학위논문(박사) -- 숭실대학교 대학원 , 정보통신공학과(일원) , 2026. 2
2025
영어
서울
122 ; 26 cm
지도교수: 신요안
I804:11044-200000949635
0
상세조회0
다운로드학원 원격 탐사 (Remote Sensing, RS) 객체 탐지는 실제 센싱 환경에서 지속적인 어려움을 겪고 있으며, 그 주요 오류 원인은 네트워크 용량 자체보다는 센서 물리 특성과 취득 기하에 의해 지배된...
학원 원격 탐사 (Remote Sensing, RS) 객체 탐지는 실제 센싱 환경에서 지속적인 어려움을 겪고 있으며, 그 주요 오류 원인은 네트워크 용량 자체보다는 센서 물리 특성과 취득 기하에 의해 지배된다. 합성 개구 레이다 (Synthetic Aperture Radar, SAR) 영상에서는 반점 노이즈 (Speckle)와 약한 객체 경계로 인해 클러터 기반 오경보 (false alarm) 와 위치 추정의 모호성이 발생한다. 무인항공기 (Unmanned Aerial Vehicle, UAV) 기반 광학 영상에서는 극소형 객체, 높은 공간 밀집도, 그리고 빈번한 가림 현상으로 인해 특징 표현이 심각하게 저하되며 탐지 실패가 증가한다. 또한 다중 센싱 모달리티 (Multi-modality sensing)를 공동으로 활용하는 경우, 서로 다른 영상 형성 메커니즘으로 인해 표현 불일치, 공간 매칭의 불완전함, 그리고 과업 수준의 불일치를 초래할 수 있으며, 단순한 초기 융합 전략은 불안정할 수 있고 심지어 성능 저하를 초래할 수 있다. 이러한 문제를 해결하기 위해 본 학위논문은 원격 탐사 객체 탐지를 위한 확장 가능하고 제약 기반의 고효율 아키텍처 방법론을 제안한다. 제안된 방법론은 탐지 성능을 모델 용량이나 융합 복잡도의 함수로 간주하는 기존 전통적인 방법과 달리, 센싱으로 유발된 불일치가 탐지 프로세스의 어느 단계에서 발생하는지를 명확하게 분석하고, 이를 세 단계를 통해 점진적으로 제거한다. 구체적으로, 모달리티별 표현 안정화, 제어된 경량 교차 모달리티 상호 작용 및 탐지 헤드에서의 과업 일관적 예측을 통해 불일치를 해결한다. 이러한 단계 인지적 (stage-aware) 설계는 단일 모달리티 최적화에서 교차 모달 통합으로 자연스럽게 확장 가능한 통합 설계 논리를 제공하며, 불필요한 계산 오버헤드 (overhead)를 초래하지 않는다. 제안된 방법론은 세 가지 대표적인 객체 탐지 프레임워크를 통해 구체화되고 검증된다. SMEP-DETR은 SAR 영상에 특화된 트랜스포머 기반 (transformer-based) 탐지 모델로서, 다중 에지 (edge) 강화와 병렬 팽창 문맥 집계를 통해 반점 노이즈 안정화와 명시적 경계 강화를 수행함으로써 해양 환경에서의 탐지 정확도를 향상시킨다. MCG- RTDETR은 UAV 영상에서의 극단적인 스케일 및 밀집 분포 상황을 고려하여, 다중 컨볼루션 (convolution)과 문맥 유도 (context-guided) 구조, 그리고 캐스케이드 그룹 어텐션 (cascade group attention)을 결합함으로써 미세 특징 보존, 문맥 기반 인코딩, 그리고 소형 객체 탐지 성능을 강화하는 동시에 실시간 적용 가능성을 유지한다. 이종 센싱 환경을 위해 제안된 교차 모달 (cross-modality) 특징 적응 탐지 프레임워크인 CMFADet은 RGB-적외선 및 SAR-광학 쌍 영상에 대해 이중 스트림 (dual-stream) 구조를 채택하고, 모달리티별 안정화, 경량 잔차 (residual) 상호작용, 그리고 적응적 과업 인지 정렬을 통합적으로 적용함으로써 표현 불일치와 분류 및 위치 추정 간의 불일치를 완화한다. 다양한 벤치마크 (benchmark) 데이터셋에 대한 광범위한 실험 결과는 제안된 방법론이 서로 다른 센싱 조건 전반에서 탐지의 강건성과 효율성을 일관되게 향상시킴을 입증한다. 구체적으로, SMEP-DETR은 SSDD 데이터셋에서 98.6% 𝑚𝐴𝑃 를 달성하였으며, MCG-RTDETR은 VisDrone2019에서 58.2% 𝐴𝑃50 을 기록하였다. 또한 CMFADet은 3.64M의 파라미터 수만으로 DroneVehicle 데이터셋에서 83.42% 𝑚𝐴𝑃50 을 달성하여, 정확도와 효율성 간의 우수한 균형을 보여준다. 이러한 결과는 센싱 유발 불일치를 확장 가능한 아키텍처 방법론을 통해 해결함으로써, 단일 모달리티 (single-modality) 및 교차 모달리티 (cross-modality) 원격 탐사 환경 전반에서 정확하고 실제 배치 가능한 객체 탐지가 가능함을 입증한다.
다국어 초록 (Multilingual Abstract)
Remote sensing (RS) object detection faces persistent challenges under real- world sensing conditions, where dominant error sources are governed by sensing physics and acquisition geometry rather than network capacity alone. In synthetic aperture rada...
Remote sensing (RS) object detection faces persistent challenges under real- world sensing conditions, where dominant error sources are governed by sensing physics and acquisition geometry rather than network capacity alone. In synthetic aperture radar (SAR) imagery, coherent speckle and weak object boundaries induce clutter-driven false alarms and localization ambiguity. In unmanned aerial vehicle (UAV) optical imagery, extremely small objects, dense spatial distributions, and frequent occlusion lead to severe feature degradation and missed detections. When multiple sensing modalities are jointly exploited, heterogeneous imaging mechanisms further introduce representation mismatch, imperfect spatial correspondence, and task-level inconsistency, rendering early fusion strategies unstable or even detrimental. To address these challenges under strict efficiency constraints, this dissertation develops a scalable, constraint-driven architectural methodology for high-efficiency object detection in remote sensing. Rather than treating detection performance as a function of model capacity or fusion complexity, the proposed methodology explicitly investigates where sensing-induced discrepancy enters the detection pipeline and resolves it progressively through three stages: stabilizing modality- specific representations, enabling controlled and lightweight cross-modality interaction, and enforcing task-consistent prediction at the detection head. This stage-aware formulation establishes a unified design logic that generalizes from single-modality optimization to cross-modality integration without incurring unnecessary computational overhead. The proposed methodology is instantiated and validated through three representative detection frameworks. SMEP-DETR is a transformer-based detector for SAR imagery with multi-edge enhancement and parallel dilated context aggregation, which introduces speckle noise stabilization and edge enhancement to improve detection accuracy in maritime scenes. MCG-RTDETR, a multi- convolution and context-guided network with cascaded group attention for UAV imagery, is proposed to enhance fine-grained feature preservation, context-guided encoding, and small-object detection while maintaining real-time viability under extreme scale and density constraints. For heterogeneous sensing scenarios, a cross- modality feature adaptive detection framework, termed CMFADet, is proposed. CMFADet adopts a dual-stream architecture for paired RGB-infrared and SAR- optical imagery, where modality-specific stabilization, lightweight residual interaction, and adaptive task-aware alignment are jointly employed to mitigate representation incompatibility and the inconsistency of classification and localization tasks. Extensive experiments on widely used benchmarks demonstrate that the proposed methodology consistently improves both detection robustness and efficiency across diverse sensing conditions. Specifically, SMEP-DETR achieves 98.6% 𝑚𝐴𝑃 on SSDD, MCG-RTDETR attains 58.2% 𝐴𝑃50 on VisDrone2019, and CMFADet reaches 83.42% 𝑚𝐴𝑃50 on DroneVehicle with only 3.64M parameters, validating a favorable accuracy and efficiency trade-off. These results confirm that resolving sensing-induced discrepancy through a scalable architectural methodology enables accurate and deployable object detection across both single- modality and cross-modality remote sensing scenarios.
목차 (Table of Contents)