RISS 학술연구정보서비스

검색

인기 검색어

    다국어 입력

    http://chineseinput.net/에서 pinyin(병음)방식으로 중국어를 변환할 수 있습니다.

    변환된 중국어를 복사하여 사용하시면 됩니다.

    예시)
    • 中文 을 입력하시려면 zhongwen을 입력하시고 space를누르시면됩니다.
    • 北京 을 입력하시려면 beijing을 입력하시고 space를 누르시면 됩니다.
    닫기
    KCI등재

    생성형 변환 모델을 통한 SAR 데이터의 의미론적 해석 및 시각적 이해 확장 = Extending Semantic Interpretation and Visual Understanding of SAR Data via Generative Translation Models

    한글로보기

    https://www.riss.kr/link?id=A110248778

    • 0

      상세조회
    • 0

      다운로드
    서지정보 열기
    • 내보내기
    • 내책장담기
    • 공유하기
    • 오류접수
    인용문이 복사되었습니다.

    부가정보

    다국어 초록 (Multilingual Abstract) kakao i 다국어 번역

    Synthetic aperture radar (SAR) offers critical all-weather observation capabilities, yet its interpretation remains challenging due to inherent speckle noise and non-intuitive scattering characteristics. Consequently, directly applying vision-language models (VLMs) trained on natural images to the SAR domain is limited by significant modality gaps and the scarcity of high-quality SAR-text datasets. To overcome these challenges, this study proposes a two-stage framework that leverages SAR-to-optical translation to bridge the domain gap. First, we introduce a conditional Brownian Bridge Diffusion Model integrated with a SAR feature guidance module. This approach transforms SAR images into optical-like representations while preserving structural fidelity, thereby addressing the geometric distortions and hallucinations common in generative adversarial network (GAN)-based methods. Second, the translated images are analyzed by a domain-adapted VLM, utilizing the GeoRSCLIP visual encoder and a LoRA-tuned LLaVA model to generate precise semantic captions. Experimental results using Sentinel-1 and Sentinel-2 datasets demonstrate that the proposed translation model outperforms existing GAN models in terms of PSNR and SSIM. Furthermore, the framework achieves significant improvements in captioning metrics, including BLEU, ROUGE-L, and BERT-Score, compared to direct SAR interpretation. This study validates that high-fidelity modality translation can effectively extend the reasoning capabilities of pre-trained VLMs to the SAR domain without requiring extensive SAR-specific annotations.
    번역하기

    Synthetic aperture radar (SAR) offers critical all-weather observation capabilities, yet its interpretation remains challenging due to inherent speckle noise and non-intuitive scattering characteristics. Consequently, directly applying vision-language...

    Synthetic aperture radar (SAR) offers critical all-weather observation capabilities, yet its interpretation remains challenging due to inherent speckle noise and non-intuitive scattering characteristics. Consequently, directly applying vision-language models (VLMs) trained on natural images to the SAR domain is limited by significant modality gaps and the scarcity of high-quality SAR-text datasets. To overcome these challenges, this study proposes a two-stage framework that leverages SAR-to-optical translation to bridge the domain gap. First, we introduce a conditional Brownian Bridge Diffusion Model integrated with a SAR feature guidance module. This approach transforms SAR images into optical-like representations while preserving structural fidelity, thereby addressing the geometric distortions and hallucinations common in generative adversarial network (GAN)-based methods. Second, the translated images are analyzed by a domain-adapted VLM, utilizing the GeoRSCLIP visual encoder and a LoRA-tuned LLaVA model to generate precise semantic captions. Experimental results using Sentinel-1 and Sentinel-2 datasets demonstrate that the proposed translation model outperforms existing GAN models in terms of PSNR and SSIM. Furthermore, the framework achieves significant improvements in captioning metrics, including BLEU, ROUGE-L, and BERT-Score, compared to direct SAR interpretation. This study validates that high-fidelity modality translation can effectively extend the reasoning capabilities of pre-trained VLMs to the SAR domain without requiring extensive SAR-specific annotations.

    더보기

    동일학술지(권/호) 다른 논문

    동일학술지 더보기

    더보기

    분석정보

    View

    상세정보조회

    0

    Usage

    원문다운로드

    0

    대출신청

    0

    복사신청

    0

    EDDS신청

    0

    동일 주제 내 활용도 TOP

    더보기

    주제

    연도별 연구동향

    연도별 활용동향

    연관논문

    연구자 네트워크맵

    공동연구자 (7)

    유사연구자 (20) 활용도상위20명

    이 자료와 함께 이용한 RISS 자료

    나만을 위한 추천자료

    해외이동버튼