대규모 언어 모델(LLM)의 확산에 따라 생성 텍스트의 출처를 판별하는 워터마킹 기술의 중요성이 대두되고 있다. 본 연구는 고정된 분할 비율을 사용하는 기존 BiMarker 방식의 탐지 성능 한계...

http://chineseinput.net/에서 pinyin(병음)방식으로 중국어를 변환할 수 있습니다.
변환된 중국어를 복사하여 사용하시면 됩니다.
https://www.riss.kr/link?id=A109984568
2025
Korean
KCI등재
학술저널
1325-1331(7쪽)
0
상세조회0
다운로드대규모 언어 모델(LLM)의 확산에 따라 생성 텍스트의 출처를 판별하는 워터마킹 기술의 중요성이 대두되고 있다. 본 연구는 고정된 분할 비율을 사용하는 기존 BiMarker 방식의 탐지 성능 한계...
대규모 언어 모델(LLM)의 확산에 따라 생성 텍스트의 출처를 판별하는 워터마킹 기술의 중요성이 대두되고 있다. 본 연구는 고정된 분할 비율을 사용하는 기존 BiMarker 방식의 탐지 성능 한계를 개선하기 위해, 통계 기반 극성 분할 비율 설정을 통해 최적 비율을 선택하는 동적 극 분할(DPP) 기법을 제안한다. 제안 기법은 OPT-1.3B 및 StarCoder 모델, C4 및 HumanEval·MBPP 데이터셋을 기반으로 탐지 성능, 삽입 일관성, 토큰 분포 다양성 측면에서 실험을 수행하였다. 특히 DPP 설정 7(γ = 0.65 / 0.50 / 0.35)은 모든 환경에서 우수한 성능을 나타냈으며, 단어 순서 변경 및 동의어 치환 공격에 대해서도 25% 이내의 성능 저하로 높은 내성을 보였다. 본 연구는 DPP의 실용성과 강건성을 실험적으로 입증하였으며, LLM 기반 텍스트의 신뢰성 확보를 위한 효과적인 대안을 제시한다.
다국어 초록 (Multilingual Abstract)
With the rapid proliferation of large language models (LLMs), watermarking techniques for identifying the origin of generated text have become increasingly important. This study proposes a Dynamic Polarity Partitioning (DPP) method that selects the op...
With the rapid proliferation of large language models (LLMs), watermarking techniques for identifying the origin of generated text have become increasingly important. This study proposes a Dynamic Polarity Partitioning (DPP) method that selects the optimal polarity ratio from predefined statistical settings, addressing the detection limitations of the fixed-ratio-based BiMarker approach. The proposed method is evaluated using the OPT-1.3B and StarCoder models with the C4, HumanEval, and MBPP datasets, focusing on detection performance, watermarking consistency, and token distribution diversity. Experimental results show that DPP setting 7 (γ = 0.65 / 0.50 / 0.35) achieves superior performance across all evaluation criteria. Furthermore, under perturbation attacks such as word order shuffling and synonym substitution, the detection performance decreased by no more than 25%, demonstrating strong robustness. These findings validate the practicality and resilience of DPP and present an effective solution for ensuring the reliability of LLM-generated text.
대규모 데이터 파일 관리를 위한 분할병합 프로그램 모듈 설계
BERTopic과 LLM을 활용한 AI 및 SW 교육 현황 및 예측 분석
딥러닝 기반 흉부 X-ray 영상 분석을 통한 폐렴 진단 모델 연구
자연어의 표지성 이론에 기반한 의존관계 중심 지식그래프 모델링