RISS 학술연구정보서비스

검색

인기 검색어

    다국어 입력

    http://chineseinput.net/에서 pinyin(병음)방식으로 중국어를 변환할 수 있습니다.

    변환된 중국어를 복사하여 사용하시면 됩니다.

    예시)
    • 中文 을 입력하시려면 zhongwen을 입력하시고 space를누르시면됩니다.
    • 北京 을 입력하시려면 beijing을 입력하시고 space를 누르시면 됩니다.
    닫기

    대형 언어모델의 허위정보 처리 능력에 대한 지시 튜닝의 영향 = The Impact of Instruction Tuning on the Misinformation Handling Ability of Large Language Models

    한글로보기

    https://www.riss.kr/link?id=T17380394

    • 0

      상세조회
    • 0

      다운로드
    서지정보 열기
    • 내보내기
    • 내책장담기
    • 공유하기
    • 오류접수

    부가정보

    국문 초록 (Abstract) kakao i 다국어 번역

    지시튜닝(Instruction-tuning)은 대형 언어 모델(LLM)이 사용자 지시를 보 다 정확하게 따를 수 있도록 하여, 사용성을 개선하고 해로운 출력을 감소 시키는 데 기여한다. 그러나 이러한 과정은 모델의 사용자 입력에 대한 의 존도를 높여, 잘못된 정보를 여과 없이 수용하거나 환각(hallucination)을 생 성할 가능성을 증가시킬 수 있다. 기존 연구들은 주로 대형 언어 모델이 자 체 파라메트릭(parametric) 지식과 상충하는 외부 정보에 수용적인 경향이 있음을 보여주었으나, 지시튜닝이 이러한 현상에 미치는 직접적인 영향에 대한 연구는 부족하다. 본 연구에서는 지시튜닝이 대형 언어 모델의 허위 정보(misinformation) 취약성에 미치는 영향을 분석하였다. 그 결과, 지시튜닝을 거친 모델은 사용 자가 제시한 허위 정보를 수용할 가능성이 유의미하게 높았다. 베이스 모델 과의 비교를 통해, 지시튜닝이 모델의 사용자 제공 정보에 대한 의존성을 강화하여, 허위 정보 취약성이 사용자 역할에서 두드러지게 나타남을 확인 하였다. 또한 프롬프트 구조 내 사용자 역할, 허위 정보의 길이, 시스템 프롬프트의 경고 존재 여부 등 허위 정보 취약성에 영향을 미치는 추가 요인들을 탐구 하였다. 연구 결과는 지시튜닝의 의도치 않은 부작용을 완화하고, 실제 응용 환경에서 대형 언어모델의 신뢰성을 향상시키기 위한 체계적 접근의 필요성 을 시사한다.
    번역하기

    지시튜닝(Instruction-tuning)은 대형 언어 모델(LLM)이 사용자 지시를 보 다 정확하게 따를 수 있도록 하여, 사용성을 개선하고 해로운 출력을 감소 시키는 데 기여한다. 그러나 이러한 과정은 모...

    지시튜닝(Instruction-tuning)은 대형 언어 모델(LLM)이 사용자 지시를 보 다 정확하게 따를 수 있도록 하여, 사용성을 개선하고 해로운 출력을 감소 시키는 데 기여한다. 그러나 이러한 과정은 모델의 사용자 입력에 대한 의 존도를 높여, 잘못된 정보를 여과 없이 수용하거나 환각(hallucination)을 생 성할 가능성을 증가시킬 수 있다. 기존 연구들은 주로 대형 언어 모델이 자 체 파라메트릭(parametric) 지식과 상충하는 외부 정보에 수용적인 경향이 있음을 보여주었으나, 지시튜닝이 이러한 현상에 미치는 직접적인 영향에 대한 연구는 부족하다. 본 연구에서는 지시튜닝이 대형 언어 모델의 허위 정보(misinformation) 취약성에 미치는 영향을 분석하였다. 그 결과, 지시튜닝을 거친 모델은 사용 자가 제시한 허위 정보를 수용할 가능성이 유의미하게 높았다. 베이스 모델 과의 비교를 통해, 지시튜닝이 모델의 사용자 제공 정보에 대한 의존성을 강화하여, 허위 정보 취약성이 사용자 역할에서 두드러지게 나타남을 확인 하였다. 또한 프롬프트 구조 내 사용자 역할, 허위 정보의 길이, 시스템 프롬프트의 경고 존재 여부 등 허위 정보 취약성에 영향을 미치는 추가 요인들을 탐구 하였다. 연구 결과는 지시튜닝의 의도치 않은 부작용을 완화하고, 실제 응용 환경에서 대형 언어모델의 신뢰성을 향상시키기 위한 체계적 접근의 필요성 을 시사한다.

    더보기

    다국어 초록 (Multilingual Abstract) kakao i 다국어 번역

    Instruction tuning enhances the usability of large language models (LLMs) by enabling them to follow user instructions more accurately and reducing harmful outputs. However, this process also increases the model’s reliance on user input, potentially making it more likely to accept incorrect information without filtering it out or to generate hallucinations. While previous studies have largely shown that LLMs tend to be receptive to external information that conflicts with their own parametric knowledge, there has been limited research on the direct impact of instruction tuning
    on this phenomenon.
    In this study, we analyze how instruction tuning affects LLMs’ vulnerability to misinformation. Our findings show that instruction-tuned models are significantly more likely to accept false information provided by users. By comparing them with base models, we confirm that instruction tuning strengthens the model’s dependence on user-supplied information, making misinformation vulnerability particularly pronounced in the user role.
    We also explore additional factors that influence susceptibility to misinformation, including the user role within the prompt structure, the length of the misinformation, and the presence of system-prompt warnings. The results suggest the need for systematic approaches to mitigate the unintended side effects of instruction tuning and to improve the reliability of LLMs in real-world applications.
    번역하기

    Instruction tuning enhances the usability of large language models (LLMs) by enabling them to follow user instructions more accurately and reducing harmful outputs. However, this process also increases the model’s reliance on user input, potentially...

    Instruction tuning enhances the usability of large language models (LLMs) by enabling them to follow user instructions more accurately and reducing harmful outputs. However, this process also increases the model’s reliance on user input, potentially making it more likely to accept incorrect information without filtering it out or to generate hallucinations. While previous studies have largely shown that LLMs tend to be receptive to external information that conflicts with their own parametric knowledge, there has been limited research on the direct impact of instruction tuning
    on this phenomenon.
    In this study, we analyze how instruction tuning affects LLMs’ vulnerability to misinformation. Our findings show that instruction-tuned models are significantly more likely to accept false information provided by users. By comparing them with base models, we confirm that instruction tuning strengthens the model’s dependence on user-supplied information, making misinformation vulnerability particularly pronounced in the user role.
    We also explore additional factors that influence susceptibility to misinformation, including the user role within the prompt structure, the length of the misinformation, and the presence of system-prompt warnings. The results suggest the need for systematic approaches to mitigate the unintended side effects of instruction tuning and to improve the reliability of LLMs in real-world applications.

    더보기

    목차 (Table of Contents)

    • 제1장 서론 1
    • 제2장 관련 연구 5
    • 제1절 지식 충돌(Knowledge Conflict). 5
    • 제2절 지시튜닝된 대형 언어 모델. 6
    • 제3장 실험 설정 7
    • 제1장 서론 1
    • 제2장 관련 연구 5
    • 제1절 지식 충돌(Knowledge Conflict). 5
    • 제2절 지시튜닝된 대형 언어 모델. 6
    • 제3장 실험 설정 7
    • 제1절 데이터셋 7
    • 제2절 실험 시나리오 9
    • 제3절 평가지표. 12
    • 제4장 실험 및 분석 14
    • 제1절 실험 모델 14
    • 제2절 지시튜닝된 대형 언어 모델의 허위 정보 취약성 14
    • 제3절 지시튜닝과 허위 정보 취약성의 상관관계 17
    • 제4절 허위 정보 취약성에 영향을 주는 요인 20
    • 제5장 결론 25
    • 참고문헌 26
    • 부록 33
    • ABSTRACT 41
    더보기

    분석정보

    View

    상세정보조회

    0

    Usage

    원문다운로드

    0

    대출신청

    0

    복사신청

    0

    EDDS신청

    0

    동일 주제 내 활용도 TOP

    더보기

    주제

    연도별 연구동향

    연도별 활용동향

    연관논문

    연구자 네트워크맵

    공동연구자 (7)

    유사연구자 (20) 활용도상위20명

    이 자료와 함께 이용한 RISS 자료

    나만을 위한 추천자료

    해외이동버튼