본 연구는 분류(Classification)와 이상 탐지(Anomaly Detection) 환경에서 대규모 언어 모델(LLM)의 성능을 분석한다. 실험 결과, 분류 환경에서는 샘플 크기가 작을 때 LLM이 높은 분류 성능을 기록했으...

http://chineseinput.net/에서 pinyin(병음)방식으로 중국어를 변환할 수 있습니다.
변환된 중국어를 복사하여 사용하시면 됩니다.
본 연구는 분류(Classification)와 이상 탐지(Anomaly Detection) 환경에서 대규모 언어 모델(LLM)의 성능을 분석한다. 실험 결과, 분류 환경에서는 샘플 크기가 작을 때 LLM이 높은 분류 성능을 기록했으...
본 연구는 분류(Classification)와 이상 탐지(Anomaly Detection) 환경에서 대규모 언어 모델(LLM)의 성능을 분석한다. 실험 결과, 분류 환경에서는 샘플 크기가 작을 때 LLM이 높은 분류 성능을 기록했으나, 데이터의 규모가 커질수록 TabPFN과 같은 모델 및 전통적인 분류 모델의 성능이 상대적으로 우세하였다. 반면 이상 탐지 환경에서는 LLM이 기존 비지도(Unsupervised) 또는 준지도(Semi-Supervised) 방법론보다 우수한 성능을 보였다. 그러나 이러한 높은 성능에도 불구하고 학습에 상당한 시간이 소요된다는 한계가 확인되었다. 따라서 비용과 성능을 종합적으로 고려할 때, LLM을 전면적으로 활용하기보다는 문제 특성에 따라 선별적으로 활용하는 것이 바람직하다는 결론에 도달하였다.
다국어 초록 (Multilingual Abstract)
This study investigates the performance of large language models (LLMs) in classification and anomaly detection. In classification, LLMs show strong performance when the sample size is small. However, as the size of the dataset increases, TabPFN-based...
This study investigates the performance of large language models (LLMs) in classification and anomaly detection.
In classification, LLMs show strong performance when the sample size is small.
However, as the size of the dataset increases, TabPFN-based models and traditional classification models achieve better performance.
In contrast, for anomaly detection, LLMs outperform existing unsupervised and semi-supervised methods.
Nevertheless, this improved performance comes with a considerable increase in training time.
Therefore, when both cost and performance are taken into account, our results suggest that LLMs should be used selectively depending on the characteristics of the problem.
목차 (Table of Contents)