서포트 벡터 머신(Support Vector Machine; SVM)은 분류문제에서 강력한 성능을 보이지만 커널공간(kernel space)을 이용하기 때문에 변수의 중요도를 책정하기 어렵고 변수 선택 역시 쉽지 않다. 기존...

http://chineseinput.net/에서 pinyin(병음)방식으로 중국어를 변환할 수 있습니다.
변환된 중국어를 복사하여 사용하시면 됩니다.
https://www.riss.kr/link?id=T17387637
서울 : 성균관대학교 일반대학원, 2026
학위논문(석사) -- 성균관대학교 일반대학원 , 통계학과 , 2026. 2
2026
한국어
서울
Efficient variable selection for support vector machine in high-dimensional data
56 p. : 삽화 ; 30 cm
지도교수: 김재직
참고문헌: p. 52-54
I804:11040-000000189877
0
상세조회0
다운로드서포트 벡터 머신(Support Vector Machine; SVM)은 분류문제에서 강력한 성능을 보이지만 커널공간(kernel space)을 이용하기 때문에 변수의 중요도를 책정하기 어렵고 변수 선택 역시 쉽지 않다. 기존...
서포트 벡터 머신(Support Vector Machine; SVM)은 분류문제에서 강력한 성능을 보이지만 커널공간(kernel space)을 이용하기 때문에 변수의 중요도를 책정하기 어렵고 변수 선택 역시 쉽지 않다. 기존의 SVM을 위한 변수 선택 방법들은 적은 변수의 수를 상정하여 교차검증(cross-validation)을 통해 예측률을 구하고 이를 기준으로 후진제거(backward elimination)를 진행하였다. 하지만 이러한 방식은 고차원 데이터에 적용하기에 계산 비용측면에서 문제가 있다. 본 연구에서는 고차원 데이터에 대한 SVM의 변수 선택 문제를 해결하기 위해 변수 거르기(variable screening)와 재귀적 변수 제거(Recursive Feature Elimination; RFE)를 결합하는 효율적인 변수 선택 방법을 제시한다. 제안하는 방법은 먼저 SVS(Sufficient Variable Screening)를 적용하여 비선형성과 변수들의 조합을 모두 고려하여 변수의 개수를 줄인 후, 축소된 변수 집합에 SVM-RFE를 적용한다. 더 나아가 본 방법에서는 SVM-RFE가 변수의 랭킹만 정하던 문제를 해결하여 변수선택이 가능하도록 개선한다. 제시한 방법론의 효과를 검증하기 다양한 상황에 대한 모의실험을 진행하고, 실제 유전자 데이터에 적용하여 그 활용성을 검증한다.
다국어 초록 (Multilingual Abstract)
Support Vector Machine (SVM) deliver strong performance in classification, but their operation in a kernel feature space makes variable importance hard to quantify and feature selection non-trivial. Many existing SVM feature selection schemes assume a...
Support Vector Machine (SVM) deliver strong performance in classification, but their operation in a kernel feature space makes variable importance hard to quantify and feature selection non-trivial. Many existing SVM feature selection schemes assume a small number of candidates: they rank variables via cross-validated predictive accuracy and then perform backward elimination. Such procedures scale poorly and become unreliable in high-dimensional regimes.
We propose an efficient two-stage method for high-dimensional SVM feature selection that couples variable screening with Recursive Feature Elimination (RFE). In the first stage, we replace the commonly used Iterative Sure Independence Screening (ISIS) with Sufficient Variable Screening (SVS), which retains both nonlinear relationships and joint effects, thereby preserving the signal structure most relevant to SVM. In the second stage, we adopt a modified SVM-RFE that goes beyond ranking: a permutation-test–based stopping rule decides when elimination should stop. This yields a data-driven determination of the final subset size while avoiding unnecessary computation.
We assess the proposed method in simulated settings to evaluate variable recovery and selection stability, and further demonstrate its practical value on real gene-expression classification tasks.
목차 (Table of Contents)