딥러닝 가속기는 인공 지능 및 기계 학습 응용 프로그램을 가속하도록 설계된 특수 하드웨어 장치이다. 딥러닝 가속기를 활용하여 YOLO 객체 검출 모델의 GPU 전력 소모를 줄이고 행동 추정 추...

http://chineseinput.net/에서 pinyin(병음)방식으로 중국어를 변환할 수 있습니다.
변환된 중국어를 복사하여 사용하시면 됩니다.
https://www.riss.kr/link?id=T16951041
청주 : 충북대학교, 2024
2024
한국어
YOLO, Jetson ; NPU ; DLA ; 경량화
004.73 판사항(5)
충청북도
A Search Strategy for Alternative Operations of DLA-unsupported Operations in YOLO Models
49p. ; 26cm
충북대학교 논문은 저작권에 의해 보호됩니다
지도교수:이건명
참고문헌: p.47-49
I804:43009-000000059172
0
상세조회0
다운로드딥러닝 가속기는 인공 지능 및 기계 학습 응용 프로그램을 가속하도록 설계된 특수 하드웨어 장치이다. 딥러닝 가속기를 활용하여 YOLO 객체 검출 모델의 GPU 전력 소모를 줄이고 행동 추정 추...
딥러닝 가속기는 인공 지능 및 기계 학습 응용 프로그램을 가속하도록 설계된 특수 하드웨어 장치이다. 딥러닝 가속기를 활용하여 YOLO 객체 검출 모델의 GPU 전력 소모를 줄이고 행동 추정 추론 환경의 성능 향상을 가져온다. 딥러닝 가속기가 지원하지 않는 블록 연산이 존재하고, 딥러닝 가속기가 연산할 수 있도록 블록 연산을 대체한다. 대체 연산이 적용된 모델을 TensorRT를 통해 딥러닝 가속기가 추론할 수 있도록 컴파일을 수행하며, 블록 연산 대체의 적용 여부를 확인한다. 대체 연산이 적용 결과 기존 모델 보다 초당 처리 프레임의 향상과 표준 편차가 크지 않는 안정적인 GPU 이용률을 기록하였다. GPU를 대체할 수 있는 딥러닝 가속기를 효율적으로 활용하는 방법을 제안하였으며, 딥러닝 추론 환경의 에너지 효율성 향상 연구에 기여 할 수 있을 것으로 생각한다.
다국어 초록 (Multilingual Abstract)
Deep learning accelerators are specialized hardware devices designed to accelerate artificial intelligence and machine learning applications. By utilizing deep learning accelerators, the GPU power consumption of the YOLO object detection model is redu...
Deep learning accelerators are specialized hardware devices designed to accelerate artificial intelligence and machine learning applications. By utilizing deep learning accelerators, the GPU power consumption of the YOLO object detection model is reduced, and the performance of the behavior estimation inference environment is enhanced. There are block operations that deep learning accelerators do not support, and block operations are replaced so that deep learning accelerators can perform operations. Compile the model to which the substitution operation is applied so that the deep learning accelerator can infer it through TensorRT, and check whether the block operation substitution is applied. The application of these alternative operations results in an improved frame rate per second compared to the original model, while maintaining a stable GPU utilization rate with minimal standard deviation. We proposed a method to efficiently utilize deep learning accelerators that can replace GPUs, and we believe that this will contribute to research on improving energy efficiency in deep learning inference environments.
목차 (Table of Contents)