RISS 학술연구정보서비스

검색

인기 검색어

    다국어 입력

    http://chineseinput.net/에서 pinyin(병음)방식으로 중국어를 변환할 수 있습니다.

    변환된 중국어를 복사하여 사용하시면 됩니다.

    예시)
    • 中文 을 입력하시려면 zhongwen을 입력하시고 space를누르시면됩니다.
    • 北京 을 입력하시려면 beijing을 입력하시고 space를 누르시면 됩니다.
    닫기

    Resource-Aware Distributed Analytics and Machine Learning for Hybrid Edge-Cloud Systems.

    한글로보기

    https://www.riss.kr/link?id=T16601248

    • 저자
    • 발행사항

      Ann Arbor : ProQuest Dissertations & Theses, 2021

    • 학위수여대학

      Rensselaer Polytechnic Institute Computer Science

    • 수여연도

      2021

    • 작성언어

      영어

    • 주제어
    • 학위

      Ph.D.

    • 페이지수

      203 p.

    • 지도교수/심사위원

      Advisor: Patterson, Stacy.

    • 0

      상세조회
    • 0

      다운로드
    서지정보 열기
    • 내보내기
    • 내책장담기
    • 공유하기
    • 오류접수

    소속기관이 구독 중이 아닌 경우 오후 4시부터 익일 오전 9시까지 원문보기가 가능합니다.

    부가정보

    다국어 초록 (Multilingual Abstract) kakao i 다국어 번역

    With more intelligent applications, data analytics and inference at the edge are proliferating as a complement to traditional computation done at a centralized cloud location. At the same time, distributed machine learning training at the edge of the network near the data producers is also gaining popularity, mainly due to benefits to security, privacy, and communication costs. However, edge devices are often resource-constrained, and further, there may be communication bottlenecks between the edge and the cloud. Successful solutions for these edge computing workloads must address challenges posed by constrained computation and communication resources.The first part of the thesis focuses on scheduling and task placement of data processing, analytics, and inference workloads. The goal is to provide some quality of service, for example, low latency or cost reduction in the context of edge-cloud architectures. We start with benchmarking the leading industry edge computing platforms that use the serverless computing paradigm as the medium of execution. We next consider serverless applications, consisting of a single-stage, and propose a framework to jointly execute such applications in the presence of an edge device and the public cloud. The aim is to decide whether to execute user jobs at the edge or the public cloud based on given latency or cost constraints. Finally, we consider a hybrid cloud scenario, where we consider a private cloud instead of a single edge device. Here, we study the problem of task placement and scheduling of multi-stage serverless applications between a private and the public cloud to minimize the cost of public cloud usage. In the second part of the thesis, we consider machine learning training workloads in edge-cloud platforms. More specifically, we study federated learning in this part of the thesis. Like the first part of the thesis, we first conduct a feasibility study of federated learning algorithms on resource-constrained devices. Next, we study an algorithm for horizontal federated learning in a hierarchical communication network. We analyze the convergence of the algorithm when there is a non-IID data distribution among the participants. Our analysis shows that the non-IID data distribution can have a significant impact on the algorithm convergence error. This insight paves the way for a more sophisticated algorithm design to diminish this performance gap. We then turn our focus towards vertical federated learning in a hierarchical network. We propose a new algorithm for model training where data is vertically partitioned across silos in the top tier and horizontally partitioned in the bottom tier among clients inside each silo. We present a theoretical analysis of our algorithm and show the dependence of the convergence rate on the number of vertical partitions, the number of local updates, and the number of clients in each hub.Lastly, we close with the summary and discussions on the future research directions and open questions of interest.
    번역하기

    With more intelligent applications, data analytics and inference at the edge are proliferating as a complement to traditional computation done at a centralized cloud location. At the same time, distributed machine learning training at the edge of the...

    With more intelligent applications, data analytics and inference at the edge are proliferating as a complement to traditional computation done at a centralized cloud location. At the same time, distributed machine learning training at the edge of the network near the data producers is also gaining popularity, mainly due to benefits to security, privacy, and communication costs. However, edge devices are often resource-constrained, and further, there may be communication bottlenecks between the edge and the cloud. Successful solutions for these edge computing workloads must address challenges posed by constrained computation and communication resources.The first part of the thesis focuses on scheduling and task placement of data processing, analytics, and inference workloads. The goal is to provide some quality of service, for example, low latency or cost reduction in the context of edge-cloud architectures. We start with benchmarking the leading industry edge computing platforms that use the serverless computing paradigm as the medium of execution. We next consider serverless applications, consisting of a single-stage, and propose a framework to jointly execute such applications in the presence of an edge device and the public cloud. The aim is to decide whether to execute user jobs at the edge or the public cloud based on given latency or cost constraints. Finally, we consider a hybrid cloud scenario, where we consider a private cloud instead of a single edge device. Here, we study the problem of task placement and scheduling of multi-stage serverless applications between a private and the public cloud to minimize the cost of public cloud usage. In the second part of the thesis, we consider machine learning training workloads in edge-cloud platforms. More specifically, we study federated learning in this part of the thesis. Like the first part of the thesis, we first conduct a feasibility study of federated learning algorithms on resource-constrained devices. Next, we study an algorithm for horizontal federated learning in a hierarchical communication network. We analyze the convergence of the algorithm when there is a non-IID data distribution among the participants. Our analysis shows that the non-IID data distribution can have a significant impact on the algorithm convergence error. This insight paves the way for a more sophisticated algorithm design to diminish this performance gap. We then turn our focus towards vertical federated learning in a hierarchical network. We propose a new algorithm for model training where data is vertically partitioned across silos in the top tier and horizontally partitioned in the bottom tier among clients inside each silo. We present a theoretical analysis of our algorithm and show the dependence of the convergence rate on the number of vertical partitions, the number of local updates, and the number of clients in each hub.Lastly, we close with the summary and discussions on the future research directions and open questions of interest.

    더보기

    분석정보

    View

    상세정보조회

    0

    Usage

    원문다운로드

    0

    대출신청

    0

    복사신청

    0

    EDDS신청

    0

    동일 주제 내 활용도 TOP

    더보기

    주제

    연도별 연구동향

    연도별 활용동향

    연관논문

    연구자 네트워크맵

    공동연구자 (7)

    유사연구자 (20) 활용도상위20명

    이 자료와 함께 이용한 RISS 자료

    나만을 위한 추천자료

    해외이동버튼