Efficient Use of Limited-Memory Accelerators for Linear Learning on   Heterogeneous Systems

Celestine D\"unner; Thomas Parnell; Martin Jaggi

arXiv:1708.05357·cs.LG·November 8, 2017·2 cites

Efficient Use of Limited-Memory Accelerators for Linear Learning on Heterogeneous Systems

Celestine D\"unner, Thomas Parnell, Martin Jaggi

PDF

Open Access 1 Repo

TL;DR

This paper introduces a flexible algorithmic approach that enables efficient training of large-scale machine learning models on heterogeneous systems with limited memory, achieving significant speedups over existing methods.

Contribution

It presents a novel primal-dual coordinate method leveraging duality gap information for dynamic data management in heterogeneous environments.

Findings

01

Order-of-magnitude speedup over existing approaches

02

Effective training of models exceeding GPU memory capacity

03

Adaptability to various system memory hierarchies

Abstract

We propose a generic algorithmic building block to accelerate training of machine learning models on heterogeneous compute systems. Our scheme allows to efficiently employ compute accelerators such as GPUs and FPGAs for the training of large-scale machine learning models, when the training data exceeds their memory capacity. Also, it provides adaptivity to any system's memory hierarchy in terms of size and processing speed. Our technique is built upon novel theoretical insights regarding primal-dual coordinate methods, and uses duality gap information to dynamically decide which part of the data should be made available for fast processing. To illustrate the power of our approach we demonstrate its performance for training of generalized linear models on a large-scale dataset exceeding the memory size of a modern GPU, showing an order-of-magnitude speedup over existing approaches.

Peer Reviews

No public reviews on file for this paper yet. If you reviewed it on a platform where reviews are public (OpenReview, ICLR, NeurIPS, ICML), you can paste yours below so the community can read it here.

Code & Models

Repositories

ElizaWszola/HTHC
none

Videos

No videos yet. Explain this paper in a talk, walkthrough, or lecture? Add one.

Taxonomy

TopicsNeural Networks and Applications · Model Reduction and Neural Networks · Gaussian Processes and Bayesian Inference