Pixelated Butterfly: Simple and Efficient Sparse training for Neural Network Models

Dao, Tri
Chen, Beidi
Liang, Kaizhao
Yang, Jiaming
Song, Zhao
Rudra, Atri
Ré, Christopher

Publication date

May 2022

Language

English

Abstract

Overparameterized neural networks generalize well but are expensive to train. Ideally, one would like to reduce their computational cost while retaining their generalization benefits. Sparse model training is a simple and promising approach to achieve this, but there remain challenges as existing methods struggle with accuracy loss, slow training runtime, or difficulty in sparsifying all model components. The core problem is that searching for a sparsity mask over a discrete set of sparse matrices is difficult and expensive. To address this, our main insight is to optimize over a continuous superset of sparse matrices with a fixed structure known as products of butterfly matrices. As butterfly matrices are not hardware efficient, we propose...

Extracted data

We use cookies to provide a better user experience.

Data Protection

Pixelated Butterfly: Simple and Efficient Sparse training for Neural Network Models

Abstract

Extracted data

Pixelated Butterfly: Simple and Efficient Sparse training for Neural Network Models

Abstract

Extracted data

Related items

Related items