TensorIR: An Abstraction for Automatic Tensorized Program Optimization

Feng, Siyuan
Hou, Bohan
Jin, Hongyi
Lin, Wuwei
Shao, Junru
Lai, Ruihang
Ye, Zihao
Zheng, Lianmin
Yu, Cody Hao
Yu, Yong
Chen, Tianqi

Publication date

October 2022

Language

English

Abstract

Deploying deep learning models on various devices has become an important topic. The wave of hardware specialization brings a diverse set of acceleration primitives for multi-dimensional tensor computations. These new acceleration primitives, along with the emerging machine learning models, bring tremendous engineering challenges. In this paper, we present TensorIR, a compiler abstraction for optimizing programs with these tensor computation primitives. TensorIR generalizes the loop nest representation used in existing machine learning compilers to bring tensor computation as the first-class citizen. Finally, we build an end-to-end framework on top of our abstraction to automatically optimize deep learning models for given tensor computatio...

Extracted data

We use cookies to provide a better user experience.

Data Protection

TensorIR: An Abstraction for Automatic Tensorized Program Optimization

Abstract

Extracted data

TensorIR: An Abstraction for Automatic Tensorized Program Optimization

Abstract

Extracted data

Related items

Related items