TensorIR: An Abstraction for Automatic Tensorized Program Optimization

Feng, Siyuan, Hou, Bohan, Jin, Hongyi, Lin, Wuwei, Shao, Junru, Lai, Ruihang, Ye, Zihao, Zheng, Lianmin, Yu, Cody Hao, Yu, Yong, Chen, Tianqi

Oct-27-2022–arXiv.org Artificial Intelligence

Deploying deep learning models on various devices has become an important topic. The wave of hardware specialization brings a diverse set of acceleration primitives for multi-dimensional tensor computations. These new acceleration primitives, along with the emerging machine learning models, bring tremendous engineering challenges. In this paper, we present TensorIR, a compiler abstraction for optimizing programs with these tensor computation primitives. TensorIR generalizes the loop nest representation used in existing machine learning compilers to bring tensor computation as the first-class citizen. Finally, we build an end-to-end framework on top of our abstraction to automatically optimize deep learning models for given tensor computation primitives. Experimental results show that TensorIR compilation automatically uses the tensor computation primitives for given hardware backends and delivers performance that is competitive to state-of-art hand-optimized systems across platforms.

artificial intelligence, computation, machine learning, (18 more...)

arXiv.org Artificial Intelligence

Oct-27-2022

arXiv.org PDF

Add feedback

Country:
- North America > United States
  - District of Columbia > Washington (0.04)
  - Texas > Harris County
    - Houston (0.04)
  - Pennsylvania > Allegheny County
    - Pittsburgh (0.04)
  - New York > New York County
    - New York City (0.14)
  - California > San Francisco County
    - San Francisco (0.14)
  - Arizona > Maricopa County
    - Phoenix (0.04)
- Europe
  - United Kingdom > England
    - Greater London > London (0.04)
  - Italy > Calabria
    - Catanzaro Province > Catanzaro (0.04)
- Asia
  - Middle East > Saudi Arabia
    - Riyadh Province > Riyadh (0.04)
  - China > Shanghai
    - Shanghai (0.04)

Genre:
- Research Report (0.70)

Technology:
- Information Technology > Artificial Intelligence > Machine Learning > Neural Networks > Deep Learning (1.00)

Duplicate Docs Excel Report

Title
None found

Similar Docs Excel Report more

Title	Similarity	Source
None found