ReportGem ReportGem

Academic paper

DeVIT: Low-Power Vision Transformer Acceleration Using Delta Computation

Authors: Reyhaneh Hosseinzadeh, Parham Zilouchian Moghaddam, Mehdi ModarressiPublished: 2026-08-02Paper ID: 2608.01343Category: cs.CVLicense: CC BY 4.0

Abstract

The emergence of transformer-based deep learning models has brought unprecedented performance across various domains, particularly in natural language processing and computer vision. However, deploying these models, especially on resource-constrained devices, poses significant challenges due to their high computational complexity and large memory size and bandwidth requirements. This complexity has led researchers to use low-bit model weights to reduce memory usage and improve efficiency. In addition to reducing processing and memory demands, quantization introduces another useful property: value locality, where the extremely large number of parameters are restricted to a limited range of values. To fully take advantage of this locality, this paper presents DeVIT, an acceleration method for vision transformers that leverages differential computation to enable multiplier-less matrix multiplication.

This public page contains bibliographic metadata and the author abstract. Use the reader for licensed document access.

Open licensed paper reader