Optimizing Deep Learning Architectures for Resource-Constrained Edge Devices

Main Article Content

Shihabul Haq. M
Joshna. M
Joshna. M
Dr. Muhamed Illyas. P

Abstract

Implementing large-scale transformer models on edge devices in precision agriculture causes high computation and memory cost. While traditional CNN has limited global contextual modeling capabilities, the large number of parameters required by a Vision Transformer makes it difficult to run on local agricultural hardware, without the use of clusters of graphics processing units (GPUs) in the cloud. This study aims at understanding how to optimize deep learning architectures for resource-constrained edge devices using sophisticated compression techniques in deep learning. In detail, the effectiveness of the Activation-aware Weight Quantization and structured pruning of Multi Head Self-Attention models is examined. These optimization techniques allow the use of a light-weight hybrid vision system, significantly reducing computing power while maintaining diagnostic performance for crop disease classification problems. The analytical results show that by using a 4-bit quantization scheme and a ½ structured pruning ratio, the memory usage and inference time in platforms like Raspberry Pi 5 and NVIDIA Jetson Orin Nano can be significantly reduced. Overall, the findings validate the feasibility of using heavy transformer models for real time, on-device crop disease detection in agriculture to enable offline and decentralized smart farming without relying on cloud connectivity at all times.

Article Details

Section

Articles

How to Cite

M, S. H., M, J., M, J., & Illyas. P, D. M. (2026). Optimizing Deep Learning Architectures for Resource-Constrained Edge Devices. International Journal of Aquatic Research and Environmental Studies, 6(S2), 1069-1075. https://doi.org/10.70102/39671g67

Similar Articles

You may also start an advanced similarity search for this article.