Near-Infrared and Low-Rank Adaptation of Vision Transformers in Remote   Sensing

Irem Ulku; O. Ozgur Tanriover; Erdem Akag\"und\"uz

arXiv:2405.17901·cs.CV·May 29, 2024

Near-Infrared and Low-Rank Adaptation of Vision Transformers in Remote Sensing

Irem Ulku, O. Ozgur Tanriover, Erdem Akag\"und\"uz

PDF

Open Access

TL;DR

This paper explores the use of vision transformers with low-rank adaptation to improve remote sensing tasks using Near-Infrared images, addressing domain shift issues and enhancing efficiency.

Contribution

It introduces a novel approach combining ViT pre-trained on RGB data with LoRA for NIR domain adaptation, which was not previously explored.

Findings

01

LoRA with ViT yields superior NIR task performance

02

Pre-trained ViT models benefit from low-rank adaptation in NIR domain

03

Efficient training without extensive fine-tuning

Abstract

Plant health can be monitored dynamically using multispectral sensors that measure Near-Infrared reflectance (NIR). Despite this potential, obtaining and annotating high-resolution NIR images poses a significant challenge for training deep neural networks. Typically, large networks pre-trained on the RGB domain are utilized to fine-tune infrared images. This practice introduces a domain shift issue because of the differing visual traits between RGB and NIR images.As an alternative to fine-tuning, a method called low-rank adaptation (LoRA) enables more efficient training by optimizing rank-decomposition matrices while keeping the original network weights frozen. However, existing parameter-efficient adaptation strategies for remote sensing images focus on RGB images and overlook domain shift issues in the NIR domain. Therefore, this study investigates the potential benefits of using…

Peer Reviews

No public reviews on file for this paper yet. If you reviewed it on a platform where reviews are public (OpenReview, ICLR, NeurIPS, ICML), you can paste yours below so the community can read it here.

Videos

No videos yet. Explain this paper in a talk, walkthrough, or lecture? Add one.

Taxonomy

TopicsInfrared Target Detection Methodologies · Remote-Sensing Image Classification

MethodsAttention Is All You Need · Dense Connections · Softmax · Focus · Layer Normalization · Linear Layer · Multi-Head Attention · Residual Connection · Vision Transformer