Advanced fine-tuning procedures to enhance DNN robustness in visual coding for machines

Autor:	Alban Marie, Karol Desnos, Alexandre Mercat, Luce Morin, Jarno Vanne, Lu Zhang
Jazyk:	angličtina
Rok vydání:	2024
Předmět:	Video Coding for Machines (VCM) Coding artifacts Image and video coding Deep neural network (DNN) Deep learning Fine-tuning Electronics TK7800-8360
Zdroj:	EURASIP Journal on Image and Video Processing, Vol 2024, Iss 1, Pp 1-23 (2024)
Druh dokumentu:	article
ISSN:	1687-5281
DOI:	10.1186/s13640-024-00650-3
Popis:	Abstract Video Coding for Machines (VCM) is gaining momentum in applications like autonomous driving, industry manufacturing, and surveillance, where the robustness of machine learning algorithms against coding artifacts is one of the key success factors. This work complements the MPEG/JVET standardization efforts in improving the resilience of deep neural network (DNN)-based machine models against such coding artifacts by proposing the following three advanced fine-tuning procedures for their training: (1) the progressive increase of the distortion strength as the training proceeds; (2) the incorporation of a regularization term in the original loss function to minimize the distance between predictions on compressed and original content; and (3) a joint training procedure that combines the proposed two approaches. These proposals were evaluated against a conventional fine-tuning anchor on two different machine tasks and datasets: image classification on ImageNet and semantic segmentation on Cityscapes. Our joint training procedure is shown to reduce the training time in both cases and still obtain a 2.4% coding gain in image classification and 7.4% in semantic segmentation, whereas a slight increase in training time can bring up to 9.4% better coding efficiency for the segmentation. All these coding gains are obtained without any additional inference or encoding time. As these advanced fine-tuning procedures are standard-compliant, they offer the potential to have a significant impact on visual coding for machine applications.
Databáze:	Directory of Open Access Journals
Externí odkaz:	https://doaj.org/article/56247cf086994635a26371fa90a1ad1e Zobrazit plný text záznamu Full text from SpringerLink View record in DOAJ