Reinforced Bit Allocation under Task-Driven Semantic Distortion Metrics

10/16/2019
by   Jun Shi, et al.
0

Rapid growing intelligent applications require optimized bit allocation in image/video coding to support specific task-driven scenarios such as detection, classification, segmentation, etc. Some learning-based frameworks have been proposed for this purpose due to their inherent end-to-end optimization mechanisms. However, it is still quite challenging to integrate these task-driven metrics seamlessly into traditional hybrid coding framework. To the best of our knowledge, this paper is the first work trying to solve this challenge based on reinforcement learning (RL) approach. Specifically, we formulate the bit allocation problem as a Markovian Decision Process (MDP) and train RL agents to automatically decide the quantization parameter (QP) of each coding tree unit (CTU) for HEVC intra coding, according to the task-driven semantic distortion metrics. This bit allocation scheme can maximize the semantic level fidelity of the task, such as classification accuracy, while minimizing the bit-rate. We also employ gradient class activation map (Grad-CAM) and Mask R-CNN tools to extract task-related importance maps to help the agents make decisions. Extensive experimental results demonstrate the superior performance of our approach by achieving 43.1 saving over the anchor of HEVC under the equivalent task-related distortions.

READ FULL TEXT

page 1

page 2

page 4

research
06/07/2021

Task-driven Semantic Coding via Reinforcement Learning

Task-driven semantic video/image coding has drawn considerable attention...
research
04/21/2021

Visual Analysis Motivated Rate-Distortion Model for Image Coding

Optimized for pixel fidelity metrics, images compressed by existing imag...
research
08/08/2022

Towards Semantic Communications: Deep Learning-Based Image Semantic Coding

Semantic communications has received growing interest since it can remar...
research
09/27/2022

Neural Frank-Wolfe Policy Optimization for Region-of-Interest Intra-Frame Coding with HEVC/H.265

This paper presents a reinforcement learning (RL) framework that utilize...
research
04/05/2021

A Dual-Critic Reinforcement Learning Framework for Frame-level Bit Allocation in HEVC/H.265

This paper introduces a dual-critic reinforcement learning (RL) framewor...
research
12/25/2018

Learning based Facial Image Compression with Semantic Fidelity Metric

Surveillance and security scenarios usually require high efficient facia...

Please sign up or login with your details

Forgot password? Click here to reset