Hierarchical Representation Network for Steganalysis of QIM Steganography in Low-Bit-Rate Speech Signals

10/10/2019
by   Hao Yang, et al.
0

With the Volume of Voice over IP (VoIP) traffic rises shapely, more and more VoIP-based steganography methods have emerged in recent years, which poses a great threat to the security of cyberspace. Low bit-rate speech codecs are widely used in the VoIP application due to its powerful compression capability. QIM steganography makes it possible to hide secret information in VoIP streams. Previous research mostly focus on capturing the inter-frame correlation or inner-frame correlation features in code-words but ignore the hierarchical structure which exists in speech frame. In this paper, motivated by the complex multi-scale structure, we design a Hierarchical Representation Network to tackle the steganalysis of QIM steganography in low-bit-rate speech signal. In the proposed model, Convolution Neural Network (CNN) is used to model the hierarchical structure in the speech frame, and three level of attention mechanisms are applied at different convolution block, enabling it to attend differentially to more and less important content in speech frame. Experiments demonstrated that the steganalysis performance of the proposed method can outperforms the state-of-the-art methods especially in detecting both short and low embeded speech samples. Moreover, our model needs less computation and has higher time efficiency to be applied to real online services.

READ FULL TEXT
research
02/05/2019

An Enhanced Interleaving Frame Loss Concealment Method for Voice Over IP Network Services

This paper focuses on AMR WB G.722.2 speech codec, and discusses the unu...
research
01/13/2021

F3SNet: A Four-Step Strategy for QIM Steganalysis of Compressed Speech Based on Hierarchical Attention Network

Traditional machine learning-based steganalysis methods on compressed sp...
research
08/09/2021

A Streamwise GAN Vocoder for Wideband Speech Coding at Very Low Bit Rate

Recently, GAN vocoders have seen rapid progress in speech synthesis, sta...
research
02/04/2019

Real-Time Steganalysis for Stream Media Based on Multi-channel Convolutional Sliding Windows

Previous VoIP steganalysis methods face great challenges in detecting sp...
research
11/02/2019

FCEM: A Novel Fast Correlation Extract Model For Real Time Steganalysis of VoIP Stream via Multi-head Attention

Extracting correlation features between codes-words with high computatio...
research
10/31/2019

Fast Steganalysis Method for VoIP Streams

In this letter, we present a novel and extremely fast steganalysis metho...
research
04/18/2023

Coded Speech Quality Measurement by a Non-Intrusive PESQ-DNN

Wideband codecs such as AMR-WB or EVS are widely used in (mobile) speech...

Please sign up or login with your details

Forgot password? Click here to reset