Audio Steganography: A Comprehensive Survey of Techniques and Emerging Challenges
DOI:
https://doi.org/10.54388/jkues.v4i2.356Keywords:
Audio steganography, Tone insertion, Payload capacity, ImperceptibilityAbstract
Recently, information-hiding techniques have become important due to the widespread use of multimedia and the Internet. Actually, information hiding falls into two large categories: steganography, which is the art and science of secure communication, and watermarking, which is used to protect copyright and data ownership. Audio has some unique features that have motivated many researchers and developers to use it as an information-hiding medium. The audio, referred to as cover audio, is used to conceal a secret message, making it unnoticeable to the observer due to steganography techniques. This paper reviews audio steganography techniques and classifies them scientifically. The strengths and limitations of each steganography technique are presented in this work, moreover, it shows how the researchers treat the weaknesses of each method. The paper highlights an audio steganography emerging challenges such as advanced hiding methods and real-time steganography systems which have special constraints discussed in this paper. Audio steganography can utilize cover audio in various domains, which we have compared in our work, along with a comparison of steganography techniques. Subjective and objective quality measurements of audio steganography are depicted in this paper
References
A. H. Ali, L. E. George, A. Zaidan, and M. R. Mokhtar, “High capacity, transparent and secure audio steganography model based on fractal coding and chaotic map in temporal domain,” Multimedia Tools and Applications, vol. 77, no. 23, pp. 31487–31516, 2018.
H. Baisuni, D. Djulaeka, and M. A. Sajjad, “Legal protection for unauthorized copying of songs on digital platforms through audio watermarking method,” JUSTISI, vol. 10, no. 3, pp. 547–564, 2024.
F. Hemeida, W. Alexan, and S. Mamdouh, “A comparative study of audio steganography schemes,” International Journal of Computing and Digital Systems, vol. 10, pp. 555–562, 2021.
N. Duarte, N. Coelho, and T. Guarda, “Social engineering: The art of attacks,” in International Conference on Advanced Research in Technologies, Information, Innovation and Sustainability, pp. 474–483, Springer, 2021.
M. Rakhra, R. Kumar, H. Walia, et al., “A review on data hiding using steganography and cryptography,” in 2021 9th International Conference on Reliability, Infocom Technologies and Optimization (Trends and Future Directions)(ICRITO), pp. 1–4, IEEE, 2021.
H. Dutta, R. K. Das, S. Nandi, and S. M. Prasanna, “An overview of digital audio steganography,” IETE Technical Review, vol. 37, no. 6, pp. 632–650, 2020.
W. R. Bender, D. Gruhl, and N. Morimoto, “Techniques for data hiding,” in Storage and Retrieval for Image and Video Databases III, vol. 2420, pp. 164–173, SPIE, 1995.
M. N. Kunchur, “The human auditory system and audio,” Applied Acoustics, vol. 211, p. 109507, 2023.
F. Q. A. Alyousuf, R. Din, and A. J. Qasim, “Analysis review on spatial and transform domain technique in digital steganography,” Bulletin of Electrical Engineering and Informatics, vol. 9, no. 2, pp. 573–581, 2020.
P. M. Reyers, “A comparative analysis of audio steganography methods and tools,” B.S. thesis, University of Twente, 2023.
K. Gopalan, “Audio steganography for information hiding and covert communication–a tutorial,” in 2018 IEEE International Conference on Electro/Information Technology (EIT), pp. 0242–0243, IEEE, 2018.
A. A. Alsabhany, F. Ridzuan, and A. Azni, “The adaptive multi-level phase coding method in audio steganography,” IEEE Access, vol. 7, pp. 129291–129306, 2019.
D. M. Ballesteros L and J. M. Moreno A, “Highly transparent steganog- raphy model of speech signals using efficient wavelet masking,” Expert Systems with Applications, vol. 39, no. 10, pp. 9141–9149, 2012.
F. A. Petitcolas, R. J. Anderson, and M. G. Kuhn, “Information hiding-a survey,” Proceedings of the IEEE, vol. 87, no. 7, pp. 1062–1078, 2002.
F. Djebbar, B. Ayad, K. A. Meraim, and H. Hamam, “Comparative study of digital audio steganography techniques,” EURASIP Journal on Audio, Speech, and Music Processing, vol. 2012, no. 1, p. 25, 2012.
M. Driss, L. Berriche, S. B. Atitallah, and S. Rekik, “Steganography in iot: A comprehensive survey on approaches, challenges, and future directions,” IEEE Access, 2025.
A. Singh and H. Singh, “An improved lsb based image steganography technique for rgb images,” in 2015 IEEE International Conference on electrical, computer and communication technologies (ICECCT), pp. 1–4, IEEE, 2015.
K. Bhowal, A. J. Pal, G. S. Tomar, and P. Sarkar, “Audio steganog- raphy using ga,” in 2010 International Conference on Computational Intelligence and Communication Networks, pp. 449–453, IEEE, 2010.
U. Ehsan Ali, E. Ali, M. Sohrawordi, and M. N. Sultan, “A lsb based image steganography using random pixel and bit selection for high payload,” 2021.
N. Cvejic and T. Seppanen, “Increasing the capacity of lsb-based audio steganography,” in 2002 IEEE Workshop on Multimedia Signal Processing., pp. 336–338, IEEE, 2002.
C. Parthasarathy and S. Srivatsa, “Increased robustness of lsb audio steganography by reduced distortion lsb coding,” Journal of Theoretical and Applied Information Technology, vol. 7, no. 1, pp. 080–086, 2005.
R. Martinez-Noriega, H. Kang, B. Kurkoski, K. Yamaguchi, K. Kobayashi, and M. Nakano, “Increasing robustness of audio wa- termarking dm using athc codes,” in Proceedings of the Mexican conference on informatics security, 2006.
B. Harjito, B. Sulistyarso, and E. Suryani, “Audio steganography using two lsb modification and rsa for security data transmission,” Journal of Telecommunication, Electronic and Computer Engineering (JTEC), vol. 10, no. 2-4, pp. 107–111, 2018.
I. Q. Abduljaleel and A. H. Khaleel, “Hiding text in speech signal using k-means, lsb techniques and chaotic maps.,” International Journal of Electrical & Computer Engineering (2088-8708), vol. 10, no. 6, 2020.
X. Sun, K. Wang, and S. Li, “Audio steganography with less modifica- tion to the optimal matching cnv-qim path with the minimal hamming distance expected value to a secret,” Multimedia Systems, vol. 27, no. 3, pp. 341–352, 2021.
Z. Wu, R. Li, and C. Li, “Adaptive speech information hiding method based on k-means,” IEEE Access, vol. 8, pp. 23308–23316, 2020.
G. P. TVS and S. Varadarajan, “A novel hybrid audio steganography for imperceptible data hiding,” in 2015 International Conference on Communications and Signal Processing (ICCSP), pp. 0634–0638, IEEE, 2015.
M. Tayel, A. Gamal, and H. Shawky, “A proposed implementation method of an audio steganography technique,” in 2016 18th interna- tional conference on advanced communication technology (ICACT), pp. 180–184, IEEE, 2016.
D. Gruhl, A. Lu, and W. Bender, “Echo hiding,” in International Workshop on Information Hiding, pp. 295–315, Springer, 1996.
M. Velayatipour, M. Mosleh, M. Y. Nejad, and M. Kheyrandish, “Quantum reversible circuits for audio watermarking based on echo hiding technique,” Quantum Information Processing, vol. 21, no. 9, p. 316, 2022.
S. A. Khayam, “The discrete cosine transform (dct): theory and application,” Michigan State University, vol. 114, no. 1, p. 31, 2003.
P. Venugopala, S. Raghavendra, B. Ashwini, et al., “Crypticcare: A strategic approach to telemedicine security using lsb and dct steganog- raphy for enhancing the patient data protection,” IEEE Access, 2024.
B. A. Dharani, B. Yashaswini, G. R. S. Reddy, and S. M. Rajagopal, “Multimodal steganography: a comparative analysis of lsb and dct methods for image and audio data concealment,” in 2024 IEEE 9th International Conference for Convergence in Technology (I2CT), pp. 1– 5, IEEE, 2024.
M. P. Jain and P. Trivedi, “Effective audio steganography by using coefficient comparison in dct domain,” International Journal of Engi- neering Research & Technology (IJERT), vol. 2, no. 8, 2013.
A. Kanhe and G. Aghila, “Dct based audio steganography in voiced and un-voiced frames,” in Proceedings of the international conference on informatics and analytics, pp. 1–4, 2016.
A. Kanhe and G. Aghila, “A dct–svd-based speech steganography in voiced frames,” Circuits, systems, and signal processing, vol. 37, no. 11, pp. 5049–5068, 2018.
S. Hemalatha, U. D. Acharya, and A. Renuka, “Wavelet transform based steganography technique to hide audio signals in image,” Pro- cedia Computer Science, vol. 47, pp. 272–281, 2015.
S. Gupta and N. Dhanda, “Audio steganography using discrete wavelet transformation (dwt) & discrete cosine transformation (dct),” IOSR Journal of Computer Engineering, vol. 17, no. 2, pp. 32–44, 2015.
M. Sheikhan, K. Asadollahi, and R. Shahnazi, “Improvement of embed- ding capacity and quality of dwt-based audio steganography systems,” World Applied Sciences Journal, vol. 13, no. 3, pp. 507–516, 2011.
S.-T. Chen, T.-W. Huang, and C.-T. Yang, “High-snr steganography for digital audio signal in the wavelet domain,” Multimedia Tools and Applications, vol. 80, no. 6, pp. 9597–9614, 2021.
L. Gang, A. N. Akansu, and M. Ramkumar, “Mp3 resistant oblivious steganography,” in 2001 IEEE International Conference on Acoustics, Speech, and Signal Processing. Proceedings (Cat. No. 01CH37221), vol. 3, pp. 1365–1368, IEEE, 2001.
X. Dong, M. F. Bocko, and Z. Ignjatovic, “Data hiding via phase manipulation of audio signals,” in 2004 IEEE International Conference on Acoustics, Speech, and Signal Processing, vol. 5, pp. V–377, IEEE, 2004.
K. Gopalan, S. Wenndt, A. Noga, D. Haddad, and S. Adams, “Covert speech communication via cover speech by tone insertion,” in 2003 IEEE Aerospace Conference Proceedings (Cat. No. 03TH8652), vol. 4, pp. 4 1647–4 1653, IEEE, 2003.
K. Gopalan and S. Wenndt, “Audio steganography for covert data transmission by imperceptible tone insertion,” in Proc. The IASTED International Conference on Communication Systems And Applications (CSA 2004), Banff, Canada, 2004.
M. H. Sayed, “Audio steganography using tone insertion technique,”
B. A. Patil and V. A. Chakkarwar, “Review of an improved audio steganographic technique over lsb through random based approach,” IOSR J Comput Eng, vol. 9, no. 1, pp. 30–34, 2013.
Z. Kexin, “Audio steganalysis of spread spectrum hiding based on statistical moment,” in 2010 2nd International Conference on Signal Processing Systems, vol. 3, pp. V3–381, IEEE, 2010.
R. Kaur, A. Thakur, H. S. Saini, and R. Kumar, “Enhanced stegano- graphic method preserving base quality of information using lsb, parity and spread spectrum technique,” in 2015 Fifth International Conference on Advanced Computing & Communication Technologies, pp. 148–152, IEEE, 2015.
R. Tanwar, K. Singh, M. Zamani, A. Verma, and P. Kumar, “An optimized approach for secure data transmission using spread spectrum audio steganography, chaos theory, and social impact theory optimizer,” Journal of computer networks and communications, vol. 2019, no. 1, p. 5124364, 2019.
B. Q. A. Ali, H. I. Shahadi, M. S. Kod, and H. R. Farhan, “Covert voip communication based on audio steganography,” International Journal of Computing and Digital Systems, vol. 11, no. 1, pp. 821–830, 2022.
H. Kheddar and D. Megias, “High capacity speech steganography for the g723. 1 coder based on quantised line spectral pairs interpolation and cnn auto-encoding,” Applied Intelligence, vol. 52, no. 8, pp. 9441– 9459, 2022.
E. A. Alsolami, “Audio steganography method using least significant bit (lsb) encoding technique,” Journal Of Theoretical And Applied Information Technology, vol. 100, p. 12, 2022.
C. Zhang, S. Jiang, and Z. Chen, “Spm: estimating payload locations of qim-based steganography in low-bit-rate compressed speeches,” Multimedia Tools and Applications, vol. 83, no. 37, pp. 85227–85252, 2024.
F. Li, B. Li, L. Liu, Y. Feng, and L. Peng, “A steganography method for g. 729a speech coding,” International Journal of Performability Engineering, vol. 16, no. 5, p. 784, 2020.
J. N. Hasoon and S. N. Al-Saad, “Speech hiding based on quantization of lpc parameters,”
N. Tahilramani and N. Bhatt, “Information hiding in proposed 10.6 kbps cs-acelp based speech codec using quantization index modu- lation,” International Journal of Speech Technology, vol. 25, no. 1, pp. 219–230, 2022.
S. Talati, P. Etezadifar, M. H. Ahangar, and M. Molazade, “Investi- gation of steganography methods in audio standard coders: Lpc, celp, melp,” Majlesi Journal of Telecommunication Devices, vol. 12, no. 1, pp. 7–15, 2023.
H. Tan, C. Liu, Y. Lyu, X. Zhang, D. Zhang, and Z. Gu, “Audio steganography with speech recognition system,” in 2021 IEEE Sixth International Conference on Data Science in Cyberspace (DSC), pp. 256–263, IEEE, 2021.
Z. Wang, O. Byrnes, H. Wang, R. Sun, C. Ma, H. Chen, Q. Wu, and M. Xue, “Data hiding with deep learning: A survey unifying digital wa- termarking and steganography,” IEEE Transactions on Computational Social Systems, vol. 10, no. 6, pp. 2985–2999, 2023.
M. Nelson and J.-L. Gailly, “The data compression book 2nd edition,” M & T Books, New York, NY, 1995.
A. H. Ali, L. E. George, and M. R. Mokhtar, “An adaptive high capacity model for secure audio communication based on fractal coding and uniform coefficient modulation,” Circuits, Systems, and Signal Processing, vol. 39, no. 10, pp. 5198–5225, 2020.
B. A. Mitras and N. F. Hassan, “Using hybird genetic algorithm in audio steganography,” Iraqi Journal of Statistical Sciences, vol. 13, no. 25, pp. 150–164, 2013.
A. Q. Raheema, “Adopt an optimal location using a genetic algorithm for audio steganography,” Periodicals of Engineering and Natural Sciences (PEN), vol. 9, no. 4, pp. 1070–1082, 2021.
M. H. N. Azam, F. H. M. Ridzuan, and M. N. S. M. Sayuti, “Opti- mized cover selection for audio steganography using multi-objective evolutionary algorithm,” Journal of Information and Communication Technology, vol. 22, no. 2, pp. 255–282, 2023.
D. Ye, S. Jiang, and J. Huang, “Heard more than heard: An audio steganography method based on gan,” arXiv preprint arXiv:1907.04986, 2019.
L. Chen, R. Wang, D. Yan, and J. Wang, “Learning to generate steganographic cover for audio steganography using gan,” IEEE Access, vol. 9, pp. 88098–88107, 2021.
R. Zhang, H. Dong, Z. Yang, W. Ying, and J. Liu, “A cnn based visual audio steganography model,” in International conference on adaptive and intelligent systems, pp. 431–442, Springer, 2022.
J. Wang, R. Wang, L. Dong, and D. Yan, “Robust, imperceptible and end-to-end audio steganography based on cnn,” in International conference on security and privacy in digital economy, pp. 427–442, Springer, 2020.
Z. Yang, X. Du, Y. Tan, Y. Huang, and Y.-J. Zhang, “Aag-stega: Automatic audio generation-based steganography,” arXiv preprint arXiv:1809.03463, 2018.
Z. Lin, Y. Huang, and J. Wang, “Rnn-sm: Fast steganalysis of voip streams using recurrent neural network,” IEEE Transactions on Infor- mation Forensics and Security, vol. 13, no. 7, pp. 1854–1868, 2018.
N. MV and M. GR, “A survey on audio stream steganography tech- niques,” in Proceedings of the International Conference on Smart Data Intelligence (ICSMDI 2021), 2021.
R. Shrivastava, M. Singh, K. S. S. R. Teja, et al., “A real-time implementation for the speech steganography using short-time fourier transformior secured mobile communication,” in Journal of Physics: Conference Series, vol. 2089, p. 012066, IOP Publishing, 2021.
S. Deepikaa and R. Saravanan, “Coverless voip steganography using hash and hash,” Cybern. Inf. Technol, vol. 20, pp. 102–115, 2020.
P. Bedi and A. Dua, “Network steganography using the overflow field of timestamp option in an ipv4 packet,” Procedia Computer Science, vol. 171, pp. 1810–1818, 2020.
W. Mazurczyk, “Lost audio packets steganography: the first practical evaluation,” Security and Communication Networks, vol. 5, no. 12, pp. 1394–1403, 2012.
H. Moodi and A. R. Naghsh-Nilchi, “A new hybrid method for voip stream steganography,” Journal of Computing and Security, vol. 3, no. 3, pp. 175–182, 2016.
T. Thiede, W. C. Treurniet, R. Bitto, C. Schmidmer, T. Sporer, J. G. Beerends, and C. Colomes, “Peaq-the itu standard for objective mea- surement of perceived audio quality,” Journal of the Audio Engineering Society, vol. 48, no. 1/2, pp. 3–29, 2000.
H. Traunm¨uller and A. Eriksson, “The frequency range of the voice fundamental in the speech of male and female adults,” Unpublished manuscript, vol. 11, 1995.
S. Shirali-Shahreza and M. Shirali-Shahreza, “Steganography in silence intervals of speech,” in 2008 International Conference on Intelligent Information Hiding and Multimedia Signal Processing, pp. 605–607, IEEE, 2008.
A. S. Hameed, “A high secure speech transmission using audio steganography and duffing oscillator,” Wireless Personal Communica- tions, vol. 120, no. 1, pp. 499–513, 2021.
D. A. Shehab and M. J. Alhaddad, “Comprehensive survey of multime- dia steganalysis: Techniques, evaluations, and trends in future research,” Symmetry, vol. 14, no. 1, p. 117, 2022.
N. R. K. Kesa, “Steganography a data hiding technique,” 2018.
B. E. Carvajal-Gamez, M. A. Castillo-Martinez, L. A. Castaneda- Briones, F. J. Gallegos-Funes, and M. A. Diaz-Casco, “Audio steganal- ysis estimation with the goertzel algorithm,” Applied Sciences, vol. 14, no. 14, p. 6000, 2024.
A. Singhal and P. Bedi, “Blind quantitative steganalysis using svd features,” in 2018 International Conference on Advances in Computing, Communications and Informatics (ICACCI), pp. 369–374, IEEE, 2018.
A. J. Al-Najjar, “The decoy: multi-level digital multimedia steganog- raphy model,” in Proceedings of the 12th WSEAS international confer- ence on Communications, pp. 445–450, 2008.
A. A. Alsabhany, F. Ridzuan, and A. Azni, “The progressive multilevel embedding method for audio steganography,” in Journal of physics: conference series, vol. 1551, p. 012011, IOP Publishing, 2020.
E. Satish, N. Sreenivasa, E. Naresh, P. R. Naidu, and A. Ramachandra, “Multimedia multilevel security by integrating steganography and cryp- tography techniques,” in ITM Web of Conferences, vol. 57, p. 01012, EDP Sciences, 2023.
M. H. Sayed and T. M. Wahbi, “Information security for audio steganography using a phase coding method,” European Journal of Theoretical and Applied Sciences, vol. 2, no. 1, pp. 634–647, 2024.
C. Sujatha, Vibration, Acoustics and Strain Measurement. Springer, 2023.
M. Dodson, “Shannon’s sampling theorem,” Current science, vol. 63, no. 5, pp. 253–260, 1992.
T. O. Hodson, “Root mean square error (rmse) or mean absolute error (mae): When to use them or not,” Geoscientific Model Development Discussions, vol. 2022, pp. 1–10, 2022.
Z. Mazdak, B. A. M. Azizah, M. A. Shahidan, and S. C. Saman, “Maz- dak technique for psnr estimation in audio steganography,” Applied Mechanics and Materials, vol. 229, pp. 2798–2803, 2012.
M. H. N. Azam, F. Ridzuan, and M. Sayuti, “A new method to estimate peak signal to noise ratio for least significant bit modification audio steganography.,” Pertanika Journal of Science & Technology, vol. 30, no. 1, 2022.
Z. Wang, A. C. Bovik, H. R. Sheikh, and E. P. Simoncelli, “Image quality assessment: from error visibility to structural similarity,” IEEE transactions on image processing, vol. 13, no. 4, pp. 600–612, 2004.
Z. Wang, E. P. Simoncelli, and A. C. Bovik, “Multiscale structural similarity for image quality assessment,” in The thrity-seventh asilomar conference on signals, systems & computers, 2003, vol. 2, pp. 1398– 1402, Ieee, 2003.
S. Kandadai, J. Hardin, and C. D. Creusere, “Audio quality assess- ment using the mean structural similarity measure,” in 2008 IEEE International Conference on Acoustics, Speech and Signal Processing, pp. 221–224, IEEE, 2008.
C. Sloan, N. Harte, D. Kelly, A. C. Kokaram, and A. Hines, “Objec- tive assessment of perceptual audio quality using visqolaudio,” IEEE Transactions on Broadcasting, vol. 63, no. 4, pp. 693–705, 2017.
X. Min, G. Zhai, J. Zhou, M. C. Farias, and A. C. Bovik, “Study of subjective and objective quality assessment of audio-visual signals,” IEEE Transactions on Image Processing, vol. 29, pp. 6054–6068, 2020.
P. Kabal, “Perceptual evaluation of audio quality (peaq),” 2004.
S. K¨ampf, J. Liebetrau, S. Schneider, and T. Sporer, “Standardization of peaq-mc: Extension of itu-r bs. 1387-1 to multichannel audio,” in Audio Engineering Society Conference: 40th International Conference: Spatial Audio: Sense the Sound of Space, Audio Engineering Society, 2010.