Home / Journals / CMC / Online First / doi:10.32604/cmc.2026.085750
Special Issues
Table of Content

Open Access

ARTICLE

EchoMark: A Practical Audio Disruption Scheme for Anti-Synthesis Protection

Hung-Jr Shiu1, Ming-Ya Tseng1, Wei-Chung Lin2,*
1 Department of Computer Science and Information Engineering, National Taipei University, New Taipei, Taiwan
2 Department of Computer Science and Information Engineering, National Taiwan University, Taipei, Taiwan
* Corresponding Author: Wei-Chung Lin. Email: email

Computers, Materials & Continua https://doi.org/10.32604/cmc.2026.085750

Received 17 May 2026; Accepted 03 August 2026; Published online 24 August 2026

Abstract

Driven by recent breakthroughs in generative artificial intelligence, modern voice cloning technologies can synthesize remarkably lifelike human speech, exacerbating security vulnerabilities associated with identity impersonation, financial fraud, and deepfake audio proliferation. To mitigate these risks, this paper introduces EchoMark, an acoustic-layer disruption framework designed to systematically undermine neural speech generation workflows. Unlike conventional digital watermarking or software-level perturbation strategies, EchoMark embeds structured, multi-tiered echo patterns directly into audio during physical playback and re-recording. This physical-layer integration severely compromises the spectral coherence essential for neural text-to-speech (TTS) modeling, resulting in degraded acoustic fidelity and impaired speech-fitting capabilities. We rigorously validate EchoMark across multiple representative TTS architectures—specifically Mockingbird, FishSpeech, and GPT-SoVITS—under diverse operating environments, dataset categories, and quantitative metrics. Experimental evaluations confirm that structured echo perturbations serve as an efficient, lightweight physical countermeasure capable of inducing performance degradation in target speech synthesis pipelines, providing a viable anti-spoofing defense for physical-layer audio applications.

Keywords

Audio security; speech synthesis; echo perturbation; deepfake voice; anti-spoofing; voice integrity; physical-layer defense
  • 90

    View

  • 21

    Download

  • 0

    Like

Share Link