Effectiveness of Wav2Vec2 in LLM-based deepfake voice detection
Determine whether self-supervised Wav2Vec2 audio representations are as effective for deepfake voice detection in LLM-based architectures as they are in standalone deepfake detection systems.
References
However, it remains an open question whether they are equally effective in LLM-based architectures.
— Textual Acoustic Grounding for Generalizable LLM-Based Deepfake Voice Detection
(2608.30622 - Kheir et al., 31 Aug 2026) in Section 1, Introduction