Complementary rPPG-Derived and Lip-Region Frequency Cues for Talking-Face Deepfake Detection
Two lightweight visual-only cues, rPPG-derived waveforms extracted by RhythmFormer and lip-region discrete cosine transform (DCT) coefficients, are studied on the seven TF methods of Celeb-DF++ under a subject-independent protocol.