Credibility attacks do not enhance the impact of deepfake warnings
摘要
Synthetic deepfake videos are a realistic form of misinformation, which appear to show someone saying or doing something they never did. The societal risks posed by such fakeries have motivated legislation mandating the labelling of AI-generated content, but initial evidence suggests that even explicit warnings are insufficient to eliminate deepfake influences on viewer perceptions. Across two experiments (total N = 2,500), we examined the impact of a deepfake video in which a witness makes a criminal accusation, and we tested whether undermining the credibility of the witness could enhance the effect of a specific warning flagging the video as a deepfake. Two methods for reducing witness credibility were tested: presenting the witness with a foreign (vs. native) accent and explicitly introducing them as an untrustworthy (vs. trustworthy) character. Deepfake exposure significantly increased guilt judgements concerning the person being accused in the video; a deepfake warning only partially reduced this impact. When an explicit witness discreditation was added to the warning, effectiveness remained unchanged. Findings demonstrate the pervasive impact of known deepfakes, highlighting the need to combine or develop new interventions to combat their influence.