SMARTACT연구소디지털미디어리터러시 아카이브스마택트 홈 ↗
← 논문 목록

EfficientNet와 Visual Transformer에 기반한 투 스트림 딥페이크 영상 탐지 알고리즘 제안

Proposing A Two-Stream Deepfake Video Detection Algorithm Based on EfficientNet and Visual Transformer

발행
2025
소속·발행
한국외국어대학교
출처
국내 KCI
DOI
10.9717/kmms.2025.28.3.470
원문등록
2025-03-31
원문
원문 보기 ↗
개념
키워드

Deep Learning, Computer Vision, Deep Fake, EfficientNetwork, Visual Transformer, Deep Learning, Computer Vision, Deep Fake, EfficientNetwork, Visual Transformer

초록

This paper proposes a model using a two-stream network to extract both global and local spatial forgery features in deepfake videos and learn temporal forgery features across frames. The spatial detection model is based on EfficientNet, and the temporal detection model is based on ViT. Spatial features are divided into two streams to extract global and local features separately, while two ViTs are used to effectively learn temporal features, making the temporal part also two-stream. This achieved an average 9.4% improvement in detection accuracy compared to existing deepfake detection models, and through Grad-CAM, we visually confirmed the regions that significantly influenced the determination of deepfakes, demonstrating the model's effective detection capabilities.

같은 개념의 다른 논문
이 개념이 나오는 강연