Speaker Inconsistency Detection in Tampered Video

Chattez avec Graph Search

Posez n’importe quelle question sur les cours, conférences, exercices, recherches, actualités, etc. de l’EPFL ou essayez les exemples de questions ci-dessous.

AVERTISSEMENT : Le chatbot Graph n'est pas programmé pour fournir des réponses explicites ou catégoriques à vos questions. Il transforme plutôt vos questions en demandes API qui sont distribuées aux différents services informatiques officiellement administrés par l'EPFL. Son but est uniquement de collecter et de recommander des références pertinentes à des contenus que vous pouvez explorer pour vous aider à répondre à vos questions.

With the increasing amount of video being consumed by people daily, there is a danger of the rise in maliciously modified video content (i.e., 'fake news') that could be used to damage innocent people or to impose a certain agenda, e.g., meddle in elections. In this paper, we consider audio manipulations in video of a person speaking to the camera. Such manipulation is easy to perform, for instance, one can just replace a part of audio, while it can dramatically change the message and the meaning of the video. With the goal to develop an automated system that can detect these audio-visual speaker inconsistencies, we consider several approaches proposed for lip-syncing and dubbing detection, based on convolutional and recurrent networks and compare them with systems that are based on more traditional classifiers. We evaluated these methods on publicly available databases VidTIMIT, AMI, and GRID, for which we generated sets of tampered data.

Speaker Inconsistency Detection in Tampered Video

Graph Chatbot

Chattez avec Graph Search

Acoustical Features as Knee Health Biomarkers: A Critical Analysis

Improving Deepfake Detectors against Real-world Perturbations with Amplitude-Phase Switch Augmentation

Multi-task Neural Network for Robust Multiple Speaker Embedding Extraction

Acoustical Features as Knee Health Biomarkers: A Critical Analysis

Multi-task Neural Network for Robust Multiple Speaker Embedding Extraction

Improving Deepfake Detectors against Real-world Perturbations with Amplitude-Phase Switch Augmentation