Deepfake: definitions, performance metrics and standards, datasets, and a meta-review

Enes Altuncu1, Virginia N L Franqueira1, Shujun Li1

  • 1Institute of Cyber Security for Society (iCSS) & School of Computing, University of Kent, Canterbury, United Kingdom.

Frontiers in Big Data
|September 19, 2024
PubMed
Summary

This paper provides a comprehensive overview of deepfake technology, covering definitions, metrics, datasets, and challenges. It synthesizes research to offer insights into the rapidly evolving field of synthetic media.

Related Concept Videos

Deindividuation00:57

Deindividuation

Deindividuation is a form of social influence on an individual’s behavior such that the individual engages in unusual or non-normal behavior while in a group setting. Why? Because in these group settings, the individual no longer sees themselves as an individual anymore, disinhibiting their behavior and personal restraint.
Midrange01:07

Midrange

A somewhat easy to compute quantitative estimate of a data set’s central tendency is its midrange, which is defined as the mean of the minimum and maximum values of an ordered data set.
Simply put, the midrange is half of the data set’s range. Similar to the mean, the midrange is sensitive to the extreme values and hence the prospective outliers. However, unlike the mean, the midrange is not sensitive to all the values of the data set that lie in the middle. Thus, it is prone to outliers and...
Review and Preview01:13

Review and Preview

Data are individual items of information obtained from a population or sample. Data may be classified as qualitative (categorical), quantitative continuous, or quantitative discrete. Because it is not practical to measure the entire population in a study, researchers use samples to represent the population. A random sample is a representative group from the population chosen by using a method that gives each individual in the population an equal chance of being included in the sample. Random...
Review and Preview01:10

Review and Preview

In statistics, several tools are used to interpret the data. Measures of central tendency represent the characteristics of the data, such as mean, median, and mode. Additionally, measures of variance like standard deviation and range are used to find the spread of data from the mean. Relative standing measures the distance between data locations. Commonly used measures of relative standings are percentile, z score, and quartiles.
Percentiles are a type of fractile that partition data into...
Bias01:22

Bias

Bias refers to any tendency that prevents a question from being considered unprejudiced. In research, bias occurs when one outcome or answer is selected or encouraged over others in sampling or testing. Bias can occur during any research phase, including study design, data collection, analysis, and publication.
In statistics, a sampling bias is created when a sample is collected from a population, and some members of the population are not as likely to be chosen as others (remember, each member...
Deep Sea Microbial Ecology01:18

Deep Sea Microbial Ecology

The deep ocean and its underlying sediments represent vast, largely unexplored microbial habitats that extend far beyond the sunlit photic zone. The photic (euphotic) zone typically spans the upper ~100–200 meters of pelagic waters in the open ocean, but its depth varies geographically and seasonally, where sufficient light supports photosynthetic life. Below this lies the deep sea, spanning roughly 1000–6000 meters (bathypelagic to abyssal zones), with deeper hadal trenches extending beyond...