New Gartner® report — Reality Defender is named a Market Shaper in deepfake detection, as of June 2026.

Get the report

\

Insight

\

How Deepfakes Are Made

Gabe Regan

VP of Human Engagement

Disclaimer

This article is for educational purposes only. Understanding how deepfakes are made helps organizations, security professionals, and individuals recognize and defend against synthetic media threats. Knowledge of the deepfake creation process encourages better detection and protection strategies against malicious actors who exploit this technology for fraud and deception.

Simple Explanation

Deepfake technology relies on artificial intelligence systems that learn to create convincing fake media by analyzing vast amounts of data. At its core, how deepfakes are made involves training AI models on hundreds or thousands of images, videos or audio samples of an individual. 

The AI system studies patterns in facial expressions, voice characteristics, speech patterns and visual features. Through machine learning algorithms, particularly GAN technology (Generative Adversarial Networks), it learns to generate new content that mimics the original person's appearance or voice. The content can be so convincing that most people can’t differentiate between it and the real thing.
Data requirements are substantial. Compelling deepfakes typically require feeding hundreds or thousands of data points into a deep learning network, training it to reconstruct visual, audio and textual patterns. But some tools now claim to produce basic deepfakes with just a few minutes of audio or a handful of photos.

The Deepfake Creation Process

  • Step 1 (Data Collection): Gathering source material, including photos, videos and audio recordings of the target person, often scraped from social media or public appearances.
  • Step 2 (Data Preprocessing): Cleaning and organizing the collected material, extracting faces from videos, isolating voice samples, and preparing data for training.
  • Step 3 (Model Training): Using GAN technology or similar AI frameworks to train neural networks on the prepared data.
  • Step 4: (Generation): Running the trained model to create new synthetic content, fine-tuning parameters to improve realism and believability.
  • Step 5 (Post-Processing): Refining the output through editing software to eliminate obvious artifacts and enhance overall quality.

Time required varies based on the desired quality and available computing power. Rudimentary deepfakes can be created in under 30 seconds using off-the-shelf tools, while high-quality results may require days to weeks of processing time on more powerful systems.

Tools Used

Deepfake software ranges from user-friendly consumer applications to sophisticated development frameworks, including mobile apps that create basic face swaps and professional-grade platforms that require technical expertise.

GAN technology powers most deepfake creation, using two competing neural networks: one generates fake content while another tries to detect it, iteratively improving until the output becomes highly convincing.

Accessibility to deepfake tools has dramatically increased, with cloud-based services and simplified interfaces bringing creative capabilities to users without technical backgrounds. “Generative artificial intelligence tools make it easy for even low-skill threat actors to create deepfakes,” according to the FS-ISAC Artificial Intelligence Risk Working Group.

Why Detection Works

Deepfakes contain inherent flaws that detection systems can identify, including visual inconsistencies, audio issues and movement irregularities. Reality Defender uses inference-based methods to detect signs of manipulation or synthetic generation. Our advantage is using proprietary models that analyze video, audio or image files from thousands of approaches — all in milliseconds.

See Reality Defender's deepfake detection in action. Get in touch with our team for a demo.

Frequently Asked Questions

Frequently asked questions

A deepfake is synthetic media (video, image, or audio) created with AI to convincingly depict a real person saying or doing something they never did. Deepfakes are generated by training machine-learning models, most often GANs, on real footage or recordings of a target until the output mimics their appearance or voice.

Deepfakes are made in five stages: collecting source images, video, or audio of the target; preprocessing that data; training an AI model (typically GAN or diffusion) on it; generating the synthetic output; and post-processing to remove artifacts. Simple face swaps take under a minute with consumer apps, while convincing results can take days.

A basic deepfake can be produced in under 30 seconds using off-the-shelf tools, while high-quality results may take days or weeks on more powerful systems. Some tools now claim to generate a passable voice clone from just a few minutes of audio or a handful of photos.

Deepfake creation tools range from consumer face-swap mobile apps to professional GAN- or diffusion-based frameworks. Most rely on generative adversarial networks: two competing neural networks that refine each other until the output looks real. Cloud-based services have made these capabilities accessible even to low-skill threat actors.

Deepfakes used to have tell-tale flaws: unnatural blinking or facial movement, mismatched lighting or shadows, audio slightly out of sync, and inconsistent texture at the edges of the face. Yet as generation improved, these visual cues became unreliable, which is why automated detection has become essential.

Yes. Even convincing deepfakes contain artifacts invisible to the human eye that detection systems can identify across visual, audio, and movement signals. Reality Defender uses inference-based, proprietary models that analyze media from thousands of approaches in milliseconds to flag manipulation or synthetic generation.

It depends on jurisdiction and use. Many U.S. states have enacted laws targeting specific harms such as non-consensual intimate imagery and election disinformation, and legislation is evolving rapidly worldwide. Using deepfakes for fraud, harassment, or impersonation is broadly unlawful, while consensual or clearly satirical uses are often permitted.