FB pixel

New Microsoft benchmark for evaluating deepfake detection prioritizes breadth

Open source dataset project sees collaboration with Northwestern, WITNESS
New Microsoft benchmark for evaluating deepfake detection prioritizes breadth
 

Juan M. Lavista Ferres, corporate vice president and chief data scientist at Microsoft’s AI for Good Lab, has announced the release of a “large-scale, open-source benchmark for evaluating deepfake and manipulated media detection systems.”

Writing on LinkedIn, he says the initiative is a collaborative effort between the lab, Northwestern University’s Security and AI Lab, and tech-focused human rights nonprofit WITNESS. In Northwestern University’s words, it is “intended to help evaluate and improve algorithms to detect AI-generated audio, video, and image content.”

Lavista Ferres says it “introduces a rigorously curated dataset designed to support robust, real-world evaluation of multimodal detection tools,” intended to provide a shared foundation for empirical comparison of detection methods.

It is only to be licensed for evaluation, and is not intended for training or commercial purposes.

The dataset includes more than 50,000 samples of real, AI-generated and manipulated audio-visual content – deepfakes and synthetic media – annotated with data from real-world use cases. Adversarial attacks allow for the testing of model robustness.

Lavista Ferres says the benchmark is intended to support research in multimodal forensics, adversarial robustness and detection in real-world media ecosystems, and invites the research community to “explore the dataset and help maintain its relevance by contributing new data and evaluation protocols over time.”

Northwestern offers more background on the project, and how it is driven by advances in generative technologies. “In the past few years, a new paradigm has emerged with the diffusion architecture, showing impressive achievements in audio, image and video generation,” it says.  “Previous approaches to detection are now obsolete and the detection scene must re-invent itself.”

The summary from Northwestern notes that, historically, the evaluation of deepfake models was based on large datasets opened up during deepfake detection challenges. “These datasets typically had a lot of depth but almost no breadth. They were suitable for the previous era (the GAN era) but are not up to the challenge brought by the new generative AI landscape and the evolving type of harm it brings: scams, non-consensual intimate image generation, disinformation, etc.”

“We argue that depth is less important than breadth and we propose the creation of an evaluation set that contains small samples of as many generators and ‘in the wild’ cases as possible – rather than millions of samples from a few generators.”

Related Posts

Article Topics

 |   |   |   |   | 

Latest Biometrics News

 

IDScan confirms breach after 170M identity documents put up for sale

New Orleans-based ID verification provider IDScan.net has acknowledged the major breach of one of its databases, that exposed more than…

 

Microsoft launches Age API, following OS-level declared age range model

Whether because of its age, its product focus or its relatively anemic branding, Microsoft has not caught much public scrutiny…

 

IEEE developing parental consent standard for online age assurance

Talking about age assurance means talking about parental consent. Many continue to maintain that parents are the best arbiters of…

 

Touch Biometrix’ TFT technology reaches market with Lakota FAP60 scanner

The new FAP60 fingerprint biometric scanner from Lakota Software Solutions works natively with iPhones and iPads, which the company says…

 

Advance.AI targets regional growth as SE Asia’s local IAD provider

Singapore-headquartered digital identity verification, compliance and credit information provider Advance.AI is positioning itself as the native regional choice for biometric…

 

Myanmar’s digital ID strategy comes amid a fraught landscape

Myanmar’s military government is accelerating construction of national digital identity infrastructure despite ongoing civil war, using a phased deployment strategy…

Comments

Leave a Reply

This site uses Akismet to reduce spam. Learn how your comment data is processed.

Market Intelligence

Featured Company

Biometric Update Podcast

Most Read This Week

White Papers

Latest Webinars

Biometrics Industry Events