AV-Deepfake1M: A Large-Scale LLM-Driven Audio-Visual Deepfake Dataset
Fuente:
arXiv
Saved in:
| Main Authors: | Cai, Zhixi, Ghosh, Shreya, Adatia, Aman Pankaj, Hayat, Munawar, Dhall, Abhinav, Gedeon, Tom, Stefanov, Kalin |
|---|---|
| Format: | Preprint |
| Published: |
2023
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
1M-Deepfakes Detection Challenge
by: Cai, Zhixi, et al.
Published: (2024)
by: Cai, Zhixi, et al.
Published: (2024)
AV-Deepfake1M++: A Large-Scale Audio-Visual Deepfake Benchmark with Real-World Perturbations
by: Cai, Zhixi, et al.
Published: (2025)
by: Cai, Zhixi, et al.
Published: (2025)
Emolysis: A Multimodal Open-Source Group Emotion Analysis and Visualization Toolkit
by: Ghosh, Shreya, et al.
Published: (2023)
by: Ghosh, Shreya, et al.
Published: (2023)
Multiverse Through Deepfakes: The MultiFakeVerse Dataset of Person-Centric Visual and Conceptual Manipulations
by: Gupta, Parul, et al.
Published: (2025)
by: Gupta, Parul, et al.
Published: (2025)
MRAC Track 1: 2nd Workshop on Multimodal, Generative and Responsible Affective Computing
by: Ghosh, Shreya, et al.
Published: (2024)
by: Ghosh, Shreya, et al.
Published: (2024)
Pavlok-Nudge: A Feedback Mechanism for Atomic Behaviour Modification with Snoring Usecase
by: Hasan, Md Rakibul, et al.
Published: (2023)
by: Hasan, Md Rakibul, et al.
Published: (2023)
SignMAE: Segmentation-Driven Self-Supervised Learning for Sign Language Recognition
by: Xie, Kunyuan, et al.
Published: (2026)
by: Xie, Kunyuan, et al.
Published: (2026)
GTA-HDR: A Large-Scale Synthetic Dataset for HDR Image Reconstruction
by: Barua, Hrishav Bakul, et al.
Published: (2024)
by: Barua, Hrishav Bakul, et al.
Published: (2024)
Gems: Group Emotion Profiling Through Multimodal Situational Understanding
by: Kataria, Anubhav, et al.
Published: (2025)
by: Kataria, Anubhav, et al.
Published: (2025)
CSGaze: Context-aware Social Gaze Prediction
by: Madan, Surbhi, et al.
Published: (2025)
by: Madan, Surbhi, et al.
Published: (2025)
Do Blind Spots Matter for Word-Referent Mapping? A Computational Study with Infant Egocentric Video
by: Shi, Zekai, et al.
Published: (2025)
by: Shi, Zekai, et al.
Published: (2025)
Investigating the Viability of Employing Multi-modal Large Language Models in the Context of Audio Deepfake Detection
by: Chuchra, Akanksha, et al.
Published: (2026)
by: Chuchra, Akanksha, et al.
Published: (2026)
Conditional Distribution Modelling for Few-Shot Image Synthesis with Diffusion Models
by: Gupta, Parul, et al.
Published: (2024)
by: Gupta, Parul, et al.
Published: (2024)
LayLens: Improving Deepfake Understanding through Simplified Explanations
by: Narang, Abhijeet, et al.
Published: (2025)
by: Narang, Abhijeet, et al.
Published: (2025)
PhysHDR: When Lighting Meets Materials and Scene Geometry in HDR Reconstruction
by: Barua, Hrishav Bakul, et al.
Published: (2025)
by: Barua, Hrishav Bakul, et al.
Published: (2025)
HistoHDR-Net: Histogram Equalization for Single LDR to HDR Image Translation
by: Barua, Hrishav Bakul, et al.
Published: (2024)
by: Barua, Hrishav Bakul, et al.
Published: (2024)
Generation and Detection of Sign Language Deepfakes - A Linguistic and Visual Analysis
by: Naeem, Shahzeb, et al.
Published: (2024)
by: Naeem, Shahzeb, et al.
Published: (2024)
DiffAugment: Diffusion based Long-Tailed Visual Relationship Recognition
by: Gupta, Parul, et al.
Published: (2024)
by: Gupta, Parul, et al.
Published: (2024)
MIP-GAF: A MLLM-annotated Benchmark for Most Important Person Localization and Group Context Understanding
by: Madan, Surbhi, et al.
Published: (2024)
by: Madan, Surbhi, et al.
Published: (2024)
SFANet: Spatial-Frequency Attention Network for Deepfake Detection
by: Ahire, Vrushank, et al.
Published: (2025)
by: Ahire, Vrushank, et al.
Published: (2025)
Codecfake: An Initial Dataset for Detecting LLM-based Deepfake Audio
by: Lu, Yi, et al.
Published: (2024)
by: Lu, Yi, et al.
Published: (2024)
Pixels Don't Lie (But Your Detector Might): Bootstrapping MLLM-as-a-Judge for Trustworthy Deepfake Detection and Reasoning Supervision
by: Kuckreja, Kartik, et al.
Published: (2026)
by: Kuckreja, Kartik, et al.
Published: (2026)
DexAvatar: 3D Sign Language Reconstruction with Hand and Body Pose Priors
by: Kundu, Kaustubh, et al.
Published: (2025)
by: Kundu, Kaustubh, et al.
Published: (2025)
Human Brain Exhibits Distinct Patterns When Listening to Fake Versus Real Audio: Preliminary Evidence
by: Salehi, Mahsa, et al.
Published: (2024)
by: Salehi, Mahsa, et al.
Published: (2024)
Audio Deepfake Attribution: An Initial Dataset and Investigation
by: Yan, Xinrui, et al.
Published: (2022)
by: Yan, Xinrui, et al.
Published: (2022)
CodecFake+: A Large-Scale Neural Audio Codec-Based Deepfake Speech Dataset
by: Chen, Xuanjun, et al.
Published: (2025)
by: Chen, Xuanjun, et al.
Published: (2025)
Audio Deepfake Verification
by: Wang, Li, et al.
Published: (2025)
by: Wang, Li, et al.
Published: (2025)
A Cycle Ride to HDR: Semantics Aware Self-Supervised Framework for Unpaired LDR-to-HDR Image Reconstruction
by: Barua, Hrishav Bakul, et al.
Published: (2024)
by: Barua, Hrishav Bakul, et al.
Published: (2024)
The Codecfake Dataset and Countermeasures for the Universally Detection of Deepfake Audio
by: Xie, Yuankun, et al.
Published: (2024)
by: Xie, Yuankun, et al.
Published: (2024)
Cross-Domain Audio Deepfake Detection: Dataset and Analysis
by: Li, Yuang, et al.
Published: (2024)
by: Li, Yuang, et al.
Published: (2024)
AUDETER: A Large-scale Dataset for Deepfake Audio Detection in Open Worlds
by: Wang, Qizhou, et al.
Published: (2025)
by: Wang, Qizhou, et al.
Published: (2025)
Audio-Visual Deepfake Detection With Local Temporal Inconsistencies
by: Astrid, Marcella, et al.
Published: (2025)
by: Astrid, Marcella, et al.
Published: (2025)
Detecting Audio-Visual Deepfakes with Fine-Grained Inconsistencies
by: Astrid, Marcella, et al.
Published: (2024)
by: Astrid, Marcella, et al.
Published: (2024)
IndieFake Dataset: A Benchmark Dataset for Audio Deepfake Detection
by: Kumar, Abhay, et al.
Published: (2025)
by: Kumar, Abhay, et al.
Published: (2025)
Hindi audio-video-Deepfake (HAV-DF): A Hindi language-based Audio-video Deepfake Dataset
by: Kaur, Sukhandeep, et al.
Published: (2024)
by: Kaur, Sukhandeep, et al.
Published: (2024)
Scaling Laws for Deepfake Detection
by: Wang, Wenhao, et al.
Published: (2025)
by: Wang, Wenhao, et al.
Published: (2025)
DFADD: The Diffusion and Flow-Matching Based Audio Deepfake Dataset
by: Du, Jiawei, et al.
Published: (2024)
by: Du, Jiawei, et al.
Published: (2024)
Human Perception of Audio Deepfakes
by: Müller, Nicolas M., et al.
Published: (2021)
by: Müller, Nicolas M., et al.
Published: (2021)
Generalizable Detection of Audio Deepfakes
by: Lopez, Jose A., et al.
Published: (2025)
by: Lopez, Jose A., et al.
Published: (2025)
Detection of Deepfake Environmental Audio
by: Ouajdi, Hafsa, et al.
Published: (2024)
by: Ouajdi, Hafsa, et al.
Published: (2024)
Similar Items
-
1M-Deepfakes Detection Challenge
by: Cai, Zhixi, et al.
Published: (2024) -
AV-Deepfake1M++: A Large-Scale Audio-Visual Deepfake Benchmark with Real-World Perturbations
by: Cai, Zhixi, et al.
Published: (2025) -
Emolysis: A Multimodal Open-Source Group Emotion Analysis and Visualization Toolkit
by: Ghosh, Shreya, et al.
Published: (2023) -
Multiverse Through Deepfakes: The MultiFakeVerse Dataset of Person-Centric Visual and Conceptual Manipulations
by: Gupta, Parul, et al.
Published: (2025) -
MRAC Track 1: 2nd Workshop on Multimodal, Generative and Responsible Affective Computing
by: Ghosh, Shreya, et al.
Published: (2024)