DEAL-300K: Diffusion-based Editing Area Localization with a 300K-Scale Dataset and Frequency-Prompted Baseline
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Zhang, Rui, Wang, Hongxia, Liu, Hangqing, Zhou, Yang, Zeng, Qiang |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2025
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
CigTime: Corrective Instruction Generation Through Inverse Motion Editing
von: Fang, Qihang, et al.
Veröffentlicht: (2024)
von: Fang, Qihang, et al.
Veröffentlicht: (2024)
Gaze into the Heart: A Multi-View Video Dataset for rPPG and Health Biomarkers Estimation
von: Egorov, Konstantin, et al.
Veröffentlicht: (2025)
von: Egorov, Konstantin, et al.
Veröffentlicht: (2025)
Efficient Vision-based Vehicle Speed Estimation
von: Macko, Andrej, et al.
Veröffentlicht: (2025)
von: Macko, Andrej, et al.
Veröffentlicht: (2025)
FACEMUG: A Multimodal Generative and Fusion Framework for Local Facial Editing
von: Lu, Wanglong, et al.
Veröffentlicht: (2024)
von: Lu, Wanglong, et al.
Veröffentlicht: (2024)
Visual Style Prompt Learning Using Diffusion Models for Blind Face Restoration
von: Lu, Wanglong, et al.
Veröffentlicht: (2024)
von: Lu, Wanglong, et al.
Veröffentlicht: (2024)
Detailed Evaluation of Modern Machine Learning Approaches for Optic Plastics Sorting
von: Maheshkar, Vaishali, et al.
Veröffentlicht: (2025)
von: Maheshkar, Vaishali, et al.
Veröffentlicht: (2025)
PolyGlotFake: A Novel Multilingual and Multimodal DeepFake Dataset
von: Hou, Yang, et al.
Veröffentlicht: (2024)
von: Hou, Yang, et al.
Veröffentlicht: (2024)
Relightable and Dynamic Gaussian Avatar Reconstruction from Monocular Video
von: Choi, Seonghwa, et al.
Veröffentlicht: (2025)
von: Choi, Seonghwa, et al.
Veröffentlicht: (2025)
Stylized Face Sketch Extraction via Generative Prior with Limited Data
von: Yun, Kwan, et al.
Veröffentlicht: (2024)
von: Yun, Kwan, et al.
Veröffentlicht: (2024)
SymFace: Additional Facial Symmetry Loss for Deep Face Recognition
von: Prakash, Pritesh, et al.
Veröffentlicht: (2024)
von: Prakash, Pritesh, et al.
Veröffentlicht: (2024)
LeGO: Leveraging a Surface Deformation Network for Animatable Stylized Face Generation with One Example
von: Yoon, Soyeon, et al.
Veröffentlicht: (2024)
von: Yoon, Soyeon, et al.
Veröffentlicht: (2024)
EditP23: 3D Editing via Propagation of Image Prompts to Multi-View
von: Bar-On, Roi, et al.
Veröffentlicht: (2025)
von: Bar-On, Roi, et al.
Veröffentlicht: (2025)
N-DriverMotion: Driver motion learning and prediction using an event-based camera and directly trained spiking neural networks on Loihi 2
von: Chung, Hyo Jong, et al.
Veröffentlicht: (2024)
von: Chung, Hyo Jong, et al.
Veröffentlicht: (2024)
Few-Class Arena: A Benchmark for Efficient Selection of Vision Models and Dataset Difficulty Measurement
von: Cao, Bryan Bo, et al.
Veröffentlicht: (2024)
von: Cao, Bryan Bo, et al.
Veröffentlicht: (2024)
Event-ECC: Asynchronous Tracking of Events with Continuous Optimization
von: Zafeiri, Maria, et al.
Veröffentlicht: (2024)
von: Zafeiri, Maria, et al.
Veröffentlicht: (2024)
HieraEdgeNet: A Multi-Scale Edge-Enhanced Framework for Automated Pollen Recognition
von: Long, Yuchong, et al.
Veröffentlicht: (2025)
von: Long, Yuchong, et al.
Veröffentlicht: (2025)
The Impact of Image Resolution on Face Detection: A Comparative Analysis of MTCNN, YOLOv XI and YOLOv XII models
von: Ömercikoğlu, Ahmet Can, et al.
Veröffentlicht: (2025)
von: Ömercikoğlu, Ahmet Can, et al.
Veröffentlicht: (2025)
SelectiveKD: A semi-supervised framework for cancer detection in DBT through Knowledge Distillation and Pseudo-labeling
von: Dillard, Laurent, et al.
Veröffentlicht: (2024)
von: Dillard, Laurent, et al.
Veröffentlicht: (2024)
Distilling foundation models for robust and efficient models in digital pathology
von: Filiot, Alexandre, et al.
Veröffentlicht: (2025)
von: Filiot, Alexandre, et al.
Veröffentlicht: (2025)
Harnessing Deep Learning and Satellite Imagery for Post-Buyout Land Cover Mapping
von: Otal, Hakan T., et al.
Veröffentlicht: (2024)
von: Otal, Hakan T., et al.
Veröffentlicht: (2024)
Fixed-Threshold Evaluation of a Hybrid CNN-ViT for AI-Generated Image Detection Across Photos and Art
von: Khan, Md Ashik, et al.
Veröffentlicht: (2025)
von: Khan, Md Ashik, et al.
Veröffentlicht: (2025)
TextDoctor: Unified Document Image Inpainting via Patch Pyramid Diffusion Models
von: Lu, Wanglong, et al.
Veröffentlicht: (2025)
von: Lu, Wanglong, et al.
Veröffentlicht: (2025)
FastFit: Accelerating Multi-Reference Virtual Try-On via Cacheable Diffusion Models
von: Chong, Zheng, et al.
Veröffentlicht: (2025)
von: Chong, Zheng, et al.
Veröffentlicht: (2025)
DeepFusionNet: Autoencoder-Based Low-Light Image Enhancement and Super-Resolution
von: Çalışkan, Halil Hüseyin, et al.
Veröffentlicht: (2025)
von: Çalışkan, Halil Hüseyin, et al.
Veröffentlicht: (2025)
Tri-Plane Mamba: Efficiently Adapting Segment Anything Model for 3D Medical Images
von: Wang, Hualiang, et al.
Veröffentlicht: (2024)
von: Wang, Hualiang, et al.
Veröffentlicht: (2024)
StatsMerging: Statistics-Guided Model Merging via Task-Specific Teacher Distillation
von: Merugu, Ranjith, et al.
Veröffentlicht: (2025)
von: Merugu, Ranjith, et al.
Veröffentlicht: (2025)
Motion-Based Sign Language Video Summarization using Curvature and Torsion
von: Sartinas, Evangelos G., et al.
Veröffentlicht: (2023)
von: Sartinas, Evangelos G., et al.
Veröffentlicht: (2023)
Eleven Primitives and Three Gates: The Universal Structure of Computational Imaging
von: Yang, Chengshuai, et al.
Veröffentlicht: (2026)
von: Yang, Chengshuai, et al.
Veröffentlicht: (2026)
Rethinking VLMs for Image Forgery Detection and Localization
von: Guo, Shaofeng, et al.
Veröffentlicht: (2026)
von: Guo, Shaofeng, et al.
Veröffentlicht: (2026)
ForensicFormer: Hierarchical Multi-Scale Reasoning for Cross-Domain Image Forgery Detection
von: Samson, Hema Hariharan
Veröffentlicht: (2026)
von: Samson, Hema Hariharan
Veröffentlicht: (2026)
BG-YOLO: A Bidirectional-Guided Method for Underwater Object Detection
von: Zhang, Jian, et al.
Veröffentlicht: (2024)
von: Zhang, Jian, et al.
Veröffentlicht: (2024)
CatVTON: Concatenation Is All You Need for Virtual Try-On with Diffusion Models
von: Chong, Zheng, et al.
Veröffentlicht: (2024)
von: Chong, Zheng, et al.
Veröffentlicht: (2024)
SPEAK: Speech-Driven Pose and Emotion-Adjustable Talking Head Generation
von: Cai, Changpeng, et al.
Veröffentlicht: (2024)
von: Cai, Changpeng, et al.
Veröffentlicht: (2024)
CatV2TON: Taming Diffusion Transformers for Vision-Based Virtual Try-On with Temporal Concatenation
von: Chong, Zheng, et al.
Veröffentlicht: (2025)
von: Chong, Zheng, et al.
Veröffentlicht: (2025)
S3Simulator: A benchmarking Side Scan Sonar Simulator dataset for Underwater Image Analysis
von: S, Kamal Basha, et al.
Veröffentlicht: (2024)
von: S, Kamal Basha, et al.
Veröffentlicht: (2024)
Mapping Tomato Cropping Systems in California Using AlphaEarth Geospatial Embeddings and Deep Learning Analysis
von: Narimani, Mohammadreza, et al.
Veröffentlicht: (2026)
von: Narimani, Mohammadreza, et al.
Veröffentlicht: (2026)
Satellite-Net: Automatic Extraction of Land Cover Indicators from Satellite Imagery by Deep Learning
von: Bernasconi, Eleonora, et al.
Veröffentlicht: (2019)
von: Bernasconi, Eleonora, et al.
Veröffentlicht: (2019)
Balanced conic rectified flow
von: Kim, Shin Seong, et al.
Veröffentlicht: (2025)
von: Kim, Shin Seong, et al.
Veröffentlicht: (2025)
CADE 2.5 - ZeResFDG: Frequency-Decoupled, Rescaled and Zero-Projected Guidance for SD/SDXL Latent Diffusion Models
von: Rychkovskiy, Denis
Veröffentlicht: (2025)
von: Rychkovskiy, Denis
Veröffentlicht: (2025)
Identity Deepfake Threats to Biometric Authentication Systems: Public and Expert Perspectives
von: He, Shijing, et al.
Veröffentlicht: (2025)
von: He, Shijing, et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
CigTime: Corrective Instruction Generation Through Inverse Motion Editing
von: Fang, Qihang, et al.
Veröffentlicht: (2024) -
Gaze into the Heart: A Multi-View Video Dataset for rPPG and Health Biomarkers Estimation
von: Egorov, Konstantin, et al.
Veröffentlicht: (2025) -
Efficient Vision-based Vehicle Speed Estimation
von: Macko, Andrej, et al.
Veröffentlicht: (2025) -
FACEMUG: A Multimodal Generative and Fusion Framework for Local Facial Editing
von: Lu, Wanglong, et al.
Veröffentlicht: (2024) -
Visual Style Prompt Learning Using Diffusion Models for Blind Face Restoration
von: Lu, Wanglong, et al.
Veröffentlicht: (2024)