DEAL-300K: Diffusion-based Editing Area Localization with a 300K-Scale Dataset and Frequency-Prompted Baseline
Fuente:
arXiv
Salvato in:
| Autori principali: | Zhang, Rui, Wang, Hongxia, Liu, Hangqing, Zhou, Yang, Zeng, Qiang |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2025
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
CigTime: Corrective Instruction Generation Through Inverse Motion Editing
di: Fang, Qihang, et al.
Pubblicazione: (2024)
di: Fang, Qihang, et al.
Pubblicazione: (2024)
Gaze into the Heart: A Multi-View Video Dataset for rPPG and Health Biomarkers Estimation
di: Egorov, Konstantin, et al.
Pubblicazione: (2025)
di: Egorov, Konstantin, et al.
Pubblicazione: (2025)
Efficient Vision-based Vehicle Speed Estimation
di: Macko, Andrej, et al.
Pubblicazione: (2025)
di: Macko, Andrej, et al.
Pubblicazione: (2025)
FACEMUG: A Multimodal Generative and Fusion Framework for Local Facial Editing
di: Lu, Wanglong, et al.
Pubblicazione: (2024)
di: Lu, Wanglong, et al.
Pubblicazione: (2024)
Visual Style Prompt Learning Using Diffusion Models for Blind Face Restoration
di: Lu, Wanglong, et al.
Pubblicazione: (2024)
di: Lu, Wanglong, et al.
Pubblicazione: (2024)
Detailed Evaluation of Modern Machine Learning Approaches for Optic Plastics Sorting
di: Maheshkar, Vaishali, et al.
Pubblicazione: (2025)
di: Maheshkar, Vaishali, et al.
Pubblicazione: (2025)
PolyGlotFake: A Novel Multilingual and Multimodal DeepFake Dataset
di: Hou, Yang, et al.
Pubblicazione: (2024)
di: Hou, Yang, et al.
Pubblicazione: (2024)
Relightable and Dynamic Gaussian Avatar Reconstruction from Monocular Video
di: Choi, Seonghwa, et al.
Pubblicazione: (2025)
di: Choi, Seonghwa, et al.
Pubblicazione: (2025)
Stylized Face Sketch Extraction via Generative Prior with Limited Data
di: Yun, Kwan, et al.
Pubblicazione: (2024)
di: Yun, Kwan, et al.
Pubblicazione: (2024)
SymFace: Additional Facial Symmetry Loss for Deep Face Recognition
di: Prakash, Pritesh, et al.
Pubblicazione: (2024)
di: Prakash, Pritesh, et al.
Pubblicazione: (2024)
LeGO: Leveraging a Surface Deformation Network for Animatable Stylized Face Generation with One Example
di: Yoon, Soyeon, et al.
Pubblicazione: (2024)
di: Yoon, Soyeon, et al.
Pubblicazione: (2024)
EditP23: 3D Editing via Propagation of Image Prompts to Multi-View
di: Bar-On, Roi, et al.
Pubblicazione: (2025)
di: Bar-On, Roi, et al.
Pubblicazione: (2025)
N-DriverMotion: Driver motion learning and prediction using an event-based camera and directly trained spiking neural networks on Loihi 2
di: Chung, Hyo Jong, et al.
Pubblicazione: (2024)
di: Chung, Hyo Jong, et al.
Pubblicazione: (2024)
Few-Class Arena: A Benchmark for Efficient Selection of Vision Models and Dataset Difficulty Measurement
di: Cao, Bryan Bo, et al.
Pubblicazione: (2024)
di: Cao, Bryan Bo, et al.
Pubblicazione: (2024)
Event-ECC: Asynchronous Tracking of Events with Continuous Optimization
di: Zafeiri, Maria, et al.
Pubblicazione: (2024)
di: Zafeiri, Maria, et al.
Pubblicazione: (2024)
HieraEdgeNet: A Multi-Scale Edge-Enhanced Framework for Automated Pollen Recognition
di: Long, Yuchong, et al.
Pubblicazione: (2025)
di: Long, Yuchong, et al.
Pubblicazione: (2025)
The Impact of Image Resolution on Face Detection: A Comparative Analysis of MTCNN, YOLOv XI and YOLOv XII models
di: Ömercikoğlu, Ahmet Can, et al.
Pubblicazione: (2025)
di: Ömercikoğlu, Ahmet Can, et al.
Pubblicazione: (2025)
SelectiveKD: A semi-supervised framework for cancer detection in DBT through Knowledge Distillation and Pseudo-labeling
di: Dillard, Laurent, et al.
Pubblicazione: (2024)
di: Dillard, Laurent, et al.
Pubblicazione: (2024)
Distilling foundation models for robust and efficient models in digital pathology
di: Filiot, Alexandre, et al.
Pubblicazione: (2025)
di: Filiot, Alexandre, et al.
Pubblicazione: (2025)
Harnessing Deep Learning and Satellite Imagery for Post-Buyout Land Cover Mapping
di: Otal, Hakan T., et al.
Pubblicazione: (2024)
di: Otal, Hakan T., et al.
Pubblicazione: (2024)
Fixed-Threshold Evaluation of a Hybrid CNN-ViT for AI-Generated Image Detection Across Photos and Art
di: Khan, Md Ashik, et al.
Pubblicazione: (2025)
di: Khan, Md Ashik, et al.
Pubblicazione: (2025)
TextDoctor: Unified Document Image Inpainting via Patch Pyramid Diffusion Models
di: Lu, Wanglong, et al.
Pubblicazione: (2025)
di: Lu, Wanglong, et al.
Pubblicazione: (2025)
FastFit: Accelerating Multi-Reference Virtual Try-On via Cacheable Diffusion Models
di: Chong, Zheng, et al.
Pubblicazione: (2025)
di: Chong, Zheng, et al.
Pubblicazione: (2025)
DeepFusionNet: Autoencoder-Based Low-Light Image Enhancement and Super-Resolution
di: Çalışkan, Halil Hüseyin, et al.
Pubblicazione: (2025)
di: Çalışkan, Halil Hüseyin, et al.
Pubblicazione: (2025)
Tri-Plane Mamba: Efficiently Adapting Segment Anything Model for 3D Medical Images
di: Wang, Hualiang, et al.
Pubblicazione: (2024)
di: Wang, Hualiang, et al.
Pubblicazione: (2024)
StatsMerging: Statistics-Guided Model Merging via Task-Specific Teacher Distillation
di: Merugu, Ranjith, et al.
Pubblicazione: (2025)
di: Merugu, Ranjith, et al.
Pubblicazione: (2025)
Motion-Based Sign Language Video Summarization using Curvature and Torsion
di: Sartinas, Evangelos G., et al.
Pubblicazione: (2023)
di: Sartinas, Evangelos G., et al.
Pubblicazione: (2023)
Eleven Primitives and Three Gates: The Universal Structure of Computational Imaging
di: Yang, Chengshuai, et al.
Pubblicazione: (2026)
di: Yang, Chengshuai, et al.
Pubblicazione: (2026)
Rethinking VLMs for Image Forgery Detection and Localization
di: Guo, Shaofeng, et al.
Pubblicazione: (2026)
di: Guo, Shaofeng, et al.
Pubblicazione: (2026)
ForensicFormer: Hierarchical Multi-Scale Reasoning for Cross-Domain Image Forgery Detection
di: Samson, Hema Hariharan
Pubblicazione: (2026)
di: Samson, Hema Hariharan
Pubblicazione: (2026)
BG-YOLO: A Bidirectional-Guided Method for Underwater Object Detection
di: Zhang, Jian, et al.
Pubblicazione: (2024)
di: Zhang, Jian, et al.
Pubblicazione: (2024)
CatVTON: Concatenation Is All You Need for Virtual Try-On with Diffusion Models
di: Chong, Zheng, et al.
Pubblicazione: (2024)
di: Chong, Zheng, et al.
Pubblicazione: (2024)
SPEAK: Speech-Driven Pose and Emotion-Adjustable Talking Head Generation
di: Cai, Changpeng, et al.
Pubblicazione: (2024)
di: Cai, Changpeng, et al.
Pubblicazione: (2024)
CatV2TON: Taming Diffusion Transformers for Vision-Based Virtual Try-On with Temporal Concatenation
di: Chong, Zheng, et al.
Pubblicazione: (2025)
di: Chong, Zheng, et al.
Pubblicazione: (2025)
S3Simulator: A benchmarking Side Scan Sonar Simulator dataset for Underwater Image Analysis
di: S, Kamal Basha, et al.
Pubblicazione: (2024)
di: S, Kamal Basha, et al.
Pubblicazione: (2024)
Mapping Tomato Cropping Systems in California Using AlphaEarth Geospatial Embeddings and Deep Learning Analysis
di: Narimani, Mohammadreza, et al.
Pubblicazione: (2026)
di: Narimani, Mohammadreza, et al.
Pubblicazione: (2026)
Satellite-Net: Automatic Extraction of Land Cover Indicators from Satellite Imagery by Deep Learning
di: Bernasconi, Eleonora, et al.
Pubblicazione: (2019)
di: Bernasconi, Eleonora, et al.
Pubblicazione: (2019)
Balanced conic rectified flow
di: Kim, Shin Seong, et al.
Pubblicazione: (2025)
di: Kim, Shin Seong, et al.
Pubblicazione: (2025)
CADE 2.5 - ZeResFDG: Frequency-Decoupled, Rescaled and Zero-Projected Guidance for SD/SDXL Latent Diffusion Models
di: Rychkovskiy, Denis
Pubblicazione: (2025)
di: Rychkovskiy, Denis
Pubblicazione: (2025)
Identity Deepfake Threats to Biometric Authentication Systems: Public and Expert Perspectives
di: He, Shijing, et al.
Pubblicazione: (2025)
di: He, Shijing, et al.
Pubblicazione: (2025)
Documenti analoghi
-
CigTime: Corrective Instruction Generation Through Inverse Motion Editing
di: Fang, Qihang, et al.
Pubblicazione: (2024) -
Gaze into the Heart: A Multi-View Video Dataset for rPPG and Health Biomarkers Estimation
di: Egorov, Konstantin, et al.
Pubblicazione: (2025) -
Efficient Vision-based Vehicle Speed Estimation
di: Macko, Andrej, et al.
Pubblicazione: (2025) -
FACEMUG: A Multimodal Generative and Fusion Framework for Local Facial Editing
di: Lu, Wanglong, et al.
Pubblicazione: (2024) -
Visual Style Prompt Learning Using Diffusion Models for Blind Face Restoration
di: Lu, Wanglong, et al.
Pubblicazione: (2024)