PixLens: A Novel Framework for Disentangled Evaluation in Diffusion-Based Image Editing with Object Detection + SAM
Fuente:
arXiv
Saved in:
| Main Authors: | Stefanache, Stefan, Pérez, Lluís Pastor, Watanabe, Julen Costa, Tejedor, Ernesto Sanchez, Hofmann, Thomas, Simsar, Enis |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
JEDI: The Force of Jensen-Shannon Divergence in Disentangling Diffusion Models
by: Bill, Eric Tillmann, et al.
Published: (2025)
by: Bill, Eric Tillmann, et al.
Published: (2025)
LIME: Localized Image Editing via Attention Regularization in Diffusion Models
by: Simsar, Enis, et al.
Published: (2023)
by: Simsar, Enis, et al.
Published: (2023)
UIP2P: Unsupervised Instruction-based Image Editing via Edit Reversibility Constraint
by: Simsar, Enis, et al.
Published: (2024)
by: Simsar, Enis, et al.
Published: (2024)
FOCUS: Optimal Control for Multi-Entity World Modeling in Text-to-Image Generation
by: Bill, Eric Tillmann, et al.
Published: (2025)
by: Bill, Eric Tillmann, et al.
Published: (2025)
LoRACLR: Contrastive Adaptation for Customization of Diffusion Models
by: Simsar, Enis, et al.
Published: (2024)
by: Simsar, Enis, et al.
Published: (2024)
MegaPortrait: Revisiting Diffusion Control for High-fidelity Portrait Generation
by: Yang, Han, et al.
Published: (2024)
by: Yang, Han, et al.
Published: (2024)
FullFlow: Upgrading Text-to-Image Flow Matching Models for Bidirectional Vision--Language Generation
by: Bill, Eric Tillmann, et al.
Published: (2026)
by: Bill, Eric Tillmann, et al.
Published: (2026)
SHYI: Action Support for Contrastive Learning in High-Fidelity Text-to-Image Generation
by: Xia, Tianxiang, et al.
Published: (2025)
by: Xia, Tianxiang, et al.
Published: (2025)
Contrastive Test-Time Composition of Multiple LoRA Models for Image Generation
by: Meral, Tuna Han Salih, et al.
Published: (2024)
by: Meral, Tuna Han Salih, et al.
Published: (2024)
ReasonPix2Pix: Instruction Reasoning Dataset for Advanced Image Editing
by: Jin, Ying, et al.
Published: (2024)
by: Jin, Ying, et al.
Published: (2024)
Shifting the Breaking Point of Flow Matching for Multi-Instance Editing
by: Zaccagnino, Carmine, et al.
Published: (2026)
by: Zaccagnino, Carmine, et al.
Published: (2026)
InstructRL4Pix: Training Diffusion for Image Editing by Reinforcement Learning
by: Li, Tiancheng, et al.
Published: (2024)
by: Li, Tiancheng, et al.
Published: (2024)
Stylebreeder: Exploring and Democratizing Artistic Styles through Text-to-Image Models
by: Zheng, Matthew, et al.
Published: (2024)
by: Zheng, Matthew, et al.
Published: (2024)
IC-Portrait: In-Context Matching for View-Consistent Personalized Portrait
by: Yang, Han, et al.
Published: (2025)
by: Yang, Han, et al.
Published: (2025)
Evaluating SAM2's Role in Camouflaged Object Detection: From SAM to SAM2
by: Tang, Lv, et al.
Published: (2024)
by: Tang, Lv, et al.
Published: (2024)
La dramatización radiofónica de contenidos educativos: Una experiencia universitaria
by: Lluís Pastor
Published: (2010)
by: Lluís Pastor
Published: (2010)
Novel Hybrid Integrated Pix2Pix and WGAN Model with Gradient Penalty for Binary Images Denoising
by: Tirel, Luca, et al.
Published: (2024)
by: Tirel, Luca, et al.
Published: (2024)
PixPerfect: Seamless Latent Diffusion Local Editing with Discriminative Pixel-Space Refinement
by: Zheng, Haitian, et al.
Published: (2025)
by: Zheng, Haitian, et al.
Published: (2025)
MULTI: Disentangling Camera Lens, Sensor, View, and Domain for Novel Image Generation
by: Godavarthy, Sonali, et al.
Published: (2026)
by: Godavarthy, Sonali, et al.
Published: (2026)
SAM-COD: SAM-guided Unified Framework for Weakly-Supervised Camouflaged Object Detection
by: Chen, Huafeng, et al.
Published: (2024)
by: Chen, Huafeng, et al.
Published: (2024)
ST-SAM: SAM-Driven Self-Training Framework for Semi-Supervised Camouflaged Object Detection
by: Hu, Xihang, et al.
Published: (2025)
by: Hu, Xihang, et al.
Published: (2025)
Identificación de sectores de servicios y de alta tecnología en la Comunidad Valenciana: ¿Un nuevo cluster mapping?
by: Lluís Miret-Pastor
Published: (2011)
by: Lluís Miret-Pastor
Published: (2011)
Hacia una gestión integral de los contenidos
by: Lluís Pastor Pérez
Published: (2006)
by: Lluís Pastor Pérez
Published: (2006)
Low-Light Image Enhancement Framework for Improved Object Detection in Fisheye Lens Datasets
by: Tran, Dai Quoc, et al.
Published: (2024)
by: Tran, Dai Quoc, et al.
Published: (2024)
Object detection in adverse weather conditions for autonomous vehicles using Instruct Pix2Pix
by: Gurbindo, Unai, et al.
Published: (2025)
by: Gurbindo, Unai, et al.
Published: (2025)
Creating a holistic excellence model adapted for technology-based companies
by: Ana Clara Pastor Tejedor
Published: (2014)
by: Ana Clara Pastor Tejedor
Published: (2014)
COMPARACIÓN DE LOS MODELOS DE EVALUACIÓN DE LA EXCELENCIA EMPRESARIAL.
by: Ana Clara Pastor Tejedor
Published: (2013)
by: Ana Clara Pastor Tejedor
Published: (2013)
Mapping New Realities: Ground Truth Image Creation with Pix2Pix Image-to-Image Translation
by: Li, Zhenglin, et al.
Published: (2024)
by: Li, Zhenglin, et al.
Published: (2024)
FlowDet: Unifying Object Detection and Generative Transport Flows
by: Baty, Enis, et al.
Published: (2025)
by: Baty, Enis, et al.
Published: (2025)
Enhanced Pix2Pix GAN for Visual Defect Removal in UAV-Captured Images
by: Rizun, Volodymyr
Published: (2024)
by: Rizun, Volodymyr
Published: (2024)
Ambient-Pix2PixGAN for Translating Medical Images from Noisy Data
by: Chen, Wentao, et al.
Published: (2024)
by: Chen, Wentao, et al.
Published: (2024)
Disentangling Instruction Influence in Diffusion Transformers for Parallel Multi-Instruction-Guided Image Editing
by: Liu, Hui, et al.
Published: (2025)
by: Liu, Hui, et al.
Published: (2025)
PixNerd: Pixel Neural Field Diffusion
by: Wang, Shuai, et al.
Published: (2025)
by: Wang, Shuai, et al.
Published: (2025)
TerraCodec: Compressing Optical Earth Observation Data
by: Costa-Watanabe, Julen, et al.
Published: (2025)
by: Costa-Watanabe, Julen, et al.
Published: (2025)
GANTASTIC: GAN-based Transfer of Interpretable Directions for Disentangled Image Editing in Text-to-Image Diffusion Models
by: Dalva, Yusuf, et al.
Published: (2024)
by: Dalva, Yusuf, et al.
Published: (2024)
Crowd-SAM: SAM as a Smart Annotator for Object Detection in Crowded Scenes
by: Cai, Zhi, et al.
Published: (2024)
by: Cai, Zhi, et al.
Published: (2024)
InstructPix2NeRF: Instructed 3D Portrait Editing from a Single Image
by: Li, Jianhui, et al.
Published: (2023)
by: Li, Jianhui, et al.
Published: (2023)
A Dual-Path Generative Framework for Zero-Day Fraud Detection in Banking Systems
by: Ismail, Nasim Abdirahman, et al.
Published: (2026)
by: Ismail, Nasim Abdirahman, et al.
Published: (2026)
MDE-Edit: Masked Dual-Editing for Multi-Object Image Editing via Diffusion Models
by: Zhu, Hongyang, et al.
Published: (2025)
by: Zhu, Hongyang, et al.
Published: (2025)
Image-to-Image Translation with Disentangled Latent Vectors for Face Editing
by: Dalva, Yusuf, et al.
Published: (2023)
by: Dalva, Yusuf, et al.
Published: (2023)
Similar Items
-
JEDI: The Force of Jensen-Shannon Divergence in Disentangling Diffusion Models
by: Bill, Eric Tillmann, et al.
Published: (2025) -
LIME: Localized Image Editing via Attention Regularization in Diffusion Models
by: Simsar, Enis, et al.
Published: (2023) -
UIP2P: Unsupervised Instruction-based Image Editing via Edit Reversibility Constraint
by: Simsar, Enis, et al.
Published: (2024) -
FOCUS: Optimal Control for Multi-Entity World Modeling in Text-to-Image Generation
by: Bill, Eric Tillmann, et al.
Published: (2025) -
LoRACLR: Contrastive Adaptation for Customization of Diffusion Models
by: Simsar, Enis, et al.
Published: (2024)