Efficient High-Resolution Image Editing with Hallucination-Aware Loss and Adaptive Tiling
Fuente:
arXiv
Guardado en:
| Autores principales: | Kwon, Young D., Mehrotra, Abhinav, Chadwick, Malcolm, Ramos, Alberto Gil, Bhattacharya, Sourav |
|---|---|
| Formato: | Preprint |
| Publicado: |
2025
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
Guidance Free Image Editing via Explicit Conditioning
por: Noroozi, Mehdi, et al.
Publicado: (2025)
por: Noroozi, Mehdi, et al.
Publicado: (2025)
FraQAT: Quantization Aware Training with Fractional bits
por: Morreale, Luca, et al.
Publicado: (2025)
por: Morreale, Luca, et al.
Publicado: (2025)
Upcycling Text-to-Image Diffusion Models for Multi-Task Capabilities
por: Chavhan, Ruchika, et al.
Publicado: (2025)
por: Chavhan, Ruchika, et al.
Publicado: (2025)
RFDM: Residual Flow Diffusion Model for Efficient Causal Video Editing
por: Salehi, Mohammadreza, et al.
Publicado: (2026)
por: Salehi, Mohammadreza, et al.
Publicado: (2026)
EDiT: Efficient Diffusion Transformers with Linear Compressed Attention
por: Becker, Philipp, et al.
Publicado: (2025)
por: Becker, Philipp, et al.
Publicado: (2025)
NanoFLUX: Distillation-Driven Compression of Large Text-to-Image Generation Models for Mobile Devices
por: Chavhan, Ruchika, et al.
Publicado: (2026)
por: Chavhan, Ruchika, et al.
Publicado: (2026)
Fast Sampling Through The Reuse Of Attention Maps In Diffusion Models
por: Hunter, Rosco, et al.
Publicado: (2023)
por: Hunter, Rosco, et al.
Publicado: (2023)
HierEdit: Region-Aware Hierarchical Diffusion for Efficient High-Resolution Editing
por: Zhang, Yuyao, et al.
Publicado: (2026)
por: Zhang, Yuyao, et al.
Publicado: (2026)
HierarchicalPrune: Position-Aware Compression for Large-Scale Diffusion Models
por: Kwon, Young D., et al.
Publicado: (2025)
por: Kwon, Young D., et al.
Publicado: (2025)
Mitigating Hallucinations in Diffusion Models through Adaptive Attention Modulation
por: Oorloff, Trevine, et al.
Publicado: (2025)
por: Oorloff, Trevine, et al.
Publicado: (2025)
Divide and Conquer: High-Resolution Industrial Anomaly Detection via Memory Efficient Tiled Ensemble
por: Rolih, Blaž, et al.
Publicado: (2024)
por: Rolih, Blaž, et al.
Publicado: (2024)
RegionE: Adaptive Region-Aware Generation for Efficient Image Editing
por: Chen, Pengtao, et al.
Publicado: (2025)
por: Chen, Pengtao, et al.
Publicado: (2025)
Image Tiling for High-Resolution Reasoning: Balancing Local Detail with Global Context
por: de Margerie, Anatole Jacquin, et al.
Publicado: (2025)
por: de Margerie, Anatole Jacquin, et al.
Publicado: (2025)
Low-Resolution Editing is All You Need for High-Resolution Editing
por: Lee, Junsung, et al.
Publicado: (2025)
por: Lee, Junsung, et al.
Publicado: (2025)
From Full Boards to Tiny Defects: Scale-Aware Tile Inference with Topology-Aware Merging for High-Resolution PCB Defect Detection
por: Shalmani, Mohammad Alijanpour, et al.
Publicado: (2026)
por: Shalmani, Mohammad Alijanpour, et al.
Publicado: (2026)
Towards Efficient Exemplar Based Image Editing with Multimodal VLMs
por: Jadhav, Avadhoot, et al.
Publicado: (2025)
por: Jadhav, Avadhoot, et al.
Publicado: (2025)
Hallucination Score: Towards Mitigating Hallucinations in Generative Image Super-Resolution
por: Ren, Weiming, et al.
Publicado: (2025)
por: Ren, Weiming, et al.
Publicado: (2025)
Frequency-Aware Autoregressive Modeling for Efficient High-Resolution Image Synthesis
por: Chen, Zhuokun, et al.
Publicado: (2025)
por: Chen, Zhuokun, et al.
Publicado: (2025)
One Look is Enough: Seamless Patchwise Refinement for Zero-Shot Monocular Depth Estimation on High-Resolution Images
por: Kwon, Byeongjun, et al.
Publicado: (2025)
por: Kwon, Byeongjun, et al.
Publicado: (2025)
AFTER: Mitigating the Object Hallucination of LVLM via Adaptive Factual-Guided Activation Editing
por: Wang, Tianbo, et al.
Publicado: (2026)
por: Wang, Tianbo, et al.
Publicado: (2026)
EfficientIML: Efficient High-Resolution Image Manipulation Localization
por: Li, Jinhan, et al.
Publicado: (2025)
por: Li, Jinhan, et al.
Publicado: (2025)
VINS-120K: Ultra High-Resolution Image Editing with A Large-Scale Dataset
por: Chen, Zhizhou, et al.
Publicado: (2026)
por: Chen, Zhizhou, et al.
Publicado: (2026)
EditCrafter: Tuning-free High-Resolution Image Editing via Pretrained Diffusion Model
por: Kim, Kunho, et al.
Publicado: (2026)
por: Kim, Kunho, et al.
Publicado: (2026)
AFRAgent : An Adaptive Feature Renormalization Based High Resolution Aware GUI agent
por: Anand, Neeraj, et al.
Publicado: (2025)
por: Anand, Neeraj, et al.
Publicado: (2025)
Fieldscale: Locality-Aware Field-based Adaptive Rescaling for Thermal Infrared Image
por: Gil, Hyeonjae, et al.
Publicado: (2024)
por: Gil, Hyeonjae, et al.
Publicado: (2024)
Training-free, Perceptually Consistent Low-Resolution Previews with High-Resolution Image for Efficient Workflows of Diffusion Models
por: Jeong, Wongi, et al.
Publicado: (2026)
por: Jeong, Wongi, et al.
Publicado: (2026)
REJEPA: A Novel Joint-Embedding Predictive Architecture for Efficient Remote Sensing Image Retrieval
por: Choudhury, Shabnam, et al.
Publicado: (2025)
por: Choudhury, Shabnam, et al.
Publicado: (2025)
HIME: Mitigating Object Hallucinations in LVLMs via Hallucination Insensitivity Model Editing
por: Akl, Ahmed, et al.
Publicado: (2026)
por: Akl, Ahmed, et al.
Publicado: (2026)
Uncertainty-Aware Source-Free Adaptive Image Super-Resolution with Wavelet Augmentation Transformer
por: Ai, Yuang, et al.
Publicado: (2023)
por: Ai, Yuang, et al.
Publicado: (2023)
GeoSynth: Contextually-Aware High-Resolution Satellite Image Synthesis
por: Sastry, Srikumar, et al.
Publicado: (2024)
por: Sastry, Srikumar, et al.
Publicado: (2024)
Efficient Image Super-Resolution with Multi-Scale Spatial Adaptive Attention Networks
por: Rao, Sushi, et al.
Publicado: (2026)
por: Rao, Sushi, et al.
Publicado: (2026)
CATANet: Efficient Content-Aware Token Aggregation for Lightweight Image Super-Resolution
por: Liu, Xin, et al.
Publicado: (2025)
por: Liu, Xin, et al.
Publicado: (2025)
Task-Aware Dynamic Transformer for Efficient Arbitrary-Scale Image Super-Resolution
por: Xu, Tianyi, et al.
Publicado: (2024)
por: Xu, Tianyi, et al.
Publicado: (2024)
ESOD: Efficient Small Object Detection on High-Resolution Images
por: Liu, Kai, et al.
Publicado: (2024)
por: Liu, Kai, et al.
Publicado: (2024)
PhysEdit: Physically-Consistent Region-Aware Image Editing via Adaptive Spatio-Temporal Reasoning
por: Li, Guandong, et al.
Publicado: (2026)
por: Li, Guandong, et al.
Publicado: (2026)
Histogram Assisted Quality Aware Generative Model for Resolution Invariant NIR Image Colorization
por: Attri, Abhinav, et al.
Publicado: (2026)
por: Attri, Abhinav, et al.
Publicado: (2026)
Tiled Prompts: Overcoming Prompt Misguidance in Image and Video Super-Resolution
por: Kim, Bryan Sangwoo, et al.
Publicado: (2026)
por: Kim, Bryan Sangwoo, et al.
Publicado: (2026)
Leveraging Adaptive Implicit Representation Mapping for Ultra High-Resolution Image Segmentation
por: Zhao, Ziyu, et al.
Publicado: (2024)
por: Zhao, Ziyu, et al.
Publicado: (2024)
Beyond Image Super-Resolution for Image Recognition with Task-Driven Perceptual Loss
por: Kim, Jaeha, et al.
Publicado: (2024)
por: Kim, Jaeha, et al.
Publicado: (2024)
DiM: Diffusion Mamba for Efficient High-Resolution Image Synthesis
por: Teng, Yao, et al.
Publicado: (2024)
por: Teng, Yao, et al.
Publicado: (2024)
Ejemplares similares
-
Guidance Free Image Editing via Explicit Conditioning
por: Noroozi, Mehdi, et al.
Publicado: (2025) -
FraQAT: Quantization Aware Training with Fractional bits
por: Morreale, Luca, et al.
Publicado: (2025) -
Upcycling Text-to-Image Diffusion Models for Multi-Task Capabilities
por: Chavhan, Ruchika, et al.
Publicado: (2025) -
RFDM: Residual Flow Diffusion Model for Efficient Causal Video Editing
por: Salehi, Mohammadreza, et al.
Publicado: (2026) -
EDiT: Efficient Diffusion Transformers with Linear Compressed Attention
por: Becker, Philipp, et al.
Publicado: (2025)