Improving Image Clustering with Artifacts Attenuation via Inference-Time Attention Engineering
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Nakamura, Kazumoto, Nozawa, Yuji, Lin, Yu-Chieh, Nakata, Kengo, Ng, Youyang |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2024
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Prompt-Guided Attention Head Selection for Focus-Oriented Image Retrieval
von: Nozawa, Yuji, et al.
Veröffentlicht: (2025)
von: Nozawa, Yuji, et al.
Veröffentlicht: (2025)
Revisiting Relevance Feedback for CLIP-based Interactive Image Retrieval
von: Nara, Ryoya, et al.
Veröffentlicht: (2024)
von: Nara, Ryoya, et al.
Veröffentlicht: (2024)
CIRCLED: A Multi-turn CIR Dataset with Consistent Dialogues across Domains
von: Takeda, Tomohisa, et al.
Veröffentlicht: (2026)
von: Takeda, Tomohisa, et al.
Veröffentlicht: (2026)
Rethinking Sparse Lexical Representations for Image Retrieval in the Age of Rising Multi-Modal Large Language Models
von: Nakata, Kengo, et al.
Veröffentlicht: (2024)
von: Nakata, Kengo, et al.
Veröffentlicht: (2024)
Clustering-friendly Representation Learning for Enhancing Salient Features
von: Oshima, Toshiyuki, et al.
Veröffentlicht: (2024)
von: Oshima, Toshiyuki, et al.
Veröffentlicht: (2024)
ArtifactLens: Hundreds of Labels Are Enough for Artifact Detection with VLMs
von: Burgess, James, et al.
Veröffentlicht: (2026)
von: Burgess, James, et al.
Veröffentlicht: (2026)
Feature-based Graph Attention Networks Improve Online Continual Learning
von: Sim, Adjovi, et al.
Veröffentlicht: (2025)
von: Sim, Adjovi, et al.
Veröffentlicht: (2025)
Improving Artifact Robustness for CT Deep Learning Models Without Labeled Artifact Images via Domain Adaptation
von: Cheung, Justin, et al.
Veröffentlicht: (2025)
von: Cheung, Justin, et al.
Veröffentlicht: (2025)
MaskBit: Embedding-free Image Generation via Bit Tokens
von: Weber, Mark, et al.
Veröffentlicht: (2024)
von: Weber, Mark, et al.
Veröffentlicht: (2024)
Constructive Distortion: Improving MLLMs with Attention-Guided Image Warping
von: Dalal, Dwip, et al.
Veröffentlicht: (2025)
von: Dalal, Dwip, et al.
Veröffentlicht: (2025)
Prominence-Aware Artifact Detection and Dataset for Image Super-Resolution
von: Molodetskikh, Ivan, et al.
Veröffentlicht: (2025)
von: Molodetskikh, Ivan, et al.
Veröffentlicht: (2025)
VIVAT: Virtuous Improving VAE Training through Artifact Mitigation
von: Novitskiy, Lev, et al.
Veröffentlicht: (2025)
von: Novitskiy, Lev, et al.
Veröffentlicht: (2025)
Atlas: Multi-Scale Attention Improves Long Context Image Modeling
von: Agrawal, Kumar Krishna, et al.
Veröffentlicht: (2025)
von: Agrawal, Kumar Krishna, et al.
Veröffentlicht: (2025)
Similarity Trajectories: Linking Sampling Process to Artifacts in Diffusion-Generated Images
von: Menn, Dennis, et al.
Veröffentlicht: (2024)
von: Menn, Dennis, et al.
Veröffentlicht: (2024)
Attention-based Shape-Deformation Networks for Artifact-Free Geometry Reconstruction of Lumbar Spine from MR Images
von: Qian, Linchen, et al.
Veröffentlicht: (2024)
von: Qian, Linchen, et al.
Veröffentlicht: (2024)
Scaling Inference-Time Search with Vision Value Model for Improved Visual Comprehension
von: Wang, Xiyao, et al.
Veröffentlicht: (2024)
von: Wang, Xiyao, et al.
Veröffentlicht: (2024)
Time Travel: A Comprehensive Benchmark to Evaluate LMMs on Historical and Cultural Artifacts
von: Ghaboura, Sara, et al.
Veröffentlicht: (2025)
von: Ghaboura, Sara, et al.
Veröffentlicht: (2025)
ODE$_t$(ODE$_l$): Shortcutting the Time and the Length in Diffusion and Flow Models for Faster Sampling
von: Gudovskiy, Denis, et al.
Veröffentlicht: (2025)
von: Gudovskiy, Denis, et al.
Veröffentlicht: (2025)
UnSegMedGAT: Unsupervised Medical Image Segmentation using Graph Attention Networks Clustering
von: Adityaja, A. Mudit, et al.
Veröffentlicht: (2024)
von: Adityaja, A. Mudit, et al.
Veröffentlicht: (2024)
Rate-In: Information-Driven Adaptive Dropout Rates for Improved Inference-Time Uncertainty Estimation
von: Zeevi, Tal, et al.
Veröffentlicht: (2024)
von: Zeevi, Tal, et al.
Veröffentlicht: (2024)
FrameBridge: Improving Image-to-Video Generation with Bridge Models
von: Wang, Yuji, et al.
Veröffentlicht: (2024)
von: Wang, Yuji, et al.
Veröffentlicht: (2024)
Aletheia: Physics-Conditioned Localized Artifact Attention (PhyLAA-X) for End-to-End Generalizable and Robust Deepfake Video Detection
von: Ghori, Devendra
Veröffentlicht: (2026)
von: Ghori, Devendra
Veröffentlicht: (2026)
QIANets: Quantum-Integrated Adaptive Networks for Reduced Latency and Improved Inference Times in CNN Models
von: Balapanov, Zhumazhan, et al.
Veröffentlicht: (2024)
von: Balapanov, Zhumazhan, et al.
Veröffentlicht: (2024)
NeRF-US: Removing Ultrasound Imaging Artifacts from Neural Radiance Fields in the Wild
von: Dagli, Rishit, et al.
Veröffentlicht: (2024)
von: Dagli, Rishit, et al.
Veröffentlicht: (2024)
SynArtifact: Classifying and Alleviating Artifacts in Synthetic Images via Vision-Language Model
von: Cao, Bin, et al.
Veröffentlicht: (2024)
von: Cao, Bin, et al.
Veröffentlicht: (2024)
Large Scale Diffusion Distillation via Score-Regularized Continuous-Time Consistency
von: Zheng, Kaiwen, et al.
Veröffentlicht: (2025)
von: Zheng, Kaiwen, et al.
Veröffentlicht: (2025)
Image Clustering via the Principle of Rate Reduction in the Age of Pretrained Models
von: Chu, Tianzhe, et al.
Veröffentlicht: (2023)
von: Chu, Tianzhe, et al.
Veröffentlicht: (2023)
EPIC: Efficient Predicate-Guided Inference-Time Control for Compositional Text-to-Image Generation
von: Mun, Sunung, et al.
Veröffentlicht: (2026)
von: Mun, Sunung, et al.
Veröffentlicht: (2026)
Performance Plateaus in Inference-Time Scaling for Text-to-Image Diffusion Without External Models
von: Choi, Changhyun, et al.
Veröffentlicht: (2025)
von: Choi, Changhyun, et al.
Veröffentlicht: (2025)
Spectral Evolution Search: Efficient Inference-Time Scaling for Reward-Aligned Image Generation
von: Ye, Jinyan, et al.
Veröffentlicht: (2026)
von: Ye, Jinyan, et al.
Veröffentlicht: (2026)
Faster Inference of Flow-Based Generative Models via Improved Data-Noise Coupling
von: Davtyan, Aram, et al.
Veröffentlicht: (2026)
von: Davtyan, Aram, et al.
Veröffentlicht: (2026)
Not All Attention Heads Are What You Need: Refining CLIP's Image Representation with Attention Ablation
von: Lin, Feng, et al.
Veröffentlicht: (2025)
von: Lin, Feng, et al.
Veröffentlicht: (2025)
Style Aligned Image Generation via Shared Attention
von: Hertz, Amir, et al.
Veröffentlicht: (2023)
von: Hertz, Amir, et al.
Veröffentlicht: (2023)
ASAP: Attention-Shift-Aware Pruning for Efficient LVLM Inference
von: Pathak, Surendra, et al.
Veröffentlicht: (2026)
von: Pathak, Surendra, et al.
Veröffentlicht: (2026)
CHAI: CacHe Attention Inference for text2video
von: Cherian, Joel Mathew, et al.
Veröffentlicht: (2026)
von: Cherian, Joel Mathew, et al.
Veröffentlicht: (2026)
Text-Guided Image Clustering
von: Stephan, Andreas, et al.
Veröffentlicht: (2024)
von: Stephan, Andreas, et al.
Veröffentlicht: (2024)
Local Clustering for Lung Cancer Image Classification via Sparse Solution Technique
von: Hamel, Jackson, et al.
Veröffentlicht: (2024)
von: Hamel, Jackson, et al.
Veröffentlicht: (2024)
Multimodal Metadata Assignment for Cultural Heritage Artifacts
von: Rei, Luis, et al.
Veröffentlicht: (2024)
von: Rei, Luis, et al.
Veröffentlicht: (2024)
Hierarchical Compact Clustering Attention (COCA) for Unsupervised Object-Centric Learning
von: Küçüksözen, Can, et al.
Veröffentlicht: (2025)
von: Küçüksözen, Can, et al.
Veröffentlicht: (2025)
Inference-Time Scaling for Flow Models via Stochastic Generation and Rollover Budget Forcing
von: Kim, Jaihoon, et al.
Veröffentlicht: (2025)
von: Kim, Jaihoon, et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
Prompt-Guided Attention Head Selection for Focus-Oriented Image Retrieval
von: Nozawa, Yuji, et al.
Veröffentlicht: (2025) -
Revisiting Relevance Feedback for CLIP-based Interactive Image Retrieval
von: Nara, Ryoya, et al.
Veröffentlicht: (2024) -
CIRCLED: A Multi-turn CIR Dataset with Consistent Dialogues across Domains
von: Takeda, Tomohisa, et al.
Veröffentlicht: (2026) -
Rethinking Sparse Lexical Representations for Image Retrieval in the Age of Rising Multi-Modal Large Language Models
von: Nakata, Kengo, et al.
Veröffentlicht: (2024) -
Clustering-friendly Representation Learning for Enhancing Salient Features
von: Oshima, Toshiyuki, et al.
Veröffentlicht: (2024)