On the Provable Importance of Gradients for Language-Assisted Image Clustering
Fuente:
arXiv
Saved in:
| Main Authors: | Peng, Bo, Lu, Jie, Zhang, Guangquan, Fang, Zhen |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Debiased Negative Mining Improves Out-of-distribution Detection with Pre-trained Vision-Language Models
by: Peng, Bo, et al.
Published: (2026)
by: Peng, Bo, et al.
Published: (2026)
Delving into Spectral Clustering with Vision-Language Representations
by: Peng, Bo, et al.
Published: (2026)
by: Peng, Bo, et al.
Published: (2026)
MiraGe: Multimodal Discriminative Representation Learning for Generalizable AI-Generated Image Detection
by: Shi, Kuo, et al.
Published: (2025)
by: Shi, Kuo, et al.
Published: (2025)
Decoder Gradient Shield: Provable and High-Fidelity Prevention of Gradient-Based Box-Free Watermark Removal
by: An, Haonan, et al.
Published: (2025)
by: An, Haonan, et al.
Published: (2025)
Efficient Adaptive Label Refinement for Label Noise Learning
by: Zhang, Wenzhen, et al.
Published: (2025)
by: Zhang, Wenzhen, et al.
Published: (2025)
Respecting Modality Gap in Post-hoc Out-of-distribution Detection with Pre-trained Vision-Language Models
by: Hu, Yuanwei, et al.
Published: (2026)
by: Hu, Yuanwei, et al.
Published: (2026)
Decoder Gradient Shields: A Family of Provable and High-Fidelity Methods Against Gradient-Based Box-Free Watermark Removal
by: An, Haonan, et al.
Published: (2026)
by: An, Haonan, et al.
Published: (2026)
Towards Omnimodal Expressions and Reasoning in Referring Audio-Visual Segmentation
by: Ying, Kaining, et al.
Published: (2025)
by: Ying, Kaining, et al.
Published: (2025)
ROSE: Retrieval-Oriented Segmentation Enhancement
by: Tang, Song, et al.
Published: (2026)
by: Tang, Song, et al.
Published: (2026)
Once for Both: Single Stage of Importance and Sparsity Search for Vision Transformer Compression
by: Ye, Hancheng, et al.
Published: (2024)
by: Ye, Hancheng, et al.
Published: (2024)
Provable Ordering and Continuity in Vision-Language Pretraining for Generalizable Embodied Agents
by: Zhang, Zhizhen, et al.
Published: (2025)
by: Zhang, Zhizhen, et al.
Published: (2025)
Ref-SAM3D: Bridging SAM3D with Text for Reference 3D Reconstruction
by: Zhou, Yun, et al.
Published: (2025)
by: Zhou, Yun, et al.
Published: (2025)
Universal Medical Image Representation Learning with Compositional Decoders
by: Wang, Kaini, et al.
Published: (2024)
by: Wang, Kaini, et al.
Published: (2024)
Gaussian Shading: Provable Performance-Lossless Image Watermarking for Diffusion Models
by: Yang, Zijin, et al.
Published: (2024)
by: Yang, Zijin, et al.
Published: (2024)
CELLO: Causal Evaluation of Large Vision-Language Models
by: Chen, Meiqi, et al.
Published: (2024)
by: Chen, Meiqi, et al.
Published: (2024)
Deep Multiview Clustering by Contrasting Cluster Assignments
by: Chen, Jie, et al.
Published: (2023)
by: Chen, Jie, et al.
Published: (2023)
On the Learnability of Out-of-distribution Detection
by: Fang, Zhen, et al.
Published: (2024)
by: Fang, Zhen, et al.
Published: (2024)
Prune Redundancy, Preserve Essence: Vision Token Compression in VLMs via Synergistic Importance-Diversity
by: Fang, Zhengyao, et al.
Published: (2026)
by: Fang, Zhengyao, et al.
Published: (2026)
NoiseDiffusion: Correcting Noise for Image Interpolation with Diffusion Models beyond Spherical Linear Interpolation
by: Zheng, PengFei, et al.
Published: (2024)
by: Zheng, PengFei, et al.
Published: (2024)
See It, Say It, Sorted: An Iterative Training-Free Framework for Visually-Grounded Multimodal Reasoning in LVLMs
by: Zhang, Yongchang, et al.
Published: (2026)
by: Zhang, Yongchang, et al.
Published: (2026)
GLAD: Generative Language-Assisted Visual Tracking for Low-Semantic Templates
by: Luo, Xingyu, et al.
Published: (2026)
by: Luo, Xingyu, et al.
Published: (2026)
Concept Regions Matter: Benchmarking CLIP with a New Cluster-Importance Approach
by: Agarwal, Aishwarya, et al.
Published: (2025)
by: Agarwal, Aishwarya, et al.
Published: (2025)
Infrared-Assisted Single-Stage Framework for Joint Restoration and Fusion of Visible and Infrared Images under Hazy Conditions
by: Li, Huafeng, et al.
Published: (2024)
by: Li, Huafeng, et al.
Published: (2024)
L-MAGIC: Language Model Assisted Generation of Images with Coherence
by: Cai, Zhipeng, et al.
Published: (2024)
by: Cai, Zhipeng, et al.
Published: (2024)
GLASS: Graph and Vision-Language Assisted Semantic Shape Correspondence
by: Xiao, Qinfeng, et al.
Published: (2026)
by: Xiao, Qinfeng, et al.
Published: (2026)
ReferSplat: Referring Segmentation in 3D Gaussian Splatting
by: He, Shuting, et al.
Published: (2025)
by: He, Shuting, et al.
Published: (2025)
Provably Uncertainty-Guided Universal Domain Adaptation
by: Wang, Yifan, et al.
Published: (2022)
by: Wang, Yifan, et al.
Published: (2022)
Out-Of-Distribution Detection with Diversification (Provably)
by: Yao, Haiyun, et al.
Published: (2024)
by: Yao, Haiyun, et al.
Published: (2024)
Decoupled Contrastive Multi-View Clustering with High-Order Random Walks
by: Lu, Yiding, et al.
Published: (2023)
by: Lu, Yiding, et al.
Published: (2023)
SAT-LDM: Provably Generalizable Image Watermarking for Latent Diffusion Models with Self-Augmented Training
by: Zhang, Lu, et al.
Published: (2024)
by: Zhang, Lu, et al.
Published: (2024)
Animate Anyone 2: High-Fidelity Character Image Animation with Environment Affordance
by: Hu, Li, et al.
Published: (2025)
by: Hu, Li, et al.
Published: (2025)
POUR: A Provably Optimal Method for Unlearning Representations via Neural Collapse
by: Le, Anjie, et al.
Published: (2025)
by: Le, Anjie, et al.
Published: (2025)
High-Precision Dichotomous Image Segmentation via Probing Diffusion Capacity
by: Yu, Qian, et al.
Published: (2024)
by: Yu, Qian, et al.
Published: (2024)
Tracking Everything in Robotic-Assisted Surgery
by: Zhan, Bohan, et al.
Published: (2024)
by: Zhan, Bohan, et al.
Published: (2024)
Dynamic Importance in Diffusion U-Net for Enhanced Image Synthesis
by: Wang, Xi, et al.
Published: (2025)
by: Wang, Xi, et al.
Published: (2025)
Through the PRISm: Importance-Aware Scene Graphs for Image Retrieval
by: Georgoulopoulos, Dimitrios, et al.
Published: (2025)
by: Georgoulopoulos, Dimitrios, et al.
Published: (2025)
Importance-Based Token Merging for Efficient Image and Video Generation
by: Wu, Haoyu, et al.
Published: (2024)
by: Wu, Haoyu, et al.
Published: (2024)
Mixing Importance with Diversity: Joint Optimization for KV Cache Compression in Large Vision-Language Models
by: Liu, Xuyang, et al.
Published: (2025)
by: Liu, Xuyang, et al.
Published: (2025)
A Causality-Inspired Model for Intima-Media Thickening Assessment in Ultrasound Videos
by: Gao, Shuo, et al.
Published: (2025)
by: Gao, Shuo, et al.
Published: (2025)
Think as Cardiac Sonographers: Marrying SAM with Left Ventricular Indicators Measurements According to Clinical Guidelines
by: Liu, Tuo, et al.
Published: (2025)
by: Liu, Tuo, et al.
Published: (2025)
Similar Items
-
Debiased Negative Mining Improves Out-of-distribution Detection with Pre-trained Vision-Language Models
by: Peng, Bo, et al.
Published: (2026) -
Delving into Spectral Clustering with Vision-Language Representations
by: Peng, Bo, et al.
Published: (2026) -
MiraGe: Multimodal Discriminative Representation Learning for Generalizable AI-Generated Image Detection
by: Shi, Kuo, et al.
Published: (2025) -
Decoder Gradient Shield: Provable and High-Fidelity Prevention of Gradient-Based Box-Free Watermark Removal
by: An, Haonan, et al.
Published: (2025) -
Efficient Adaptive Label Refinement for Label Noise Learning
by: Zhang, Wenzhen, et al.
Published: (2025)