Focus, Don't Prune: Identifying Instruction-Relevant Regions for Information-Rich Image Understanding
Fuente:
arXiv
Saved in:
| Main Authors: | Kwon, Mincheol, Lee, Minseung, Choi, Seonga, Choi, Miso, Oh, Kyeong-Jin, Lee, Hyunyoung, Park, Cheonyoung, Song, Yongho, Park, Seunghyun, Kim, Jinkyu |
|---|---|
| Format: | Preprint |
| Published: |
2026
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Watermarking for Factuality: Guiding Vision-Language Models Toward Truth via Tri-layer Contrastive Decoding
by: Back, Kyungryul, et al.
Published: (2025)
by: Back, Kyungryul, et al.
Published: (2025)
Semantic Communication Challenges: Understanding Dos and Avoiding Don'ts
by: Choi, Jinho, et al.
Published: (2024)
by: Choi, Jinho, et al.
Published: (2024)
SpatiO: Adaptive Test-Time Orchestration of Vision-Language Agents for Spatial Reasoning
by: Hwang, Chan Yeong, et al.
Published: (2026)
by: Hwang, Chan Yeong, et al.
Published: (2026)
Temporal Linear Item-Item Model for Sequential Recommendation
by: Park, Seongmin, et al.
Published: (2024)
by: Park, Seongmin, et al.
Published: (2024)
Finetuning Pre-trained Model with Limited Data for LiDAR-based 3D Object Detection by Bridging Domain Gaps
by: Jang, Jiyun, et al.
Published: (2024)
by: Jang, Jiyun, et al.
Published: (2024)
Causality-Aware End-to-End Autonomous Driving via Ego-Centric Joint Scene Modeling
by: Moon, Seokha, et al.
Published: (2026)
by: Moon, Seokha, et al.
Published: (2026)
Just Add $100 More: Augmenting NeRF-based Pseudo-LiDAR Point Cloud for Resolving Class-imbalance Problem
by: Chang, Mincheol, et al.
Published: (2024)
by: Chang, Mincheol, et al.
Published: (2024)
InstructBooth: Instruction-following Personalized Text-to-Image Generation
by: Chae, Daewon, et al.
Published: (2023)
by: Chae, Daewon, et al.
Published: (2023)
Image-Guided Semantic Pseudo-LiDAR Point Generation for 3D Object Detection
by: Lee, Minseung, et al.
Published: (2024)
by: Lee, Minseung, et al.
Published: (2024)
Deep Discrete Encoders: Identifiable Deep Generative Models for Rich Data with Discrete Latent Layers
by: Lee, Seunghyun, et al.
Published: (2025)
by: Lee, Seunghyun, et al.
Published: (2025)
Transferable Model-agnostic Vision-Language Model Adaptation for Efficient Weak-to-Strong Generalization
by: Park, Jihwan, et al.
Published: (2025)
by: Park, Jihwan, et al.
Published: (2025)
FC-TTS: Style and Timbre Control in Zero-Shot Text-to-Speech with Disentangled Speech Representations
by: Lee, Yoonhyung, et al.
Published: (2026)
by: Lee, Yoonhyung, et al.
Published: (2026)
Visually Dehallucinative Instruction Generation: Know What You Don't Know
by: Cha, Sungguk, et al.
Published: (2024)
by: Cha, Sungguk, et al.
Published: (2024)
Do's and Don'ts: Learning Desirable Skills with Instruction Videos
by: Kim, Hyunseung, et al.
Published: (2024)
by: Kim, Hyunseung, et al.
Published: (2024)
All‐Polarized Elastic Wave Attenuation and Harvesting via Chiral Mechanical Metamaterials
by: Jeonghoon Park, et al.
Published: (2024)
by: Jeonghoon Park, et al.
Published: (2024)
A theoretical study for the linear free energy relationship of CH bond activation and the role of the axial ligand in cytochrome P450 model complexes
by: Soobin Kwon, et al.
Published: (2024)
by: Soobin Kwon, et al.
Published: (2024)
Mechano‐Electrochemical Behavior of Nanostructured Li‐ and Mn‐Rich Layered Oxides with Superior Capacity Retention and Voltage Decay for Sulfide‐Based All‐Solid‐State Batteries
by: Gawon Song, et al.
Published: (2024)
by: Gawon Song, et al.
Published: (2024)
Mechano‐Electrochemical Behavior of Nanostructured Li‐ and Mn‐Rich Layered Oxides with Superior Capacity Retention and Voltage Decay for Sulfide‐Based All‐Solid‐State Batteries (Adv. Energy Mater. 47/2024)
by: Gawon Song, et al.
Published: (2024)
by: Gawon Song, et al.
Published: (2024)
REPrune: Channel Pruning via Kernel Representative Selection
by: Park, Mincheol, et al.
Published: (2024)
by: Park, Mincheol, et al.
Published: (2024)
Learning Temporal Cues by Predicting Objects Move for Multi-camera 3D Object Detection
by: Moon, Seokha, et al.
Published: (2024)
by: Moon, Seokha, et al.
Published: (2024)
Don't Use LLMs to Make Relevance Judgments
by: Soboroff, Ian
Published: (2024)
by: Soboroff, Ian
Published: (2024)
ALIGN: Advanced Query Initialization with LiDAR-Image Guidance for Occlusion-Robust 3D Object Detection
by: Baek, Janghyun, et al.
Published: (2025)
by: Baek, Janghyun, et al.
Published: (2025)
Progressive Weight Loading: Accelerating Initial Inference and Gradually Boosting Performance on Resource-Constrained Environments
by: Kim, Hyunwoo, et al.
Published: (2025)
by: Kim, Hyunwoo, et al.
Published: (2025)
SurgMLLMBench: A Multimodal Large Language Model Benchmark Dataset for Surgical Scene Understanding
by: Choi, Tae-Min, et al.
Published: (2025)
by: Choi, Tae-Min, et al.
Published: (2025)
MUFFIN: Mixture of User-Adaptive Frequency Filtering for Sequential Recommendation
by: Baek, Ilwoong, et al.
Published: (2025)
by: Baek, Ilwoong, et al.
Published: (2025)
Identifying the Source of Generation for Large Language Models
by: Park, Bumjin, et al.
Published: (2024)
by: Park, Bumjin, et al.
Published: (2024)
OmniDrop: Layer-wise Token Pruning for Omni-modal LLMs via Query-Guidance
by: Park, Yeo Jeong, et al.
Published: (2026)
by: Park, Yeo Jeong, et al.
Published: (2026)
Bayesian Function-on-Function Regression for Spatial Functional Data
by: Lee, Heesang, et al.
Published: (2024)
by: Lee, Heesang, et al.
Published: (2024)
Complete intersection hyperkähler fourfolds with respect to equivariant vector bundles over rational homogeneous varieties of Picard number one
by: Lee, Eunjeong, et al.
Published: (2021)
by: Lee, Eunjeong, et al.
Published: (2021)
Identifying Key Terms in Prompts for Relevance Evaluation with GPT Models
by: Jaekeol, Choi
Published: (2026)
by: Jaekeol, Choi
Published: (2026)
Identifying Key Terms in Prompts for Relevance Evaluation with GPT Models
by: Choi, Jaekeol
Published: (2024)
by: Choi, Jaekeol
Published: (2024)
Design of reliable technology valuation model with calibrated machine learning of patent indicators
by: Lee, Seunghyun, et al.
Published: (2024)
by: Lee, Seunghyun, et al.
Published: (2024)
Better Safe Than Sorry? Overreaction Problem of Vision Language Models in Visual Emergency Recognition
by: Choi, Dasol, et al.
Published: (2025)
by: Choi, Dasol, et al.
Published: (2025)
Enhancing Time Awareness in Generative Recommendation
by: Lee, Sunkyung, et al.
Published: (2025)
by: Lee, Sunkyung, et al.
Published: (2025)
If You Don't Understand It, Don't Use It: Eliminating Trojans with Filters Between Layers
by: Hernandez, Adriano
Published: (2024)
by: Hernandez, Adriano
Published: (2024)
Citrus sunki Peel Extract, Containing Nobiletin and Tangeretin, Enhances Proliferation and Differentiation in 3T3-L1 Adipocytes
by: Sehyuk Oh, et al.
Published: (2024)
by: Sehyuk Oh, et al.
Published: (2024)
DataFreeShield: Defending Adversarial Attacks without Training Data
by: Lee, Hyeyoon, et al.
Published: (2024)
by: Lee, Hyeyoon, et al.
Published: (2024)
Editorial: Variceal Bleeding in Patients Receiving Atezolizumab–Bevacizumab for Hepatocellular Carcinoma—Don't Ignore the Risk Factors! Authors' Reply
by: Jonggi Choi
Published: (2025)
by: Jonggi Choi
Published: (2025)
On Theoretical Identifiability of Discrete Latent Causal Graphical Models
by: Lee, Seunghyun, et al.
Published: (2025)
by: Lee, Seunghyun, et al.
Published: (2025)
Recent Progress on Quantum Dot Patterning Technologies for Commercialization of QD‐LEDs: Current Status, Future Prospects, and Exploratory Approaches
by: Jaeyeop Lee, et al.
Published: (2024)
by: Jaeyeop Lee, et al.
Published: (2024)
Similar Items
-
Watermarking for Factuality: Guiding Vision-Language Models Toward Truth via Tri-layer Contrastive Decoding
by: Back, Kyungryul, et al.
Published: (2025) -
Semantic Communication Challenges: Understanding Dos and Avoiding Don'ts
by: Choi, Jinho, et al.
Published: (2024) -
SpatiO: Adaptive Test-Time Orchestration of Vision-Language Agents for Spatial Reasoning
by: Hwang, Chan Yeong, et al.
Published: (2026) -
Temporal Linear Item-Item Model for Sequential Recommendation
by: Park, Seongmin, et al.
Published: (2024) -
Finetuning Pre-trained Model with Limited Data for LiDAR-based 3D Object Detection by Bridging Domain Gaps
by: Jang, Jiyun, et al.
Published: (2024)