Saved in:
| Main Authors: | Liang, Yuxuan, Li, Xu, Chen, Xiaolei, Chen, Haotian, Zheng, Yi, Lai, Chenghang, Li, Bin, Xue, Xiangyang |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | https://arxiv.org/abs/2501.14276 |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Instruction-Guided Fusion of Multi-Layer Visual Features in Large Vision-Language Models
by: Li, Xu, et al.
Published: (2024)
by: Li, Xu, et al.
Published: (2024)
Pyramid Token Pruning for High-Resolution Large Vision-Language Models via Region, Token, and Instruction-Guided Importance
by: Liang, Yuxuan, et al.
Published: (2025)
by: Liang, Yuxuan, et al.
Published: (2025)
HERO: Rethinking Visual Token Early Dropping in High-Resolution Large Vision-Language Models
by: Li, Xu, et al.
Published: (2025)
by: Li, Xu, et al.
Published: (2025)
ResPrune: Text-Conditioned Subspace Reconstruction for Visual Token Pruning in Large Vision-Language Models
by: Li, Xu, et al.
Published: (2026)
by: Li, Xu, et al.
Published: (2026)
When Large Vision-Language Models Meet Person Re-Identification
by: Wang, Qizao, et al.
Published: (2024)
by: Wang, Qizao, et al.
Published: (2024)
FedRA: A Random Allocation Strategy for Federated Tuning to Unleash the Power of Heterogeneous Clients
by: Su, Shangchao, et al.
Published: (2023)
by: Su, Shangchao, et al.
Published: (2023)
HLGFA: High-Low Resolution Guided Feature Alignment for Unsupervised Anomaly Detection
by: Zhou, Han, et al.
Published: (2026)
by: Zhou, Han, et al.
Published: (2026)
Distribution Aligned Semantics Adaption for Lifelong Person Re-Identification
by: Wang, Qizao, et al.
Published: (2024)
by: Wang, Qizao, et al.
Published: (2024)
Weakly Supervised Gaussian Contrastive Grounding with Large Multimodal Models for Video Question Answering
by: Wang, Haibo, et al.
Published: (2024)
by: Wang, Haibo, et al.
Published: (2024)
Combination therapy for colorectal cancer with anti-PD-L1 and cancer vaccine: A multiscale mathematical model of tumor-immune interactions
by: Li, Chenghang, et al.
Published: (2026)
by: Li, Chenghang, et al.
Published: (2026)
Mathematical modeling of tumor-immune interactions: methods, applications, and future perspectives
by: Li, Chenghang, et al.
Published: (2025)
by: Li, Chenghang, et al.
Published: (2025)
Content and Salient Semantics Collaboration for Cloth-Changing Person Re-Identification
by: Wang, Qizao, et al.
Published: (2024)
by: Wang, Qizao, et al.
Published: (2024)
One-Shot Heterogeneous Federated Learning with Local Model-Guided Diffusion Models
by: Yang, Mingzhao, et al.
Published: (2023)
by: Yang, Mingzhao, et al.
Published: (2023)
Learning Global Object-Centric Representations via Disentangled Slot Attention
by: Chen, Tonglin, et al.
Published: (2024)
by: Chen, Tonglin, et al.
Published: (2024)
Vision-Language Feature Alignment for Road Anomaly Segmentation
by: He, Zhuolin, et al.
Published: (2026)
by: He, Zhuolin, et al.
Published: (2026)
ReasonGrounder: LVLM-Guided Hierarchical Feature Splatting for Open-Vocabulary 3D Visual Grounding and Reasoning
by: Liu, Zhenyang, et al.
Published: (2025)
by: Liu, Zhenyang, et al.
Published: (2025)
Towards Generative Abstract Reasoning: Completing Raven's Progressive Matrix via Rule Abstraction and Selection
by: Shi, Fan, et al.
Published: (2024)
by: Shi, Fan, et al.
Published: (2024)
Beyond Task-Specific Reasoning: A Unified Conditional Generative Framework for Abstract Visual Reasoning
by: Shi, Fan, et al.
Published: (2025)
by: Shi, Fan, et al.
Published: (2025)
Towards a Vision-Language Episodic Memory Framework: Large-scale Pretrained Model-Augmented Hippocampal Attractor Dynamics
by: Li, Chong, et al.
Published: (2025)
by: Li, Chong, et al.
Published: (2025)
Modeling tumor cell heterogeneity and plasticity in adaptive therapy
by: Yue, Rui, et al.
Published: (2026)
by: Yue, Rui, et al.
Published: (2026)
Quantitative cancer-immunity cycle modeling to optimize bevacizumab and atezolizumab combination therapy for advanced renal cell carcinoma
by: Du, Lei, et al.
Published: (2026)
by: Du, Lei, et al.
Published: (2026)
Automated Label Unification for Multi-Dataset Semantic Segmentation with GNNs
by: Ma, Rong, et al.
Published: (2024)
by: Ma, Rong, et al.
Published: (2024)
Applying Unsupervised Semantic Segmentation to High-Resolution UAV Imagery for Enhanced Road Scene Parsing
by: Ma, Zihan, et al.
Published: (2024)
by: Ma, Zihan, et al.
Published: (2024)
Global Compression Commander: Plug-and-Play Inference Acceleration for High-Resolution Large Vision-Language Models
by: Liu, Xuyang, et al.
Published: (2025)
by: Liu, Xuyang, et al.
Published: (2025)
Sub‐Nanogram Resolution Measurement of Inertial Mass and Density Using Magnetic‐Field‐Guided Bubble Microthruster
by: Leilei Wang, et al.
Published: (2024)
by: Leilei Wang, et al.
Published: (2024)
DA-VPT: Semantic-Guided Visual Prompt Tuning for Vision Transformers
by: Ren, Li, et al.
Published: (2025)
by: Ren, Li, et al.
Published: (2025)
Unsupervised Object-Centric Learning from Multiple Unspecified Viewpoints
by: Yuan, Jinyang, et al.
Published: (2024)
by: Yuan, Jinyang, et al.
Published: (2024)
Similarity-Guided Layer-Adaptive Vision Transformer for UAV Tracking
by: Xue, Chaocan, et al.
Published: (2025)
by: Xue, Chaocan, et al.
Published: (2025)
VEGAS: Mitigating Hallucinations in Large Vision-Language Models via Vision-Encoder Attention Guided Adaptive Steering
by: Wang, Zihu, et al.
Published: (2025)
by: Wang, Zihu, et al.
Published: (2025)
CrowdTrack: A Benchmark for Difficult Multiple Pedestrian Tracking in Real Scenarios
by: Fu, Teng, et al.
Published: (2025)
by: Fu, Teng, et al.
Published: (2025)
SemanticHuman-HD: High-Resolution Semantic Disentangled 3D Human Generation
by: Zheng, Peng, et al.
Published: (2024)
by: Zheng, Peng, et al.
Published: (2024)
FuXi-Ocean: A Global Ocean Forecasting System with Sub-Daily Resolution
by: Huang, Qiusheng, et al.
Published: (2025)
by: Huang, Qiusheng, et al.
Published: (2025)
Sparse Refinement for Efficient High-Resolution Semantic Segmentation
by: Liu, Zhijian, et al.
Published: (2024)
by: Liu, Zhijian, et al.
Published: (2024)
Make a Strong Teacher with Label Assistance: A Novel Knowledge Distillation Approach for Semantic Segmentation
by: Qiu, Shoumeng, et al.
Published: (2024)
by: Qiu, Shoumeng, et al.
Published: (2024)
TSGaussian: Semantic and Depth-Guided Target-Specific Gaussian Splatting from Sparse Views
by: Zhao, Liang, et al.
Published: (2024)
by: Zhao, Liang, et al.
Published: (2024)
A Global-Local Cross-Attention Network for Ultra-high Resolution Remote Sensing Image Semantic Segmentation
by: Yi, Chen, et al.
Published: (2025)
by: Yi, Chen, et al.
Published: (2025)
When Large Vision-Language Model Meets Large Remote Sensing Imagery: Coarse-to-Fine Text-Guided Token Pruning
by: Luo, Junwei, et al.
Published: (2025)
by: Luo, Junwei, et al.
Published: (2025)
FlexAttention for Efficient High-Resolution Vision-Language Models
by: Li, Junyan, et al.
Published: (2024)
by: Li, Junyan, et al.
Published: (2024)
Improving Viewpoint-Independent Object-Centric Representations through Active Viewpoint Selection
by: Huang, Yinxuan, et al.
Published: (2024)
by: Huang, Yinxuan, et al.
Published: (2024)
Unleashing the Potential of Tracklets for Unsupervised Video Person Re-Identification
by: Meng, Nanxing, et al.
Published: (2024)
by: Meng, Nanxing, et al.
Published: (2024)
Similar Items
-
Instruction-Guided Fusion of Multi-Layer Visual Features in Large Vision-Language Models
by: Li, Xu, et al.
Published: (2024) -
Pyramid Token Pruning for High-Resolution Large Vision-Language Models via Region, Token, and Instruction-Guided Importance
by: Liang, Yuxuan, et al.
Published: (2025) -
HERO: Rethinking Visual Token Early Dropping in High-Resolution Large Vision-Language Models
by: Li, Xu, et al.
Published: (2025) -
ResPrune: Text-Conditioned Subspace Reconstruction for Visual Token Pruning in Large Vision-Language Models
by: Li, Xu, et al.
Published: (2026) -
When Large Vision-Language Models Meet Person Re-Identification
by: Wang, Qizao, et al.
Published: (2024)