Hyperbolic and Evidence-Prioritized Experts for Large Vision-Language Models
Fuente:
arXiv
Salvato in:
| Autori principali: | Zhou, Zijie, Zhu, Dandan, Wang, Hangxiangpan, Zhang, Heng, Jiao, Huishen, Zhao, Yi |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2026
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
CARPE: Context-Aware Image Representation Prioritization via Ensemble for Large Vision-Language Models
di: Lee, Donghee, et al.
Pubblicazione: (2026)
di: Lee, Donghee, et al.
Pubblicazione: (2026)
Assessing Color Vision Test in Large Vision-language Models
di: Ye, Hongfei, et al.
Pubblicazione: (2025)
di: Ye, Hongfei, et al.
Pubblicazione: (2025)
From Captions to Rewards (CAREVL): Leveraging Large Language Model Experts for Enhanced Reward Modeling in Large Vision-Language Models
di: Dai, Muzhi, et al.
Pubblicazione: (2025)
di: Dai, Muzhi, et al.
Pubblicazione: (2025)
OODBench: Out-of-Distribution Benchmark for Large Vision-Language Models
di: Lin, Ling, et al.
Pubblicazione: (2026)
di: Lin, Ling, et al.
Pubblicazione: (2026)
Quant Experts: Token-aware Adaptive Error Reconstruction with Mixture of Experts for Large Vision-Language Models Quantization
di: Jia, Chenwei, et al.
Pubblicazione: (2026)
di: Jia, Chenwei, et al.
Pubblicazione: (2026)
Rethinking Token Reduction for Large Vision-Language Models
di: Wang, Yi, et al.
Pubblicazione: (2026)
di: Wang, Yi, et al.
Pubblicazione: (2026)
Mitigating Entangled Steering in Large Vision-Language Models for Hallucination Reduction
di: Zhang, Yuanhong, et al.
Pubblicazione: (2026)
di: Zhang, Yuanhong, et al.
Pubblicazione: (2026)
Medical Large Vision Language Models with Multi-Image Visual Ability
di: Yang, Xikai, et al.
Pubblicazione: (2025)
di: Yang, Xikai, et al.
Pubblicazione: (2025)
Large Vision-Language Models Get Lost in Attention
di: Xi, Gongli, et al.
Pubblicazione: (2026)
di: Xi, Gongli, et al.
Pubblicazione: (2026)
AutoBench-V: Can Large Vision-Language Models Benchmark Themselves?
di: Bao, Han, et al.
Pubblicazione: (2024)
di: Bao, Han, et al.
Pubblicazione: (2024)
Vision-R1: Evolving Human-Free Alignment in Large Vision-Language Models via Vision-Guided Reinforcement Learning
di: Zhan, Yufei, et al.
Pubblicazione: (2025)
di: Zhan, Yufei, et al.
Pubblicazione: (2025)
Distilled Large Language Model-Driven Dynamic Sparse Expert Activation Mechanism
di: Chen, Qinghui, et al.
Pubblicazione: (2026)
di: Chen, Qinghui, et al.
Pubblicazione: (2026)
IOSVLM: A 3D Vision-Language Model for Unified Dental Diagnosis from Intraoral Scans
di: Xiong, Huimin, et al.
Pubblicazione: (2026)
di: Xiong, Huimin, et al.
Pubblicazione: (2026)
Swarm Intelligence in Geo-Localization: A Multi-Agent Large Vision-Language Model Collaborative Framework
di: Han, Xiao, et al.
Pubblicazione: (2024)
di: Han, Xiao, et al.
Pubblicazione: (2024)
Self-Introspective Decoding: Alleviating Hallucinations for Large Vision-Language Models
di: Huo, Fushuo, et al.
Pubblicazione: (2024)
di: Huo, Fushuo, et al.
Pubblicazione: (2024)
Breaking Modality Heterogeneity in Low-Bit Quantization for Large Vision-Language Models
di: Zhong, Yi, et al.
Pubblicazione: (2026)
di: Zhong, Yi, et al.
Pubblicazione: (2026)
A Structured Review of Underwater Object Detection Challenges and Solutions: From Traditional to Large Vision Language Models
di: Nabahirwa, Edwine, et al.
Pubblicazione: (2025)
di: Nabahirwa, Edwine, et al.
Pubblicazione: (2025)
Multimodal Causal Reasoning Benchmark: Challenging Vision Large Language Models to Discern Causal Links Across Modalities
di: Li, Zhiyuan, et al.
Pubblicazione: (2024)
di: Li, Zhiyuan, et al.
Pubblicazione: (2024)
RoDE: Linear Rectified Mixture of Diverse Experts for Food Large Multi-Modal Models
di: Jiao, Pengkun, et al.
Pubblicazione: (2024)
di: Jiao, Pengkun, et al.
Pubblicazione: (2024)
Compositional Entailment Learning for Hyperbolic Vision-Language Models
di: Pal, Avik, et al.
Pubblicazione: (2024)
di: Pal, Avik, et al.
Pubblicazione: (2024)
DL$^3$M: A Vision-to-Language Framework for Expert-Level Medical Reasoning through Deep Learning and Large Language Models
di: Hasan, Md. Najib, et al.
Pubblicazione: (2025)
di: Hasan, Md. Najib, et al.
Pubblicazione: (2025)
INTER: Mitigating Hallucination in Large Vision-Language Models by Interaction Guidance Sampling
di: Dong, Xin, et al.
Pubblicazione: (2025)
di: Dong, Xin, et al.
Pubblicazione: (2025)
Fine-Grained Post-Training Quantization for Large Vision Language Models with Quantization-Aware Integrated Gradients
di: Xiang, Ziwei, et al.
Pubblicazione: (2026)
di: Xiang, Ziwei, et al.
Pubblicazione: (2026)
Hyperbolic Safety-Aware Vision-Language Models
di: Poppi, Tobia, et al.
Pubblicazione: (2025)
di: Poppi, Tobia, et al.
Pubblicazione: (2025)
Uncertainty-guided Compositional Alignment with Part-to-Whole Semantic Representativeness in Hyperbolic Vision-Language Models
di: Kim, Hayeon, et al.
Pubblicazione: (2026)
di: Kim, Hayeon, et al.
Pubblicazione: (2026)
AgroGPT: Efficient Agricultural Vision-Language Model with Expert Tuning
di: Awais, Muhammad, et al.
Pubblicazione: (2024)
di: Awais, Muhammad, et al.
Pubblicazione: (2024)
Probing Perceptual Constancy in Large Vision-Language Models
di: Sun, Haoran, et al.
Pubblicazione: (2025)
di: Sun, Haoran, et al.
Pubblicazione: (2025)
An archaeological Catalog Collection Method Based on Large Vision-Language Models
di: Pang, Honglin, et al.
Pubblicazione: (2024)
di: Pang, Honglin, et al.
Pubblicazione: (2024)
White-box Multimodal Jailbreaks Against Large Vision-Language Models
di: Wang, Ruofan, et al.
Pubblicazione: (2024)
di: Wang, Ruofan, et al.
Pubblicazione: (2024)
GEASS: Gated Evidence-Adaptive Selective Caption Trust for Vision-Language Models
di: Li, Zeshang, et al.
Pubblicazione: (2026)
di: Li, Zeshang, et al.
Pubblicazione: (2026)
EchoBench: Benchmarking Sycophancy in Medical Large Vision-Language Models
di: Yuan, Botai, et al.
Pubblicazione: (2025)
di: Yuan, Botai, et al.
Pubblicazione: (2025)
Harnessing Large Vision and Language Models in Agriculture: A Review
di: Zhu, Hongyan, et al.
Pubblicazione: (2024)
di: Zhu, Hongyan, et al.
Pubblicazione: (2024)
LL-ICM: Image Compression for Low-level Machine Vision via Large Vision-Language Model
di: Xue, Yuan, et al.
Pubblicazione: (2024)
di: Xue, Yuan, et al.
Pubblicazione: (2024)
Attention Prompting on Image for Large Vision-Language Models
di: Yu, Runpeng, et al.
Pubblicazione: (2024)
di: Yu, Runpeng, et al.
Pubblicazione: (2024)
Select Less, Reason More: Prioritizing Evidence Purity for Video Reasoning
di: Li, Xuchen, et al.
Pubblicazione: (2025)
di: Li, Xuchen, et al.
Pubblicazione: (2025)
Enhancing Advanced Visual Reasoning Ability of Large Language Models
di: Li, Zhiyuan, et al.
Pubblicazione: (2024)
di: Li, Zhiyuan, et al.
Pubblicazione: (2024)
ZipVL: Efficient Large Vision-Language Models with Dynamic Token Sparsification
di: He, Yefei, et al.
Pubblicazione: (2024)
di: He, Yefei, et al.
Pubblicazione: (2024)
GeoRSMLLM: A Multimodal Large Language Model for Vision-Language Tasks in Geoscience and Remote Sensing
di: Zhang, Zilun, et al.
Pubblicazione: (2025)
di: Zhang, Zilun, et al.
Pubblicazione: (2025)
XHand: Real-time Expressive Hand Avatar
di: Gan, Qijun, et al.
Pubblicazione: (2024)
di: Gan, Qijun, et al.
Pubblicazione: (2024)
HarmoCLIP: Harmonizing Global and Regional Representations in Contrastive Vision-Language Models
di: Zeng, Haoxi, et al.
Pubblicazione: (2025)
di: Zeng, Haoxi, et al.
Pubblicazione: (2025)
Documenti analoghi
-
CARPE: Context-Aware Image Representation Prioritization via Ensemble for Large Vision-Language Models
di: Lee, Donghee, et al.
Pubblicazione: (2026) -
Assessing Color Vision Test in Large Vision-language Models
di: Ye, Hongfei, et al.
Pubblicazione: (2025) -
From Captions to Rewards (CAREVL): Leveraging Large Language Model Experts for Enhanced Reward Modeling in Large Vision-Language Models
di: Dai, Muzhi, et al.
Pubblicazione: (2025) -
OODBench: Out-of-Distribution Benchmark for Large Vision-Language Models
di: Lin, Ling, et al.
Pubblicazione: (2026) -
Quant Experts: Token-aware Adaptive Error Reconstruction with Mixture of Experts for Large Vision-Language Models Quantization
di: Jia, Chenwei, et al.
Pubblicazione: (2026)