Salvato in:
| Autori principali: | Li, Yang, Chen, Xing, Liu, Yutao, Qi, Gege, BI, Yanxian, Wang, Zizhe, Zhang, Yunjian, Zhu, Yao |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2026
|
| Soggetti: | |
| Accesso online: | https://arxiv.org/abs/2603.09337 |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
Towards Understanding How Knowledge Evolves in Large Vision-Language Models
di: Wang, Sudong, et al.
Pubblicazione: (2025)
di: Wang, Sudong, et al.
Pubblicazione: (2025)
SpatialBench: Benchmarking Multimodal Large Language Models for Spatial Cognition
di: Xu, Peiran, et al.
Pubblicazione: (2025)
di: Xu, Peiran, et al.
Pubblicazione: (2025)
RS-OOD: A Vision-Language Augmented Framework for Out-of-Distribution Detection in Remote Sensing
di: Wang, Chenhao, et al.
Pubblicazione: (2025)
di: Wang, Chenhao, et al.
Pubblicazione: (2025)
SOPSeg: Prompt-based Small Object Instance Segmentation in Remote Sensing Imagery
di: Wang, Chenhao, et al.
Pubblicazione: (2025)
di: Wang, Chenhao, et al.
Pubblicazione: (2025)
An Efficient Framework for Enhancing Discriminative Models via Diffusion Techniques
di: Li, Chunxiao, et al.
Pubblicazione: (2024)
di: Li, Chunxiao, et al.
Pubblicazione: (2024)
CAPO: Reinforcing Consistent Reasoning in Medical Decision-Making
di: Jiang, Songtao, et al.
Pubblicazione: (2025)
di: Jiang, Songtao, et al.
Pubblicazione: (2025)
Do LLMs Understand Visual Anomalies? Uncovering LLM's Capabilities in Zero-shot Anomaly Detection
di: Zhu, Jiaqi, et al.
Pubblicazione: (2024)
di: Zhu, Jiaqi, et al.
Pubblicazione: (2024)
Think 360°: Evaluating the Width-centric Reasoning Capability of MLLMs Beyond Depth
di: Chen, Mingrui, et al.
Pubblicazione: (2026)
di: Chen, Mingrui, et al.
Pubblicazione: (2026)
Noise Diffusion for Enhancing Semantic Faithfulness in Text-to-Image Synthesis
di: Miao, Boming, et al.
Pubblicazione: (2024)
di: Miao, Boming, et al.
Pubblicazione: (2024)
Exploring Decision-Making Capabilities of LLM Agents: An Experimental Study on Jump-Jump Game
di: Li, Juwu
Pubblicazione: (2025)
di: Li, Juwu
Pubblicazione: (2025)
Sliding-Window Merging for Compacting Patch-Redundant Layers in LLMs
di: Ding, Xuan, et al.
Pubblicazione: (2025)
di: Ding, Xuan, et al.
Pubblicazione: (2025)
Angle of Arrival Estimation with Transformer: A Sparse and Gridless Method with Zero-Shot Capability
di: Zhu, Zhaoxuan, et al.
Pubblicazione: (2024)
di: Zhu, Zhaoxuan, et al.
Pubblicazione: (2024)
AdvLogo: Adversarial Patch Attack against Object Detectors based on Diffusion Models
di: Miao, Boming, et al.
Pubblicazione: (2024)
di: Miao, Boming, et al.
Pubblicazione: (2024)
HAZARD Challenge: Embodied Decision Making in Dynamically Changing Environments
di: Zhou, Qinhong, et al.
Pubblicazione: (2024)
di: Zhou, Qinhong, et al.
Pubblicazione: (2024)
Strat-Reasoner: Reinforcing Strategic Reasoning of LLMs in Multi-Agent Games
di: He, Yidong, et al.
Pubblicazione: (2026)
di: He, Yidong, et al.
Pubblicazione: (2026)
TOMATO: Assessing Visual Temporal Reasoning Capabilities in Multimodal Foundation Models
di: Shangguan, Ziyao, et al.
Pubblicazione: (2024)
di: Shangguan, Ziyao, et al.
Pubblicazione: (2024)
Unlocking Multilingual Reasoning Capability of LLMs and LVLMs through Representation Engineering
di: Li, Qiming, et al.
Pubblicazione: (2025)
di: Li, Qiming, et al.
Pubblicazione: (2025)
D2-Mamba: Dual-Scale Fusion and Dual-Path Scanning with SSMs for Shadow Removal
di: Li, Linhao, et al.
Pubblicazione: (2025)
di: Li, Linhao, et al.
Pubblicazione: (2025)
Making Large Language Models Better Planners with Reasoning-Decision Alignment
di: Huang, Zhijian, et al.
Pubblicazione: (2024)
di: Huang, Zhijian, et al.
Pubblicazione: (2024)
Unsupervised Domain Adaptive Lane Detection via Contextual Contrast and Aggregation
di: Zhou, Kunyang, et al.
Pubblicazione: (2024)
di: Zhou, Kunyang, et al.
Pubblicazione: (2024)
Beyond Specialization: Assessing the Capabilities of MLLMs in Age and Gender Estimation
di: Kuprashevich, Maksim, et al.
Pubblicazione: (2024)
di: Kuprashevich, Maksim, et al.
Pubblicazione: (2024)
UCAgents: Unidirectional Convergence for Visual Evidence Anchored Multi-Agent Medical Decision-Making
di: Feng, Qianhan, et al.
Pubblicazione: (2025)
di: Feng, Qianhan, et al.
Pubblicazione: (2025)
Open-Nav: Exploring Zero-Shot Vision-and-Language Navigation in Continuous Environment with Open-Source LLMs
di: Qiao, Yanyuan, et al.
Pubblicazione: (2024)
di: Qiao, Yanyuan, et al.
Pubblicazione: (2024)
RemoteZero: Geospatial Reasoning with Zero Human Annotations
di: Yao, Liang, et al.
Pubblicazione: (2026)
di: Yao, Liang, et al.
Pubblicazione: (2026)
Chain-of-Thought Degrades Visual Spatial Reasoning Capabilities of Multimodal LLMs
di: Kancheti, Sai Srinivas, et al.
Pubblicazione: (2026)
di: Kancheti, Sai Srinivas, et al.
Pubblicazione: (2026)
V-Zero: Self-Improving Multimodal Reasoning with Zero Annotation
di: Wang, Han, et al.
Pubblicazione: (2026)
di: Wang, Han, et al.
Pubblicazione: (2026)
DoubleTake: Contrastive Reasoning for Faithful Decision-Making in Medical Imaging
di: Patel, Daivik, et al.
Pubblicazione: (2026)
di: Patel, Daivik, et al.
Pubblicazione: (2026)
Decoding Decision Reasoning: A Counterfactual-Powered Model for Knowledge Discovery
di: Fang, Yingying, et al.
Pubblicazione: (2024)
di: Fang, Yingying, et al.
Pubblicazione: (2024)
AGIR: Assessing 3D Gait Impairment with Reasoning based on LLMs
di: Wang, Diwei, et al.
Pubblicazione: (2025)
di: Wang, Diwei, et al.
Pubblicazione: (2025)
ShredBench: Evaluating the Semantic Reasoning Capabilities of Multimodal LLMs in Document Reconstruction
di: Guo, Zichun, et al.
Pubblicazione: (2026)
di: Guo, Zichun, et al.
Pubblicazione: (2026)
Quantum Conflict Measurement in Decision Making for Out-of-Distribution Detection
di: Dong, Yilin, et al.
Pubblicazione: (2025)
di: Dong, Yilin, et al.
Pubblicazione: (2025)
Bridging the Gap Between Ideal and Real-world Evaluation: Benchmarking AI-Generated Image Detection in Challenging Scenarios
di: Li, Chunxiao, et al.
Pubblicazione: (2025)
di: Li, Chunxiao, et al.
Pubblicazione: (2025)
CEOs, Information, and Decision Making: Scanning the Environment for Strategic Advantage.
di: Auster, Ethel, et al.
Pubblicazione: (1994)
di: Auster, Ethel, et al.
Pubblicazione: (1994)
Mitigating Hallucinations in Large Vision-Language Models via DPO: On-Policy Data Hold the Key
di: Yang, Zhihe, et al.
Pubblicazione: (2025)
di: Yang, Zhihe, et al.
Pubblicazione: (2025)
SGDM: Static-Guided Dynamic Module Make Stronger Visual Models
di: Xing, Wenjie, et al.
Pubblicazione: (2024)
di: Xing, Wenjie, et al.
Pubblicazione: (2024)
RobuSTereo: Robust Zero-Shot Stereo Matching under Adverse Weather
di: Wang, Yuran, et al.
Pubblicazione: (2025)
di: Wang, Yuran, et al.
Pubblicazione: (2025)
Prism: A Framework for Decoupling and Assessing the Capabilities of VLMs
di: Qiao, Yuxuan, et al.
Pubblicazione: (2024)
di: Qiao, Yuxuan, et al.
Pubblicazione: (2024)
Object Navigation with Structure-Semantic Reasoning-Based Multi-level Map and Multimodal Decision-Making LLM
di: Yan, Chongshang, et al.
Pubblicazione: (2025)
di: Yan, Chongshang, et al.
Pubblicazione: (2025)
Multi-Agent Reinforcement Learning and Real-Time Decision-Making in Robotic Soccer for Virtual Environments
di: Taourirte, Aya, et al.
Pubblicazione: (2025)
di: Taourirte, Aya, et al.
Pubblicazione: (2025)
Video-MSR: Benchmarking Multi-hop Spatial Reasoning Capabilities of MLLMs
di: Zhu, Rui, et al.
Pubblicazione: (2026)
di: Zhu, Rui, et al.
Pubblicazione: (2026)
Documenti analoghi
-
Towards Understanding How Knowledge Evolves in Large Vision-Language Models
di: Wang, Sudong, et al.
Pubblicazione: (2025) -
SpatialBench: Benchmarking Multimodal Large Language Models for Spatial Cognition
di: Xu, Peiran, et al.
Pubblicazione: (2025) -
RS-OOD: A Vision-Language Augmented Framework for Out-of-Distribution Detection in Remote Sensing
di: Wang, Chenhao, et al.
Pubblicazione: (2025) -
SOPSeg: Prompt-based Small Object Instance Segmentation in Remote Sensing Imagery
di: Wang, Chenhao, et al.
Pubblicazione: (2025) -
An Efficient Framework for Enhancing Discriminative Models via Diffusion Techniques
di: Li, Chunxiao, et al.
Pubblicazione: (2024)