Look Before You Decide: Prompting Active Deduction of MLLMs for Assumptive Reasoning
Fuente:
arXiv
Salvato in:
| Autori principali: | Li, Yian, Tian, Wentao, Jiao, Yang, Chen, Jingjing, Qian, Tianwen, Zhu, Bin, Zhao, Na, Jiang, Yu-Gang |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2024
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
SpatialImaginer: Towards Adaptive Visual Imagination for Spatial Reasoning
di: Li, Yian, et al.
Pubblicazione: (2026)
di: Li, Yian, et al.
Pubblicazione: (2026)
NuScenes-QA: A Multi-modal Visual Question Answering Benchmark for Autonomous Driving Scenario
di: Qian, Tianwen, et al.
Pubblicazione: (2023)
di: Qian, Tianwen, et al.
Pubblicazione: (2023)
Look Before You Leap: Problem Elaboration Prompting Improves Mathematical Reasoning in Large Language Models
di: Liao, Haoran, et al.
Pubblicazione: (2024)
di: Liao, Haoran, et al.
Pubblicazione: (2024)
Look Before You Leap: Autonomous Exploration for LLM Agents
di: Ye, Ziang, et al.
Pubblicazione: (2026)
di: Ye, Ziang, et al.
Pubblicazione: (2026)
OmniGenBench: A Benchmark for Omnipotent Multimodal Generation across 50+ Tasks
di: Wang, Jiayu, et al.
Pubblicazione: (2025)
di: Wang, Jiayu, et al.
Pubblicazione: (2025)
Domain Expansion and Boundary Growth for Open-Set Single-Source Domain Generalization
di: Jiao, Pengkun, et al.
Pubblicazione: (2024)
di: Jiao, Pengkun, et al.
Pubblicazione: (2024)
Unlocking Textual and Visual Wisdom: Open-Vocabulary 3D Object Detection Enhanced by Comprehensive Guidance from Text and Image
di: Jiao, Pengkun, et al.
Pubblicazione: (2024)
di: Jiao, Pengkun, et al.
Pubblicazione: (2024)
Before You Decide: The Decision Series — Ten Papers on Artificial Intelligence for People Who Have to Decide
di: Kyungae, Ahn
Pubblicazione: (2026)
di: Kyungae, Ahn
Pubblicazione: (2026)
Benchmarking Gaslighting Negation Attacks Against Reasoning Models
di: Zhu, Bin, et al.
Pubblicazione: (2025)
di: Zhu, Bin, et al.
Pubblicazione: (2025)
Hypothesis Testing Prompting Improves Deductive Reasoning in Large Language Models
di: Li, Yitian, et al.
Pubblicazione: (2024)
di: Li, Yitian, et al.
Pubblicazione: (2024)
From Canteen Food to Daily Meals: Generalizing Food Recognition to More Practical Scenarios
di: Liu, Guoshan, et al.
Pubblicazione: (2024)
di: Liu, Guoshan, et al.
Pubblicazione: (2024)
Enhancing Action and Ingredient Modeling for Semantically Grounded Recipe Generation
di: Liu, Guoshan, et al.
Pubblicazione: (2026)
di: Liu, Guoshan, et al.
Pubblicazione: (2026)
EAGLE: Towards Efficient Arbitrary Referring Visual Prompts Comprehension for Multimodal Large Language Models
di: Zhang, Jiacheng, et al.
Pubblicazione: (2024)
di: Zhang, Jiacheng, et al.
Pubblicazione: (2024)
SnapKV: LLM Knows What You are Looking for Before Generation
di: Li, Yuhong, et al.
Pubblicazione: (2024)
di: Li, Yuhong, et al.
Pubblicazione: (2024)
Look Before You Leap: Enhancing Attention and Vigilance Regarding Harmful Content with GuidelineLLM
di: Zhang, Shaoqing, et al.
Pubblicazione: (2024)
di: Zhang, Shaoqing, et al.
Pubblicazione: (2024)
Enhancing Multimodal In-Context Learning via Inductive-Deductive Reasoning
di: Wang, Haoyu, et al.
Pubblicazione: (2026)
di: Wang, Haoyu, et al.
Pubblicazione: (2026)
FGeo-DRL: Deductive Reasoning for Geometric Problems through Deep Reinforcement Learning
di: Zou, Jia, et al.
Pubblicazione: (2024)
di: Zou, Jia, et al.
Pubblicazione: (2024)
Omni-o3: Deep Nested Omnimodal Deduction for Deliberative Audio-Visual Reasoning
di: Zhang, Zhicheng, et al.
Pubblicazione: (2026)
di: Zhang, Zhicheng, et al.
Pubblicazione: (2026)
Look-Around Before You Leap: High-Frequency Injected Transformer for Image Restoration
di: Zhou, Shihao, et al.
Pubblicazione: (2024)
di: Zhou, Shihao, et al.
Pubblicazione: (2024)
Look Before You Fuse: 2D-Guided Cross-Modal Alignment for Robust 3D Detection
di: Li, Xiang, et al.
Pubblicazione: (2025)
di: Li, Xiang, et al.
Pubblicazione: (2025)
Look Before You Leap: An Exploratory Study of Uncertainty Measurement for Large Language Models
di: Huang, Yuheng, et al.
Pubblicazione: (2023)
di: Huang, Yuheng, et al.
Pubblicazione: (2023)
Actor-Critic for Continuous Action Chunks: A Reinforcement Learning Framework for Long-Horizon Robotic Manipulation with Sparse Reward
di: Yang, Jiarui, et al.
Pubblicazione: (2025)
di: Yang, Jiarui, et al.
Pubblicazione: (2025)
Don't Deceive Me: Mitigating Gaslighting through Attention Reallocation in LMMs
di: Jiao, Pengkun, et al.
Pubblicazione: (2025)
di: Jiao, Pengkun, et al.
Pubblicazione: (2025)
From Holistic to Localized: Local Enhanced Adapters for Efficient Visual Instruction Fine-Tuning
di: Jiao, Pengkun, et al.
Pubblicazione: (2024)
di: Jiao, Pengkun, et al.
Pubblicazione: (2024)
Expose Before You Defend: Unifying and Enhancing Backdoor Defenses via Exposed Models
di: Li, Yige, et al.
Pubblicazione: (2024)
di: Li, Yige, et al.
Pubblicazione: (2024)
Think Before You Move: Latent Motion Reasoning for Text-to-Motion Generation
di: Qian, Yijie, et al.
Pubblicazione: (2025)
di: Qian, Yijie, et al.
Pubblicazione: (2025)
Scaling Spatial Reasoning in MLLMs through Programmatic Data Synthesis
di: Helu, Zhi, et al.
Pubblicazione: (2025)
di: Helu, Zhi, et al.
Pubblicazione: (2025)
When Seeing Is not Enough: Revealing the Limits of Active Reasoning in MLLMs
di: Liu, Hongcheng, et al.
Pubblicazione: (2025)
di: Liu, Hongcheng, et al.
Pubblicazione: (2025)
ControlThinker: Unveiling Latent Semantics for Controllable Image Generation through Visual Reasoning
di: Han, Feng, et al.
Pubblicazione: (2025)
di: Han, Feng, et al.
Pubblicazione: (2025)
Prompt as Free Lunch: Enhancing Diversity in Source-Free Cross-domain Few-shot Learning through Semantic-Guided Prompting
di: Zhuo, Linhai, et al.
Pubblicazione: (2024)
di: Zhuo, Linhai, et al.
Pubblicazione: (2024)
Think Thrice Before You Speak: Dual knowledge-enhanced Theory-of-Mind Reasoning for Persuasive Agents
di: Ma, Minghui, et al.
Pubblicazione: (2026)
di: Ma, Minghui, et al.
Pubblicazione: (2026)
Spatiotemporal Sycophancy: Negation-Based Gaslighting in Video Large Language Models
di: Tang, Ziyao, et al.
Pubblicazione: (2026)
di: Tang, Ziyao, et al.
Pubblicazione: (2026)
CLiViS: Unleashing Cognitive Map through Linguistic-Visual Synergy for Embodied Visual Reasoning
di: Li, Kailing, et al.
Pubblicazione: (2025)
di: Li, Kailing, et al.
Pubblicazione: (2025)
The Role of Deductive and Inductive Reasoning in Large Language Models
di: Cai, Chengkun, et al.
Pubblicazione: (2024)
di: Cai, Chengkun, et al.
Pubblicazione: (2024)
Understanding Before Reasoning: Enhancing Chain-of-Thought with Iterative Summarization Pre-Prompting
di: Zhu, Dong-Hai, et al.
Pubblicazione: (2025)
di: Zhu, Dong-Hai, et al.
Pubblicazione: (2025)
Reasoning Portability: Guiding Continual Learning for MLLMs in the RLVR Era
di: Hong, Qiuhe, et al.
Pubblicazione: (2026)
di: Hong, Qiuhe, et al.
Pubblicazione: (2026)
Deductive Beam Search: Decoding Deducible Rationale for Chain-of-Thought Reasoning
di: Zhu, Tinghui, et al.
Pubblicazione: (2024)
di: Zhu, Tinghui, et al.
Pubblicazione: (2024)
Look, Compare, Decide: Alleviating Hallucination in Large Vision-Language Models via Multi-View Multi-Path Reasoning
di: Qu, Xiaoye, et al.
Pubblicazione: (2024)
di: Qu, Xiaoye, et al.
Pubblicazione: (2024)
RC-NF: Robot-Conditioned Normalizing Flow for Real-Time Anomaly Detection in Robotic Manipulation
di: Zhou, Shijie, et al.
Pubblicazione: (2026)
di: Zhou, Shijie, et al.
Pubblicazione: (2026)
Collaborative AI Teaming in Unknown Environments via Active Goal Deduction
di: Zhang, Zuyuan, et al.
Pubblicazione: (2024)
di: Zhang, Zuyuan, et al.
Pubblicazione: (2024)
Documenti analoghi
-
SpatialImaginer: Towards Adaptive Visual Imagination for Spatial Reasoning
di: Li, Yian, et al.
Pubblicazione: (2026) -
NuScenes-QA: A Multi-modal Visual Question Answering Benchmark for Autonomous Driving Scenario
di: Qian, Tianwen, et al.
Pubblicazione: (2023) -
Look Before You Leap: Problem Elaboration Prompting Improves Mathematical Reasoning in Large Language Models
di: Liao, Haoran, et al.
Pubblicazione: (2024) -
Look Before You Leap: Autonomous Exploration for LLM Agents
di: Ye, Ziang, et al.
Pubblicazione: (2026) -
OmniGenBench: A Benchmark for Omnipotent Multimodal Generation across 50+ Tasks
di: Wang, Jiayu, et al.
Pubblicazione: (2025)