Learning Concept-Based Causal Transition and Symbolic Reasoning for Visual Planning
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Qian, Yilue, Yu, Peiyu, Wu, Ying Nian, Su, Yao, Wang, Wei, Fan, Lifeng |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2023
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Object-Conditioned Energy-Based Attention Map Alignment in Text-to-Image Diffusion Models
von: Zhang, Yasi, et al.
Veröffentlicht: (2024)
von: Zhang, Yasi, et al.
Veröffentlicht: (2024)
MedVR: Annotation-Free Medical Visual Reasoning via Agentic Reinforcement Learning
von: Jiang, Zheng, et al.
Veröffentlicht: (2026)
von: Jiang, Zheng, et al.
Veröffentlicht: (2026)
R^3-VQA: "Read the Room" by Video Social Reasoning
von: Niu, Lixing, et al.
Veröffentlicht: (2025)
von: Niu, Lixing, et al.
Veröffentlicht: (2025)
Multi-Grained Compositional Visual Clue Learning for Image Intent Recognition
von: Tang, Yin, et al.
Veröffentlicht: (2025)
von: Tang, Yin, et al.
Veröffentlicht: (2025)
Weakly Supervised Concept Learning for Object-centric Visual Reasoning
von: Tiwari, Sparsh, et al.
Veröffentlicht: (2026)
von: Tiwari, Sparsh, et al.
Veröffentlicht: (2026)
A Multi-Modal Neuro-Symbolic Approach for Spatial Reasoning-Based Visual Grounding in Robotics
von: Jahangard, Simindokht, et al.
Veröffentlicht: (2025)
von: Jahangard, Simindokht, et al.
Veröffentlicht: (2025)
Visual Fourier Prompt Tuning
von: Zeng, Runjia, et al.
Veröffentlicht: (2024)
von: Zeng, Runjia, et al.
Veröffentlicht: (2024)
AdaCorrection: Adaptive Offset Cache Correction for Accurate Diffusion Transformers
von: Liu, Dong, et al.
Veröffentlicht: (2026)
von: Liu, Dong, et al.
Veröffentlicht: (2026)
Neuro-Symbolic Concepts
von: Mao, Jiayuan, et al.
Veröffentlicht: (2025)
von: Mao, Jiayuan, et al.
Veröffentlicht: (2025)
Mixture-of-Visual-Thoughts: Exploring Context-Adaptive Reasoning Mode Selection for General Visual Reasoning
von: Li, Zejun, et al.
Veröffentlicht: (2025)
von: Li, Zejun, et al.
Veröffentlicht: (2025)
Generating by Understanding: Neural Visual Generation with Logical Symbol Groundings
von: Peng, Yifei, et al.
Veröffentlicht: (2023)
von: Peng, Yifei, et al.
Veröffentlicht: (2023)
EmambaIR: Efficient Visual State Space Model for Event-guided Image Reconstruction
von: Yu, Wei, et al.
Veröffentlicht: (2026)
von: Yu, Wei, et al.
Veröffentlicht: (2026)
Symbolic Grounding Reveals Representational Bottlenecks in Abstract Visual Reasoning
von: Vaishnav, Mohit, et al.
Veröffentlicht: (2026)
von: Vaishnav, Mohit, et al.
Veröffentlicht: (2026)
Analogical Reasoning as a Doctor: A Foundation Model for Gastrointestinal Endoscopy Diagnosis
von: Peng, Peixi, et al.
Veröffentlicht: (2026)
von: Peng, Peixi, et al.
Veröffentlicht: (2026)
VisualPredicator: Learning Abstract World Models with Neuro-Symbolic Predicates for Robot Planning
von: Liang, Yichao, et al.
Veröffentlicht: (2024)
von: Liang, Yichao, et al.
Veröffentlicht: (2024)
Designing Instance-Level Sampling Schedules via REINFORCE with James-Stein Shrinkage
von: Yu, Peiyu, et al.
Veröffentlicht: (2025)
von: Yu, Peiyu, et al.
Veröffentlicht: (2025)
Bridging Neural and Symbolic Representations with Transitional Dictionary Learning
von: Cheng, Junyan, et al.
Veröffentlicht: (2023)
von: Cheng, Junyan, et al.
Veröffentlicht: (2023)
SATORI-R1: Incentivizing Multimodal Reasoning through Explicit Visual Anchoring
von: Shen, Chuming, et al.
Veröffentlicht: (2025)
von: Shen, Chuming, et al.
Veröffentlicht: (2025)
Blind Spot Navigation: Evolutionary Discovery of Sensitive Semantic Concepts for LVLMs
von: Pan, Zihao, et al.
Veröffentlicht: (2025)
von: Pan, Zihao, et al.
Veröffentlicht: (2025)
MILR: Improving Multimodal Image Generation via Test-Time Latent Reasoning
von: Mi, Yapeng, et al.
Veröffentlicht: (2025)
von: Mi, Yapeng, et al.
Veröffentlicht: (2025)
FastV-RAG: Towards Fast and Fine-Grained Video QA with Retrieval-Augmented Generation
von: Li, Gen, et al.
Veröffentlicht: (2026)
von: Li, Gen, et al.
Veröffentlicht: (2026)
Robust-R1: Degradation-Aware Reasoning for Robust Visual Understanding
von: Tang, Jiaqi, et al.
Veröffentlicht: (2025)
von: Tang, Jiaqi, et al.
Veröffentlicht: (2025)
Beyond Task-Specific Reasoning: A Unified Conditional Generative Framework for Abstract Visual Reasoning
von: Shi, Fan, et al.
Veröffentlicht: (2025)
von: Shi, Fan, et al.
Veröffentlicht: (2025)
ConceptSeg-R1: Segment Any Concept via Meta-Reinforcement Learning
von: Zhao, Yuan, et al.
Veröffentlicht: (2026)
von: Zhao, Yuan, et al.
Veröffentlicht: (2026)
GeoSym127K: Scalable Symbolically-verifiable Synthesis for Multimodal Geometric Reasoning
von: Jing, Jinhao, et al.
Veröffentlicht: (2026)
von: Jing, Jinhao, et al.
Veröffentlicht: (2026)
Discovering Hidden Visual Concepts Beyond Linguistic Input in Infant Learning
von: Ke, Xueyi, et al.
Veröffentlicht: (2025)
von: Ke, Xueyi, et al.
Veröffentlicht: (2025)
Causal-SAM-LLM: Large Language Models as Causal Reasoners for Robust Medical Segmentation
von: Tang, Tao, et al.
Veröffentlicht: (2025)
von: Tang, Tao, et al.
Veröffentlicht: (2025)
Reinforcing Spatial Reasoning in Vision-Language Models with Interwoven Thinking and Visual Drawing
von: Wu, Junfei, et al.
Veröffentlicht: (2025)
von: Wu, Junfei, et al.
Veröffentlicht: (2025)
AUVIC: Adversarial Unlearning of Visual Concepts for Multi-modal Large Language Models
von: Chen, Haokun, et al.
Veröffentlicht: (2025)
von: Chen, Haokun, et al.
Veröffentlicht: (2025)
The Role of Visual Modality in Multimodal Mathematical Reasoning: Challenges and Insights
von: Liu, Yufang, et al.
Veröffentlicht: (2025)
von: Liu, Yufang, et al.
Veröffentlicht: (2025)
ConceptScope: Characterizing Dataset Bias via Disentangled Visual Concepts
von: Choi, Jinho, et al.
Veröffentlicht: (2025)
von: Choi, Jinho, et al.
Veröffentlicht: (2025)
NEUSIS: A Compositional Neuro-Symbolic Framework for Autonomous Perception, Reasoning, and Planning in Complex UAV Search Missions
von: Cai, Zhixi, et al.
Veröffentlicht: (2024)
von: Cai, Zhixi, et al.
Veröffentlicht: (2024)
OmniPrism: Learning Disentangled Visual Concept for Image Generation
von: Li, Yangyang, et al.
Veröffentlicht: (2024)
von: Li, Yangyang, et al.
Veröffentlicht: (2024)
UniPixel: Unified Object Referring and Segmentation for Pixel-Level Visual Reasoning
von: Liu, Ye, et al.
Veröffentlicht: (2025)
von: Liu, Ye, et al.
Veröffentlicht: (2025)
ReasonMap: Towards Fine-Grained Visual Reasoning from Transit Maps
von: Feng, Sicheng, et al.
Veröffentlicht: (2025)
von: Feng, Sicheng, et al.
Veröffentlicht: (2025)
Predictive Reasoning with Augmented Anomaly Contrastive Learning for Compositional Visual Relations
von: Li, Chengtai, et al.
Veröffentlicht: (2026)
von: Li, Chengtai, et al.
Veröffentlicht: (2026)
Reasoning Path and Latent State Analysis for Multi-view Visual Spatial Reasoning: A Cognitive Science Perspective
von: Xue, Qiyao, et al.
Veröffentlicht: (2025)
von: Xue, Qiyao, et al.
Veröffentlicht: (2025)
CSVQA: A Chinese Multimodal Benchmark for Evaluating STEM Reasoning Capabilities of VLMs
von: Jian, Ai, et al.
Veröffentlicht: (2025)
von: Jian, Ai, et al.
Veröffentlicht: (2025)
Point-It-Out: Benchmarking Embodied Reasoning for Vision Language Models in Multi-Stage Visual Grounding
von: Xue, Haotian, et al.
Veröffentlicht: (2025)
von: Xue, Haotian, et al.
Veröffentlicht: (2025)
Eliminating the Language Bias for Visual Question Answering with fine-grained Causal Intervention
von: Liu, Ying, et al.
Veröffentlicht: (2024)
von: Liu, Ying, et al.
Veröffentlicht: (2024)
Ähnliche Einträge
-
Object-Conditioned Energy-Based Attention Map Alignment in Text-to-Image Diffusion Models
von: Zhang, Yasi, et al.
Veröffentlicht: (2024) -
MedVR: Annotation-Free Medical Visual Reasoning via Agentic Reinforcement Learning
von: Jiang, Zheng, et al.
Veröffentlicht: (2026) -
R^3-VQA: "Read the Room" by Video Social Reasoning
von: Niu, Lixing, et al.
Veröffentlicht: (2025) -
Multi-Grained Compositional Visual Clue Learning for Image Intent Recognition
von: Tang, Yin, et al.
Veröffentlicht: (2025) -
Weakly Supervised Concept Learning for Object-centric Visual Reasoning
von: Tiwari, Sparsh, et al.
Veröffentlicht: (2026)