Aligning Vision to Language: Annotation-Free Multimodal Knowledge Graph Construction for Enhanced LLMs Reasoning
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Liu, Junming, Meng, Siyuan, Gao, Yanting, Mao, Song, Cai, Pinlong, Yan, Guohang, Chen, Yirong, Bian, Zilin, Wang, Ding, Shi, Botian |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2025
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
From Ranking to Selection: A Simple but Efficient Dynamic Passage Selector for Retrieval Augmented Generation
von: Meng, Siyuan, et al.
Veröffentlicht: (2025)
von: Meng, Siyuan, et al.
Veröffentlicht: (2025)
DeepWriter: A Fact-Grounded Multimodal Writing Assistant Based On Offline Knowledge Base
von: Mao, Song, et al.
Veröffentlicht: (2025)
von: Mao, Song, et al.
Veröffentlicht: (2025)
RAKG:Document-level Retrieval Augmented Knowledge Graph Construction
von: Zhang, Hairong, et al.
Veröffentlicht: (2025)
von: Zhang, Hairong, et al.
Veröffentlicht: (2025)
LeanRAG: Knowledge-Graph-Based Generation with Semantic Aggregation and Hierarchical Retrieval
von: Zhang, Yaoze, et al.
Veröffentlicht: (2025)
von: Zhang, Yaoze, et al.
Veröffentlicht: (2025)
Investigating Redundancy in Multimodal Large Language Models with Multiple Vision Encoders
von: Wang, Yizhou, et al.
Veröffentlicht: (2025)
von: Wang, Yizhou, et al.
Veröffentlicht: (2025)
HetaRAG: Hybrid Deep Retrieval-Augmented Generation across Heterogeneous Data Stores
von: Yan, Guohang, et al.
Veröffentlicht: (2025)
von: Yan, Guohang, et al.
Veröffentlicht: (2025)
LimSim++: A Closed-Loop Platform for Deploying Multimodal LLMs in Autonomous Driving
von: Fu, Daocheng, et al.
Veröffentlicht: (2024)
von: Fu, Daocheng, et al.
Veröffentlicht: (2024)
Mosaic: Data-Free Knowledge Distillation via Mixture-of-Experts for Heterogeneous Distributed Environments
von: Liu, Junming, et al.
Veröffentlicht: (2025)
von: Liu, Junming, et al.
Veröffentlicht: (2025)
KG-TRACES: Enhancing Large Language Models with Knowledge Graph-constrained Trajectory Reasoning and Attribution Supervision
von: Wu, Rong, et al.
Veröffentlicht: (2025)
von: Wu, Rong, et al.
Veröffentlicht: (2025)
TrafficMCTS: A Closed-Loop Traffic Flow Generation Framework with Group-Based Monte Carlo Tree Search
von: Fu, Ze, et al.
Veröffentlicht: (2023)
von: Fu, Ze, et al.
Veröffentlicht: (2023)
MemVerse: Multimodal Memory for Lifelong Learning Agents
von: Liu, Junming, et al.
Veröffentlicht: (2025)
von: Liu, Junming, et al.
Veröffentlicht: (2025)
MGA: Memory-Driven GUI Agent for Observation-Centric Interaction
von: Cheng, Weihua, et al.
Veröffentlicht: (2025)
von: Cheng, Weihua, et al.
Veröffentlicht: (2025)
TimeMKG: Knowledge-Infused Causal Reasoning for Multivariate Time Series Modeling
von: Sun, Yifei, et al.
Veröffentlicht: (2025)
von: Sun, Yifei, et al.
Veröffentlicht: (2025)
LimSim Series: An Autonomous Driving Simulation Platform for Validation and Enhancement
von: Fu, Daocheng, et al.
Veröffentlicht: (2025)
von: Fu, Daocheng, et al.
Veröffentlicht: (2025)
GDI-Bench: A Benchmark for General Document Intelligence with Vision and Reasoning Decoupling
von: Li, Siqi, et al.
Veröffentlicht: (2025)
von: Li, Siqi, et al.
Veröffentlicht: (2025)
Self-Evolving Spatial Reasoning in Vision Language Models via Geometric Logic Consistency
von: Liu, Junming, et al.
Veröffentlicht: (2026)
von: Liu, Junming, et al.
Veröffentlicht: (2026)
DiLu: A Knowledge-Driven Approach to Autonomous Driving with Large Language Models
von: Wen, Licheng, et al.
Veröffentlicht: (2023)
von: Wen, Licheng, et al.
Veröffentlicht: (2023)
Can We Trust LLMs? Mitigate Overconfidence Bias in LLMs through Knowledge Transfer
von: Yang, Haoyan, et al.
Veröffentlicht: (2024)
von: Yang, Haoyan, et al.
Veröffentlicht: (2024)
FedRecon: Missing Modality Reconstruction in Heterogeneous Distributed Environments
von: Liu, Junming, et al.
Veröffentlicht: (2025)
von: Liu, Junming, et al.
Veröffentlicht: (2025)
UR-Bench: A Benchmark for Multi-Hop Reasoning over Ultra-High-Resolution Images
von: Li, Siqi, et al.
Veröffentlicht: (2025)
von: Li, Siqi, et al.
Veröffentlicht: (2025)
RADAR: Reasoning as Discrimination with Aligned Representations for LLM-based Knowledge Graph Reasoning
von: Xue, Bo, et al.
Veröffentlicht: (2026)
von: Xue, Bo, et al.
Veröffentlicht: (2026)
Reason-Align-Respond: Aligning LLM Reasoning with Knowledge Graphs for KGQA
von: Shen, Xiangqing, et al.
Veröffentlicht: (2025)
von: Shen, Xiangqing, et al.
Veröffentlicht: (2025)
SymDrive: Realistic and Controllable Driving Simulator via Symmetric Auto-regressive Online Restoration
von: Liu, Zhiyuan, et al.
Veröffentlicht: (2025)
von: Liu, Zhiyuan, et al.
Veröffentlicht: (2025)
LLMs for Knowledge Graph Construction and Reasoning: Recent Capabilities and Future Opportunities
von: Zhu, Yuqi, et al.
Veröffentlicht: (2023)
von: Zhu, Yuqi, et al.
Veröffentlicht: (2023)
Fine-tuning Pre-trained Vision-Language Models in a Human-Annotation-Free Manner
von: Wang, Qian-Wei, et al.
Veröffentlicht: (2026)
von: Wang, Qian-Wei, et al.
Veröffentlicht: (2026)
MELLA: Bridging Linguistic Capability and Cultural Groundedness for Low-Resource Language MLLMs
von: Gao, Yufei, et al.
Veröffentlicht: (2025)
von: Gao, Yufei, et al.
Veröffentlicht: (2025)
Vision Enhancing LLMs: Empowering Multimodal Knowledge Storage and Sharing in LLMs
von: Li, Yunxin, et al.
Veröffentlicht: (2023)
von: Li, Yunxin, et al.
Veröffentlicht: (2023)
Scene-Driven Multimodal Knowledge Graph Construction for Embodied AI
von: Yaoxian, Song, et al.
Veröffentlicht: (2023)
von: Yaoxian, Song, et al.
Veröffentlicht: (2023)
The Agent's First Day: Benchmarking Learning, Exploration, and Scheduling in the Workplace Scenarios
von: Fu, Daocheng, et al.
Veröffentlicht: (2026)
von: Fu, Daocheng, et al.
Veröffentlicht: (2026)
Knowledge Graph for Intelligent Generation of Artistic Image Creation: Constructing a New Annotation Hierarchy
von: Kaixin, Jia, et al.
Veröffentlicht: (2025)
von: Kaixin, Jia, et al.
Veröffentlicht: (2025)
HM-RAG: Hierarchical Multi-Agent Multimodal Retrieval Augmented Generation
von: Liu, Pei, et al.
Veröffentlicht: (2025)
von: Liu, Pei, et al.
Veröffentlicht: (2025)
Multimodal Reasoning with Multimodal Knowledge Graph
von: Lee, Junlin, et al.
Veröffentlicht: (2024)
von: Lee, Junlin, et al.
Veröffentlicht: (2024)
A Survey of Knowledge Graph Reasoning on Graph Types: Static, Dynamic, and Multimodal
von: Liang, Ke, et al.
Veröffentlicht: (2022)
von: Liang, Ke, et al.
Veröffentlicht: (2022)
Continual Multimodal Knowledge Graph Construction
von: Chen, Xiang, et al.
Veröffentlicht: (2023)
von: Chen, Xiang, et al.
Veröffentlicht: (2023)
MemCoT: Test-Time Scaling through Memory-Driven Chain-of-Thought
von: Lei, Haodong, et al.
Veröffentlicht: (2026)
von: Lei, Haodong, et al.
Veröffentlicht: (2026)
CodeGraph: Enhancing Graph Reasoning of LLMs with Code
von: Cai, Qiaolong, et al.
Veröffentlicht: (2024)
von: Cai, Qiaolong, et al.
Veröffentlicht: (2024)
SceneAlign: Aligning Multimodal Reasoning to Scene Graphs in Complex Visual Scenes
von: Wang, Chuhan, et al.
Veröffentlicht: (2026)
von: Wang, Chuhan, et al.
Veröffentlicht: (2026)
Encoder-Free Knowledge-Graph Reasoning with LLMs via Hyperdimensional Path Retrieval
von: Liu, Yezi, et al.
Veröffentlicht: (2025)
von: Liu, Yezi, et al.
Veröffentlicht: (2025)
OASim: an Open and Adaptive Simulator based on Neural Rendering for Autonomous Driving
von: Yan, Guohang, et al.
Veröffentlicht: (2024)
von: Yan, Guohang, et al.
Veröffentlicht: (2024)
Construct, Align, and Reason: Large Ontology Models for Enterprise Knowledge Management
von: Zhang, Yao, et al.
Veröffentlicht: (2026)
von: Zhang, Yao, et al.
Veröffentlicht: (2026)
Ähnliche Einträge
-
From Ranking to Selection: A Simple but Efficient Dynamic Passage Selector for Retrieval Augmented Generation
von: Meng, Siyuan, et al.
Veröffentlicht: (2025) -
DeepWriter: A Fact-Grounded Multimodal Writing Assistant Based On Offline Knowledge Base
von: Mao, Song, et al.
Veröffentlicht: (2025) -
RAKG:Document-level Retrieval Augmented Knowledge Graph Construction
von: Zhang, Hairong, et al.
Veröffentlicht: (2025) -
LeanRAG: Knowledge-Graph-Based Generation with Semantic Aggregation and Hierarchical Retrieval
von: Zhang, Yaoze, et al.
Veröffentlicht: (2025) -
Investigating Redundancy in Multimodal Large Language Models with Multiple Vision Encoders
von: Wang, Yizhou, et al.
Veröffentlicht: (2025)