Towards Verifiable Multimodal Deep Research: A Multi-Agent Harness for Interleaved Report Generation
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Zhang, Chenghao, Dong, Guanting, Liu, Yufan, Zhao, Tong, Dou, Zhicheng |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2026
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Towards Mixed-Modal Retrieval for Universal Retrieval-Augmented Generation
von: Zhang, Chenghao, et al.
Veröffentlicht: (2025)
von: Zhang, Chenghao, et al.
Veröffentlicht: (2025)
Progressive Multimodal Reasoning via Active Retrieval
von: Dong, Guanting, et al.
Veröffentlicht: (2024)
von: Dong, Guanting, et al.
Veröffentlicht: (2024)
Understand What LLM Needs: Dual Preference Alignment for Retrieval-Augmented Generation
von: Dong, Guanting, et al.
Veröffentlicht: (2024)
von: Dong, Guanting, et al.
Veröffentlicht: (2024)
ATIR: Towards Audio-Text Interleaved Contextual Retrieval
von: Zhao, Tong, et al.
Veröffentlicht: (2026)
von: Zhao, Tong, et al.
Veröffentlicht: (2026)
Multimodal DeepResearcher: Generating Text-Chart Interleaved Reports From Scratch with Agentic Framework
von: Yang, Zhaorui, et al.
Veröffentlicht: (2025)
von: Yang, Zhaorui, et al.
Veröffentlicht: (2025)
AgentCPM-Report: Interleaving Drafting and Deepening for Open-Ended Deep Research
von: Li, Yishan, et al.
Veröffentlicht: (2026)
von: Li, Yishan, et al.
Veröffentlicht: (2026)
Toward General Instruction-Following Alignment for Retrieval-Augmented Generation
von: Dong, Guanting, et al.
Veröffentlicht: (2024)
von: Dong, Guanting, et al.
Veröffentlicht: (2024)
Generate, Not Recommend: Personalized Multimodal Content Generation
von: Liu, Jiongnan, et al.
Veröffentlicht: (2025)
von: Liu, Jiongnan, et al.
Veröffentlicht: (2025)
WebThinker: Empowering Large Reasoning Models with Deep Research Capability
von: Li, Xiaoxi, et al.
Veröffentlicht: (2025)
von: Li, Xiaoxi, et al.
Veröffentlicht: (2025)
EnvScaler: Scaling Tool-Interactive Environments for LLM Agent via Programmatic Synthesis
von: Song, Xiaoshuai, et al.
Veröffentlicht: (2026)
von: Song, Xiaoshuai, et al.
Veröffentlicht: (2026)
HiRA: A Hierarchical Reasoning Framework for Decoupled Planning and Execution in Deep Search
von: Jin, Jiajie, et al.
Veröffentlicht: (2025)
von: Jin, Jiajie, et al.
Veröffentlicht: (2025)
DeepAgent: A General Reasoning Agent with Scalable Toolsets
von: Li, Xiaoxi, et al.
Veröffentlicht: (2025)
von: Li, Xiaoxi, et al.
Veröffentlicht: (2025)
AgentV-RL: Scaling Reward Modeling with Agentic Verifier
von: Zhang, Jiazheng, et al.
Veröffentlicht: (2026)
von: Zhang, Jiazheng, et al.
Veröffentlicht: (2026)
ET-Agent: Incentivizing Effective Tool-Integrated Reasoning Agent via Behavior Calibration
von: Chen, Yifei, et al.
Veröffentlicht: (2026)
von: Chen, Yifei, et al.
Veröffentlicht: (2026)
AgentFugue: Agent Scaling for Long-Horizon Tasks through Collective Reasoning
von: Hu, Yuyang, et al.
Veröffentlicht: (2026)
von: Hu, Yuyang, et al.
Veröffentlicht: (2026)
FinSight: Towards Real-World Financial Deep Research
von: Jin, Jiajie, et al.
Veröffentlicht: (2025)
von: Jin, Jiajie, et al.
Veröffentlicht: (2025)
MSI-Agent: Incorporating Multi-Scale Insight into Embodied Agents for Superior Planning and Decision-Making
von: Fu, Dayuan, et al.
Veröffentlicht: (2024)
von: Fu, Dayuan, et al.
Veröffentlicht: (2024)
SmartSearch: Process Reward-Guided Query Refinement for Search Agents
von: Wen, Tongyu, et al.
Veröffentlicht: (2026)
von: Wen, Tongyu, et al.
Veröffentlicht: (2026)
Toward Effective Tool-Integrated Reasoning via Self-Evolved Preference Learning
von: Chen, Yifei, et al.
Veröffentlicht: (2025)
von: Chen, Yifei, et al.
Veröffentlicht: (2025)
TextBind: Multi-turn Interleaved Multimodal Instruction-following in the Wild
von: Li, Huayang, et al.
Veröffentlicht: (2023)
von: Li, Huayang, et al.
Veröffentlicht: (2023)
Agent-World: Scaling Real-World Environment Synthesis for Evolving General Agent Intelligence
von: Dong, Guanting, et al.
Veröffentlicht: (2026)
von: Dong, Guanting, et al.
Veröffentlicht: (2026)
Deep-Reporter: Deep Research for Grounded Multimodal Long-Form Generation
von: Ye, Fangda, et al.
Veröffentlicht: (2026)
von: Ye, Fangda, et al.
Veröffentlicht: (2026)
Towards Trustworthy Report Generation: A Deep Research Agent with Progressive Confidence Estimation and Calibration
von: Yuan, Yi, et al.
Veröffentlicht: (2026)
von: Yuan, Yi, et al.
Veröffentlicht: (2026)
Beyond Single-shot Writing: Deep Research Agents are Unreliable at Multi-turn Report Revision
von: Chen, Bingsen, et al.
Veröffentlicht: (2026)
von: Chen, Bingsen, et al.
Veröffentlicht: (2026)
Search-o1: Agentic Search-Enhanced Large Reasoning Models
von: Li, Xiaoxi, et al.
Veröffentlicht: (2025)
von: Li, Xiaoxi, et al.
Veröffentlicht: (2025)
GPG: Generalized Policy Gradient Theorem for Transformer-based Policies
von: Mao, Hangyu, et al.
Veröffentlicht: (2025)
von: Mao, Hangyu, et al.
Veröffentlicht: (2025)
ProRAG: Process-Supervised Reinforcement Learning for Retrieval-Augmented Generation
von: Wang, Zhao, et al.
Veröffentlicht: (2026)
von: Wang, Zhao, et al.
Veröffentlicht: (2026)
OmniGAIA: Towards Native Omni-Modal AI Agents
von: Li, Xiaoxi, et al.
Veröffentlicht: (2026)
von: Li, Xiaoxi, et al.
Veröffentlicht: (2026)
Code as Agent Harness
von: Ning, Xuying, et al.
Veröffentlicht: (2026)
von: Ning, Xuying, et al.
Veröffentlicht: (2026)
Toward Verifiable Misinformation Detection: A Multi-Tool LLM Agent Framework
von: Cui, Zikun, et al.
Veröffentlicht: (2025)
von: Cui, Zikun, et al.
Veröffentlicht: (2025)
Dingtalk DeepResearch: A Unified Multi Agent Framework for Adaptive Intelligence in Enterprise Environments
von: Chen, Mengyuan, et al.
Veröffentlicht: (2025)
von: Chen, Mengyuan, et al.
Veröffentlicht: (2025)
Solution-oriented Agent-based Models Generation with Verifier-assisted Iterative In-context Learning
von: Niu, Tong, et al.
Veröffentlicht: (2024)
von: Niu, Tong, et al.
Veröffentlicht: (2024)
MiroEval: Benchmarking Multimodal Deep Research Agents in Process and Outcome
von: Ye, Fangda, et al.
Veröffentlicht: (2026)
von: Ye, Fangda, et al.
Veröffentlicht: (2026)
Memory Matters More: Event-Centric Memory as a Logic Map for Agent Searching and Reasoning
von: Hu, Yuyang, et al.
Veröffentlicht: (2026)
von: Hu, Yuyang, et al.
Veröffentlicht: (2026)
Tool-Star: Empowering LLM-Brained Multi-Tool Reasoner via Reinforcement Learning
von: Dong, Guanting, et al.
Veröffentlicht: (2025)
von: Dong, Guanting, et al.
Veröffentlicht: (2025)
TimelineReasoner: Advancing Timeline Summarization with Large Reasoning Models
von: Zhang, Liancheng, et al.
Veröffentlicht: (2026)
von: Zhang, Liancheng, et al.
Veröffentlicht: (2026)
Leveraging LLM-Assisted Query Understanding for Live Retrieval-Augmented Generation
von: Dong, Guanting, et al.
Veröffentlicht: (2025)
von: Dong, Guanting, et al.
Veröffentlicht: (2025)
HM-RAG: Hierarchical Multi-Agent Multimodal Retrieval Augmented Generation
von: Liu, Pei, et al.
Veröffentlicht: (2025)
von: Liu, Pei, et al.
Veröffentlicht: (2025)
AgentCTG: Harnessing Multi-Agent Collaboration for Fine-Grained Precise Control in Text Generation
von: Zhou, Xinxu, et al.
Veröffentlicht: (2025)
von: Zhou, Xinxu, et al.
Veröffentlicht: (2025)
Piecing It All Together: Verifying Multi-Hop Multimodal Claims
von: Wang, Haoran, et al.
Veröffentlicht: (2024)
von: Wang, Haoran, et al.
Veröffentlicht: (2024)
Ähnliche Einträge
-
Towards Mixed-Modal Retrieval for Universal Retrieval-Augmented Generation
von: Zhang, Chenghao, et al.
Veröffentlicht: (2025) -
Progressive Multimodal Reasoning via Active Retrieval
von: Dong, Guanting, et al.
Veröffentlicht: (2024) -
Understand What LLM Needs: Dual Preference Alignment for Retrieval-Augmented Generation
von: Dong, Guanting, et al.
Veröffentlicht: (2024) -
ATIR: Towards Audio-Text Interleaved Contextual Retrieval
von: Zhao, Tong, et al.
Veröffentlicht: (2026) -
Multimodal DeepResearcher: Generating Text-Chart Interleaved Reports From Scratch with Agentic Framework
von: Yang, Zhaorui, et al.
Veröffentlicht: (2025)