Towards Verifiable Multimodal Deep Research: A Multi-Agent Harness for Interleaved Report Generation
Fuente:
arXiv
Salvato in:
| Autori principali: | Zhang, Chenghao, Dong, Guanting, Liu, Yufan, Zhao, Tong, Dou, Zhicheng |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2026
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
Towards Mixed-Modal Retrieval for Universal Retrieval-Augmented Generation
di: Zhang, Chenghao, et al.
Pubblicazione: (2025)
di: Zhang, Chenghao, et al.
Pubblicazione: (2025)
Progressive Multimodal Reasoning via Active Retrieval
di: Dong, Guanting, et al.
Pubblicazione: (2024)
di: Dong, Guanting, et al.
Pubblicazione: (2024)
Understand What LLM Needs: Dual Preference Alignment for Retrieval-Augmented Generation
di: Dong, Guanting, et al.
Pubblicazione: (2024)
di: Dong, Guanting, et al.
Pubblicazione: (2024)
ATIR: Towards Audio-Text Interleaved Contextual Retrieval
di: Zhao, Tong, et al.
Pubblicazione: (2026)
di: Zhao, Tong, et al.
Pubblicazione: (2026)
Multimodal DeepResearcher: Generating Text-Chart Interleaved Reports From Scratch with Agentic Framework
di: Yang, Zhaorui, et al.
Pubblicazione: (2025)
di: Yang, Zhaorui, et al.
Pubblicazione: (2025)
AgentCPM-Report: Interleaving Drafting and Deepening for Open-Ended Deep Research
di: Li, Yishan, et al.
Pubblicazione: (2026)
di: Li, Yishan, et al.
Pubblicazione: (2026)
Toward General Instruction-Following Alignment for Retrieval-Augmented Generation
di: Dong, Guanting, et al.
Pubblicazione: (2024)
di: Dong, Guanting, et al.
Pubblicazione: (2024)
Generate, Not Recommend: Personalized Multimodal Content Generation
di: Liu, Jiongnan, et al.
Pubblicazione: (2025)
di: Liu, Jiongnan, et al.
Pubblicazione: (2025)
WebThinker: Empowering Large Reasoning Models with Deep Research Capability
di: Li, Xiaoxi, et al.
Pubblicazione: (2025)
di: Li, Xiaoxi, et al.
Pubblicazione: (2025)
EnvScaler: Scaling Tool-Interactive Environments for LLM Agent via Programmatic Synthesis
di: Song, Xiaoshuai, et al.
Pubblicazione: (2026)
di: Song, Xiaoshuai, et al.
Pubblicazione: (2026)
HiRA: A Hierarchical Reasoning Framework for Decoupled Planning and Execution in Deep Search
di: Jin, Jiajie, et al.
Pubblicazione: (2025)
di: Jin, Jiajie, et al.
Pubblicazione: (2025)
DeepAgent: A General Reasoning Agent with Scalable Toolsets
di: Li, Xiaoxi, et al.
Pubblicazione: (2025)
di: Li, Xiaoxi, et al.
Pubblicazione: (2025)
AgentV-RL: Scaling Reward Modeling with Agentic Verifier
di: Zhang, Jiazheng, et al.
Pubblicazione: (2026)
di: Zhang, Jiazheng, et al.
Pubblicazione: (2026)
ET-Agent: Incentivizing Effective Tool-Integrated Reasoning Agent via Behavior Calibration
di: Chen, Yifei, et al.
Pubblicazione: (2026)
di: Chen, Yifei, et al.
Pubblicazione: (2026)
AgentFugue: Agent Scaling for Long-Horizon Tasks through Collective Reasoning
di: Hu, Yuyang, et al.
Pubblicazione: (2026)
di: Hu, Yuyang, et al.
Pubblicazione: (2026)
FinSight: Towards Real-World Financial Deep Research
di: Jin, Jiajie, et al.
Pubblicazione: (2025)
di: Jin, Jiajie, et al.
Pubblicazione: (2025)
MSI-Agent: Incorporating Multi-Scale Insight into Embodied Agents for Superior Planning and Decision-Making
di: Fu, Dayuan, et al.
Pubblicazione: (2024)
di: Fu, Dayuan, et al.
Pubblicazione: (2024)
SmartSearch: Process Reward-Guided Query Refinement for Search Agents
di: Wen, Tongyu, et al.
Pubblicazione: (2026)
di: Wen, Tongyu, et al.
Pubblicazione: (2026)
Toward Effective Tool-Integrated Reasoning via Self-Evolved Preference Learning
di: Chen, Yifei, et al.
Pubblicazione: (2025)
di: Chen, Yifei, et al.
Pubblicazione: (2025)
TextBind: Multi-turn Interleaved Multimodal Instruction-following in the Wild
di: Li, Huayang, et al.
Pubblicazione: (2023)
di: Li, Huayang, et al.
Pubblicazione: (2023)
Agent-World: Scaling Real-World Environment Synthesis for Evolving General Agent Intelligence
di: Dong, Guanting, et al.
Pubblicazione: (2026)
di: Dong, Guanting, et al.
Pubblicazione: (2026)
Deep-Reporter: Deep Research for Grounded Multimodal Long-Form Generation
di: Ye, Fangda, et al.
Pubblicazione: (2026)
di: Ye, Fangda, et al.
Pubblicazione: (2026)
Towards Trustworthy Report Generation: A Deep Research Agent with Progressive Confidence Estimation and Calibration
di: Yuan, Yi, et al.
Pubblicazione: (2026)
di: Yuan, Yi, et al.
Pubblicazione: (2026)
Beyond Single-shot Writing: Deep Research Agents are Unreliable at Multi-turn Report Revision
di: Chen, Bingsen, et al.
Pubblicazione: (2026)
di: Chen, Bingsen, et al.
Pubblicazione: (2026)
Search-o1: Agentic Search-Enhanced Large Reasoning Models
di: Li, Xiaoxi, et al.
Pubblicazione: (2025)
di: Li, Xiaoxi, et al.
Pubblicazione: (2025)
GPG: Generalized Policy Gradient Theorem for Transformer-based Policies
di: Mao, Hangyu, et al.
Pubblicazione: (2025)
di: Mao, Hangyu, et al.
Pubblicazione: (2025)
ProRAG: Process-Supervised Reinforcement Learning for Retrieval-Augmented Generation
di: Wang, Zhao, et al.
Pubblicazione: (2026)
di: Wang, Zhao, et al.
Pubblicazione: (2026)
OmniGAIA: Towards Native Omni-Modal AI Agents
di: Li, Xiaoxi, et al.
Pubblicazione: (2026)
di: Li, Xiaoxi, et al.
Pubblicazione: (2026)
Code as Agent Harness
di: Ning, Xuying, et al.
Pubblicazione: (2026)
di: Ning, Xuying, et al.
Pubblicazione: (2026)
Toward Verifiable Misinformation Detection: A Multi-Tool LLM Agent Framework
di: Cui, Zikun, et al.
Pubblicazione: (2025)
di: Cui, Zikun, et al.
Pubblicazione: (2025)
Dingtalk DeepResearch: A Unified Multi Agent Framework for Adaptive Intelligence in Enterprise Environments
di: Chen, Mengyuan, et al.
Pubblicazione: (2025)
di: Chen, Mengyuan, et al.
Pubblicazione: (2025)
Solution-oriented Agent-based Models Generation with Verifier-assisted Iterative In-context Learning
di: Niu, Tong, et al.
Pubblicazione: (2024)
di: Niu, Tong, et al.
Pubblicazione: (2024)
MiroEval: Benchmarking Multimodal Deep Research Agents in Process and Outcome
di: Ye, Fangda, et al.
Pubblicazione: (2026)
di: Ye, Fangda, et al.
Pubblicazione: (2026)
Memory Matters More: Event-Centric Memory as a Logic Map for Agent Searching and Reasoning
di: Hu, Yuyang, et al.
Pubblicazione: (2026)
di: Hu, Yuyang, et al.
Pubblicazione: (2026)
Tool-Star: Empowering LLM-Brained Multi-Tool Reasoner via Reinforcement Learning
di: Dong, Guanting, et al.
Pubblicazione: (2025)
di: Dong, Guanting, et al.
Pubblicazione: (2025)
TimelineReasoner: Advancing Timeline Summarization with Large Reasoning Models
di: Zhang, Liancheng, et al.
Pubblicazione: (2026)
di: Zhang, Liancheng, et al.
Pubblicazione: (2026)
Leveraging LLM-Assisted Query Understanding for Live Retrieval-Augmented Generation
di: Dong, Guanting, et al.
Pubblicazione: (2025)
di: Dong, Guanting, et al.
Pubblicazione: (2025)
HM-RAG: Hierarchical Multi-Agent Multimodal Retrieval Augmented Generation
di: Liu, Pei, et al.
Pubblicazione: (2025)
di: Liu, Pei, et al.
Pubblicazione: (2025)
AgentCTG: Harnessing Multi-Agent Collaboration for Fine-Grained Precise Control in Text Generation
di: Zhou, Xinxu, et al.
Pubblicazione: (2025)
di: Zhou, Xinxu, et al.
Pubblicazione: (2025)
Piecing It All Together: Verifying Multi-Hop Multimodal Claims
di: Wang, Haoran, et al.
Pubblicazione: (2024)
di: Wang, Haoran, et al.
Pubblicazione: (2024)
Documenti analoghi
-
Towards Mixed-Modal Retrieval for Universal Retrieval-Augmented Generation
di: Zhang, Chenghao, et al.
Pubblicazione: (2025) -
Progressive Multimodal Reasoning via Active Retrieval
di: Dong, Guanting, et al.
Pubblicazione: (2024) -
Understand What LLM Needs: Dual Preference Alignment for Retrieval-Augmented Generation
di: Dong, Guanting, et al.
Pubblicazione: (2024) -
ATIR: Towards Audio-Text Interleaved Contextual Retrieval
di: Zhao, Tong, et al.
Pubblicazione: (2026) -
Multimodal DeepResearcher: Generating Text-Chart Interleaved Reports From Scratch with Agentic Framework
di: Yang, Zhaorui, et al.
Pubblicazione: (2025)