Multimodal DeepResearcher: Generating Text-Chart Interleaved Reports From Scratch with Agentic Framework
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Yang, Zhaorui, Pan, Bo, Wang, Han, Wang, Yiyao, Liu, Xingyu, Weng, Luoxuan, Feng, Yingchaojie, Feng, Haozhe, Zhu, Minfeng, Zhang, Bo, Chen, Wei |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2025
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
InsightLens: Augmenting LLM-Powered Data Analysis with Interactive Insight Management and Navigation
von: Weng, Luoxuan, et al.
Veröffentlicht: (2024)
von: Weng, Luoxuan, et al.
Veröffentlicht: (2024)
IGenBench: Benchmarking the Reliability of Text-to-Infographic Generation
von: Tang, Yinghao, et al.
Veröffentlicht: (2026)
von: Tang, Yinghao, et al.
Veröffentlicht: (2026)
Skywork-R1V4: Toward Agentic Multimodal Intelligence through Interleaved Thinking with Images and DeepResearch
von: Zhang, Yifan, et al.
Veröffentlicht: (2025)
von: Zhang, Yifan, et al.
Veröffentlicht: (2025)
Vision-DeepResearch: Incentivizing DeepResearch Capability in Multimodal Large Language Models
von: Huang, Wenxuan, et al.
Veröffentlicht: (2026)
von: Huang, Wenxuan, et al.
Veröffentlicht: (2026)
Exploring Multimodal Prompt for Visualization Authoring with Large Language Models
von: Wen, Zhen, et al.
Veröffentlicht: (2025)
von: Wen, Zhen, et al.
Veröffentlicht: (2025)
MM-DeepResearch: A Simple and Effective Multimodal Agentic Search Baseline
von: Yao, Huanjin, et al.
Veröffentlicht: (2026)
von: Yao, Huanjin, et al.
Veröffentlicht: (2026)
Tongyi DeepResearch Technical Report
von: Tongyi DeepResearch Team, et al.
Veröffentlicht: (2025)
von: Tongyi DeepResearch Team, et al.
Veröffentlicht: (2025)
Step-DeepResearch Technical Report
von: Hu, Chen, et al.
Veröffentlicht: (2025)
von: Hu, Chen, et al.
Veröffentlicht: (2025)
XGraphRAG: Interactive Visual Analysis for Graph-based Retrieval-Augmented Generation
von: Wang, Ke, et al.
Veröffentlicht: (2025)
von: Wang, Ke, et al.
Veröffentlicht: (2025)
Yunque DeepResearch Technical Report
von: Cai, Yuxuan, et al.
Veröffentlicht: (2026)
von: Cai, Yuxuan, et al.
Veröffentlicht: (2026)
Understanding DeepResearch via Reports
von: Fan, Tianyu, et al.
Veröffentlicht: (2025)
von: Fan, Tianyu, et al.
Veröffentlicht: (2025)
Mind DeepResearch Technical Report
von: MindDR Team, et al.
Veröffentlicht: (2026)
von: MindDR Team, et al.
Veröffentlicht: (2026)
Self-Distillation Bridges Distribution Gap in Language Model Fine-Tuning
von: Yang, Zhaorui, et al.
Veröffentlicht: (2024)
von: Yang, Zhaorui, et al.
Veröffentlicht: (2024)
AgentCoord: Visually Exploring Coordination Strategy for LLM-based Multi-Agent Collaboration
von: Pan, Bo, et al.
Veröffentlicht: (2024)
von: Pan, Bo, et al.
Veröffentlicht: (2024)
IDRBench: Interactive Deep Research Benchmark
von: Feng, Yingchaojie, et al.
Veröffentlicht: (2026)
von: Feng, Yingchaojie, et al.
Veröffentlicht: (2026)
DEEPMED: Building a Medical DeepResearch Agent via Multi-hop Med-Search Data and Turn-Controlled Agentic Training & Inference
von: Wang, Zihan, et al.
Veröffentlicht: (2026)
von: Wang, Zihan, et al.
Veröffentlicht: (2026)
Marco DeepResearch: Unlocking Efficient Deep Research Agents via Verification-Centric Design
von: Zhu, Bin, et al.
Veröffentlicht: (2026)
von: Zhu, Bin, et al.
Veröffentlicht: (2026)
DeepResearch Bench: A Comprehensive Benchmark for Deep Research Agents
von: Du, Mingxuan, et al.
Veröffentlicht: (2025)
von: Du, Mingxuan, et al.
Veröffentlicht: (2025)
DeepResearch-Slice: Bridging the Retrieval-Utilization Gap via Explicit Text Slicing
von: Lu, Shuo, et al.
Veröffentlicht: (2025)
von: Lu, Shuo, et al.
Veröffentlicht: (2025)
DeepResearch Bench II: Diagnosing Deep Research Agents via Rubrics from Expert Report
von: Li, Ruizhe, et al.
Veröffentlicht: (2026)
von: Li, Ruizhe, et al.
Veröffentlicht: (2026)
LiteResearcher: A Scalable Agentic RL Training Framework for Deep Research Agent
von: Li, Wanli, et al.
Veröffentlicht: (2026)
von: Li, Wanli, et al.
Veröffentlicht: (2026)
TVIR: Building Deep Research Agents Towards Text--Visual Interleaved Report Generation
von: Ma, Xinkai, et al.
Veröffentlicht: (2026)
von: Ma, Xinkai, et al.
Veröffentlicht: (2026)
Dingtalk DeepResearch: A Unified Multi Agent Framework for Adaptive Intelligence in Enterprise Environments
von: Chen, Mengyuan, et al.
Veröffentlicht: (2025)
von: Chen, Mengyuan, et al.
Veröffentlicht: (2025)
DeepResearch-9K: A Challenging Benchmark Dataset of Deep-Research Agent
von: Wu, Tongzhou, et al.
Veröffentlicht: (2026)
von: Wu, Tongzhou, et al.
Veröffentlicht: (2026)
Learning Query-Specific Rubrics from Human Preferences for DeepResearch Report Generation
von: Lv, Changze, et al.
Veröffentlicht: (2026)
von: Lv, Changze, et al.
Veröffentlicht: (2026)
Don't Reinvent the Wheel: Efficient Instruction-Following Text Embedding based on Guided Space Transformation
von: Feng, Yingchaojie, et al.
Veröffentlicht: (2025)
von: Feng, Yingchaojie, et al.
Veröffentlicht: (2025)
RAGExplorer: A Visual Analytics System for the Comparative Diagnosis of RAG Systems
von: Tian, Haoyu, et al.
Veröffentlicht: (2026)
von: Tian, Haoyu, et al.
Veröffentlicht: (2026)
Vision-DeepResearch Benchmark: Rethinking Visual and Textual Search for Multimodal Large Language Models
von: Zeng, Yu, et al.
Veröffentlicht: (2026)
von: Zeng, Yu, et al.
Veröffentlicht: (2026)
DeepResearch$^{\text{Eco}}$: A Recursive Agentic Workflow for Complex Scientific Question Answering in Ecology
von: D'Souza, Jennifer, et al.
Veröffentlicht: (2025)
von: D'Souza, Jennifer, et al.
Veröffentlicht: (2025)
InterDeepResearch: Enabling Human-Agent Collaborative Information Seeking through Interactive Deep Research
von: Pan, Bo, et al.
Veröffentlicht: (2026)
von: Pan, Bo, et al.
Veröffentlicht: (2026)
DataLab: A Unified Platform for LLM-Powered Business Intelligence
von: Weng, Luoxuan, et al.
Veröffentlicht: (2024)
von: Weng, Luoxuan, et al.
Veröffentlicht: (2024)
IntuiTF: MLLM-Guided Transfer Function Optimization for Direct Volume Rendering
von: Wang, Yiyao, et al.
Veröffentlicht: (2025)
von: Wang, Yiyao, et al.
Veröffentlicht: (2025)
COHERENCE: Benchmarking Fine-Grained Image-Text Alignment in Interleaved Multimodal Contexts
von: Wang, Bingli, et al.
Veröffentlicht: (2026)
von: Wang, Bingli, et al.
Veröffentlicht: (2026)
ChartVerse: Scaling Chart Reasoning via Reliable Programmatic Synthesis from Scratch
von: Liu, Zheng, et al.
Veröffentlicht: (2026)
von: Liu, Zheng, et al.
Veröffentlicht: (2026)
DeepResearcher: Scaling Deep Research via Reinforcement Learning in Real-world Environments
von: Zheng, Yuxiang, et al.
Veröffentlicht: (2025)
von: Zheng, Yuxiang, et al.
Veröffentlicht: (2025)
JailbreakLens: Visual Analysis of Jailbreak Attacks Against Large Language Models
von: Feng, Yingchaojie, et al.
Veröffentlicht: (2024)
von: Feng, Yingchaojie, et al.
Veröffentlicht: (2024)
AgentLens: Visual Analysis for Agent Behaviors in LLM-based Autonomous Systems
von: Lu, Jiaying, et al.
Veröffentlicht: (2024)
von: Lu, Jiaying, et al.
Veröffentlicht: (2024)
STICKERCONV: Generating Multimodal Empathetic Responses from Scratch
von: Zhang, Yiqun, et al.
Veröffentlicht: (2024)
von: Zhang, Yiqun, et al.
Veröffentlicht: (2024)
DeepResearch Arena: The First Exam of LLMs' Research Abilities via Seminar-Grounded Tasks
von: Wan, Haiyuan, et al.
Veröffentlicht: (2025)
von: Wan, Haiyuan, et al.
Veröffentlicht: (2025)
Towards Verifiable Multimodal Deep Research: A Multi-Agent Harness for Interleaved Report Generation
von: Zhang, Chenghao, et al.
Veröffentlicht: (2026)
von: Zhang, Chenghao, et al.
Veröffentlicht: (2026)
Ähnliche Einträge
-
InsightLens: Augmenting LLM-Powered Data Analysis with Interactive Insight Management and Navigation
von: Weng, Luoxuan, et al.
Veröffentlicht: (2024) -
IGenBench: Benchmarking the Reliability of Text-to-Infographic Generation
von: Tang, Yinghao, et al.
Veröffentlicht: (2026) -
Skywork-R1V4: Toward Agentic Multimodal Intelligence through Interleaved Thinking with Images and DeepResearch
von: Zhang, Yifan, et al.
Veröffentlicht: (2025) -
Vision-DeepResearch: Incentivizing DeepResearch Capability in Multimodal Large Language Models
von: Huang, Wenxuan, et al.
Veröffentlicht: (2026) -
Exploring Multimodal Prompt for Visualization Authoring with Large Language Models
von: Wen, Zhen, et al.
Veröffentlicht: (2025)