Real Deep Research for AI, Robotics and Beyond
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Zou, Xueyan, Ye, Jianglong, Zhang, Hao, Xiang, Xiaoyu, Ding, Mingyu, Yang, Zhaojing, Lee, Yong Jae, Tu, Zhuowen, Liu, Sifei, Wang, Xiaolong |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2025
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
M3: 3D-Spatial MultiModal Memory
von: Zou, Xueyan, et al.
Veröffentlicht: (2025)
von: Zou, Xueyan, et al.
Veröffentlicht: (2025)
NaVILA: Legged Robot Vision-Language-Action Model for Navigation
von: Cheng, An-Chieh, et al.
Veröffentlicht: (2024)
von: Cheng, An-Chieh, et al.
Veröffentlicht: (2024)
HOIDiffusion: Generating Realistic 3D Hand-Object Interaction Data
von: Zhang, Mengqi, et al.
Veröffentlicht: (2024)
von: Zhang, Mengqi, et al.
Veröffentlicht: (2024)
Interfacing Foundation Models' Embeddings
von: Zou, Xueyan, et al.
Veröffentlicht: (2023)
von: Zou, Xueyan, et al.
Veröffentlicht: (2023)
Characterizing Learning Curves During Language Model Pre-Training: Learning, Forgetting, and Stability
von: Chang, Tyler A., et al.
Veröffentlicht: (2023)
von: Chang, Tyler A., et al.
Veröffentlicht: (2023)
Goldfish: Monolingual Language Models for 350 Languages
von: Chang, Tyler A., et al.
Veröffentlicht: (2024)
von: Chang, Tyler A., et al.
Veröffentlicht: (2024)
Interaction as Intelligence: Deep Research With Human-AI Partnership
von: Ye, Lyumanshan, et al.
Veröffentlicht: (2025)
von: Ye, Lyumanshan, et al.
Veröffentlicht: (2025)
Beyond Turn Limits: Training Deep Search Agents with Dynamic Context Window
von: Tang, Qiaoyu, et al.
Veröffentlicht: (2025)
von: Tang, Qiaoyu, et al.
Veröffentlicht: (2025)
$\textbf{PLUM}$: Improving Code LMs with Execution-Guided On-Policy Preference Learning Driven By Synthetic Test Cases
von: Zhang, Dylan, et al.
Veröffentlicht: (2024)
von: Zhang, Dylan, et al.
Veröffentlicht: (2024)
PhyGrasp: Generalizing Robotic Grasping with Physics-informed Large Multimodal Models
von: Guo, Dingkun, et al.
Veröffentlicht: (2024)
von: Guo, Dingkun, et al.
Veröffentlicht: (2024)
FeatureNeRF: Learning Generalizable NeRFs by Distilling Foundation Models
von: Ye, Jianglong, et al.
Veröffentlicht: (2023)
von: Ye, Jianglong, et al.
Veröffentlicht: (2023)
DeepResearcher: Scaling Deep Research via Reinforcement Learning in Real-world Environments
von: Zheng, Yuxiang, et al.
Veröffentlicht: (2025)
von: Zheng, Yuxiang, et al.
Veröffentlicht: (2025)
How Multimodal LLMs Solve Image Tasks: A Lens on Visual Grounding, Task Reasoning, and Answer Decoding
von: Yu, Zhuoran, et al.
Veröffentlicht: (2025)
von: Yu, Zhuoran, et al.
Veröffentlicht: (2025)
Cross-modal Context Fusion and Adaptive Graph Convolutional Network for Multimodal Conversational Emotion Recognition
von: Feng, Junwei, et al.
Veröffentlicht: (2025)
von: Feng, Junwei, et al.
Veröffentlicht: (2025)
Exploring the Limitations of Large Language Models in Compositional Relation Reasoning
von: Zhao, Jinman, et al.
Veröffentlicht: (2024)
von: Zhao, Jinman, et al.
Veröffentlicht: (2024)
Beyond Simulations: What 20,000 Real Conversations Reveal About Mental Health AI Safety
von: Stamatis, Caitlin A., et al.
Veröffentlicht: (2026)
von: Stamatis, Caitlin A., et al.
Veröffentlicht: (2026)
Beyond Single-shot Writing: Deep Research Agents are Unreliable at Multi-turn Report Revision
von: Chen, Bingsen, et al.
Veröffentlicht: (2026)
von: Chen, Bingsen, et al.
Veröffentlicht: (2026)
Beyond English: Unveiling Multilingual Bias in LLM Copyright Compliance
von: Chen, Yupeng, et al.
Veröffentlicht: (2025)
von: Chen, Yupeng, et al.
Veröffentlicht: (2025)
AI Urban Scientist: Multi-Agent Collaborative Automation for Urban Research
von: Xia, Tong, et al.
Veröffentlicht: (2025)
von: Xia, Tong, et al.
Veröffentlicht: (2025)
Efficient Scaling of Diffusion Transformers for Text-to-Image Generation
von: Li, Hao, et al.
Veröffentlicht: (2024)
von: Li, Hao, et al.
Veröffentlicht: (2024)
Language Models Don't Know What You Want: Evaluating Personalization in Deep Research Needs Real Users
von: Balepur, Nishant, et al.
Veröffentlicht: (2026)
von: Balepur, Nishant, et al.
Veröffentlicht: (2026)
DeepResearchEval: An Automated Framework for Deep Research Task Construction and Agentic Evaluation
von: Wang, Yibo, et al.
Veröffentlicht: (2026)
von: Wang, Yibo, et al.
Veröffentlicht: (2026)
CogDual: Enhancing Dual Cognition of LLMs via Reinforcement Learning with Implicit Rule-Based Rewards
von: Liu, Cheng, et al.
Veröffentlicht: (2025)
von: Liu, Cheng, et al.
Veröffentlicht: (2025)
QeRL: Beyond Efficiency -- Quantization-enhanced Reinforcement Learning for LLMs
von: Huang, Wei, et al.
Veröffentlicht: (2025)
von: Huang, Wei, et al.
Veröffentlicht: (2025)
A Method for the Architecture of a Medical Vertical Large Language Model Based on Deepseek R1
von: Zhang, Mingda, et al.
Veröffentlicht: (2025)
von: Zhang, Mingda, et al.
Veröffentlicht: (2025)
RGBD Objects in the Wild: Scaling Real-World 3D Object Learning from RGB-D Videos
von: Xia, Hongchi, et al.
Veröffentlicht: (2024)
von: Xia, Hongchi, et al.
Veröffentlicht: (2024)
DEER: A Benchmark for Evaluating Deep Research Agents on Expert Report Generation
von: Han, Janghoon, et al.
Veröffentlicht: (2025)
von: Han, Janghoon, et al.
Veröffentlicht: (2025)
FinDeepResearch: Evaluating Deep Research Agents in Rigorous Financial Analysis
von: Zhu, Fengbin, et al.
Veröffentlicht: (2025)
von: Zhu, Fengbin, et al.
Veröffentlicht: (2025)
REMAC: Self-Reflective and Self-Evolving Multi-Agent Collaboration for Long-Horizon Robot Manipulation
von: Yuan, Puzhen, et al.
Veröffentlicht: (2025)
von: Yuan, Puzhen, et al.
Veröffentlicht: (2025)
Dex1B: Learning with 1B Demonstrations for Dexterous Manipulation
von: Ye, Jianglong, et al.
Veröffentlicht: (2025)
von: Ye, Jianglong, et al.
Veröffentlicht: (2025)
Deep Research as Rubric for Reinforcement Learning
von: Mei, Wangyi, et al.
Veröffentlicht: (2026)
von: Mei, Wangyi, et al.
Veröffentlicht: (2026)
Musketeer: Joint Training for Multi-task Vision Language Model with Task Explanation Prompts
von: Zhang, Zhaoyang, et al.
Veröffentlicht: (2023)
von: Zhang, Zhaoyang, et al.
Veröffentlicht: (2023)
Beyond Isolated Dots: Benchmarking Structured Table Construction as Deep Knowledge Extraction
von: Zhong, Tianyun, et al.
Veröffentlicht: (2025)
von: Zhong, Tianyun, et al.
Veröffentlicht: (2025)
An Interactive Paradigm for Deep Research
von: Ai, Lin, et al.
Veröffentlicht: (2026)
von: Ai, Lin, et al.
Veröffentlicht: (2026)
A Systematic Study of Compositional Syntactic Transformer Language Models
von: Zhao, Yida, et al.
Veröffentlicht: (2025)
von: Zhao, Yida, et al.
Veröffentlicht: (2025)
Deep Researcher with Test-Time Diffusion
von: Han, Rujun, et al.
Veröffentlicht: (2025)
von: Han, Rujun, et al.
Veröffentlicht: (2025)
Survey of Design Paradigms for Social Robots
von: Frieske, Rita, et al.
Veröffentlicht: (2024)
von: Frieske, Rita, et al.
Veröffentlicht: (2024)
OccuBench: Evaluating AI Agents on Real-World Professional Tasks via Language Environment Simulation
von: Hu, Xiaomeng, et al.
Veröffentlicht: (2026)
von: Hu, Xiaomeng, et al.
Veröffentlicht: (2026)
Beyond Passive Critical Thinking: Fostering Proactive Questioning to Enhance Human-AI Collaboration
von: Wang, Ante, et al.
Veröffentlicht: (2025)
von: Wang, Ante, et al.
Veröffentlicht: (2025)
DocKD: Knowledge Distillation from LLMs for Open-World Document Understanding Models
von: Kim, Sungnyun, et al.
Veröffentlicht: (2024)
von: Kim, Sungnyun, et al.
Veröffentlicht: (2024)
Ähnliche Einträge
-
M3: 3D-Spatial MultiModal Memory
von: Zou, Xueyan, et al.
Veröffentlicht: (2025) -
NaVILA: Legged Robot Vision-Language-Action Model for Navigation
von: Cheng, An-Chieh, et al.
Veröffentlicht: (2024) -
HOIDiffusion: Generating Realistic 3D Hand-Object Interaction Data
von: Zhang, Mengqi, et al.
Veröffentlicht: (2024) -
Interfacing Foundation Models' Embeddings
von: Zou, Xueyan, et al.
Veröffentlicht: (2023) -
Characterizing Learning Curves During Language Model Pre-Training: Learning, Forgetting, and Stability
von: Chang, Tyler A., et al.
Veröffentlicht: (2023)