Dont Stop Early: Scalable Enterprise Deep Research with Controlled Information Flow and Evidence-Aware Termination
Fuente:
arXiv
Guardado en:
| Autores principales: | Choubey, Prafulla Kumar, Huang, Kung-Hsiang, Venkit, Pranav Narayanan, Zhang, Jiaxin, Vats, Vaibhav, Li, Yu, Peng, Xiangyu, Wu, Chien-Sheng |
|---|---|
| Formato: | Preprint |
| Publicado: |
2026
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
Benchmarking Deep Search over Heterogeneous Enterprise Data
por: Choubey, Prafulla Kumar, et al.
Publicado: (2025)
por: Choubey, Prafulla Kumar, et al.
Publicado: (2025)
DeepTRACE: Auditing Deep Research AI Systems for Tracking Reliability Across Citations and Evidence
por: Venkit, Pranav Narayanan, et al.
Publicado: (2025)
por: Venkit, Pranav Narayanan, et al.
Publicado: (2025)
Agentic Uncertainty Quantification
por: Zhang, Jiaxin, et al.
Publicado: (2026)
por: Zhang, Jiaxin, et al.
Publicado: (2026)
MMPersuade: A Dataset and Evaluation Framework for Multimodal Persuasion
por: Qiu, Haoyi, et al.
Publicado: (2025)
por: Qiu, Haoyi, et al.
Publicado: (2025)
InterviewSim: A Scalable Framework for Interview-Grounded Personality Simulation
por: Li, Yu, et al.
Publicado: (2026)
por: Li, Yu, et al.
Publicado: (2026)
Nudging the Boundaries of LLM Reasoning
por: Chen, Justin Chih-Yao, et al.
Publicado: (2025)
por: Chen, Justin Chih-Yao, et al.
Publicado: (2025)
Unanswerability Evaluation for Retrieval Augmented Generation
por: Peng, Xiangyu, et al.
Publicado: (2024)
por: Peng, Xiangyu, et al.
Publicado: (2024)
GTA: Generating Long-Horizon Tasks for Web Agents at Scale
por: Huang, Tenghao, et al.
Publicado: (2026)
por: Huang, Tenghao, et al.
Publicado: (2026)
Terminal Agents Suffice for Enterprise Automation
por: Bechard, Patrice, et al.
Publicado: (2026)
por: Bechard, Patrice, et al.
Publicado: (2026)
Embrace Divergence for Richer Insights: A Multi-document Summarization Benchmark and a Case Study on Summarizing Diverse Information from News Articles
por: Huang, Kung-Hsiang, et al.
Publicado: (2023)
por: Huang, Kung-Hsiang, et al.
Publicado: (2023)
The Need for a Socially-Grounded Persona Framework for User Simulation
por: Venkit, Pranav Narayanan, et al.
Publicado: (2026)
por: Venkit, Pranav Narayanan, et al.
Publicado: (2026)
Turning Conversations into Workflows: A Framework to Extract and Evaluate Dialog Workflows for Service AI Agents
por: Choubey, Prafulla Kumar, et al.
Publicado: (2025)
por: Choubey, Prafulla Kumar, et al.
Publicado: (2025)
EarlyStopping: Implicit Regularization for Iterative Learning Procedures in Python
por: Ziebell, Eric, et al.
Publicado: (2025)
por: Ziebell, Eric, et al.
Publicado: (2025)
EET: Experience-Driven Early Termination for Cost-Efficient Software Engineering Agents
por: Guo, Yaoqi, et al.
Publicado: (2026)
por: Guo, Yaoqi, et al.
Publicado: (2026)
Don't Confuse! Redrawing GUI Navigation Flow in Mobile Apps for Visually Impaired Users
por: Zhang, Mengxi, et al.
Publicado: (2025)
por: Zhang, Mengxi, et al.
Publicado: (2025)
Search Engines in an AI Era: The False Promise of Factual and Verifiable Source-Cited Responses
por: Venkit, Pranav Narayanan, et al.
Publicado: (2024)
por: Venkit, Pranav Narayanan, et al.
Publicado: (2024)
CRMArena-Pro: Holistic Assessment of LLM Agents Across Diverse Business Scenarios and Interactions
por: Huang, Kung-Hsiang, et al.
Publicado: (2025)
por: Huang, Kung-Hsiang, et al.
Publicado: (2025)
Enhancing Inventory Management with Progressive Web Applications (PWAs): A Scalable Solution for Small and Large Enterprises
por: Desai, Abhi
Publicado: (2025)
por: Desai, Abhi
Publicado: (2025)
Multi-CoLoR: Context-Aware Localization and Reasoning across Multi-Language Codebases
por: Vats, Indira, et al.
Publicado: (2026)
por: Vats, Indira, et al.
Publicado: (2026)
Do RAG Systems Cover What Matters? Evaluating and Optimizing Responses with Sub-Question Coverage
por: Xie, Kaige, et al.
Publicado: (2024)
por: Xie, Kaige, et al.
Publicado: (2024)
Zoom, Don't Wander: Why Regional Search Outperforms Pareto Reasoning and Global Optimization in Budget-Constrained SBSE
por: Ganguly, Kishan Kumar, et al.
Publicado: (2026)
por: Ganguly, Kishan Kumar, et al.
Publicado: (2026)
PlanCompiler: A Deterministic Compilation Architecture for Structured Multi-Step LLM Pipelines
por: Harikumar, Pranav
Publicado: (2026)
por: Harikumar, Pranav
Publicado: (2026)
Studying the Impact of Early Test Termination Due to Assertion Failure on Code Coverage and Spectrum-based Fault Localization
por: Uddin, Md. Ashraf, et al.
Publicado: (2025)
por: Uddin, Md. Ashraf, et al.
Publicado: (2025)
From Melting Pots to Misrepresentations: Exploring Harms in Generative AI
por: Gautam, Sanjana, et al.
Publicado: (2024)
por: Gautam, Sanjana, et al.
Publicado: (2024)
Scalable Similarity-Aware Test Suite Minimization with Reinforcement Learning
por: Gu, Sijia, et al.
Publicado: (2024)
por: Gu, Sijia, et al.
Publicado: (2024)
Coding Agents Don't Know When to Act
por: Gloaguen, Thibaud, et al.
Publicado: (2026)
por: Gloaguen, Thibaud, et al.
Publicado: (2026)
DeltaMCP: Incremental Regeneration via Spec-Aware Transformation for MCP servers
por: Pujara, Aditya, et al.
Publicado: (2026)
por: Pujara, Aditya, et al.
Publicado: (2026)
EditFlow: Benchmarking and Optimizing Code Edit Recommendation Systems via Reconstruction of Developer Flows
por: Liu, Chenyan, et al.
Publicado: (2026)
por: Liu, Chenyan, et al.
Publicado: (2026)
AI Agentic workflows and Enterprise APIs: Adapting API architectures for the age of AI agents
por: Tupe, Vaibhav, et al.
Publicado: (2025)
por: Tupe, Vaibhav, et al.
Publicado: (2025)
How I Learned to Stop Worrying and Love ChatGPT
por: Przymus, Piotr, et al.
Publicado: (2025)
por: Przymus, Piotr, et al.
Publicado: (2025)
UTFix: Change Aware Unit Test Repairing using LLM
por: Rahman, Shanto, et al.
Publicado: (2025)
por: Rahman, Shanto, et al.
Publicado: (2025)
Challenges and Experiences of Iranian Developers with MLOps at Enterprise
por: Heydari, Mohammad, et al.
Publicado: (2024)
por: Heydari, Mohammad, et al.
Publicado: (2024)
Detecting Metadata-Related Bugs in Enterprise Applications
por: Kabir, Md Mahir Asef, et al.
Publicado: (2025)
por: Kabir, Md Mahir Asef, et al.
Publicado: (2025)
Intuition to Evidence: Measuring AI's True Impact on Developer Productivity
por: Kumar, Anand, et al.
Publicado: (2025)
por: Kumar, Anand, et al.
Publicado: (2025)
FLOW-BENCH: Towards Conversational Generation of Enterprise Workflows
por: Duesterwald, Evelyn, et al.
Publicado: (2025)
por: Duesterwald, Evelyn, et al.
Publicado: (2025)
SieveFL: Hierarchical Runtime-Aware Pruning for Scalable LLM-Based Fault Localization
por: Farzandway, Mahdi, et al.
Publicado: (2026)
por: Farzandway, Mahdi, et al.
Publicado: (2026)
SECRET: Towards Scalable and Efficient Code Retrieval via Segmented Deep Hashing
por: Gu, Wenchao, et al.
Publicado: (2024)
por: Gu, Wenchao, et al.
Publicado: (2024)
Design and Implementation of a Multiuser Enterprise Resource Planning Solution for Higher Education Institutions: Enhancing Accreditation and Administration
por: Divyansh Bansal, et al.
Publicado: (2026)
por: Divyansh Bansal, et al.
Publicado: (2026)
Efficient Computation of Collatz Sequence Stopping Times: A Novel Algorithmic Approach
por: Getachew, Eyob Solomon, et al.
Publicado: (2025)
por: Getachew, Eyob Solomon, et al.
Publicado: (2025)
When to Stop? Towards Efficient Code Generation in LLMs with Excess Token Prevention
por: Guo, Lianghong, et al.
Publicado: (2024)
por: Guo, Lianghong, et al.
Publicado: (2024)
Ejemplares similares
-
Benchmarking Deep Search over Heterogeneous Enterprise Data
por: Choubey, Prafulla Kumar, et al.
Publicado: (2025) -
DeepTRACE: Auditing Deep Research AI Systems for Tracking Reliability Across Citations and Evidence
por: Venkit, Pranav Narayanan, et al.
Publicado: (2025) -
Agentic Uncertainty Quantification
por: Zhang, Jiaxin, et al.
Publicado: (2026) -
MMPersuade: A Dataset and Evaluation Framework for Multimodal Persuasion
por: Qiu, Haoyi, et al.
Publicado: (2025) -
InterviewSim: A Scalable Framework for Interview-Grounded Personality Simulation
por: Li, Yu, et al.
Publicado: (2026)