Toward Agents That Reason About Their Computation
Fuente:
arXiv
Guardado en:
| Autores principales: | Orenstein, Adrian, Chen, Jessica, Santos, Gwyneth Anne Delos, Sapara, Bayley, Bowling, Michael |
|---|---|
| Formato: | Preprint |
| Publicado: |
2025
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
A Method for Evaluating Hyperparameter Sensitivity in Reinforcement Learning
por: Adkins, Jacob, et al.
Publicado: (2024)
por: Adkins, Jacob, et al.
Publicado: (2024)
On the Interplay Between Sparsity and Training in Deep Reinforcement Learning
por: Davelouis, Fatima, et al.
Publicado: (2025)
por: Davelouis, Fatima, et al.
Publicado: (2025)
Proper Laplacian Representation Learning
por: Gomez, Diego, et al.
Publicado: (2023)
por: Gomez, Diego, et al.
Publicado: (2023)
Rethinking the Foundations for Continual Reinforcement Learning
por: Elelimy, Esraa, et al.
Publicado: (2025)
por: Elelimy, Esraa, et al.
Publicado: (2025)
Real-Time Recurrent Learning using Trace Units in Reinforcement Learning
por: Elelimy, Esraa, et al.
Publicado: (2024)
por: Elelimy, Esraa, et al.
Publicado: (2024)
Meta-Gradient Search Control: A Method for Improving the Efficiency of Dyna-style Planning
por: Burega, Bradley, et al.
Publicado: (2024)
por: Burega, Bradley, et al.
Publicado: (2024)
Learning to Be Cautious
por: Mohammedalamen, Montaser, et al.
Publicado: (2021)
por: Mohammedalamen, Montaser, et al.
Publicado: (2021)
SheetAgent: Towards A Generalist Agent for Spreadsheet Reasoning and Manipulation via Large Language Models
por: Chen, Yibin, et al.
Publicado: (2024)
por: Chen, Yibin, et al.
Publicado: (2024)
Posts of Peril: Detecting Information About Hazards in Text
por: Burghardt, Keith, et al.
Publicado: (2024)
por: Burghardt, Keith, et al.
Publicado: (2024)
From Entropy to Calibrated Uncertainty: Training Language Models to Reason About Uncertainty
por: Jenane, Azza, et al.
Publicado: (2026)
por: Jenane, Azza, et al.
Publicado: (2026)
Towards General Computer Control with Hierarchical Agents and Multi-Level Action Spaces
por: Dong, Zihan, et al.
Publicado: (2025)
por: Dong, Zihan, et al.
Publicado: (2025)
GR-Agent: Adaptive Graph Reasoning Agent under Incomplete Knowledge
por: Zhou, Dongzhuoran, et al.
Publicado: (2025)
por: Zhou, Dongzhuoran, et al.
Publicado: (2025)
Operational Robustness of LLMs on Code Generation
por: Paul, Debalina Ghosh, et al.
Publicado: (2026)
por: Paul, Debalina Ghosh, et al.
Publicado: (2026)
Jump Start or False Start? A Theoretical and Empirical Evaluation of LLM-initialized Bandits
por: Bayley, Adam, et al.
Publicado: (2026)
por: Bayley, Adam, et al.
Publicado: (2026)
Mitigating Value Hallucination in Dyna Planning via Multistep Predecessor Models
por: Aminmansour, Farzane, et al.
Publicado: (2020)
por: Aminmansour, Farzane, et al.
Publicado: (2020)
Simulating Environments with Reasoning Models for Agent Training
por: Li, Yuetai, et al.
Publicado: (2025)
por: Li, Yuetai, et al.
Publicado: (2025)
Beyond Binary Rewards: Training LMs to Reason About Their Uncertainty
por: Damani, Mehul, et al.
Publicado: (2025)
por: Damani, Mehul, et al.
Publicado: (2025)
Neural Bayesian Filtering
por: Solinas, Christopher, et al.
Publicado: (2025)
por: Solinas, Christopher, et al.
Publicado: (2025)
Agent Q: Advanced Reasoning and Learning for Autonomous AI Agents
por: Putta, Pranav, et al.
Publicado: (2024)
por: Putta, Pranav, et al.
Publicado: (2024)
From Imperative to Declarative: Towards LLM-friendly OS Interfaces for Boosted Computer-Use Agents
por: Wang, Yuan, et al.
Publicado: (2025)
por: Wang, Yuan, et al.
Publicado: (2025)
ML-Master: Towards AI-for-AI via Integration of Exploration and Reasoning
por: Liu, Zexi, et al.
Publicado: (2025)
por: Liu, Zexi, et al.
Publicado: (2025)
A Discordance-Aware Multimodal Framework with Multi-Agent Clinical Reasoning
por: Ahadian, Pegah, et al.
Publicado: (2026)
por: Ahadian, Pegah, et al.
Publicado: (2026)
Towards Reasonable Concept Bottleneck Models
por: Kalampalikis, Nektarios, et al.
Publicado: (2025)
por: Kalampalikis, Nektarios, et al.
Publicado: (2025)
Towards Large Reasoning Models for Agriculture
por: Zaremehrjerdi, Hossein, et al.
Publicado: (2025)
por: Zaremehrjerdi, Hossein, et al.
Publicado: (2025)
Towards a Mechanistic Understanding of Propositional Logical Reasoning in Large Language Models
por: Chen, Danchun, et al.
Publicado: (2026)
por: Chen, Danchun, et al.
Publicado: (2026)
SpecReason: Fast and Accurate Inference-Time Compute via Speculative Reasoning
por: Pan, Rui, et al.
Publicado: (2025)
por: Pan, Rui, et al.
Publicado: (2025)
AgentCaster: Reasoning-Guided Tornado Forecasting
por: Chen, Michael
Publicado: (2025)
por: Chen, Michael
Publicado: (2025)
Spatial Reasoning and Planning for Deep Embodied Agents
por: Ishida, Shu
Publicado: (2024)
por: Ishida, Shu
Publicado: (2024)
Learning to Remember: End-to-End Training of Memory Agents for Long-Context Reasoning
por: Zhang, Kehao, et al.
Publicado: (2026)
por: Zhang, Kehao, et al.
Publicado: (2026)
TxAgent: An AI Agent for Therapeutic Reasoning Across a Universe of Tools
por: Gao, Shanghua, et al.
Publicado: (2025)
por: Gao, Shanghua, et al.
Publicado: (2025)
FinSTaR: Towards Financial Reasoning with Time Series Reasoning Models
por: Lee, Seunghan, et al.
Publicado: (2026)
por: Lee, Seunghan, et al.
Publicado: (2026)
DRIVE: Modeling Skills at the Reasoning and Interaction Levels for Web Agents under Continual Learning
por: Liu, Xirui, et al.
Publicado: (2026)
por: Liu, Xirui, et al.
Publicado: (2026)
Towards Autonomous Mechanistic Reasoning in Virtual Cells
por: Jang, Yunhui, et al.
Publicado: (2026)
por: Jang, Yunhui, et al.
Publicado: (2026)
TS-Reasoner: Domain-Oriented Time Series Inference Agents for Reasoning and Automated Analysis
por: Ye, Wen, et al.
Publicado: (2024)
por: Ye, Wen, et al.
Publicado: (2024)
RLVMR: Reinforcement Learning with Verifiable Meta-Reasoning Rewards for Robust Long-Horizon Agents
por: Zhang, Zijing, et al.
Publicado: (2025)
por: Zhang, Zijing, et al.
Publicado: (2025)
Grounding Computer Use Agents on Human Demonstrations
por: Feizi, Aarash, et al.
Publicado: (2025)
por: Feizi, Aarash, et al.
Publicado: (2025)
AgentRxiv: Towards Collaborative Autonomous Research
por: Schmidgall, Samuel, et al.
Publicado: (2025)
por: Schmidgall, Samuel, et al.
Publicado: (2025)
DABstep: Data Agent Benchmark for Multi-step Reasoning
por: Egg, Alex, et al.
Publicado: (2025)
por: Egg, Alex, et al.
Publicado: (2025)
Latent State Estimation Helps UI Agents to Reason
por: Bishop, William E, et al.
Publicado: (2024)
por: Bishop, William E, et al.
Publicado: (2024)
MALT: Improving Reasoning with Multi-Agent LLM Training
por: Motwani, Sumeet Ramesh, et al.
Publicado: (2024)
por: Motwani, Sumeet Ramesh, et al.
Publicado: (2024)
Ejemplares similares
-
A Method for Evaluating Hyperparameter Sensitivity in Reinforcement Learning
por: Adkins, Jacob, et al.
Publicado: (2024) -
On the Interplay Between Sparsity and Training in Deep Reinforcement Learning
por: Davelouis, Fatima, et al.
Publicado: (2025) -
Proper Laplacian Representation Learning
por: Gomez, Diego, et al.
Publicado: (2023) -
Rethinking the Foundations for Continual Reinforcement Learning
por: Elelimy, Esraa, et al.
Publicado: (2025) -
Real-Time Recurrent Learning using Trace Units in Reinforcement Learning
por: Elelimy, Esraa, et al.
Publicado: (2024)