One if by Land, Two if by Sea, Three if by Four Seas, and More to Come -- Values of Perception, Prediction, Communication, and Common Sense in Decision Making
Fuente:
arXiv
Saved in:
| Main Author: | Xu, Aolin |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Capabilities and Fundamental Limits of Latent Chain-of-Thought
by: Zou, Jiaxuan, et al.
Published: (2026)
by: Zou, Jiaxuan, et al.
Published: (2026)
On the optimization dynamics of RLVR: Gradient gap and step size thresholds
by: Suk, Joe, et al.
Published: (2025)
by: Suk, Joe, et al.
Published: (2025)
Comparing Few to Rank Many: Active Human Preference Learning using Randomized Frank-Wolfe
by: Thekumparampil, Kiran Koshy, et al.
Published: (2024)
by: Thekumparampil, Kiran Koshy, et al.
Published: (2024)
Accelerating Convergence of Score-Based Diffusion Models, Provably
by: Li, Gen, et al.
Published: (2024)
by: Li, Gen, et al.
Published: (2024)
Multi-head Transformers Provably Learn Symbolic Multi-step Reasoning via Gradient Descent
by: Yang, Tong, et al.
Published: (2025)
by: Yang, Tong, et al.
Published: (2025)
Multi-beam Beamforming in RIS-aided MIMO Subject to Reradiation Mask Constraints -- Optimization and Machine Learning Design
by: Wang, Shumin, et al.
Published: (2025)
by: Wang, Shumin, et al.
Published: (2025)
More is Less: Inducing Sparsity via Overparameterization
by: Chou, Hung-Hsu, et al.
Published: (2021)
by: Chou, Hung-Hsu, et al.
Published: (2021)
Deep Learning for Sequential Decision Making under Uncertainty: Foundations, Frameworks, and Frontiers
by: Buyuktahtakin, I. Esra
Published: (2026)
by: Buyuktahtakin, I. Esra
Published: (2026)
Queueing-Aware Optimization of Reasoning Tokens for Accuracy-Latency Trade-offs in LLM Servers
by: Ozbas, Emre, et al.
Published: (2026)
by: Ozbas, Emre, et al.
Published: (2026)
Tight Regret Bounds for Bayesian Optimization in One Dimension
by: Scarlett, Jonathan
Published: (2018)
by: Scarlett, Jonathan
Published: (2018)
Feedback Control via Integrated Sensing and Communication: Uncertainty Optimisation
by: Soleymani, Touraj, et al.
Published: (2026)
by: Soleymani, Touraj, et al.
Published: (2026)
Status Updating via Integrated Sensing and Communication: Freshness Optimisation
by: Soleymani, Touraj, et al.
Published: (2026)
by: Soleymani, Touraj, et al.
Published: (2026)
OPO: Making Decision-Focused Data Acquisition Decisions
by: Peršak, Egon, et al.
Published: (2025)
by: Peršak, Egon, et al.
Published: (2025)
One Model, Two Roles: Emergent Specialization in a Shared Recurrent Transformer
by: Shen, Jucheng, et al.
Published: (2026)
by: Shen, Jucheng, et al.
Published: (2026)
From Automation to Autonomy in Smart Manufacturing: A Bayesian Optimization Framework for Modeling Multi-Objective Experimentation and Sequential Decision Making
by: Asru, Avijit Saha, et al.
Published: (2025)
by: Asru, Avijit Saha, et al.
Published: (2025)
Span-Based Optimal Sample Complexity for Weakly Communicating and General Average Reward MDPs
by: Zurek, Matthew, et al.
Published: (2024)
by: Zurek, Matthew, et al.
Published: (2024)
Exploring Applications of State Space Models and Advanced Training Techniques in Sequential Recommendations: A Comparative Study on Efficiency and Performance
by: Obozov, Mark, et al.
Published: (2024)
by: Obozov, Mark, et al.
Published: (2024)
Distributionally Robust Free Energy Principle for Decision-Making
by: Shafiei, Allahkaram, et al.
Published: (2025)
by: Shafiei, Allahkaram, et al.
Published: (2025)
Sven: Singular Value Descent as a Computationally Efficient Natural Gradient Method
by: Bright-Thonney, Samuel, et al.
Published: (2026)
by: Bright-Thonney, Samuel, et al.
Published: (2026)
One-Layer Transformer Provably Learns One-Nearest Neighbor In Context
by: Li, Zihao, et al.
Published: (2024)
by: Li, Zihao, et al.
Published: (2024)
Convexity in Disguise: A Theoretical Framework for Nonconvex Low-Rank Matrix Estimation
by: Cui, Chengyu, et al.
Published: (2026)
by: Cui, Chengyu, et al.
Published: (2026)
Improving the Validity of Decision Trees as Explanations
by: Nemecek, Jiri, et al.
Published: (2023)
by: Nemecek, Jiri, et al.
Published: (2023)
Finite-Time Minimax Bounds and an Optimal Lyapunov Policy in Queueing Control
by: Liu, Yujie, et al.
Published: (2025)
by: Liu, Yujie, et al.
Published: (2025)
MLorc: Momentum Low-rank Compression for Memory Efficient Large Language Model Adaptation
by: Shen, Wei, et al.
Published: (2025)
by: Shen, Wei, et al.
Published: (2025)
Data Uniformity Improves Training Efficiency and More, with a Convergence Framework Beyond the NTK Regime
by: Wang, Yuqing, et al.
Published: (2025)
by: Wang, Yuqing, et al.
Published: (2025)
Exact Dual Geometry of SOC-ICNN Value Functions
by: Liu, Kang, et al.
Published: (2026)
by: Liu, Kang, et al.
Published: (2026)
Relation between Value and Age of Information in Feedback Control
by: Soleymani, Touraj, et al.
Published: (2024)
by: Soleymani, Touraj, et al.
Published: (2024)
Teaching LLMs to Think Mathematically: A Critical Study of Decision-Making via Optimization
by: Abdel-Rahman, Mohammad J., et al.
Published: (2025)
by: Abdel-Rahman, Mohammad J., et al.
Published: (2025)
Decision-Focused Learning: Foundations, State of the Art, Benchmark and Future Opportunities
by: Mandi, Jayanta, et al.
Published: (2023)
by: Mandi, Jayanta, et al.
Published: (2023)
Value of Information-based Deceptive Path Planning Under Adversarial Interventions
by: Suttle, Wesley A., et al.
Published: (2025)
by: Suttle, Wesley A., et al.
Published: (2025)
Implicit Regularization of Gradient Flow on One-Layer Softmax Attention
by: Sheen, Heejune, et al.
Published: (2024)
by: Sheen, Heejune, et al.
Published: (2024)
Auto-Calibration and Biconvex Compressive Sensing with Applications to Parallel MRI
by: Ni, Yuan, et al.
Published: (2024)
by: Ni, Yuan, et al.
Published: (2024)
Beamforming Design for Integrated Sensing and Communications Using Uplink-Downlink Duality
by: Attiah, Kareem M., et al.
Published: (2024)
by: Attiah, Kareem M., et al.
Published: (2024)
Decoupled Continuous-Time Reinforcement Learning via Hamiltonian Flow
by: Nguyen, Minh
Published: (2026)
by: Nguyen, Minh
Published: (2026)
Byzantine Machine Learning: MultiKrum and an optimal notion of robustness
by: Bareilles, Gilles, et al.
Published: (2026)
by: Bareilles, Gilles, et al.
Published: (2026)
Sinkhorn Based Associative Memory Retrieval Using Spherical Hellinger Kantorovich Dynamics
by: Mustafi, Aratrika, et al.
Published: (2026)
by: Mustafi, Aratrika, et al.
Published: (2026)
Straight-Through meets Sparse Recovery: the Support Exploration Algorithm
by: Mohamed, Mimoun, et al.
Published: (2023)
by: Mohamed, Mimoun, et al.
Published: (2023)
Inverse Mixed-Integer Programming: Learning Constraints then Objective Functions
by: Kitaoka, Akira
Published: (2025)
by: Kitaoka, Akira
Published: (2025)
A Differential and Pointwise Control Approach to Reinforcement Learning
by: Nguyen, Minh, et al.
Published: (2024)
by: Nguyen, Minh, et al.
Published: (2024)
Precise gradient descent training dynamics for finite-width multi-layer neural networks
by: Han, Qiyang, et al.
Published: (2025)
by: Han, Qiyang, et al.
Published: (2025)
Similar Items
-
Capabilities and Fundamental Limits of Latent Chain-of-Thought
by: Zou, Jiaxuan, et al.
Published: (2026) -
On the optimization dynamics of RLVR: Gradient gap and step size thresholds
by: Suk, Joe, et al.
Published: (2025) -
Comparing Few to Rank Many: Active Human Preference Learning using Randomized Frank-Wolfe
by: Thekumparampil, Kiran Koshy, et al.
Published: (2024) -
Accelerating Convergence of Score-Based Diffusion Models, Provably
by: Li, Gen, et al.
Published: (2024) -
Multi-head Transformers Provably Learn Symbolic Multi-step Reasoning via Gradient Descent
by: Yang, Tong, et al.
Published: (2025)