One if by Land, Two if by Sea, Three if by Four Seas, and More to Come -- Values of Perception, Prediction, Communication, and Common Sense in Decision Making
Fuente:
arXiv
Guardado en:
| Autor principal: | Xu, Aolin |
|---|---|
| Formato: | Preprint |
| Publicado: |
2025
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
Capabilities and Fundamental Limits of Latent Chain-of-Thought
por: Zou, Jiaxuan, et al.
Publicado: (2026)
por: Zou, Jiaxuan, et al.
Publicado: (2026)
On the optimization dynamics of RLVR: Gradient gap and step size thresholds
por: Suk, Joe, et al.
Publicado: (2025)
por: Suk, Joe, et al.
Publicado: (2025)
Comparing Few to Rank Many: Active Human Preference Learning using Randomized Frank-Wolfe
por: Thekumparampil, Kiran Koshy, et al.
Publicado: (2024)
por: Thekumparampil, Kiran Koshy, et al.
Publicado: (2024)
Accelerating Convergence of Score-Based Diffusion Models, Provably
por: Li, Gen, et al.
Publicado: (2024)
por: Li, Gen, et al.
Publicado: (2024)
Multi-head Transformers Provably Learn Symbolic Multi-step Reasoning via Gradient Descent
por: Yang, Tong, et al.
Publicado: (2025)
por: Yang, Tong, et al.
Publicado: (2025)
Multi-beam Beamforming in RIS-aided MIMO Subject to Reradiation Mask Constraints -- Optimization and Machine Learning Design
por: Wang, Shumin, et al.
Publicado: (2025)
por: Wang, Shumin, et al.
Publicado: (2025)
More is Less: Inducing Sparsity via Overparameterization
por: Chou, Hung-Hsu, et al.
Publicado: (2021)
por: Chou, Hung-Hsu, et al.
Publicado: (2021)
Deep Learning for Sequential Decision Making under Uncertainty: Foundations, Frameworks, and Frontiers
por: Buyuktahtakin, I. Esra
Publicado: (2026)
por: Buyuktahtakin, I. Esra
Publicado: (2026)
Queueing-Aware Optimization of Reasoning Tokens for Accuracy-Latency Trade-offs in LLM Servers
por: Ozbas, Emre, et al.
Publicado: (2026)
por: Ozbas, Emre, et al.
Publicado: (2026)
Tight Regret Bounds for Bayesian Optimization in One Dimension
por: Scarlett, Jonathan
Publicado: (2018)
por: Scarlett, Jonathan
Publicado: (2018)
Feedback Control via Integrated Sensing and Communication: Uncertainty Optimisation
por: Soleymani, Touraj, et al.
Publicado: (2026)
por: Soleymani, Touraj, et al.
Publicado: (2026)
Status Updating via Integrated Sensing and Communication: Freshness Optimisation
por: Soleymani, Touraj, et al.
Publicado: (2026)
por: Soleymani, Touraj, et al.
Publicado: (2026)
OPO: Making Decision-Focused Data Acquisition Decisions
por: Peršak, Egon, et al.
Publicado: (2025)
por: Peršak, Egon, et al.
Publicado: (2025)
One Model, Two Roles: Emergent Specialization in a Shared Recurrent Transformer
por: Shen, Jucheng, et al.
Publicado: (2026)
por: Shen, Jucheng, et al.
Publicado: (2026)
From Automation to Autonomy in Smart Manufacturing: A Bayesian Optimization Framework for Modeling Multi-Objective Experimentation and Sequential Decision Making
por: Asru, Avijit Saha, et al.
Publicado: (2025)
por: Asru, Avijit Saha, et al.
Publicado: (2025)
Span-Based Optimal Sample Complexity for Weakly Communicating and General Average Reward MDPs
por: Zurek, Matthew, et al.
Publicado: (2024)
por: Zurek, Matthew, et al.
Publicado: (2024)
Exploring Applications of State Space Models and Advanced Training Techniques in Sequential Recommendations: A Comparative Study on Efficiency and Performance
por: Obozov, Mark, et al.
Publicado: (2024)
por: Obozov, Mark, et al.
Publicado: (2024)
Distributionally Robust Free Energy Principle for Decision-Making
por: Shafiei, Allahkaram, et al.
Publicado: (2025)
por: Shafiei, Allahkaram, et al.
Publicado: (2025)
Sven: Singular Value Descent as a Computationally Efficient Natural Gradient Method
por: Bright-Thonney, Samuel, et al.
Publicado: (2026)
por: Bright-Thonney, Samuel, et al.
Publicado: (2026)
One-Layer Transformer Provably Learns One-Nearest Neighbor In Context
por: Li, Zihao, et al.
Publicado: (2024)
por: Li, Zihao, et al.
Publicado: (2024)
Convexity in Disguise: A Theoretical Framework for Nonconvex Low-Rank Matrix Estimation
por: Cui, Chengyu, et al.
Publicado: (2026)
por: Cui, Chengyu, et al.
Publicado: (2026)
Improving the Validity of Decision Trees as Explanations
por: Nemecek, Jiri, et al.
Publicado: (2023)
por: Nemecek, Jiri, et al.
Publicado: (2023)
Finite-Time Minimax Bounds and an Optimal Lyapunov Policy in Queueing Control
por: Liu, Yujie, et al.
Publicado: (2025)
por: Liu, Yujie, et al.
Publicado: (2025)
MLorc: Momentum Low-rank Compression for Memory Efficient Large Language Model Adaptation
por: Shen, Wei, et al.
Publicado: (2025)
por: Shen, Wei, et al.
Publicado: (2025)
Data Uniformity Improves Training Efficiency and More, with a Convergence Framework Beyond the NTK Regime
por: Wang, Yuqing, et al.
Publicado: (2025)
por: Wang, Yuqing, et al.
Publicado: (2025)
Exact Dual Geometry of SOC-ICNN Value Functions
por: Liu, Kang, et al.
Publicado: (2026)
por: Liu, Kang, et al.
Publicado: (2026)
Relation between Value and Age of Information in Feedback Control
por: Soleymani, Touraj, et al.
Publicado: (2024)
por: Soleymani, Touraj, et al.
Publicado: (2024)
Teaching LLMs to Think Mathematically: A Critical Study of Decision-Making via Optimization
por: Abdel-Rahman, Mohammad J., et al.
Publicado: (2025)
por: Abdel-Rahman, Mohammad J., et al.
Publicado: (2025)
Decision-Focused Learning: Foundations, State of the Art, Benchmark and Future Opportunities
por: Mandi, Jayanta, et al.
Publicado: (2023)
por: Mandi, Jayanta, et al.
Publicado: (2023)
Value of Information-based Deceptive Path Planning Under Adversarial Interventions
por: Suttle, Wesley A., et al.
Publicado: (2025)
por: Suttle, Wesley A., et al.
Publicado: (2025)
Implicit Regularization of Gradient Flow on One-Layer Softmax Attention
por: Sheen, Heejune, et al.
Publicado: (2024)
por: Sheen, Heejune, et al.
Publicado: (2024)
Auto-Calibration and Biconvex Compressive Sensing with Applications to Parallel MRI
por: Ni, Yuan, et al.
Publicado: (2024)
por: Ni, Yuan, et al.
Publicado: (2024)
Beamforming Design for Integrated Sensing and Communications Using Uplink-Downlink Duality
por: Attiah, Kareem M., et al.
Publicado: (2024)
por: Attiah, Kareem M., et al.
Publicado: (2024)
Decoupled Continuous-Time Reinforcement Learning via Hamiltonian Flow
por: Nguyen, Minh
Publicado: (2026)
por: Nguyen, Minh
Publicado: (2026)
Byzantine Machine Learning: MultiKrum and an optimal notion of robustness
por: Bareilles, Gilles, et al.
Publicado: (2026)
por: Bareilles, Gilles, et al.
Publicado: (2026)
Sinkhorn Based Associative Memory Retrieval Using Spherical Hellinger Kantorovich Dynamics
por: Mustafi, Aratrika, et al.
Publicado: (2026)
por: Mustafi, Aratrika, et al.
Publicado: (2026)
Straight-Through meets Sparse Recovery: the Support Exploration Algorithm
por: Mohamed, Mimoun, et al.
Publicado: (2023)
por: Mohamed, Mimoun, et al.
Publicado: (2023)
Inverse Mixed-Integer Programming: Learning Constraints then Objective Functions
por: Kitaoka, Akira
Publicado: (2025)
por: Kitaoka, Akira
Publicado: (2025)
A Differential and Pointwise Control Approach to Reinforcement Learning
por: Nguyen, Minh, et al.
Publicado: (2024)
por: Nguyen, Minh, et al.
Publicado: (2024)
Precise gradient descent training dynamics for finite-width multi-layer neural networks
por: Han, Qiyang, et al.
Publicado: (2025)
por: Han, Qiyang, et al.
Publicado: (2025)
Ejemplares similares
-
Capabilities and Fundamental Limits of Latent Chain-of-Thought
por: Zou, Jiaxuan, et al.
Publicado: (2026) -
On the optimization dynamics of RLVR: Gradient gap and step size thresholds
por: Suk, Joe, et al.
Publicado: (2025) -
Comparing Few to Rank Many: Active Human Preference Learning using Randomized Frank-Wolfe
por: Thekumparampil, Kiran Koshy, et al.
Publicado: (2024) -
Accelerating Convergence of Score-Based Diffusion Models, Provably
por: Li, Gen, et al.
Publicado: (2024) -
Multi-head Transformers Provably Learn Symbolic Multi-step Reasoning via Gradient Descent
por: Yang, Tong, et al.
Publicado: (2025)