Online Decision Making with Generative Action Sets
Fuente:
arXiv
Saved in:
| Main Authors: | Xu, Jianyu, Jain, Vidhi, Wilder, Bryan, Singh, Aarti |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Frequency Principle: Fourier Analysis Sheds Light on Deep Neural Networks
by: Xu, Zhi-Qin John, et al.
Published: (2019)
by: Xu, Zhi-Qin John, et al.
Published: (2019)
Regret Bounds for Robust Online Decision Making
by: Appel, Alexander, et al.
Published: (2025)
by: Appel, Alexander, et al.
Published: (2025)
Artificial Neural Networks on Graded Vector Spaces
by: Shaska, Tony
Published: (2024)
by: Shaska, Tony
Published: (2024)
A Reduction from Delayed to Immediate Feedback for Online Convex Optimization with Improved Guarantees
by: Ryabchenko, Alexander, et al.
Published: (2026)
by: Ryabchenko, Alexander, et al.
Published: (2026)
Multipole Semantic Attention: A Fast Approximation of Softmax Attention for Pretraining
by: Mitchell, Rupert, et al.
Published: (2025)
by: Mitchell, Rupert, et al.
Published: (2025)
The Current and Future Perspectives of Zinc Oxide Nanoparticles in the Treatment of Diabetes Mellitus
by: Yousaf, Iqra
Published: (2024)
by: Yousaf, Iqra
Published: (2024)
Ambiguous Online Learning
by: Kosoy, Vanessa
Published: (2025)
by: Kosoy, Vanessa
Published: (2025)
Backpropagation Through Time For Networks With Long-Term Dependencies
by: Bird, George, et al.
Published: (2021)
by: Bird, George, et al.
Published: (2021)
Adaptive Discretization in Online Reinforcement Learning
by: Sinclair, Sean R., et al.
Published: (2021)
by: Sinclair, Sean R., et al.
Published: (2021)
Decision Making under Imperfect Recall: Algorithms and Benchmarks
by: Tewolde, Emanuel, et al.
Published: (2026)
by: Tewolde, Emanuel, et al.
Published: (2026)
A First Runtime Analysis of the PAES-25: An Enhanced Variant of the Pareto Archived Evolution Strategy
by: Opris, Andre
Published: (2025)
by: Opris, Andre
Published: (2025)
Promoting Fair Online Resource Allocation with Indivisible Units
by: Averbakh, Igor, et al.
Published: (2026)
by: Averbakh, Igor, et al.
Published: (2026)
Understanding the Uncertainty of LLM Explanations: A Perspective Based on Reasoning Topology
by: Da, Longchao, et al.
Published: (2025)
by: Da, Longchao, et al.
Published: (2025)
Position: Uncertainty Quantification in LLMs is Just Unsupervised Clustering
by: Chen, Tiejin, et al.
Published: (2026)
by: Chen, Tiejin, et al.
Published: (2026)
Minimizing Maximum Dissatisfaction in the Allocation of Indivisible Items under a Common Preference Graph
by: Chiarelli, Nina, et al.
Published: (2023)
by: Chiarelli, Nina, et al.
Published: (2023)
NeurOptimisation: The Spiking Way to Evolve
by: Cruz-Duarte, Jorge Mario, et al.
Published: (2025)
by: Cruz-Duarte, Jorge Mario, et al.
Published: (2025)
Superior Scoring Rules for Probabilistic Evaluation of Single-Label Multi-Class Classification Tasks
by: Ahmadian, Rouhollah, et al.
Published: (2024)
by: Ahmadian, Rouhollah, et al.
Published: (2024)
On the Edge of Core (Non-)Emptiness: An Automated Reasoning Approach to Approval-Based Multi-Winner Voting
by: Berker, Ratip Emin, et al.
Published: (2025)
by: Berker, Ratip Emin, et al.
Published: (2025)
Aligning Inductive Bias for Data-Efficient Generalization in State Space Models
by: Chen, Qiyu, et al.
Published: (2025)
by: Chen, Qiyu, et al.
Published: (2025)
Socio-cognitive agent-oriented evolutionary algorithm with trust-based optimization
by: Urbańczyk, Aleksandra, et al.
Published: (2025)
by: Urbańczyk, Aleksandra, et al.
Published: (2025)
New Theoretical Insights and Algorithmic Solutions for Reconstructing Score Sequences from Tournament Score Sets
by: Liu, Bowen
Published: (2025)
by: Liu, Bowen
Published: (2025)
EvoPref: Multi-Objective Evolutionary Optimization Discovers Diverse LLM Alignments Beyond Gradient Descent
by: Guo, Dongxin, et al.
Published: (2026)
by: Guo, Dongxin, et al.
Published: (2026)
Parameter-Efficient Neuroevolution for Diverse LLM Generation: Quality-Diversity Optimization via Prompt Embedding Evolution
by: Guo, Dongxin, et al.
Published: (2026)
by: Guo, Dongxin, et al.
Published: (2026)
Separate Before You Compress: The WWHO Tokenization Architecture
by: Darshana, Kusal
Published: (2026)
by: Darshana, Kusal
Published: (2026)
Extracting and Validating Explanatory Word Archipelagoes using Dual Entropy
by: Ohsawa, Yukio
Published: (2020)
by: Ohsawa, Yukio
Published: (2020)
Can Synthetic Data Improve Symbolic Regression Extrapolation Performance?
by: Ramlan, Fitria Wulandari, et al.
Published: (2025)
by: Ramlan, Fitria Wulandari, et al.
Published: (2025)
Boosting Test Performance with Importance Sampling--a Subpopulation Perspective
by: Shen, Hongyu, et al.
Published: (2024)
by: Shen, Hongyu, et al.
Published: (2024)
Retrieval-Augmented Memory for Online Learning
by: Du, Wenzhang
Published: (2025)
by: Du, Wenzhang
Published: (2025)
Optimistic Feasible Search for Closed-Loop Fair Threshold Decision-Making
by: Du, Wenzhang
Published: (2025)
by: Du, Wenzhang
Published: (2025)
Diagnosing Multi-step Reasoning Failures in Black-box LLMs via Stepwise Confidence Attribution
by: Liu, Xiaoou, et al.
Published: (2026)
by: Liu, Xiaoou, et al.
Published: (2026)
InhibiDistilbert: Knowledge Distillation for a ReLU and Addition-based Transformer
by: Zhang, Tony, et al.
Published: (2025)
by: Zhang, Tony, et al.
Published: (2025)
CELI: Controller-Embedded Language Model Interactions
by: Wagner, Jan-Samuel, et al.
Published: (2024)
by: Wagner, Jan-Samuel, et al.
Published: (2024)
Random feature-based double Vovk-Azoury-Warmuth algorithm for online multi-kernel learning
by: Rokhlin, Dmitry B., et al.
Published: (2025)
by: Rokhlin, Dmitry B., et al.
Published: (2025)
A hierarchical Vovk-Azoury-Warmuth forecaster with discounting for online regression in RKHS
by: Rokhlin, Dmitry B.
Published: (2025)
by: Rokhlin, Dmitry B.
Published: (2025)
Generating Realistic Safety-Critical Scenarios for Vehicle-Pedestrian Interactions
by: Pu, Qingwen, et al.
Published: (2026)
by: Pu, Qingwen, et al.
Published: (2026)
Optimal phase change for a generalized Grover's algorithm
by: Cardullo, Christopher, et al.
Published: (2025)
by: Cardullo, Christopher, et al.
Published: (2025)
A Theoretical Study of (Hyper) Self-Attention through the Lens of Interactions: Representation, Training, Generalization
by: Ustaomeroglu, Muhammed, et al.
Published: (2025)
by: Ustaomeroglu, Muhammed, et al.
Published: (2025)
Understanding the Nature of Generative AI as Threshold Logic in High-Dimensional Space
by: Levin, Ilya
Published: (2026)
by: Levin, Ilya
Published: (2026)
Navigational Thinking as an Emerging Paradigm of Computer Science in the Age of Generative AI
by: Levin, Ilya
Published: (2026)
by: Levin, Ilya
Published: (2026)
Sparse Knowledge Distillation: A Mathematical Framework for Probability-Domain Temperature Scaling and Multi-Stage Compression
by: Flouro, Aaron R., et al.
Published: (2026)
by: Flouro, Aaron R., et al.
Published: (2026)
Similar Items
-
Frequency Principle: Fourier Analysis Sheds Light on Deep Neural Networks
by: Xu, Zhi-Qin John, et al.
Published: (2019) -
Regret Bounds for Robust Online Decision Making
by: Appel, Alexander, et al.
Published: (2025) -
Artificial Neural Networks on Graded Vector Spaces
by: Shaska, Tony
Published: (2024) -
A Reduction from Delayed to Immediate Feedback for Online Convex Optimization with Improved Guarantees
by: Ryabchenko, Alexander, et al.
Published: (2026) -
Multipole Semantic Attention: A Fast Approximation of Softmax Attention for Pretraining
by: Mitchell, Rupert, et al.
Published: (2025)