Plan Before You Trade: Inference-Time Optimization for RL Trading Agents
Fuente:
arXiv
Saved in:
| Main Authors: | Go, Eun, Deb, Rohan, Banerjee, Arindam |
|---|---|
| Format: | Preprint |
| Published: |
2026
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Inference Time Policy Optimization for Offline RL with Differentiable World Models
by: Deb, Rohan, et al.
Published: (2026)
by: Deb, Rohan, et al.
Published: (2026)
Conservative Contextual Bandits: Beyond Linear Representations
by: Deb, Rohan, et al.
Published: (2024)
by: Deb, Rohan, et al.
Published: (2024)
Beyond Johnson-Lindenstrauss: Uniform Bounds for Sketched Bilinear Forms
by: Deb, Rohan, et al.
Published: (2025)
by: Deb, Rohan, et al.
Published: (2025)
Pyramid MoA: A Probabilistic Framework for Cost-Optimized Anytime Inference
by: Khaled, Arindam
Published: (2026)
by: Khaled, Arindam
Published: (2026)
Finding the Sweet Spot: Trading Quality, Cost, and Speed During Inference-Time LLM Reflection
by: Butler, Jack, et al.
Published: (2025)
by: Butler, Jack, et al.
Published: (2025)
Scout Before You Attend: Sketch-and-Walk Sparse Attention for Efficient LLM Inference
by: Le, Hoang Anh Duy, et al.
Published: (2026)
by: Le, Hoang Anh Duy, et al.
Published: (2026)
Accuracy-Privacy Trade-off in the Mitigation of Membership Inference Attack in Federated Learning
by: Ahamed, Sayyed Farid, et al.
Published: (2024)
by: Ahamed, Sayyed Farid, et al.
Published: (2024)
Ensemble RL through Classifier Models: Enhancing Risk-Return Trade-offs in Trading Strategies
by: Xiong, Zheli
Published: (2025)
by: Xiong, Zheli
Published: (2025)
Optimizing Quantile-based Trading Strategies in Electricity Arbitrage
by: O'Connor, Ciaran, et al.
Published: (2024)
by: O'Connor, Ciaran, et al.
Published: (2024)
Finding Optimal Trading History in Reinforcement Learning for Stock Market Trading
by: Montazeri, Sina, et al.
Published: (2025)
by: Montazeri, Sina, et al.
Published: (2025)
Can You Break RLVER? Probing Adversarial Robustness of RL-Trained Empathetic Agents
by: K, Deeraj S, et al.
Published: (2026)
by: K, Deeraj S, et al.
Published: (2026)
Revisiting Ensemble Methods for Stock Trading and Crypto Trading Tasks at ACM ICAIF FinRL Contest 2023-2024
by: Holzer, Nikolaus, et al.
Published: (2025)
by: Holzer, Nikolaus, et al.
Published: (2025)
Pruning Extensions and Efficiency Trade-Offs for Sustainable Time Series Classification
by: Fischer, Raphael, et al.
Published: (2026)
by: Fischer, Raphael, et al.
Published: (2026)
Optimization and Generalization Guarantees for Weight Normalization
by: Cisneros-Velarde, Pedro, et al.
Published: (2024)
by: Cisneros-Velarde, Pedro, et al.
Published: (2024)
MARK: Memory Augmented Refinement of Knowledge
by: Ganguli, Anish, et al.
Published: (2025)
by: Ganguli, Anish, et al.
Published: (2025)
KMM-CP: Practical Conformal Prediction under Covariate Shift via Selective Kernel Mean Matching
by: Laghuvarapu, Siddhartha, et al.
Published: (2026)
by: Laghuvarapu, Siddhartha, et al.
Published: (2026)
TradingAgents: Multi-Agents LLM Financial Trading Framework
by: Xiao, Yijia, et al.
Published: (2024)
by: Xiao, Yijia, et al.
Published: (2024)
A Framework for Predictive Directional Trading Based on Volatility and Causal Inference
by: Letteri, Ivan
Published: (2025)
by: Letteri, Ivan
Published: (2025)
What Can You Do When You Have Zero Rewards During RL?
by: Prakash, Jatin, et al.
Published: (2025)
by: Prakash, Jatin, et al.
Published: (2025)
Why Keep Your Doubts to Yourself? Trading Visual Uncertainties in Multi-Agent Bandit Systems
by: Zhang, Jusheng, et al.
Published: (2026)
by: Zhang, Jusheng, et al.
Published: (2026)
A Framework for Empowering Reinforcement Learning Agents with Causal Analysis: Enhancing Automated Cryptocurrency Trading
by: Amirzadeh, Rasoul, et al.
Published: (2023)
by: Amirzadeh, Rasoul, et al.
Published: (2023)
Orchestration Framework for Financial Agents: From Algorithmic Trading to Agentic Trading
by: Li, Jifeng, et al.
Published: (2025)
by: Li, Jifeng, et al.
Published: (2025)
MatKV: Trading Compute for Flash Storage in LLM Inference
by: Shin, Kun-Woo, et al.
Published: (2025)
by: Shin, Kun-Woo, et al.
Published: (2025)
A Tractable Inference Perspective of Offline RL
by: Liu, Xuejie, et al.
Published: (2023)
by: Liu, Xuejie, et al.
Published: (2023)
PhyPlan: Generalizable and Rapid Physical Task Planning with Physics Informed Skill Networks for Robot Manipulators
by: Chopra, Mudit, et al.
Published: (2024)
by: Chopra, Mudit, et al.
Published: (2024)
Capacity-Aware Planning and Scheduling in Budget-Constrained Multi-Agent MDPs: A Meta-RL Approach
by: Vora, Manav, et al.
Published: (2024)
by: Vora, Manav, et al.
Published: (2024)
Fairness-Accuracy Trade-Offs: A Causal Perspective
by: Plecko, Drago, et al.
Published: (2024)
by: Plecko, Drago, et al.
Published: (2024)
Trading-off Accuracy and Communication Cost in Federated Learning
by: Villani, Mattia Jacopo, et al.
Published: (2025)
by: Villani, Mattia Jacopo, et al.
Published: (2025)
Zero-Shot LLMs in Human-in-the-Loop RL: Replacing Human Feedback for Reward Shaping
by: Nazir, Mohammad Saif, et al.
Published: (2025)
by: Nazir, Mohammad Saif, et al.
Published: (2025)
Replicable Bandits with UCB based Exploration
by: Deb, Rohan, et al.
Published: (2026)
by: Deb, Rohan, et al.
Published: (2026)
Look Before Leap: Look-Ahead Planning with Uncertainty in Reinforcement Learning
by: Liu, Yongshuai, et al.
Published: (2025)
by: Liu, Yongshuai, et al.
Published: (2025)
Think Before You Lie: How Reasoning Leads to Honesty
by: Yuan, Ann, et al.
Published: (2026)
by: Yuan, Ann, et al.
Published: (2026)
Think Before You Act: Decision Transformers with Working Memory
by: Kang, Jikun, et al.
Published: (2023)
by: Kang, Jikun, et al.
Published: (2023)
Agent JIT Compilation for Latency-Optimizing Web Agent Planning and Scheduling
by: Winston, Caleb, et al.
Published: (2026)
by: Winston, Caleb, et al.
Published: (2026)
AgentFlux: Decoupled Fine-Tuning & Inference for On-Device Agentic Systems
by: Kadekodi, Rohan, et al.
Published: (2025)
by: Kadekodi, Rohan, et al.
Published: (2025)
PLUM: Improving Inference Efficiency By Leveraging Repetition-Sparsity Trade-Off
by: Kuhar, Sachit, et al.
Published: (2023)
by: Kuhar, Sachit, et al.
Published: (2023)
Auditing and Generating Synthetic Data with Controllable Trust Trade-offs
by: Belgodere, Brian, et al.
Published: (2023)
by: Belgodere, Brian, et al.
Published: (2023)
Mixtraining: A Better Trade-Off Between Compute and Performance
by: Li, Zexin, et al.
Published: (2025)
by: Li, Zexin, et al.
Published: (2025)
Learning Heterogeneous Performance-Fairness Trade-offs in Federated Learning
by: Ye, Rongguang, et al.
Published: (2025)
by: Ye, Rongguang, et al.
Published: (2025)
Utilizing RNN for Real-time Cryptocurrency Price Prediction and Trading Strategy Optimization
by: Tumpa, Shamima Nasrin, et al.
Published: (2024)
by: Tumpa, Shamima Nasrin, et al.
Published: (2024)
Similar Items
-
Inference Time Policy Optimization for Offline RL with Differentiable World Models
by: Deb, Rohan, et al.
Published: (2026) -
Conservative Contextual Bandits: Beyond Linear Representations
by: Deb, Rohan, et al.
Published: (2024) -
Beyond Johnson-Lindenstrauss: Uniform Bounds for Sketched Bilinear Forms
by: Deb, Rohan, et al.
Published: (2025) -
Pyramid MoA: A Probabilistic Framework for Cost-Optimized Anytime Inference
by: Khaled, Arindam
Published: (2026) -
Finding the Sweet Spot: Trading Quality, Cost, and Speed During Inference-Time LLM Reflection
by: Butler, Jack, et al.
Published: (2025)