Deb, R., Wright, S. J., & Banerjee, A. (2026). Inference Time Policy Optimization for Offline RL with Differentiable World Models.
Chicago Style (17th ed.) CitationDeb, Rohan, Stephen J. Wright, and Arindam Banerjee. Inference Time Policy Optimization for Offline RL with Differentiable World Models. 2026.
MLA (9th ed.) CitationDeb, Rohan, et al. Inference Time Policy Optimization for Offline RL with Differentiable World Models. 2026.
Warning: These citations may not always be 100% accurate.