Using Forwards-Backwards Models to Approximate MDP Homomorphisms
Fuente:
arXiv
Saved in:
| Main Authors: | Mavor-Parker, Augustine N., Sargent, Matthew J., Pehle, Christian, Banino, Andrea, Griffin, Lewis D., Barry, Caswell |
|---|---|
| Format: | Preprint |
| Published: |
2022
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Frequency and Generalisation of Periodic Activation Functions in Reinforcement Learning
by: Mavor-Parker, Augustine N., et al.
Published: (2024)
by: Mavor-Parker, Augustine N., et al.
Published: (2024)
How to Stay Curious while Avoiding Noisy TVs using Aleatoric Uncertainty Estimation
by: Mavor-Parker, Augustine N., et al.
Published: (2021)
by: Mavor-Parker, Augustine N., et al.
Published: (2021)
PopuLoRA: Co-Evolving LLM Populations for Reasoning Self-Play
by: Castanyer, Roger Creus, et al.
Published: (2026)
by: Castanyer, Roger Creus, et al.
Published: (2026)
unix-ctf: Procedural Environments for Unix-Competence Reinforcement Learning
by: Bradway, Geoffrey, et al.
Published: (2026)
by: Bradway, Geoffrey, et al.
Published: (2026)
Thinking Forward and Backward: Effective Backward Planning with Large Language Models
by: Ren, Allen Z., et al.
Published: (2024)
by: Ren, Allen Z., et al.
Published: (2024)
An MDP Model for Censoring in Harvesting Sensors: Optimal and Approximated Solutions
by: Fernandez-Bes, Jesus, et al.
Published: (2025)
by: Fernandez-Bes, Jesus, et al.
Published: (2025)
Stability Bounds for the Unfolded Forward-Backward Algorithm
by: Chouzenoux, Emilie, et al.
Published: (2024)
by: Chouzenoux, Emilie, et al.
Published: (2024)
FB-RAG: Improving RAG with Forward and Backward Lookup
by: Chawla, Kushal, et al.
Published: (2025)
by: Chawla, Kushal, et al.
Published: (2025)
Forward-Backward Reasoning in Large Language Models for Mathematical Verification
by: Jiang, Weisen, et al.
Published: (2023)
by: Jiang, Weisen, et al.
Published: (2023)
Adult learners recall and recognition performance and affective feedback when learning from an AI-generated synthetic video
by: Li, Zoe Ruo-Yu, et al.
Published: (2024)
by: Li, Zoe Ruo-Yu, et al.
Published: (2024)
Watson & Holmes: A Naturalistic Benchmark for Comparing Human and LLM Reasoning
by: Leelawat, Thatchawin, et al.
Published: (2026)
by: Leelawat, Thatchawin, et al.
Published: (2026)
Spectral Alignment in Forward-Backward Representations via Temporal Abstraction
by: Azad, Seyed Mahdi B., et al.
Published: (2026)
by: Azad, Seyed Mahdi B., et al.
Published: (2026)
Forward versus Backward: Comparing Reasoning Objectives in Direct Preference Optimization
by: Nikzad, Murtaza, et al.
Published: (2026)
by: Nikzad, Murtaza, et al.
Published: (2026)
Align Forward, Adapt Backward: Closing the Discretization Gap in Logic Gate Networks
by: Kim, Youngsung
Published: (2026)
by: Kim, Youngsung
Published: (2026)
Breaking the Conventional Forward-Backward Tie in Neural Networks: Activation Functions
by: Troiano, Luigi, et al.
Published: (2025)
by: Troiano, Luigi, et al.
Published: (2025)
Privacy-Preserving Diffusion Model Using Homomorphic Encryption
by: Chen, Yaojian, et al.
Published: (2024)
by: Chen, Yaojian, et al.
Published: (2024)
Thinking Forward and Backward: Multi-Objective Reinforcement Learning for Retrieval-Augmented Reasoning
by: Wei, Wenda, et al.
Published: (2025)
by: Wei, Wenda, et al.
Published: (2025)
Beyond Recommendations: From Backward to Forward AI Support of Pilots' Decision-Making Process
by: Zhang, Zelun Tony, et al.
Published: (2024)
by: Zhang, Zelun Tony, et al.
Published: (2024)
A-LAMP: Agentic LLM-Based Framework for Automated MDP Modeling and Policy Generation
by: Je-Gal, Hong, et al.
Published: (2025)
by: Je-Gal, Hong, et al.
Published: (2025)
Transcript of GPT-4 playing a rogue AGI in a Matrix Game
by: Griffin, Lewis D, et al.
Published: (2024)
by: Griffin, Lewis D, et al.
Published: (2024)
Predictive representations: building blocks of intelligence
by: Carvalho, Wilka, et al.
Published: (2024)
by: Carvalho, Wilka, et al.
Published: (2024)
Backward-Friendly Optimization: Training Large Language Models with Approximate Gradients under Memory Constraints
by: Yang, Jing, et al.
Published: (2025)
by: Yang, Jing, et al.
Published: (2025)
Optimized Layerwise Approximation for Efficient Private Inference on Fully Homomorphic Encryption
by: Lee, Junghyun, et al.
Published: (2023)
by: Lee, Junghyun, et al.
Published: (2023)
A Forward and Backward Compatible Framework for Few-shot Class-incremental Pill Recognition
by: Zhang, Jinghua, et al.
Published: (2023)
by: Zhang, Jinghua, et al.
Published: (2023)
StoryEnsemble: Enabling Dynamic Exploration & Iteration in the Design Process with AI and Forward-Backward Propagation
by: Suh, Sangho, et al.
Published: (2025)
by: Suh, Sangho, et al.
Published: (2025)
Homomorphisms and Embeddings of STRIPS Planning Models
by: Lequen, Arnaud, et al.
Published: (2024)
by: Lequen, Arnaud, et al.
Published: (2024)
Synth$^2$: Boosting Visual-Language Models with Synthetic Captions and Image Embeddings
by: Sharifzadeh, Sahand, et al.
Published: (2024)
by: Sharifzadeh, Sahand, et al.
Published: (2024)
Reinforcement Learning in a Safety-Embedded MDP with Trajectory Optimization
by: Yang, Fan, et al.
Published: (2023)
by: Yang, Fan, et al.
Published: (2023)
MDP: Multidimensional Vision Model Pruning with Latency Constraint
by: Sun, Xinglong, et al.
Published: (2025)
by: Sun, Xinglong, et al.
Published: (2025)
A Context Engineering Framework for Improving Enterprise AI Agents based on Digital-Twin MDP
by: Yang, Xi, et al.
Published: (2026)
by: Yang, Xi, et al.
Published: (2026)
Learning Using a Single Forward Pass
by: Somasundaram, Aditya, et al.
Published: (2024)
by: Somasundaram, Aditya, et al.
Published: (2024)
Deep reinforcement learning for weakly coupled MDP's with continuous actions
by: Robledo, Francisco, et al.
Published: (2024)
by: Robledo, Francisco, et al.
Published: (2024)
GPT, But Backwards: Exactly Inverting Language Model Outputs
by: Skapars, Adrians, et al.
Published: (2025)
by: Skapars, Adrians, et al.
Published: (2025)
How Overconfidence in Initial Choices and Underconfidence Under Criticism Modulate Change of Mind in Large Language Models
by: Kumaran, Dharshan, et al.
Published: (2025)
by: Kumaran, Dharshan, et al.
Published: (2025)
A Factored MDP Approach To Moving Target Defense With Dynamic Threat Modeling and Cost Efficiency
by: Bose, Megha, et al.
Published: (2024)
by: Bose, Megha, et al.
Published: (2024)
Structured Extraction from Business Process Diagrams Using Vision-Language Models
by: Deka, Pritam, et al.
Published: (2025)
by: Deka, Pritam, et al.
Published: (2025)
Your Learned Constraint is Secretly a Backward Reachable Tube
by: Qadri, Mohamad, et al.
Published: (2025)
by: Qadri, Mohamad, et al.
Published: (2025)
Scaling Homomorphic Applications in Deployment
by: Marinelli, Ryan, et al.
Published: (2025)
by: Marinelli, Ryan, et al.
Published: (2025)
Backward Learning for Goal-Conditioned Policies
by: Höftmann, Marc, et al.
Published: (2023)
by: Höftmann, Marc, et al.
Published: (2023)
Forward-Cooperation-Backward (FCB) learning in a Multi-Encoding Uni-Decoding neural network architecture
by: Dutta, Prasun, et al.
Published: (2025)
by: Dutta, Prasun, et al.
Published: (2025)
Similar Items
-
Frequency and Generalisation of Periodic Activation Functions in Reinforcement Learning
by: Mavor-Parker, Augustine N., et al.
Published: (2024) -
How to Stay Curious while Avoiding Noisy TVs using Aleatoric Uncertainty Estimation
by: Mavor-Parker, Augustine N., et al.
Published: (2021) -
PopuLoRA: Co-Evolving LLM Populations for Reasoning Self-Play
by: Castanyer, Roger Creus, et al.
Published: (2026) -
unix-ctf: Procedural Environments for Unix-Competence Reinforcement Learning
by: Bradway, Geoffrey, et al.
Published: (2026) -
Thinking Forward and Backward: Effective Backward Planning with Large Language Models
by: Ren, Allen Z., et al.
Published: (2024)