Highway Value Iteration Networks
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Wang, Yuhui, Li, Weida, Faccio, Francesco, Wu, Qingyuan, Schmidhuber, Jürgen |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2024
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Scaling Value Iteration Networks to 5000 Layers for Extreme Long-Term Planning
von: Wang, Yuhui, et al.
Veröffentlicht: (2024)
von: Wang, Yuhui, et al.
Veröffentlicht: (2024)
Highway Reinforcement Learning
von: Wang, Yuhui, et al.
Veröffentlicht: (2024)
von: Wang, Yuhui, et al.
Veröffentlicht: (2024)
Curious Causality-Seeking Agents Learn Meta Causal World
von: Zhao, Zhiyu, et al.
Veröffentlicht: (2025)
von: Zhao, Zhiyu, et al.
Veröffentlicht: (2025)
Language Agents as Optimizable Graphs
von: Zhuge, Mingchen, et al.
Veröffentlicht: (2024)
von: Zhuge, Mingchen, et al.
Veröffentlicht: (2024)
Upside Down Reinforcement Learning with Policy Generators
von: Di Ventura, Jacopo, et al.
Veröffentlicht: (2025)
von: Di Ventura, Jacopo, et al.
Veröffentlicht: (2025)
Learning Useful Representations of Recurrent Neural Network Weight Matrices
von: Herrmann, Vincent, et al.
Veröffentlicht: (2024)
von: Herrmann, Vincent, et al.
Veröffentlicht: (2024)
Boosting Reinforcement Learning with Strongly Delayed Feedback Through Auxiliary Short Delays
von: Wu, Qingyuan, et al.
Veröffentlicht: (2024)
von: Wu, Qingyuan, et al.
Veröffentlicht: (2024)
Towards a Robust Soft Baby Robot With Rich Interaction Ability for Advanced Machine Learning Algorithms
von: Alhakami, Mohannad, et al.
Veröffentlicht: (2024)
von: Alhakami, Mohannad, et al.
Veröffentlicht: (2024)
FACTS: A Factored State-Space Framework For World Modelling
von: Nanbo, Li, et al.
Veröffentlicht: (2024)
von: Nanbo, Li, et al.
Veröffentlicht: (2024)
Interestingness as an Inductive Heuristic for Future Compression Progress
von: Herrmann, Vincent, et al.
Veröffentlicht: (2026)
von: Herrmann, Vincent, et al.
Veröffentlicht: (2026)
MeSH: Memory-as-State-Highways for Recursive Transformers
von: Yu, Chengting, et al.
Veröffentlicht: (2025)
von: Yu, Chengting, et al.
Veröffentlicht: (2025)
Efficient Morphology-Control Co-Design via Stackelberg Proximal Policy Optimization
von: Dai, Yanning, et al.
Veröffentlicht: (2026)
von: Dai, Yanning, et al.
Veröffentlicht: (2026)
Sequence Compression Speeds Up Credit Assignment in Reinforcement Learning
von: Ramesh, Aditya A., et al.
Veröffentlicht: (2024)
von: Ramesh, Aditya A., et al.
Veröffentlicht: (2024)
Pessimistic Value Iteration for Multi-Task Data Sharing in Offline Reinforcement Learning
von: Bai, Chenjia, et al.
Veröffentlicht: (2024)
von: Bai, Chenjia, et al.
Veröffentlicht: (2024)
Variational Delayed Policy Optimization
von: Wu, Qingyuan, et al.
Veröffentlicht: (2024)
von: Wu, Qingyuan, et al.
Veröffentlicht: (2024)
Cross-Modal Reconstruction Pretraining for Ramp Flow Prediction at Highway Interchanges
von: Li, Yongchao, et al.
Veröffentlicht: (2025)
von: Li, Yongchao, et al.
Veröffentlicht: (2025)
On the Convergence and Stability of Upside-Down Reinforcement Learning, Goal-Conditioned Supervised Learning, and Online Decision Transformers
von: Štrupl, Miroslav, et al.
Veröffentlicht: (2025)
von: Štrupl, Miroslav, et al.
Veröffentlicht: (2025)
Fast and scalable retrosynthetic planning with a transformer neural network and speculative beam search
von: Andronov, Mikhail, et al.
Veröffentlicht: (2025)
von: Andronov, Mikhail, et al.
Veröffentlicht: (2025)
Fairness Overfitting in Machine Learning: An Information-Theoretic Perspective
von: Laakom, Firas, et al.
Veröffentlicht: (2025)
von: Laakom, Firas, et al.
Veröffentlicht: (2025)
How to Correctly do Semantic Backpropagation on Language-based Agentic Systems
von: Wang, Wenyi, et al.
Veröffentlicht: (2024)
von: Wang, Wenyi, et al.
Veröffentlicht: (2024)
A Unified Framework for Rethinking Policy Divergence Measures in GRPO
von: Wu, Qingyuan, et al.
Veröffentlicht: (2026)
von: Wu, Qingyuan, et al.
Veröffentlicht: (2026)
Stop-RAG: Value-Based Retrieval Control for Iterative RAG
von: Park, Jaewan, et al.
Veröffentlicht: (2025)
von: Park, Jaewan, et al.
Veröffentlicht: (2025)
Hybrid LSTM-Transformer Models for Profiling Highway-Railway Grade Crossings
von: Chatterjee, Kaustav, et al.
Veröffentlicht: (2025)
von: Chatterjee, Kaustav, et al.
Veröffentlicht: (2025)
Decoupling the "What" and "Where" With Polar Coordinate Positional Embeddings
von: Gopalakrishnan, Anand, et al.
Veröffentlicht: (2025)
von: Gopalakrishnan, Anand, et al.
Veröffentlicht: (2025)
Evolutionary Guided Decoding: Iterative Value Refinement for LLMs
von: Liu, Zhenhua, et al.
Veröffentlicht: (2025)
von: Liu, Zhenhua, et al.
Veröffentlicht: (2025)
Concurrent Learning with Aggregated States via Randomized Least Squares Value Iteration
von: Chen, Yan, et al.
Veröffentlicht: (2025)
von: Chen, Yan, et al.
Veröffentlicht: (2025)
Multi-View Subgraph Neural Networks: Self-Supervised Learning with Scarce Labeled Data
von: Wang, Zhenzhong, et al.
Veröffentlicht: (2024)
von: Wang, Zhenzhong, et al.
Veröffentlicht: (2024)
PhysGym: Benchmarking LLMs in Interactive Physics Discovery with Controlled Priors
von: Chen, Yimeng, et al.
Veröffentlicht: (2025)
von: Chen, Yimeng, et al.
Veröffentlicht: (2025)
Counterfactual explainability and analysis of variance
von: Gao, Zijun, et al.
Veröffentlicht: (2024)
von: Gao, Zijun, et al.
Veröffentlicht: (2024)
Advancing Autonomous VLM Agents via Variational Subgoal-Conditioned Reinforcement Learning
von: Wu, Qingyuan, et al.
Veröffentlicht: (2025)
von: Wu, Qingyuan, et al.
Veröffentlicht: (2025)
Resolve Highway Conflict in Multi-Autonomous Vehicle Controls with Local State Attention
von: Ta, Xuan Duy, et al.
Veröffentlicht: (2025)
von: Ta, Xuan Duy, et al.
Veröffentlicht: (2025)
Accelerating the inference of string generation-based chemical reaction models for industrial applications
von: Andronov, Mikhail, et al.
Veröffentlicht: (2024)
von: Andronov, Mikhail, et al.
Veröffentlicht: (2024)
Set-Valued Sensitivity Analysis of Deep Neural Networks
von: Wang, Xin, et al.
Veröffentlicht: (2024)
von: Wang, Xin, et al.
Veröffentlicht: (2024)
Heterogeneous Self-Play for Realistic Highway Traffic Simulation
von: Qiu, Jinkai, et al.
Veröffentlicht: (2026)
von: Qiu, Jinkai, et al.
Veröffentlicht: (2026)
Significativity Indices for Agreement Values
von: Casagrande, Alberto, et al.
Veröffentlicht: (2025)
von: Casagrande, Alberto, et al.
Veröffentlicht: (2025)
VC-Soup: Value-Consistency Guided Multi-Value Alignment for Large Language Models
von: Xu, Hefei, et al.
Veröffentlicht: (2026)
von: Xu, Hefei, et al.
Veröffentlicht: (2026)
Shift-Invariant Attribute Scoring for Kolmogorov-Arnold Networks via Shapley Value
von: Fan, Wangxuan, et al.
Veröffentlicht: (2025)
von: Fan, Wangxuan, et al.
Veröffentlicht: (2025)
Iterative Inference in a Chess-Playing Neural Network
von: Sandmann, Elias, et al.
Veröffentlicht: (2025)
von: Sandmann, Elias, et al.
Veröffentlicht: (2025)
Multi-Scenario Highway Lane-Change Intention Prediction: A Physics-Informed AI Framework for Three-Class Classification
von: Shi, Jiazhao, et al.
Veröffentlicht: (2025)
von: Shi, Jiazhao, et al.
Veröffentlicht: (2025)
Highway Networks for Improved Surface Reconstruction: The Role of Residuals and Weight Updates
von: Noorizadegan, A., et al.
Veröffentlicht: (2024)
von: Noorizadegan, A., et al.
Veröffentlicht: (2024)
Ähnliche Einträge
-
Scaling Value Iteration Networks to 5000 Layers for Extreme Long-Term Planning
von: Wang, Yuhui, et al.
Veröffentlicht: (2024) -
Highway Reinforcement Learning
von: Wang, Yuhui, et al.
Veröffentlicht: (2024) -
Curious Causality-Seeking Agents Learn Meta Causal World
von: Zhao, Zhiyu, et al.
Veröffentlicht: (2025) -
Language Agents as Optimizable Graphs
von: Zhuge, Mingchen, et al.
Veröffentlicht: (2024) -
Upside Down Reinforcement Learning with Policy Generators
von: Di Ventura, Jacopo, et al.
Veröffentlicht: (2025)