Enhancing Human-Likeness in Reinforcement Learning Agents via Hierarchical Macro Action Quantization
Fuente:
arXiv
Salvato in:
| Autori principali: | Nizamani, Usman, Luqman, M. Shaheer, Fateh, Fawad Javed, Ali, Ali Shah, Popattia, Murad, Zia, M. Zeeshan, Tran, Quoc-Huy |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2026
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
A Hierarchical Spatiotemporal Action Tokenizer for In-Context Imitation Learning in Robotics
di: Fateh, Fawad Javed, et al.
Pubblicazione: (2026)
di: Fateh, Fawad Javed, et al.
Pubblicazione: (2026)
Unsupervised Skeleton-Based Action Segmentation via Hierarchical Spatiotemporal Vector Quantization
di: Ahmed, Umer, et al.
Pubblicazione: (2026)
di: Ahmed, Umer, et al.
Pubblicazione: (2026)
Procedure Learning via Regularized Gromov-Wasserstein Optimal Transport
di: Mahmood, Syed Ahmed, et al.
Pubblicazione: (2025)
di: Mahmood, Syed Ahmed, et al.
Pubblicazione: (2025)
TemporalVLM: Video LLMs for Temporal Reasoning in Long Videos
di: Fateh, Fawad Javed, et al.
Pubblicazione: (2024)
di: Fateh, Fawad Javed, et al.
Pubblicazione: (2024)
Learning by Aligning 2D Skeleton Sequences and Multi-Modality Fusion
di: Tran, Quoc-Huy, et al.
Pubblicazione: (2023)
di: Tran, Quoc-Huy, et al.
Pubblicazione: (2023)
Action Segmentation Using 2D Skeleton Heatmaps and Multi-Modality Fusion
di: Hyder, Syed Waleed, et al.
Pubblicazione: (2023)
di: Hyder, Syed Waleed, et al.
Pubblicazione: (2023)
Joint Self-Supervised Video Alignment and Action Segmentation
di: Ali, Ali Shah, et al.
Pubblicazione: (2025)
di: Ali, Ali Shah, et al.
Pubblicazione: (2025)
Permutation-Aware Action Segmentation via Unsupervised Frame-to-Segment Alignment
di: Tran, Quoc-Huy, et al.
Pubblicazione: (2023)
di: Tran, Quoc-Huy, et al.
Pubblicazione: (2023)
Hierarchical Meta-Reinforcement Learning via Automated Macro-Action Discovery
di: Cho, Minjae, et al.
Pubblicazione: (2024)
di: Cho, Minjae, et al.
Pubblicazione: (2024)
Learning Human-Like RL Agents Through Trajectory Optimization With Action Quantization
di: Guo, Jian-Ting, et al.
Pubblicazione: (2025)
di: Guo, Jian-Ting, et al.
Pubblicazione: (2025)
MA-RLHF: Reinforcement Learning from Human Feedback with Macro Actions
di: Chai, Yekun, et al.
Pubblicazione: (2024)
di: Chai, Yekun, et al.
Pubblicazione: (2024)
Governed By Agents: A Survey On The Role Of Agentic AI In Future Computing Environments
di: Murad, Nauman Ali, et al.
Pubblicazione: (2025)
di: Murad, Nauman Ali, et al.
Pubblicazione: (2025)
Attenuation of Vascular Dementia Associated Neuroinflammation by Inhibition of the JNK Pathway in HFD / STZ ‐Induced Diabetic Rat Model
di: Sundas Firdoos, et al.
Pubblicazione: (2026)
di: Sundas Firdoos, et al.
Pubblicazione: (2026)
Split-Fuse-Transport: Annotation-Free Saliency via Dual Clustering and Optimal Transport Alignment
di: Ramzan, Muhammad Umer, et al.
Pubblicazione: (2025)
di: Ramzan, Muhammad Umer, et al.
Pubblicazione: (2025)
Continuous Robust Formation Tracking Controllers for Underactuated Planar Agents
di: Quoc Van Tran
Pubblicazione: (2026)
di: Quoc Van Tran
Pubblicazione: (2026)
Test-Time Adaptation for Anomaly Segmentation via Topology-Aware Optimal Transport Chaining
di: Zia, Ali, et al.
Pubblicazione: (2026)
di: Zia, Ali, et al.
Pubblicazione: (2026)
Optimized Local Updates in Federated Learning via Reinforcement Learning
di: Murad, Ali, et al.
Pubblicazione: (2025)
di: Murad, Ali, et al.
Pubblicazione: (2025)
Hypergraph Contrastive Sensor Fusion for Multimodal Fault Diagnosis in Induction Motors
di: Ali, Usman, et al.
Pubblicazione: (2025)
di: Ali, Usman, et al.
Pubblicazione: (2025)
Component-Aware Sketch-to-Image Generation Using Self-Attention Encoding and Coordinate-Preserving Fusion
di: Zia, Ali, et al.
Pubblicazione: (2026)
di: Zia, Ali, et al.
Pubblicazione: (2026)
Geometry-Aware Semantic Reasoning for Training Free Video Anomaly Detection
di: Zia, Ali, et al.
Pubblicazione: (2026)
di: Zia, Ali, et al.
Pubblicazione: (2026)
Locally-Focused Face Representation for Sketch-to-Image Generation Using Noise-Induced Refinement
di: Ramzan, Muhammad Umer, et al.
Pubblicazione: (2024)
di: Ramzan, Muhammad Umer, et al.
Pubblicazione: (2024)
Novel Multi-Agent Action Masked Deep Reinforcement Learning for General Industrial Assembly Lines Balancing Problems
di: Ali, Ali Mohamed, et al.
Pubblicazione: (2025)
di: Ali, Ali Mohamed, et al.
Pubblicazione: (2025)
Assessing the Impact of Nitrogen Trade‐Off Between Coffee Tree Productivity and Bean Sensory Attributes to Produce Australian ‘Specialty Coffee’
di: Fawad Ali, et al.
Pubblicazione: (2026)
di: Fawad Ali, et al.
Pubblicazione: (2026)
Towards Fault Diagnosis in Induction Motor using Fractional Fourier Transform
di: Ali, Usman
Pubblicazione: (2024)
di: Ali, Usman
Pubblicazione: (2024)
A Multimodal Lightweight Approach to Fault Diagnosis of Induction Motors in High-Dimensional Dataset
di: Ali, Usman
Pubblicazione: (2025)
di: Ali, Usman
Pubblicazione: (2025)
Viper-F1: Fast and Fine-Grained Multimodal Understanding with Cross-Modal State-Space Modulation
di: Trinh, Quoc-Huy
Pubblicazione: (2025)
di: Trinh, Quoc-Huy
Pubblicazione: (2025)
PERFORMANCE EVALUATION OF HOT MIX ASPHALT (HMA) BY USING NANO SILICA AND FLY ASH
di: Usman Ali Khan, Sawood Sadiq, Ali Akbar Khan, Fahad Mehmood, Zeeshan Ali Khan, Ayesha Noreen
Pubblicazione: (2025)
di: Usman Ali Khan, Sawood Sadiq, Ali Akbar Khan, Fahad Mehmood, Zeeshan Ali Khan, Ayesha Noreen
Pubblicazione: (2025)
Instruments And Effects Of Monetary And Fiscal Policy: The Relationship Between Inflation, Vat, And Deposit Interest Rate
di: Dogdu, Ali, et al.
Pubblicazione: (2024)
di: Dogdu, Ali, et al.
Pubblicazione: (2024)
Multi-Scale Reversible Chaos Game Representation: A Unified Framework for Sequence Classification
di: Ali, Sarwan, et al.
Pubblicazione: (2026)
di: Ali, Sarwan, et al.
Pubblicazione: (2026)
The Impact Of Industrial Production On Economic Growth: New Empirical Evidence For Turkiye In A Material Development Framework
di: Doğdu, Ali, et al.
Pubblicazione: (2025)
di: Doğdu, Ali, et al.
Pubblicazione: (2025)
Murmur2Vec: A Hashing Based Solution For Embedding Generation Of COVID-19 Spike Sequences
di: Ali, Sarwan, et al.
Pubblicazione: (2025)
di: Ali, Sarwan, et al.
Pubblicazione: (2025)
Topological Deep Learning: A Review of an Emerging Paradigm
di: Zia, Ali, et al.
Pubblicazione: (2023)
di: Zia, Ali, et al.
Pubblicazione: (2023)
Deep Reinforcement Learning Based Navigation with Macro Actions and Topological Maps
di: Hakenes, Simon, et al.
Pubblicazione: (2025)
di: Hakenes, Simon, et al.
Pubblicazione: (2025)
Deep Reinforcement Learning for Decentralized Multi-Robot Exploration With Macro Actions
di: Tan, Aaron Hao, et al.
Pubblicazione: (2021)
di: Tan, Aaron Hao, et al.
Pubblicazione: (2021)
Representation Geometry as a Diagnostic for Out-of-Distribution Robustness
di: Zia, Ali, et al.
Pubblicazione: (2026)
di: Zia, Ali, et al.
Pubblicazione: (2026)
Pragmatic motivation and willingness to communicate as predictors of L2 pragmatic knowledge
di: Zia Tajeddin, et al.
Pubblicazione: (2024)
di: Zia Tajeddin, et al.
Pubblicazione: (2024)
Experimental investigation of two solar air heaters, with and without employing PCM
di: Yousif Fateh Midhat, et al.
Pubblicazione: (2024)
di: Yousif Fateh Midhat, et al.
Pubblicazione: (2024)
Hierarchical Vector Quantization for Unsupervised Action Segmentation
di: Spurio, Federico, et al.
Pubblicazione: (2024)
di: Spurio, Federico, et al.
Pubblicazione: (2024)
Balancing Bank Profits With Sustainable Development Goals: Examining the Pivotal Role of Financial Stability
di: Nguyen Quoc Huy, et al.
Pubblicazione: (2025)
di: Nguyen Quoc Huy, et al.
Pubblicazione: (2025)
AgentForge: Execution-Grounded Multi-Agent LLM Framework for Autonomous Software Engineering
di: Kumar, Rajesh, et al.
Pubblicazione: (2026)
di: Kumar, Rajesh, et al.
Pubblicazione: (2026)
Documenti analoghi
-
A Hierarchical Spatiotemporal Action Tokenizer for In-Context Imitation Learning in Robotics
di: Fateh, Fawad Javed, et al.
Pubblicazione: (2026) -
Unsupervised Skeleton-Based Action Segmentation via Hierarchical Spatiotemporal Vector Quantization
di: Ahmed, Umer, et al.
Pubblicazione: (2026) -
Procedure Learning via Regularized Gromov-Wasserstein Optimal Transport
di: Mahmood, Syed Ahmed, et al.
Pubblicazione: (2025) -
TemporalVLM: Video LLMs for Temporal Reasoning in Long Videos
di: Fateh, Fawad Javed, et al.
Pubblicazione: (2024) -
Learning by Aligning 2D Skeleton Sequences and Multi-Modality Fusion
di: Tran, Quoc-Huy, et al.
Pubblicazione: (2023)