Compositional Learning of Visually-Grounded Concepts Using Reinforcement
Fuente:
arXiv
Guardado en:
| Autores principales: | Lin, Zijun, Azaman, Haidi, Kumar, M Ganesh, Tan, Cheston |
|---|---|
| Formato: | Preprint |
| Publicado: |
2023
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
STUPD: A Synthetic Dataset for Spatial and Temporal Relation Reasoning
por: Agrawal, Palaash, et al.
Publicado: (2023)
por: Agrawal, Palaash, et al.
Publicado: (2023)
Learning to Reason Iteratively and Parallelly for Complex Visual Reasoning Scenarios
por: Jaiswal, Shantanu, et al.
Publicado: (2024)
por: Jaiswal, Shantanu, et al.
Publicado: (2024)
GRAIL: Autonomous Concept Grounding for Neuro-Symbolic Reinforcement Learning
por: Shindo, Hikaru, et al.
Publicado: (2026)
por: Shindo, Hikaru, et al.
Publicado: (2026)
CORE: Concept-Oriented Reinforcement for Bridging the Definition-Application Gap in Mathematical Reasoning
por: Gao, Zijun, et al.
Publicado: (2025)
por: Gao, Zijun, et al.
Publicado: (2025)
Zero-Shot Visual Reasoning by Vision-Language Models: Benchmarking and Analysis
por: Nagar, Aishik, et al.
Publicado: (2024)
por: Nagar, Aishik, et al.
Publicado: (2024)
Position Paper: Rethinking Privacy in RL for Sequential Decision-making in the Age of LLMs
por: Fan, Flint Xiaofeng, et al.
Publicado: (2025)
por: Fan, Flint Xiaofeng, et al.
Publicado: (2025)
Social Learning through Interactions with Other Agents: A Survey
por: Hillier, Dylan, et al.
Publicado: (2024)
por: Hillier, Dylan, et al.
Publicado: (2024)
FedHPD: Heterogeneous Federated Reinforcement Learning via Policy Distillation
por: Jiang, Wenzheng, et al.
Publicado: (2025)
por: Jiang, Wenzheng, et al.
Publicado: (2025)
Efficient $Q$-Learning and Actor-Critic Methods for Robust Average Reward Reinforcement Learning
por: Xu, Yang, et al.
Publicado: (2025)
por: Xu, Yang, et al.
Publicado: (2025)
Selecting Decision-Relevant Concepts in Reinforcement Learning
por: Raman, Naveen, et al.
Publicado: (2026)
por: Raman, Naveen, et al.
Publicado: (2026)
On Predictability of Reinforcement Learning Dynamics for Large Language Models
por: Cai, Yuchen, et al.
Publicado: (2025)
por: Cai, Yuchen, et al.
Publicado: (2025)
Offline Safe Reinforcement Learning Using Trajectory Classification
por: Gong, Ze, et al.
Publicado: (2024)
por: Gong, Ze, et al.
Publicado: (2024)
Multi-Ontology Integration with Dual-Axis Propagation for Medical Concept Representation
por: Kerdabadi, Mohsen Nayebi, et al.
Publicado: (2025)
por: Kerdabadi, Mohsen Nayebi, et al.
Publicado: (2025)
SPIRAL: Self-Play on Zero-Sum Games Incentivizes Reasoning via Multi-Agent Multi-Turn Reinforcement Learning
por: Liu, Bo, et al.
Publicado: (2025)
por: Liu, Bo, et al.
Publicado: (2025)
Grounding by Trying: LLMs with Reinforcement Learning-Enhanced Retrieval
por: Hsu, Sheryl, et al.
Publicado: (2024)
por: Hsu, Sheryl, et al.
Publicado: (2024)
Augmented Reinforcement Learning Framework For Enhancing Decision-Making In Machine Learning Models Using External Agents
por: Singh, Sandesh Kumar
Publicado: (2025)
por: Singh, Sandesh Kumar
Publicado: (2025)
A Survey on Explainable Reinforcement Learning: Concepts, Algorithms, Challenges
por: Qing, Yunpeng, et al.
Publicado: (2022)
por: Qing, Yunpeng, et al.
Publicado: (2022)
LICORICE: Label-Efficient Concept-Based Interpretable Reinforcement Learning
por: Ye, Zhuorui, et al.
Publicado: (2024)
por: Ye, Zhuorui, et al.
Publicado: (2024)
Learning Quantifiable Visual Explanations Without Ground-Truth
por: Singh, Amritpal, et al.
Publicado: (2026)
por: Singh, Amritpal, et al.
Publicado: (2026)
In-Context Compositional Q-Learning for Offline Reinforcement Learning
por: Xu, Qiushui, et al.
Publicado: (2025)
por: Xu, Qiushui, et al.
Publicado: (2025)
Concept Drift Adaptation Using Self-Supervised and Reinforcement Learning In Android Malware Detection
por: Sabbah, Ahmed, et al.
Publicado: (2026)
por: Sabbah, Ahmed, et al.
Publicado: (2026)
CAESAR: Enhancing Federated RL in Heterogeneous MDPs through Convergence-Aware Sampling with Screening
por: Mak, Hei Yi, et al.
Publicado: (2024)
por: Mak, Hei Yi, et al.
Publicado: (2024)
Neurosymbolic Grounding for Compositional World Models
por: Sehgal, Atharva, et al.
Publicado: (2023)
por: Sehgal, Atharva, et al.
Publicado: (2023)
FairPO: Robust Preference Optimization for Fair Multi-Label Learning
por: Mondal, Soumen Kumar, et al.
Publicado: (2025)
por: Mondal, Soumen Kumar, et al.
Publicado: (2025)
Composite Flow Matching for Reinforcement Learning with Shifted-Dynamics Data
por: Kong, Lingkai, et al.
Publicado: (2025)
por: Kong, Lingkai, et al.
Publicado: (2025)
Explainable Molecular Property Prediction: Aligning Chemical Concepts with Predictions via Language Models
por: Wang, Zhenzhong, et al.
Publicado: (2024)
por: Wang, Zhenzhong, et al.
Publicado: (2024)
Prototype-Grounded Concept Models for Verifiable Concept Alignment
por: Colamonaco, Stefano, et al.
Publicado: (2026)
por: Colamonaco, Stefano, et al.
Publicado: (2026)
Highway Reinforcement Learning
por: Wang, Yuhui, et al.
Publicado: (2024)
por: Wang, Yuhui, et al.
Publicado: (2024)
Safe Reinforcement Learning with Learned Non-Markovian Safety Constraints
por: Low, Siow Meng, et al.
Publicado: (2024)
por: Low, Siow Meng, et al.
Publicado: (2024)
Robotic Manipulation Datasets for Offline Compositional Reinforcement Learning
por: Hussing, Marcel, et al.
Publicado: (2023)
por: Hussing, Marcel, et al.
Publicado: (2023)
Survival Concept-Based Learning Models
por: Kirpichenko, Stanislav R., et al.
Publicado: (2025)
por: Kirpichenko, Stanislav R., et al.
Publicado: (2025)
From Grunts to Lexicons: Emergent Language from Cooperative Foraging
por: Piriyajitakonkij, Maytus, et al.
Publicado: (2025)
por: Piriyajitakonkij, Maytus, et al.
Publicado: (2025)
Revisiting Plasticity in Visual Reinforcement Learning: Data, Modules and Training Stages
por: Ma, Guozheng, et al.
Publicado: (2023)
por: Ma, Guozheng, et al.
Publicado: (2023)
Concept Heterogeneity-aware Representation Steering
por: Abdullaev, Laziz U., et al.
Publicado: (2026)
por: Abdullaev, Laziz U., et al.
Publicado: (2026)
SEBA: Sample-Efficient Black-Box Attacks on Visual Reinforcement Learning
por: Huang, Tairan, et al.
Publicado: (2025)
por: Huang, Tairan, et al.
Publicado: (2025)
Enhanced Pruning Strategy for Multi-Component Neural Architectures Using Component-Aware Graph Analysis
por: Sundaram, Ganesh, et al.
Publicado: (2025)
por: Sundaram, Ganesh, et al.
Publicado: (2025)
Satisficing Exploration for Deep Reinforcement Learning
por: Arumugam, Dilip, et al.
Publicado: (2024)
por: Arumugam, Dilip, et al.
Publicado: (2024)
Efficient Reinforcement Learning in Probabilistic Reward Machines
por: Lin, Xiaofeng, et al.
Publicado: (2024)
por: Lin, Xiaofeng, et al.
Publicado: (2024)
HER: Human-like Reasoning and Reinforcement Learning for LLM Role-playing
por: Du, Chengyu, et al.
Publicado: (2026)
por: Du, Chengyu, et al.
Publicado: (2026)
Exploring Accurate and Transparent Domain Adaptation in Predictive Healthcare via Concept-Grounded Orthogonal Inference
por: Hu, Pengfei, et al.
Publicado: (2026)
por: Hu, Pengfei, et al.
Publicado: (2026)
Ejemplares similares
-
STUPD: A Synthetic Dataset for Spatial and Temporal Relation Reasoning
por: Agrawal, Palaash, et al.
Publicado: (2023) -
Learning to Reason Iteratively and Parallelly for Complex Visual Reasoning Scenarios
por: Jaiswal, Shantanu, et al.
Publicado: (2024) -
GRAIL: Autonomous Concept Grounding for Neuro-Symbolic Reinforcement Learning
por: Shindo, Hikaru, et al.
Publicado: (2026) -
CORE: Concept-Oriented Reinforcement for Bridging the Definition-Application Gap in Mathematical Reasoning
por: Gao, Zijun, et al.
Publicado: (2025) -
Zero-Shot Visual Reasoning by Vision-Language Models: Benchmarking and Analysis
por: Nagar, Aishik, et al.
Publicado: (2024)