Gespeichert in:
| Hauptverfasser: | Liang, Dayang, Liu, Ruihan, Wan, Lipeng, Liu, Yunlong, An, Bo |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2026
|
| Schlagworte: | |
| Online-Zugang: | https://arxiv.org/abs/2602.18724 |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Intrinsic Dynamics-Driven Generalizable Scene Representations for Vision-Oriented Decision-Making Applications
von: Liang, Dayang, et al.
Veröffentlicht: (2024)
von: Liang, Dayang, et al.
Veröffentlicht: (2024)
Episodic Reinforcement Learning with Expanded State-reward Space
von: Liang, Dayang, et al.
Veröffentlicht: (2024)
von: Liang, Dayang, et al.
Veröffentlicht: (2024)
InterReal: A Unified Physics-Based Imitation Framework for Learning Human-Object Interaction Skills
von: Liang, Dayang, et al.
Veröffentlicht: (2026)
von: Liang, Dayang, et al.
Veröffentlicht: (2026)
Imagine, Initialize, and Explore: An Effective Exploration Method in Multi-Agent Reinforcement Learning
von: Liu, Zeyang, et al.
Veröffentlicht: (2024)
von: Liu, Zeyang, et al.
Veröffentlicht: (2024)
ThanoRA: Task Heterogeneity-Aware Multi-Task Low-Rank Adaptation
von: Liang, Jian, et al.
Veröffentlicht: (2025)
von: Liang, Jian, et al.
Veröffentlicht: (2025)
A Generalized Bisimulation Metric of State Similarity between Markov Decision Processes: From Theoretical Propositions to Applications
von: Tao, Zhenyu, et al.
Veröffentlicht: (2025)
von: Tao, Zhenyu, et al.
Veröffentlicht: (2025)
LASIL: Learner-Aware Supervised Imitation Learning For Long-term Microscopic Traffic Simulation
von: Guo, Ke, et al.
Veröffentlicht: (2024)
von: Guo, Ke, et al.
Veröffentlicht: (2024)
Game Generation via Large Language Models
von: Hu, Chengpeng, et al.
Veröffentlicht: (2024)
von: Hu, Chengpeng, et al.
Veröffentlicht: (2024)
Stable Offline Value Function Learning with Bisimulation-based Representations
von: Pavse, Brahma S., et al.
Veröffentlicht: (2024)
von: Pavse, Brahma S., et al.
Veröffentlicht: (2024)
DeepThink3D: Enhancing Large Language Models with Programmatic Reasoning in Complex 3D Situated Reasoning Tasks
von: Song, Jiayi, et al.
Veröffentlicht: (2025)
von: Song, Jiayi, et al.
Veröffentlicht: (2025)
Boosting Meta-Learning for Few-Shot Text Classification via Label-guided Distance Scaling
von: Gao, Yunlong, et al.
Veröffentlicht: (2026)
von: Gao, Yunlong, et al.
Veröffentlicht: (2026)
SD-Net: Symmetric-Aware Keypoint Prediction and Domain Adaptation for 6D Pose Estimation In Bin-picking Scenarios
von: Huang, Ding-Tao, et al.
Veröffentlicht: (2024)
von: Huang, Ding-Tao, et al.
Veröffentlicht: (2024)
Toward Automated and Trustworthy Scientific Analysis and Visualization with LLM-Generated Code
von: Chakroborti, Apu Kumar, et al.
Veröffentlicht: (2025)
von: Chakroborti, Apu Kumar, et al.
Veröffentlicht: (2025)
Scaling Synthetic Task Generation for Agents via Exploration
von: Ramrakhya, Ram, et al.
Veröffentlicht: (2025)
von: Ramrakhya, Ram, et al.
Veröffentlicht: (2025)
Selection, Reflection and Self-Refinement: Revisit Reasoning Tasks via a Causal Lens
von: Deng, Yunlong, et al.
Veröffentlicht: (2025)
von: Deng, Yunlong, et al.
Veröffentlicht: (2025)
RDEx-CSOP: Feasibility-Aware Reconstructed Differential Evolution with Adaptive epsilon-Constraint Ranking
von: Tao, Sichen, et al.
Veröffentlicht: (2026)
von: Tao, Sichen, et al.
Veröffentlicht: (2026)
CALM: Consensus-Aware Localized Merging for Multi-Task Learning
von: Yan, Kunda, et al.
Veröffentlicht: (2025)
von: Yan, Kunda, et al.
Veröffentlicht: (2025)
Measuring and Analyzing Intelligence via Contextual Uncertainty in Large Language Models using Information-Theoretic Metrics
von: Shim, Jae Wan
Veröffentlicht: (2025)
von: Shim, Jae Wan
Veröffentlicht: (2025)
RDEx-CMOP: Feasibility-Aware Indicator-Guided Differential Evolution for Fixed-Budget Constrained Multiobjective Optimization
von: Tao, Sichen, et al.
Veröffentlicht: (2026)
von: Tao, Sichen, et al.
Veröffentlicht: (2026)
Video-XL-2: Towards Very Long-Video Understanding Through Task-Aware KV Sparsification
von: Qin, Minghao, et al.
Veröffentlicht: (2025)
von: Qin, Minghao, et al.
Veröffentlicht: (2025)
Autonomous Implicit Indoor Scene Reconstruction with Frontier Exploration
von: Zeng, Jing, et al.
Veröffentlicht: (2024)
von: Zeng, Jing, et al.
Veröffentlicht: (2024)
Provably Efficient Exploration in Inverse Constrained Reinforcement Learning
von: Yue, Bo, et al.
Veröffentlicht: (2024)
von: Yue, Bo, et al.
Veröffentlicht: (2024)
DRT: Deep Reasoning Translation via Long Chain-of-Thought
von: Wang, Jiaan, et al.
Veröffentlicht: (2024)
von: Wang, Jiaan, et al.
Veröffentlicht: (2024)
Scene-Aware Explainable Multimodal Trajectory Prediction
von: Liu, Pei, et al.
Veröffentlicht: (2024)
von: Liu, Pei, et al.
Veröffentlicht: (2024)
On the Benefits of Free Exploration for Regret Minimization in Multi-Armed Bandits
von: Hou, Yunlong, et al.
Veröffentlicht: (2026)
von: Hou, Yunlong, et al.
Veröffentlicht: (2026)
Grounded Answers for Multi-agent Decision-making Problem through Generative World Model
von: Liu, Zeyang, et al.
Veröffentlicht: (2024)
von: Liu, Zeyang, et al.
Veröffentlicht: (2024)
Transport-Hub-Aware Spatial-Temporal Adaptive Graph Transformer for Traffic Flow Prediction
von: Xu, Xiao, et al.
Veröffentlicht: (2023)
von: Xu, Xiao, et al.
Veröffentlicht: (2023)
Fake News Detection and Manipulation Reasoning via Large Vision-Language Models
von: Jin, Ruihan, et al.
Veröffentlicht: (2024)
von: Jin, Ruihan, et al.
Veröffentlicht: (2024)
Towards Provably Unlearnable Examples via Bayes Error Optimisation
von: Zhang, Ruihan, et al.
Veröffentlicht: (2025)
von: Zhang, Ruihan, et al.
Veröffentlicht: (2025)
SeqUDA-Rec: Sequential User Behavior Enhanced Recommendation via Global Unsupervised Data Augmentation for Personalized Content Marketing
von: Luo, Ruihan, et al.
Veröffentlicht: (2025)
von: Luo, Ruihan, et al.
Veröffentlicht: (2025)
One-Shot Sensitivity-Aware Mixed Sparsity Pruning for Large Language Models
von: Shao, Hang, et al.
Veröffentlicht: (2023)
von: Shao, Hang, et al.
Veröffentlicht: (2023)
BEE: Metric-Adapted Explanations via Baseline Exploration-Exploitation
von: Barkan, Oren, et al.
Veröffentlicht: (2024)
von: Barkan, Oren, et al.
Veröffentlicht: (2024)
Lifecycle-Aware code generation: Leveraging Software Engineering Phases in LLMs
von: Xing, Xing, et al.
Veröffentlicht: (2025)
von: Xing, Xing, et al.
Veröffentlicht: (2025)
How to Evaluate Semantic Communications for Images with ViTScore Metric?
von: Zhu, Tingting, et al.
Veröffentlicht: (2023)
von: Zhu, Tingting, et al.
Veröffentlicht: (2023)
Lion Secretly Solves Constrained Optimization: As Lyapunov Predicts
von: Chen, Lizhang, et al.
Veröffentlicht: (2023)
von: Chen, Lizhang, et al.
Veröffentlicht: (2023)
Joint Agent Memory and Exploration Learning via Novelty Signals
von: Tian, Shizuo, et al.
Veröffentlicht: (2026)
von: Tian, Shizuo, et al.
Veröffentlicht: (2026)
CrossLinear: Plug-and-Play Cross-Correlation Embedding for Time Series Forecasting with Exogenous Variables
von: Zhou, Pengfei, et al.
Veröffentlicht: (2025)
von: Zhou, Pengfei, et al.
Veröffentlicht: (2025)
OPRIDE: Offline Preference-based Reinforcement Learning via In-Dataset Exploration
von: Yang, Yiqin, et al.
Veröffentlicht: (2026)
von: Yang, Yiqin, et al.
Veröffentlicht: (2026)
Learning to Explore: Scaling Agentic Reasoning via Exploration-Aware Policy Optimization
von: Hua, Xingyuan, et al.
Veröffentlicht: (2026)
von: Hua, Xingyuan, et al.
Veröffentlicht: (2026)
Learning to Adapt: Self-Improving Web Agent via Cognitive-Aware Exploration
von: Chen, Weile, et al.
Veröffentlicht: (2026)
von: Chen, Weile, et al.
Veröffentlicht: (2026)
Ähnliche Einträge
-
Intrinsic Dynamics-Driven Generalizable Scene Representations for Vision-Oriented Decision-Making Applications
von: Liang, Dayang, et al.
Veröffentlicht: (2024) -
Episodic Reinforcement Learning with Expanded State-reward Space
von: Liang, Dayang, et al.
Veröffentlicht: (2024) -
InterReal: A Unified Physics-Based Imitation Framework for Learning Human-Object Interaction Skills
von: Liang, Dayang, et al.
Veröffentlicht: (2026) -
Imagine, Initialize, and Explore: An Effective Exploration Method in Multi-Agent Reinforcement Learning
von: Liu, Zeyang, et al.
Veröffentlicht: (2024) -
ThanoRA: Task Heterogeneity-Aware Multi-Task Low-Rank Adaptation
von: Liang, Jian, et al.
Veröffentlicht: (2025)