Saved in:
| Main Authors: | Wang, Fan, Shao, Pengtao, Zhang, Yiming, Yu, Bo, Liu, Shaoshan, Ding, Ning, Cao, Yang, Kang, Yu, Wang, Haifeng |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | https://arxiv.org/abs/2502.02869 |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Context and Diversity Matter: The Emergence of In-Context Learning in World Models
by: Wang, Fan, et al.
Published: (2025)
by: Wang, Fan, et al.
Published: (2025)
StateLinFormer: Stateful Training Enhancing Long-term Memory in Navigation
by: Chen, Zhiyuan, et al.
Published: (2026)
by: Chen, Zhiyuan, et al.
Published: (2026)
In-Context Learning can Perform Continual Learning Like Humans
by: Kang, Liuwang, et al.
Published: (2025)
by: Kang, Liuwang, et al.
Published: (2025)
Conceptual Framework Toward Embodied Collective Adaptive Intelligence
by: Wang, Fan, et al.
Published: (2025)
by: Wang, Fan, et al.
Published: (2025)
VAP: The Vulnerability-Adaptive Protection Paradigm Toward Reliable Autonomous Machines
by: Wan, Zishen, et al.
Published: (2024)
by: Wan, Zishen, et al.
Published: (2024)
Benchmarking General-Purpose In-Context Learning
by: Wang, Fan, et al.
Published: (2024)
by: Wang, Fan, et al.
Published: (2024)
Training Cross-Morphology Embodied AI Agents: From Practical Challenges to Theoretical Foundations
by: Liu, Shaoshan, et al.
Published: (2025)
by: Liu, Shaoshan, et al.
Published: (2025)
Dadu‐E : Rethinking the Role of Large Language Model in Robotic Computing Pipelines
by: Wenhao Sun, et al.
Published: (2025)
by: Wenhao Sun, et al.
Published: (2025)
DaDu-E: Rethinking the Role of Large Language Model in Robotic Computing Pipeline
by: Sun, Wenhao, et al.
Published: (2024)
by: Sun, Wenhao, et al.
Published: (2024)
Thales: Formulating and Estimating Architectural Vulnerability Factors for DNN Accelerators
by: Tyagi, Abhishek, et al.
Published: (2022)
by: Tyagi, Abhishek, et al.
Published: (2022)
Towards Monotonic Improvement in In-Context Reinforcement Learning
by: Zhang, Wenhao, et al.
Published: (2025)
by: Zhang, Wenhao, et al.
Published: (2025)
EfficientNav: Towards On-Device Object-Goal Navigation with Navigation Map Caching and Retrieval
by: Yang, Zebin, et al.
Published: (2025)
by: Yang, Zebin, et al.
Published: (2025)
DuMapper: Towards Automatic Verification of Large-Scale POIs with Street Views at Baidu Maps
by: Fan, Miao, et al.
Published: (2024)
by: Fan, Miao, et al.
Published: (2024)
A Theoretical Understanding of Gradient Bias in Meta-Reinforcement Learning
by: Feng, Xidong, et al.
Published: (2021)
by: Feng, Xidong, et al.
Published: (2021)
SimpleVLA-RL: Scaling VLA Training via Reinforcement Learning
by: Li, Haozhan, et al.
Published: (2025)
by: Li, Haozhan, et al.
Published: (2025)
Towards an Information Theoretic Framework of Context-Based Offline Meta-Reinforcement Learning
by: Li, Lanqing, et al.
Published: (2024)
by: Li, Lanqing, et al.
Published: (2024)
ATLAS: Agentic Test-time Learning-to-Allocate Scaling
by: Qin, Peijia, et al.
Published: (2026)
by: Qin, Peijia, et al.
Published: (2026)
The Biggest Risk of Embodied AI is Governance Lag
by: Liu, Shaoshan
Published: (2026)
by: Liu, Shaoshan
Published: (2026)
Large‐Scale Multi‐Objective Optimization Algorithms: A Decade Survey
by: Pengtao Wang, et al.
Published: (2025)
by: Pengtao Wang, et al.
Published: (2025)
DreamPRM-1.5: Unlocking the Potential of Each Instance for Multimodal Process Reward Model Training
by: Cao, Qi, et al.
Published: (2025)
by: Cao, Qi, et al.
Published: (2025)
Several functional capacities and Carleson type embeddings of fractional Sobolev sapces on stratified Lie groups
by: Wang, Zhiyong, et al.
Published: (2024)
by: Wang, Zhiyong, et al.
Published: (2024)
WebWorld: A Large-Scale World Model for Web Agent Training
by: Xiao, Zikai, et al.
Published: (2026)
by: Xiao, Zikai, et al.
Published: (2026)
KEEP: A KV-Cache-Centric Memory Management System for Efficient Embodied Planning
by: Yang, Zebin, et al.
Published: (2026)
by: Yang, Zebin, et al.
Published: (2026)
DySL-VLA: Efficient Vision-Language-Action Model Inference via Dynamic-Static Layer-Skipping for Robot Manipulation
by: Yang, Zebin, et al.
Published: (2026)
by: Yang, Zebin, et al.
Published: (2026)
Curious Causality-Seeking Agents Learn Meta Causal World
by: Zhao, Zhiyu, et al.
Published: (2025)
by: Zhao, Zhiyu, et al.
Published: (2025)
Vulnerable Agent Identification in Large-Scale Multi-Agent Reinforcement Learning
by: Li, Simin, et al.
Published: (2025)
by: Li, Simin, et al.
Published: (2025)
Relax: An Asynchronous Reinforcement Learning Engine for Omni-Modal Post-Training at Scale
by: Zhang, Liujie, et al.
Published: (2026)
by: Zhang, Liujie, et al.
Published: (2026)
LoongRL: Reinforcement Learning for Advanced Reasoning over Long Contexts
by: Wang, Siyuan, et al.
Published: (2025)
by: Wang, Siyuan, et al.
Published: (2025)
MGM as a Large‐Scale Pretrained Foundation Model for Microbiome Analyses in Diverse Contexts
by: Haohong Zhang, et al.
Published: (2026)
by: Haohong Zhang, et al.
Published: (2026)
MicLog: Towards Accurate and Efficient LLM-based Log Parsing via Progressive Meta In-Context Learning
by: Yu, Jianbo, et al.
Published: (2026)
by: Yu, Jianbo, et al.
Published: (2026)
Jailbreaking Large Language Models through Iterative Tool-Disguised Attacks via Reinforcement Learning
by: Wang, Zhaoqi, et al.
Published: (2026)
by: Wang, Zhaoqi, et al.
Published: (2026)
Boosting the Generalization and Reasoning of Vision Language Models with Curriculum Reinforcement Learning
by: Deng, Huilin, et al.
Published: (2025)
by: Deng, Huilin, et al.
Published: (2025)
Reinforcement Learning Optimization for Large-Scale Learning: An Efficient and User-Friendly Scaling Library
by: Wang, Weixun, et al.
Published: (2025)
by: Wang, Weixun, et al.
Published: (2025)
Meta-Inverse Reinforcement Learning for Mean Field Games via Probabilistic Context Variables
by: Chen, Yang, et al.
Published: (2025)
by: Chen, Yang, et al.
Published: (2025)
Collaborative Computing in Non-Terrestrial Networks: A Multi-Time-Scale Deep Reinforcement Learning Approach
by: Cao, Yang, et al.
Published: (2024)
by: Cao, Yang, et al.
Published: (2024)
Task-Stratified Knowledge Scaling Laws for Post-Training Quantized Large Language Models
by: Zhou, Chenxi, et al.
Published: (2025)
by: Zhou, Chenxi, et al.
Published: (2025)
Global Convergence in Training Large-Scale Transformers
by: Gao, Cheng, et al.
Published: (2024)
by: Gao, Cheng, et al.
Published: (2024)
The magnetic $Z_2$ topological insulator on the AA-stacked bilayer graphene
by: Liu, Yu-Bo, et al.
Published: (2024)
by: Liu, Yu-Bo, et al.
Published: (2024)
Unconventional superfluidity of superconductivity on Penrose lattice
by: Liu, Yu-Bo, et al.
Published: (2023)
by: Liu, Yu-Bo, et al.
Published: (2023)
Improving the Language Understanding Capabilities of Large Language Models Using Reinforcement Learning
by: Hu, Bokai, et al.
Published: (2024)
by: Hu, Bokai, et al.
Published: (2024)
Similar Items
-
Context and Diversity Matter: The Emergence of In-Context Learning in World Models
by: Wang, Fan, et al.
Published: (2025) -
StateLinFormer: Stateful Training Enhancing Long-term Memory in Navigation
by: Chen, Zhiyuan, et al.
Published: (2026) -
In-Context Learning can Perform Continual Learning Like Humans
by: Kang, Liuwang, et al.
Published: (2025) -
Conceptual Framework Toward Embodied Collective Adaptive Intelligence
by: Wang, Fan, et al.
Published: (2025) -
VAP: The Vulnerability-Adaptive Protection Paradigm Toward Reliable Autonomous Machines
by: Wan, Zishen, et al.
Published: (2024)