Locality Sensitive Sparse Encoding for Learning World Models Online
Fuente:
arXiv
Guardado en:
| Autores principales: | Liu, Zichen, Du, Chao, Lee, Wee Sun, Lin, Min |
|---|---|
| Formato: | Preprint |
| Publicado: |
2024
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
Continual Reinforcement Learning by Planning with Online World Models
por: Liu, Zichen, et al.
Publicado: (2025)
por: Liu, Zichen, et al.
Publicado: (2025)
Sample-Efficient Alignment for LLMs
por: Liu, Zichen, et al.
Publicado: (2024)
por: Liu, Zichen, et al.
Publicado: (2024)
Optimizing Anytime Reasoning via Budget Relative Policy Optimization
por: Qi, Penghui, et al.
Publicado: (2025)
por: Qi, Penghui, et al.
Publicado: (2025)
Defeating the Training-Inference Mismatch via FP16
por: Qi, Penghui, et al.
Publicado: (2025)
por: Qi, Penghui, et al.
Publicado: (2025)
Understanding R1-Zero-Like Training: A Critical Perspective
por: Liu, Zichen, et al.
Publicado: (2025)
por: Liu, Zichen, et al.
Publicado: (2025)
Language Models Can Learn from Verbal Feedback Without Scalar Rewards
por: Luo, Renjie, et al.
Publicado: (2025)
por: Luo, Renjie, et al.
Publicado: (2025)
Differentiable Tree Search Network
por: Mittal, Dixant, et al.
Publicado: (2024)
por: Mittal, Dixant, et al.
Publicado: (2024)
Variational Reasoning for Language Models
por: Zhou, Xiangxin, et al.
Publicado: (2025)
por: Zhou, Xiangxin, et al.
Publicado: (2025)
SPIRAL: Self-Play on Zero-Sum Games Incentivizes Reasoning via Multi-Agent Multi-Turn Reinforcement Learning
por: Liu, Bo, et al.
Publicado: (2025)
por: Liu, Bo, et al.
Publicado: (2025)
Explaining Time Series via Contrastive and Locally Sparse Perturbations
por: Liu, Zichuan, et al.
Publicado: (2024)
por: Liu, Zichuan, et al.
Publicado: (2024)
SofT-GRPO: Surpassing Discrete-Token LLM Reinforcement Learning via Gumbel-Reparameterized Soft-Thinking Policy Optimization
por: Zheng, Zhi, et al.
Publicado: (2025)
por: Zheng, Zhi, et al.
Publicado: (2025)
Complementary Learning System Empowers Online Continual Learning of Vehicle Motion Forecasting in Smart Cities
por: Li, Zirui, et al.
Publicado: (2025)
por: Li, Zirui, et al.
Publicado: (2025)
S2O: Early Stopping for Sparse Attention via Online Permutation
por: Zhang, Yu, et al.
Publicado: (2026)
por: Zhang, Yu, et al.
Publicado: (2026)
Provably Efficient Algorithm for Best Scoring Rule Identification in Online Principal-Agent Information Acquisition
por: Wang, Zichen, et al.
Publicado: (2025)
por: Wang, Zichen, et al.
Publicado: (2025)
LSHFed: Robust and Communication-Efficient Federated Learning with Locally-Sensitive Hashing Gradient Mapping
por: Cheng, Guanjie, et al.
Publicado: (2025)
por: Cheng, Guanjie, et al.
Publicado: (2025)
PF-GNN: Differentiable particle filtering based approximation of universal graph representations
por: Dupty, Mohammed Haroon, et al.
Publicado: (2024)
por: Dupty, Mohammed Haroon, et al.
Publicado: (2024)
Misspecified $Q$-Learning with Sparse Linear Function Approximation: Tight Bounds on Approximation Error
por: Du, Ally Yalei, et al.
Publicado: (2024)
por: Du, Ally Yalei, et al.
Publicado: (2024)
Coupled Distributional Random Expert Distillation for World Model Online Imitation Learning
por: Li, Shangzhe, et al.
Publicado: (2025)
por: Li, Shangzhe, et al.
Publicado: (2025)
Global Low-Rank, Local Full-Rank: The Holographic Encoding of Learned Algorithms
por: Xu, Yongzhong
Publicado: (2026)
por: Xu, Yongzhong
Publicado: (2026)
On the Empirical Complexity of Reasoning and Planning in LLMs
por: Kang, Liwei, et al.
Publicado: (2024)
por: Kang, Liwei, et al.
Publicado: (2024)
Learning to Model the World with Language
por: Lin, Jessy, et al.
Publicado: (2023)
por: Lin, Jessy, et al.
Publicado: (2023)
Identifying Sparsely Active Circuits Through Local Loss Landscape Decomposition
por: Chrisman, Brianna, et al.
Publicado: (2025)
por: Chrisman, Brianna, et al.
Publicado: (2025)
Imagine Beyond! Distributionally Robust Auto-Encoding for State Space Coverage in Online Reinforcement Learning
por: Castanet, Nicolas, et al.
Publicado: (2025)
por: Castanet, Nicolas, et al.
Publicado: (2025)
Learning the Model While Learning Q: Finite-Time Sample Complexity of Online SyncMBQ
por: Lim, Han-Dong, et al.
Publicado: (2024)
por: Lim, Han-Dong, et al.
Publicado: (2024)
Sparse Adapter Fusion for Continual Learning in NLP
por: Zeng, Min, et al.
Publicado: (2026)
por: Zeng, Min, et al.
Publicado: (2026)
Identifying Sensitive Weights via Post-quantization Integral
por: Hu, Yuezhou, et al.
Publicado: (2025)
por: Hu, Yuezhou, et al.
Publicado: (2025)
A Unified Theory of Sparse Dictionary Learning in Mechanistic Interpretability: Piecewise Biconvexity and Spurious Minima
por: Tang, Yiming, et al.
Publicado: (2025)
por: Tang, Yiming, et al.
Publicado: (2025)
BiSHop: Bi-Directional Cellular Learning for Tabular Data with Generalized Sparse Modern Hopfield Model
por: Xu, Chenwei, et al.
Publicado: (2024)
por: Xu, Chenwei, et al.
Publicado: (2024)
Meta Additive Model: Interpretable Sparse Learning With Auto Weighting
por: Zhang, Xuelin, et al.
Publicado: (2026)
por: Zhang, Xuelin, et al.
Publicado: (2026)
Learning Latent Dynamic Robust Representations for World Models
por: Sun, Ruixiang, et al.
Publicado: (2024)
por: Sun, Ruixiang, et al.
Publicado: (2024)
Enhancing Inverse Reinforcement Learning through Encoding Dynamic Information in Reward Shaping
por: Zhan, Simon Sinong, et al.
Publicado: (2024)
por: Zhan, Simon Sinong, et al.
Publicado: (2024)
Learning from Yesterday's Error: An Efficient Online Learning Method for Traffic Demand Prediction
por: Huang, Xiannan, et al.
Publicado: (2026)
por: Huang, Xiannan, et al.
Publicado: (2026)
Is Prompt Selection Necessary for Task-Free Online Continual Learning?
por: Park, Seoyoung, et al.
Publicado: (2026)
por: Park, Seoyoung, et al.
Publicado: (2026)
Learning to Theorize the World from Observation
por: Baek, Doojin, et al.
Publicado: (2026)
por: Baek, Doojin, et al.
Publicado: (2026)
Improving Molecular Force Fields with Minimal Temporal Information
por: Mollahosseini, Ali, et al.
Publicado: (2026)
por: Mollahosseini, Ali, et al.
Publicado: (2026)
When Attention Sink Emerges in Language Models: An Empirical View
por: Gu, Xiangming, et al.
Publicado: (2024)
por: Gu, Xiangming, et al.
Publicado: (2024)
Adaptive Normalization Mamba with Multi Scale Trend Decomposition and Patch MoE Encoding
por: Jeon, MinCheol
Publicado: (2025)
por: Jeon, MinCheol
Publicado: (2025)
Extending Epistemic Uncertainty Beyond Parameters Would Assist in Designing Reliable LLMs
por: Nguyen-Hien, T. Duy, et al.
Publicado: (2025)
por: Nguyen-Hien, T. Duy, et al.
Publicado: (2025)
Combining Reinforcement Learning and Optimal Transport for the Traveling Salesman Problem
por: Goh, Yong Liang, et al.
Publicado: (2022)
por: Goh, Yong Liang, et al.
Publicado: (2022)
VCWorld: A Biological World Model for Virtual Cell Simulation
por: Wei, Zhijian, et al.
Publicado: (2025)
por: Wei, Zhijian, et al.
Publicado: (2025)
Ejemplares similares
-
Continual Reinforcement Learning by Planning with Online World Models
por: Liu, Zichen, et al.
Publicado: (2025) -
Sample-Efficient Alignment for LLMs
por: Liu, Zichen, et al.
Publicado: (2024) -
Optimizing Anytime Reasoning via Budget Relative Policy Optimization
por: Qi, Penghui, et al.
Publicado: (2025) -
Defeating the Training-Inference Mismatch via FP16
por: Qi, Penghui, et al.
Publicado: (2025) -
Understanding R1-Zero-Like Training: A Critical Perspective
por: Liu, Zichen, et al.
Publicado: (2025)