The Road Less Traveled: Enhancing Exploration in LLMs via Sequential Sampling
Fuente:
arXiv
Saved in:
| Main Authors: | Kang, Shijia, Zhang, Muhan |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
On the Completeness of Invariant Geometric Deep Learning Models
by: Li, Zian, et al.
Published: (2024)
by: Li, Zian, et al.
Published: (2024)
SPGNN: Recognizing Salient Subgraph Patterns via Enhanced Graph Convolution and Pooling
by: Dong, Zehao, et al.
Published: (2024)
by: Dong, Zehao, et al.
Published: (2024)
Sample Efficient Preference Alignment in LLMs via Active Exploration
by: Mehta, Viraj, et al.
Published: (2023)
by: Mehta, Viraj, et al.
Published: (2023)
The Road Less Scheduled
by: Defazio, Aaron, et al.
Published: (2024)
by: Defazio, Aaron, et al.
Published: (2024)
GREPO: A Benchmark for Graph Neural Networks on Repository-Level Bug Localization
by: Wang, Juntong, et al.
Published: (2026)
by: Wang, Juntong, et al.
Published: (2026)
An Empirical Study of Realized GNN Expressiveness
by: Wang, Yanbo, et al.
Published: (2023)
by: Wang, Yanbo, et al.
Published: (2023)
Enhancing Instruction Following of LLMs via Activation Steering with Dynamic Rejection
by: Kang, Minjae, et al.
Published: (2026)
by: Kang, Minjae, et al.
Published: (2026)
Enhanced Diffusion Sampling via Extrapolation with Multiple ODE Solutions
by: Choi, Jinyoung, et al.
Published: (2025)
by: Choi, Jinyoung, et al.
Published: (2025)
OCN: Effectively Utilizing Higher-Order Common Neighbors for Better Link Prediction
by: Wang, Juntong, et al.
Published: (2025)
by: Wang, Juntong, et al.
Published: (2025)
PiSSA: Principal Singular Values and Singular Vectors Adaptation of Large Language Models
by: Meng, Fanxu, et al.
Published: (2024)
by: Meng, Fanxu, et al.
Published: (2024)
Relational In-Context Learning via Synthetic Pre-training with Structural Prior
by: Wang, Yanbo, et al.
Published: (2026)
by: Wang, Yanbo, et al.
Published: (2026)
Offline Multi-Agent Reinforcement Learning via In-Sample Sequential Policy Optimization
by: Liu, Zongkai, et al.
Published: (2024)
by: Liu, Zongkai, et al.
Published: (2024)
Thompson Sampling via Fine-Tuning of LLMs
by: Menet, Nicolas, et al.
Published: (2025)
by: Menet, Nicolas, et al.
Published: (2025)
SAGE: Shaping Anchors for Guided Exploration in RLVR of LLMs
by: Lee, Chanuk, et al.
Published: (2026)
by: Lee, Chanuk, et al.
Published: (2026)
More Efficient Randomized Exploration for Reinforcement Learning via Approximate Sampling
by: Ishfaq, Haque, et al.
Published: (2024)
by: Ishfaq, Haque, et al.
Published: (2024)
Enhancing Temporal Modeling of Video LLMs via Time Gating
by: Hu, Zi-Yuan, et al.
Published: (2024)
by: Hu, Zi-Yuan, et al.
Published: (2024)
Enhanced Importance Sampling through Latent Space Exploration in Normalizing Flows
by: Kruse, Liam A., et al.
Published: (2025)
by: Kruse, Liam A., et al.
Published: (2025)
Scalable Equilibrium Sampling with Sequential Boltzmann Generators
by: Tan, Charlie B., et al.
Published: (2025)
by: Tan, Charlie B., et al.
Published: (2025)
ShiftAddLLM: Accelerating Pretrained LLMs via Post-Training Multiplication-Less Reparameterization
by: You, Haoran, et al.
Published: (2024)
by: You, Haoran, et al.
Published: (2024)
Sampling for Quality: Training-Free Reward-Guided LLM Decoding via Sequential Monte Carlo
by: Markovic-Voronov, Jelena, et al.
Published: (2026)
by: Markovic-Voronov, Jelena, et al.
Published: (2026)
The Expressivity Boundary of Probabilistic Circuits: A Comparison with Large Language Models
by: Zhao, Zhiyu, et al.
Published: (2026)
by: Zhao, Zhiyu, et al.
Published: (2026)
CLOVER: Cross-Layer Orthogonal Vectors Pruning and Fine-Tuning
by: Meng, Fanxu, et al.
Published: (2024)
by: Meng, Fanxu, et al.
Published: (2024)
Reconsidering the Performance of GAE in Link Prediction
by: Ma, Weishuo, et al.
Published: (2024)
by: Ma, Weishuo, et al.
Published: (2024)
Is Distance Matrix Enough for Geometric Deep Learning?
by: Li, Zian, et al.
Published: (2023)
by: Li, Zian, et al.
Published: (2023)
Dynamic Hypergraph-Enhanced Prediction of Sequential Medical Visits
by: Yang, Wangying, et al.
Published: (2024)
by: Yang, Wangying, et al.
Published: (2024)
Knapsack RL: Unlocking Exploration of LLMs via Optimizing Budget Allocation
by: Li, Ziniu, et al.
Published: (2025)
by: Li, Ziniu, et al.
Published: (2025)
When are LLMs Sufficient Policy Optimizers for Sequential RL Tasks?
by: Hatgis-Kessell, Stephane, et al.
Published: (2026)
by: Hatgis-Kessell, Stephane, et al.
Published: (2026)
Conditional Factuality Controlled LLMs with Generalization Certificates via Conformal Sampling
by: Ye, Kai, et al.
Published: (2026)
by: Ye, Kai, et al.
Published: (2026)
Long Context, Less Focus: A Scaling Gap in LLMs Revealed through Privacy and Personalization
by: Gu, Shangding
Published: (2026)
by: Gu, Shangding
Published: (2026)
Metacognitive Capabilities of LLMs: An Exploration in Mathematical Problem Solving
by: Didolkar, Aniket, et al.
Published: (2024)
by: Didolkar, Aniket, et al.
Published: (2024)
EVA-GAN: Enhanced Various Audio Generation via Scalable Generative Adversarial Networks
by: Liao, Shijia, et al.
Published: (2024)
by: Liao, Shijia, et al.
Published: (2024)
Empowering Contrastive Federated Sequential Recommendation with LLMs
by: Nguyen, Thi Minh Chau, et al.
Published: (2026)
by: Nguyen, Thi Minh Chau, et al.
Published: (2026)
RulE: Knowledge Graph Reasoning with Rule Embedding
by: Tang, Xiaojuan, et al.
Published: (2022)
by: Tang, Xiaojuan, et al.
Published: (2022)
SOE: Sample-Efficient Robot Policy Self-Improvement via On-Manifold Exploration
by: Jin, Yang, et al.
Published: (2025)
by: Jin, Yang, et al.
Published: (2025)
Meta Pruning via Graph Metanetworks : A Universal Meta Learning Framework for Network Pruning
by: Liu, Yewei, et al.
Published: (2025)
by: Liu, Yewei, et al.
Published: (2025)
Unlocking Reasoning Capabilities in LLMs via Reinforcement Learning Exploration
by: Deng, Wenhao, et al.
Published: (2025)
by: Deng, Wenhao, et al.
Published: (2025)
Learning More with Less: A Dynamic Dual-Level Down-Sampling Framework for Efficient Policy Optimization
by: Wang, Chao, et al.
Published: (2025)
by: Wang, Chao, et al.
Published: (2025)
Knowledge is Not Enough: Injecting RL Skills for Continual Adaptation
by: Tang, Pingzhi, et al.
Published: (2026)
by: Tang, Pingzhi, et al.
Published: (2026)
Easy Samples Are All You Need: Self-Evolving LLMs via Data-Efficient Reinforcement Learning
by: Yu, Zhiyin, et al.
Published: (2026)
by: Yu, Zhiyin, et al.
Published: (2026)
When More is Less: Understanding Chain-of-Thought Length in LLMs
by: Wu, Yuyang, et al.
Published: (2025)
by: Wu, Yuyang, et al.
Published: (2025)
Similar Items
-
On the Completeness of Invariant Geometric Deep Learning Models
by: Li, Zian, et al.
Published: (2024) -
SPGNN: Recognizing Salient Subgraph Patterns via Enhanced Graph Convolution and Pooling
by: Dong, Zehao, et al.
Published: (2024) -
Sample Efficient Preference Alignment in LLMs via Active Exploration
by: Mehta, Viraj, et al.
Published: (2023) -
The Road Less Scheduled
by: Defazio, Aaron, et al.
Published: (2024) -
GREPO: A Benchmark for Graph Neural Networks on Repository-Level Bug Localization
by: Wang, Juntong, et al.
Published: (2026)