Guardado en:
| Autores principales: | Adewuyi, Israel, Okibe, Solomon, Ivanov, Vladmir |
|---|---|
| Formato: | Preprint |
| Publicado: |
2026
|
| Materias: | |
| Acceso en línea: | https://arxiv.org/abs/2602.01599 |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
On the Sparsity of the Strong Lottery Ticket Hypothesis
por: Natale, Emanuele, et al.
Publicado: (2024)
por: Natale, Emanuele, et al.
Publicado: (2024)
The Strong Lottery Ticket Hypothesis for Multi-Head Attention Mechanisms
por: Otsuka, Hikari, et al.
Publicado: (2025)
por: Otsuka, Hikari, et al.
Publicado: (2025)
Investigating the Lottery Ticket Hypothesis for Variational Quantum Circuits
por: Kölle, Michael, et al.
Publicado: (2025)
por: Kölle, Michael, et al.
Publicado: (2025)
Reinforcement Learning Fine-Tunes a Sparse Subnetwork in Large Language Models
por: Balashov, Andrii
Publicado: (2025)
por: Balashov, Andrii
Publicado: (2025)
Partially Frozen Random Networks Contain Compact Strong Lottery Tickets
por: Otsuka, Hikari, et al.
Publicado: (2024)
por: Otsuka, Hikari, et al.
Publicado: (2024)
Instilling Inductive Biases with Subnetworks
por: Zhang, Enyan, et al.
Publicado: (2023)
por: Zhang, Enyan, et al.
Publicado: (2023)
Model Parallelism With Subnetwork Data Parallelism
por: Singh, Vaibhav, et al.
Publicado: (2025)
por: Singh, Vaibhav, et al.
Publicado: (2025)
IRDS: Interpretable RLVR Data Selection via Verifier-Coupled Sparse Autoencoder Coverage
por: Li, Yuhan, et al.
Publicado: (2026)
por: Li, Yuhan, et al.
Publicado: (2026)
SSFL: Discovering Sparse Unified Subnetworks at Initialization for Efficient Federated Learning
por: Ohib, Riyasat, et al.
Publicado: (2024)
por: Ohib, Riyasat, et al.
Publicado: (2024)
Random Masking Finds Winning Tickets for Parameter Efficient Fine-tuning
por: Xu, Jing, et al.
Publicado: (2024)
por: Xu, Jing, et al.
Publicado: (2024)
LLM-Generated Explanations Do Not Suffice for Ultra-Strong Machine Learning
por: Ai, Lun, et al.
Publicado: (2025)
por: Ai, Lun, et al.
Publicado: (2025)
Memory Constrained Dynamic Subnetwork Update for Transfer Learning
por: Quélennec, Aël, et al.
Publicado: (2025)
por: Quélennec, Aël, et al.
Publicado: (2025)
Linear Bellman Completeness Suffices for Efficient Online Reinforcement Learning with Few Actions
por: Golowich, Noah, et al.
Publicado: (2024)
por: Golowich, Noah, et al.
Publicado: (2024)
Sparse Tokens Suffice: Jailbreaking Audio Language Models via Token-Aware Gradient Optimization
por: Fang, Zheng, et al.
Publicado: (2026)
por: Fang, Zheng, et al.
Publicado: (2026)
A Combinatorial Theory of Dropout: Subnetworks, Graph Geometry, and Generalization
por: Dhayalkar, Sahil Rajesh
Publicado: (2025)
por: Dhayalkar, Sahil Rajesh
Publicado: (2025)
FedSI: Federated Subnetwork Inference for Efficient Uncertainty Quantification
por: Chen, Hui, et al.
Publicado: (2024)
por: Chen, Hui, et al.
Publicado: (2024)
Sparse but Critical: A Token-Level Analysis of Distributional Shifts in RLVR Fine-Tuning of LLMs
por: Meng, Haoming, et al.
Publicado: (2026)
por: Meng, Haoming, et al.
Publicado: (2026)
Quantifying Empirical Compute-Supervision Tradeoffs in RLVR
por: Mitsuhashi, Ryo, et al.
Publicado: (2026)
por: Mitsuhashi, Ryo, et al.
Publicado: (2026)
VL Norm: Rethink Loss Aggregation in RLVR
por: He, Zhiyuan, et al.
Publicado: (2025)
por: He, Zhiyuan, et al.
Publicado: (2025)
Spurious Rewards: Rethinking Training Signals in RLVR
por: Shao, Rulin, et al.
Publicado: (2025)
por: Shao, Rulin, et al.
Publicado: (2025)
Grokking as Structural Inference: Transformers Need Bayesian Lottery Tickets
por: Hidajat, Kai, et al.
Publicado: (2026)
por: Hidajat, Kai, et al.
Publicado: (2026)
Continual Deep Learning on the Edge via Stochastic Local Competition among Subnetworks
por: Christophides, Theodoros, et al.
Publicado: (2024)
por: Christophides, Theodoros, et al.
Publicado: (2024)
Discovering Knowledge-Critical Subnetworks in Pretrained Language Models
por: Bayazit, Deniz, et al.
Publicado: (2023)
por: Bayazit, Deniz, et al.
Publicado: (2023)
On the Direction of RLVR Updates for LLM Reasoning: Identification and Exploitation
por: Huang, Kexin, et al.
Publicado: (2026)
por: Huang, Kexin, et al.
Publicado: (2026)
On the Implicit Reward Overfitting and the Low-rank Dynamics in RLVR
por: Ye, Hao, et al.
Publicado: (2026)
por: Ye, Hao, et al.
Publicado: (2026)
The Path Not Taken: RLVR Provably Learns Off the Principals
por: Zhu, Hanqing, et al.
Publicado: (2025)
por: Zhu, Hanqing, et al.
Publicado: (2025)
RLVR-World: Training World Models with Reinforcement Learning
por: Wu, Jialong, et al.
Publicado: (2025)
por: Wu, Jialong, et al.
Publicado: (2025)
Rethinking Entropy Interventions in RLVR: An Entropy Change Perspective
por: Hao, Zhezheng, et al.
Publicado: (2025)
por: Hao, Zhezheng, et al.
Publicado: (2025)
Quantile Advantage Estimation: Stabilizing RLVR for LLM Reasoning
por: Wu, Junkang, et al.
Publicado: (2025)
por: Wu, Junkang, et al.
Publicado: (2025)
Task-specific Subnetwork Discovery in Reinforcement Learning for Autonomous Underwater Navigation
por: Liu, Yi-Ling, et al.
Publicado: (2026)
por: Liu, Yi-Ling, et al.
Publicado: (2026)
Uncovering a Winning Lottery Ticket with Continuously Relaxed Bernoulli Gates
por: Tsayag, Itamar, et al.
Publicado: (2026)
por: Tsayag, Itamar, et al.
Publicado: (2026)
Exploring Subnetwork Interactions in Heterogeneous Brain Network via Prior-Informed Graph Learning
por: Liu, Siyu, et al.
Publicado: (2026)
por: Liu, Siyu, et al.
Publicado: (2026)
Stochastic Collapse: How Gradient Noise Attracts SGD Dynamics Towards Simpler Subnetworks
por: Chen, Feng, et al.
Publicado: (2023)
por: Chen, Feng, et al.
Publicado: (2023)
Controllable Exploration in Hybrid-Policy RLVR for Multi-Modal Reasoning
por: Huang, Zhuoxu, et al.
Publicado: (2026)
por: Huang, Zhuoxu, et al.
Publicado: (2026)
LLMs Gaming Verifiers: RLVR can Lead to Reward Hacking
por: Helff, Lukas, et al.
Publicado: (2026)
por: Helff, Lukas, et al.
Publicado: (2026)
GeoRA: Geometry-Aware Low-Rank Adaptation for RLVR
por: Zhang, Jiaying, et al.
Publicado: (2026)
por: Zhang, Jiaying, et al.
Publicado: (2026)
Beyond Uniform Credit Assignment: Selective Eligibility Traces for RLVR
por: Mou, Chaoli, et al.
Publicado: (2026)
por: Mou, Chaoli, et al.
Publicado: (2026)
Learning Rate Matters: Vanilla LoRA May Suffice for LLM Fine-tuning
por: Lee, Yu-Ang, et al.
Publicado: (2026)
por: Lee, Yu-Ang, et al.
Publicado: (2026)
Routing the Lottery: Adaptive Subnetworks for Heterogeneous Data
por: Stefanski, Grzegorz, et al.
Publicado: (2026)
por: Stefanski, Grzegorz, et al.
Publicado: (2026)
Do Sparse Subnetworks Exhibit Cognitively Aligned Attention? Effects of Pruning on Saliency Map Fidelity, Sparsity, and Concept Coherence
por: Suwal, Sanish, et al.
Publicado: (2025)
por: Suwal, Sanish, et al.
Publicado: (2025)
Ejemplares similares
-
On the Sparsity of the Strong Lottery Ticket Hypothesis
por: Natale, Emanuele, et al.
Publicado: (2024) -
The Strong Lottery Ticket Hypothesis for Multi-Head Attention Mechanisms
por: Otsuka, Hikari, et al.
Publicado: (2025) -
Investigating the Lottery Ticket Hypothesis for Variational Quantum Circuits
por: Kölle, Michael, et al.
Publicado: (2025) -
Reinforcement Learning Fine-Tunes a Sparse Subnetwork in Large Language Models
por: Balashov, Andrii
Publicado: (2025) -
Partially Frozen Random Networks Contain Compact Strong Lottery Tickets
por: Otsuka, Hikari, et al.
Publicado: (2024)