Selective Steering: Norm-Preserving Control Through Discriminative Layer Selection
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Dang, Quy-Anh, Ngo, Chris |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2026
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
MoD: A Distribution-Based Approach for Merging Large Language Models
von: Dang, Quy-Anh, et al.
Veröffentlicht: (2024)
von: Dang, Quy-Anh, et al.
Veröffentlicht: (2024)
Reinforcement Learning for Reasoning in Small LLMs: What Works and What Doesn't
von: Dang, Quy-Anh, et al.
Veröffentlicht: (2025)
von: Dang, Quy-Anh, et al.
Veröffentlicht: (2025)
Adaptive Acquisition Selection for Bayesian Optimization with Large Language Models
von: Ngo, Giang, et al.
Veröffentlicht: (2026)
von: Ngo, Giang, et al.
Veröffentlicht: (2026)
Global Evolutionary Steering: Refining Activation Steering Control via Cross-Layer Consistency
von: Jiang, Xinyan, et al.
Veröffentlicht: (2026)
von: Jiang, Xinyan, et al.
Veröffentlicht: (2026)
Diversity Progress for Goal Selection in Discriminability-Motivated RL
von: Lintunen, Erik M., et al.
Veröffentlicht: (2024)
von: Lintunen, Erik M., et al.
Veröffentlicht: (2024)
SAEs Are Good for Steering -- If You Select the Right Features
von: Arad, Dana, et al.
Veröffentlicht: (2025)
von: Arad, Dana, et al.
Veröffentlicht: (2025)
The Signal is in the Steps: Local Scoring for Reasoning Data Selection
von: Just, Hoang Anh, et al.
Veröffentlicht: (2025)
von: Just, Hoang Anh, et al.
Veröffentlicht: (2025)
Noise-aware Client Selection for carbon-efficient Federated Learning via Gradient Norm Thresholding
von: Wilhelm, Patrick, et al.
Veröffentlicht: (2026)
von: Wilhelm, Patrick, et al.
Veröffentlicht: (2026)
Minimizing Collateral Damage in Activation Steering
von: Nguyen, Tam, et al.
Veröffentlicht: (2026)
von: Nguyen, Tam, et al.
Veröffentlicht: (2026)
LPLgrad: Optimizing Active Learning Through Gradient Norm Sample Selection and Auxiliary Model Training
von: Gul, Shreen, et al.
Veröffentlicht: (2024)
von: Gul, Shreen, et al.
Veröffentlicht: (2024)
The Hidden Signal of Verifier Strictness: Controlling and Improving Step-Wise Verification via Selective Latent Steering
von: Zhou, Yefan, et al.
Veröffentlicht: (2026)
von: Zhou, Yefan, et al.
Veröffentlicht: (2026)
Privacy-Preserving Dynamic Assortment Selection
von: Cho, Young Hyun, et al.
Veröffentlicht: (2024)
von: Cho, Young Hyun, et al.
Veröffentlicht: (2024)
Selective Conformal Risk Control
von: Xu, Yunpeng, et al.
Veröffentlicht: (2025)
von: Xu, Yunpeng, et al.
Veröffentlicht: (2025)
Permutation-Invariant Representation Learning for Robust and Privacy-Preserving Feature Selection
von: Liu, Rui, et al.
Veröffentlicht: (2025)
von: Liu, Rui, et al.
Veröffentlicht: (2025)
Steering LLM Reasoning Through Bias-Only Adaptation
von: Sinii, Viacheslav, et al.
Veröffentlicht: (2025)
von: Sinii, Viacheslav, et al.
Veröffentlicht: (2025)
Test-time Diverse Reasoning by Riemannian Activation Steering
von: Khanh, Ly Tran Ho, et al.
Veröffentlicht: (2025)
von: Khanh, Ly Tran Ho, et al.
Veröffentlicht: (2025)
K-means Derived Unsupervised Feature Selection using Improved ADMM
von: Sun, Ziheng, et al.
Veröffentlicht: (2024)
von: Sun, Ziheng, et al.
Veröffentlicht: (2024)
UPCORE: Utility-Preserving Coreset Selection for Balanced Unlearning
von: Patil, Vaidehi, et al.
Veröffentlicht: (2025)
von: Patil, Vaidehi, et al.
Veröffentlicht: (2025)
Diverse Subset Selection via Norm-Based Sampling and Orthogonality
von: Bar, Noga, et al.
Veröffentlicht: (2024)
von: Bar, Noga, et al.
Veröffentlicht: (2024)
SLaNC: Static LayerNorm Calibration
von: Salmani, Mahsa, et al.
Veröffentlicht: (2024)
von: Salmani, Mahsa, et al.
Veröffentlicht: (2024)
Shaping Up SHAP: Enhancing Stability through Layer-Wise Neighbor Selection
von: Kelodjou, Gwladys, et al.
Veröffentlicht: (2023)
von: Kelodjou, Gwladys, et al.
Veröffentlicht: (2023)
Adaptive Layer Selection for Layer-Wise Token Pruning in LLM Inference
von: Taniguchi, Rei, et al.
Veröffentlicht: (2026)
von: Taniguchi, Rei, et al.
Veröffentlicht: (2026)
Active Level Set Estimation for Continuous Search Space with Theoretical Guarantee
von: Ngo, Giang, et al.
Veröffentlicht: (2024)
von: Ngo, Giang, et al.
Veröffentlicht: (2024)
PHEATPRUNER: Interpretable Data-centric Feature Selection for Multivariate Time Series Classification through Persistent Homology
von: Pham, Anh-Duy, et al.
Veröffentlicht: (2025)
von: Pham, Anh-Duy, et al.
Veröffentlicht: (2025)
QUOKA: Query-Oriented KV Selection For Efficient LLM Prefill
von: Jones, Dalton, et al.
Veröffentlicht: (2026)
von: Jones, Dalton, et al.
Veröffentlicht: (2026)
Where to Steer: Input-Dependent Layer Selection for Steering Improves LLM Alignment
von: Gadgil, Soham, et al.
Veröffentlicht: (2026)
von: Gadgil, Soham, et al.
Veröffentlicht: (2026)
KnapSpec: Self-Speculative Decoding via Adaptive Layer Selection as a Knapsack Problem
von: Cha, Seongjin, et al.
Veröffentlicht: (2026)
von: Cha, Seongjin, et al.
Veröffentlicht: (2026)
Adaptive Selection of LoRA Components in Privacy-Preserving Federated Learning
von: Kim, Myoungjun, et al.
Veröffentlicht: (2026)
von: Kim, Myoungjun, et al.
Veröffentlicht: (2026)
Automatic Unsupervised Ensemble Outlier Model Selection--Extended Version
von: Phan, Hong-Phuc, et al.
Veröffentlicht: (2026)
von: Phan, Hong-Phuc, et al.
Veröffentlicht: (2026)
PNCS:Power-Norm Cosine Similarity for Diverse Client Selection in Federated Learning
von: Li, Liangyan, et al.
Veröffentlicht: (2025)
von: Li, Liangyan, et al.
Veröffentlicht: (2025)
In-context Vectors: Making In Context Learning More Effective and Controllable Through Latent Space Steering
von: Liu, Sheng, et al.
Veröffentlicht: (2023)
von: Liu, Sheng, et al.
Veröffentlicht: (2023)
Angular Steering: Behavior Control via Rotation in Activation Space
von: Vu, Hieu M., et al.
Veröffentlicht: (2025)
von: Vu, Hieu M., et al.
Veröffentlicht: (2025)
Metric-DST: Mitigating Selection Bias Through Diversity-Guided Semi-Supervised Metric Learning
von: Tepeli, Yasin I., et al.
Veröffentlicht: (2024)
von: Tepeli, Yasin I., et al.
Veröffentlicht: (2024)
Utilizing Data Fingerprints for Privacy-Preserving Algorithm Selection in Time Series Classification: Performance and Uncertainty Estimation on Unseen Datasets
von: Böcking, Lars, et al.
Veröffentlicht: (2024)
von: Böcking, Lars, et al.
Veröffentlicht: (2024)
T-SHIRT: Token-Selective Hierarchical Data Selection for Instruction Tuning
von: Fu, Yanjun, et al.
Veröffentlicht: (2025)
von: Fu, Yanjun, et al.
Veröffentlicht: (2025)
Data-Free Privacy-Preserving for LLMs via Model Inversion and Selective Unlearning
von: Zhou, Xinjie, et al.
Veröffentlicht: (2026)
von: Zhou, Xinjie, et al.
Veröffentlicht: (2026)
MID-L: Matrix-Interpolated Dropout Layer with Layer-wise Neuron Selection
von: Shaeri, Pouya, et al.
Veröffentlicht: (2025)
von: Shaeri, Pouya, et al.
Veröffentlicht: (2025)
Learning Privacy-Preserving Student Networks via Discriminative-Generative Distillation
von: Ge, Shiming, et al.
Veröffentlicht: (2024)
von: Ge, Shiming, et al.
Veröffentlicht: (2024)
wav2graph: A Framework for Supervised Learning Knowledge Graph from Speech
von: Le-Duc, Khai, et al.
Veröffentlicht: (2024)
von: Le-Duc, Khai, et al.
Veröffentlicht: (2024)
Conformal Selective Acting: Anytime-Valid Risk Control for RLVR-Trained LLMs
von: Khosravi, Hamed, et al.
Veröffentlicht: (2026)
von: Khosravi, Hamed, et al.
Veröffentlicht: (2026)
Ähnliche Einträge
-
MoD: A Distribution-Based Approach for Merging Large Language Models
von: Dang, Quy-Anh, et al.
Veröffentlicht: (2024) -
Reinforcement Learning for Reasoning in Small LLMs: What Works and What Doesn't
von: Dang, Quy-Anh, et al.
Veröffentlicht: (2025) -
Adaptive Acquisition Selection for Bayesian Optimization with Large Language Models
von: Ngo, Giang, et al.
Veröffentlicht: (2026) -
Global Evolutionary Steering: Refining Activation Steering Control via Cross-Layer Consistency
von: Jiang, Xinyan, et al.
Veröffentlicht: (2026) -
Diversity Progress for Goal Selection in Discriminability-Motivated RL
von: Lintunen, Erik M., et al.
Veröffentlicht: (2024)