Gespeichert in:
| Hauptverfasser: | Feng, Zijian, Li, Tianjiao, Zhu, Zixiao, Zhou, Hanzhang, Qian, Junlang, Zhang, Li, Chua, Jia Jim Deryl, Mak, Lee Onn, Ng, Gee Wah, Mao, Kezhi |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2026
|
| Schlagworte: | |
| Online-Zugang: | https://arxiv.org/abs/2602.04428 |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Restoring Pruned Large Language Models via Lost Component Compensation
von: Feng, Zijian, et al.
Veröffentlicht: (2025)
von: Feng, Zijian, et al.
Veröffentlicht: (2025)
Rethinking Prompt Optimizers: From Prompt Merits to Optimization
von: Zhu, Zixiao, et al.
Veröffentlicht: (2025)
von: Zhu, Zixiao, et al.
Veröffentlicht: (2025)
Unveiling and Manipulating Prompt Influence in Large Language Models
von: Feng, Zijian, et al.
Veröffentlicht: (2024)
von: Feng, Zijian, et al.
Veröffentlicht: (2024)
UniBias: Unveiling and Mitigating LLM Bias through Internal Attention and FFN Manipulation
von: Zhou, Hanzhang, et al.
Veröffentlicht: (2024)
von: Zhou, Hanzhang, et al.
Veröffentlicht: (2024)
Logit Separability-Driven Samples and Multiple Class-Related Words Selection for Advancing In-Context Learning
von: Zixiao, Zhu, et al.
Veröffentlicht: (2024)
von: Zixiao, Zhu, et al.
Veröffentlicht: (2024)
LLMs Learn Task Heuristics from Demonstrations: A Heuristic-Driven Prompting Strategy for Document-Level Event Argument Extraction
von: Zhou, Hanzhang, et al.
Veröffentlicht: (2023)
von: Zhou, Hanzhang, et al.
Veröffentlicht: (2023)
Beyond the Next Token: Towards Prompt-Robust Zero-Shot Classification via Efficient Multi-Token Prediction
von: Qian, Junlang, et al.
Veröffentlicht: (2025)
von: Qian, Junlang, et al.
Veröffentlicht: (2025)
FreeCtrl: Constructing Control Centers with Feedforward Layers for Learning-Free Controllable Text Generation
von: Feng, Zijian, et al.
Veröffentlicht: (2024)
von: Feng, Zijian, et al.
Veröffentlicht: (2024)
QCaption: Video Captioning and Q&A through Fusion of Large Multimodal Models
von: Wang, Jiale, et al.
Veröffentlicht: (2026)
von: Wang, Jiale, et al.
Veröffentlicht: (2026)
EmoSteer-TTS: Fine-Grained and Training-Free Emotion-Controllable Text-to-Speech via Activation Steering
von: Xie, Tianxin, et al.
Veröffentlicht: (2025)
von: Xie, Tianxin, et al.
Veröffentlicht: (2025)
Domain Lexical Knowledge-based Word Embedding Learning for Text Classification under Small Data
von: Zhu, Zixiao, et al.
Veröffentlicht: (2025)
von: Zhu, Zixiao, et al.
Veröffentlicht: (2025)
QMAVIS: Long Video-Audio Understanding using Fusion of Large Multimodal Models
von: Lin, Zixing, et al.
Veröffentlicht: (2026)
von: Lin, Zixing, et al.
Veröffentlicht: (2026)
FineSteer: A Unified Framework for Fine-Grained Inference-Time Steering in Large Language Models
von: Weng, Zixuan, et al.
Veröffentlicht: (2026)
von: Weng, Zixuan, et al.
Veröffentlicht: (2026)
Personalized Text Generation with Contrastive Activation Steering
von: Zhang, Jinghao, et al.
Veröffentlicht: (2025)
von: Zhang, Jinghao, et al.
Veröffentlicht: (2025)
HyperSteer: Activation Steering at Scale with Hypernetworks
von: Sun, Jiuding, et al.
Veröffentlicht: (2025)
von: Sun, Jiuding, et al.
Veröffentlicht: (2025)
Mitigating Content Effects on Reasoning in Language Models through Fine-Grained Activation Steering
von: Valentino, Marco, et al.
Veröffentlicht: (2025)
von: Valentino, Marco, et al.
Veröffentlicht: (2025)
SteerX: Disentangled Steering for LLM Personalization
von: Zhao, Xiaoyan, et al.
Veröffentlicht: (2025)
von: Zhao, Xiaoyan, et al.
Veröffentlicht: (2025)
FairSteer: Inference Time Debiasing for LLMs with Dynamic Activation Steering
von: Li, Yichen, et al.
Veröffentlicht: (2025)
von: Li, Yichen, et al.
Veröffentlicht: (2025)
Steer Like the LLM: Activation Steering that Mimics Prompting
von: Heyman, Geert, et al.
Veröffentlicht: (2026)
von: Heyman, Geert, et al.
Veröffentlicht: (2026)
RankSteer: Activation Steering for Pointwise LLM Ranking
von: Wang, Yumeng, et al.
Veröffentlicht: (2026)
von: Wang, Yumeng, et al.
Veröffentlicht: (2026)
Steering Awareness: Detecting Activation Steering from Within
von: Rivera, Joshua Fonseca, et al.
Veröffentlicht: (2025)
von: Rivera, Joshua Fonseca, et al.
Veröffentlicht: (2025)
On Optimal Steering to Achieve Exact Fairness
von: Sharma, Mohit, et al.
Veröffentlicht: (2025)
von: Sharma, Mohit, et al.
Veröffentlicht: (2025)
AlphaSteer: Learning Refusal Steering with Principled Null-Space Constraint
von: Sheng, Leheng, et al.
Veröffentlicht: (2025)
von: Sheng, Leheng, et al.
Veröffentlicht: (2025)
LF-Steering: Latent Feature Activation Steering for Enhancing Semantic Consistency in Large Language Models
von: Yang, Jingyuan, et al.
Veröffentlicht: (2025)
von: Yang, Jingyuan, et al.
Veröffentlicht: (2025)
Steer2Edit: From Activation Steering to Component-Level Editing
von: Sun, Chung-En, et al.
Veröffentlicht: (2026)
von: Sun, Chung-En, et al.
Veröffentlicht: (2026)
Dynamically Scaled Activation Steering
von: Ferrando, Alex, et al.
Veröffentlicht: (2025)
von: Ferrando, Alex, et al.
Veröffentlicht: (2025)
Steering Video Diffusion Transformers with Massive Activations
von: Cheng, Xianhang, et al.
Veröffentlicht: (2026)
von: Cheng, Xianhang, et al.
Veröffentlicht: (2026)
FoleyDirector: Fine-Grained Temporal Steering for Video-to-Audio Generation via Structured Scripts
von: Li, You, et al.
Veröffentlicht: (2026)
von: Li, You, et al.
Veröffentlicht: (2026)
Steering to Say No: Configurable Refusal via Activation Steering in Vision Language Models
von: Yang, Jiaxi, et al.
Veröffentlicht: (2026)
von: Yang, Jiaxi, et al.
Veröffentlicht: (2026)
MedSteer: Counterfactual Endoscopic Synthesis via Training-Free Activation Steering
von: Pham, Trong-Thang, et al.
Veröffentlicht: (2026)
von: Pham, Trong-Thang, et al.
Veröffentlicht: (2026)
Beyond Steering Vector: Flow-based Activation Steering for Inference-Time Intervention
von: Jin, Zehao, et al.
Veröffentlicht: (2026)
von: Jin, Zehao, et al.
Veröffentlicht: (2026)
SteerConf: Steering LLMs for Confidence Elicitation
von: Zhou, Ziang, et al.
Veröffentlicht: (2025)
von: Zhou, Ziang, et al.
Veröffentlicht: (2025)
Minimizing Collateral Damage in Activation Steering
von: Nguyen, Tam, et al.
Veröffentlicht: (2026)
von: Nguyen, Tam, et al.
Veröffentlicht: (2026)
Activation Steering with a Feedback Controller
von: Nguyen, Dung V., et al.
Veröffentlicht: (2025)
von: Nguyen, Dung V., et al.
Veröffentlicht: (2025)
Programming Refusal with Conditional Activation Steering
von: Lee, Bruce W., et al.
Veröffentlicht: (2024)
von: Lee, Bruce W., et al.
Veröffentlicht: (2024)
SAKE: Steering Activations for Knowledge Editing
von: Scialanga, Marco, et al.
Veröffentlicht: (2025)
von: Scialanga, Marco, et al.
Veröffentlicht: (2025)
Activation Steering for Chain-of-Thought Compression
von: Azizi, Seyedarmin, et al.
Veröffentlicht: (2025)
von: Azizi, Seyedarmin, et al.
Veröffentlicht: (2025)
Steering Language Models With Activation Engineering
von: Turner, Alexander Matt, et al.
Veröffentlicht: (2023)
von: Turner, Alexander Matt, et al.
Veröffentlicht: (2023)
Steered LLM Activations are Non-Surjective
von: Mishra, Aayush, et al.
Veröffentlicht: (2026)
von: Mishra, Aayush, et al.
Veröffentlicht: (2026)
MLLMEraser: Achieving Test-Time Unlearning in Multimodal Large Language Models through Activation Steering
von: Ding, Chenlu, et al.
Veröffentlicht: (2025)
von: Ding, Chenlu, et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
Restoring Pruned Large Language Models via Lost Component Compensation
von: Feng, Zijian, et al.
Veröffentlicht: (2025) -
Rethinking Prompt Optimizers: From Prompt Merits to Optimization
von: Zhu, Zixiao, et al.
Veröffentlicht: (2025) -
Unveiling and Manipulating Prompt Influence in Large Language Models
von: Feng, Zijian, et al.
Veröffentlicht: (2024) -
UniBias: Unveiling and Mitigating LLM Bias through Internal Attention and FFN Manipulation
von: Zhou, Hanzhang, et al.
Veröffentlicht: (2024) -
Logit Separability-Driven Samples and Multiple Class-Related Words Selection for Advancing In-Context Learning
von: Zixiao, Zhu, et al.
Veröffentlicht: (2024)