A Systematic Analysis of the Impact of Persona Steering on LLM Capabilities
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Chen, Jiaqi, Wang, Ming, Xie, Tingna, Feng, Shi, Liu, Yongkang |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2026
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
The Impact of Steering Large Language Models with Persona Vectors in Educational Applications
von: Wu, Yongchao, et al.
Veröffentlicht: (2026)
von: Wu, Yongchao, et al.
Veröffentlicht: (2026)
Controllable LLM Reasoning via Sparse Autoencoder-Based Steering
von: Fang, Yi, et al.
Veröffentlicht: (2026)
von: Fang, Yi, et al.
Veröffentlicht: (2026)
Steering at the Source: Style Modulation Heads for Robust Persona Control
von: Izawa, Yoshihiro, et al.
Veröffentlicht: (2026)
von: Izawa, Yoshihiro, et al.
Veröffentlicht: (2026)
BILLY: Steering Large Language Models via Merging Persona Vectors for Creative Generation
von: Pai, Tsung-Min, et al.
Veröffentlicht: (2025)
von: Pai, Tsung-Min, et al.
Veröffentlicht: (2025)
SteeringSafety: A Systematic Safety Evaluation Framework of Representation Steering in LLMs
von: Siu, Vincent, et al.
Veröffentlicht: (2025)
von: Siu, Vincent, et al.
Veröffentlicht: (2025)
Systematic Analysis of LLM Contributions to Planning: Solver, Verifier, Heuristic
von: Li, Haoming, et al.
Veröffentlicht: (2024)
von: Li, Haoming, et al.
Veröffentlicht: (2024)
EasySteer: A Unified Framework for High-Performance and Extensible LLM Steering
von: Xu, Haolei, et al.
Veröffentlicht: (2025)
von: Xu, Haolei, et al.
Veröffentlicht: (2025)
PersonaMatrix: A Recipe for Persona-Aware Evaluation of Legal Summarization
von: Pang, Tsz Fung, et al.
Veröffentlicht: (2025)
von: Pang, Tsz Fung, et al.
Veröffentlicht: (2025)
Steering LLM Thinking with Budget Guidance
von: Li, Junyan, et al.
Veröffentlicht: (2025)
von: Li, Junyan, et al.
Veröffentlicht: (2025)
Improving LLM Reasoning through Interpretable Role-Playing Steering
von: Wang, Anyi, et al.
Veröffentlicht: (2025)
von: Wang, Anyi, et al.
Veröffentlicht: (2025)
Tracing Persona Vectors Through LLM Pretraining
von: Moskvoretskii, Viktor, et al.
Veröffentlicht: (2026)
von: Moskvoretskii, Viktor, et al.
Veröffentlicht: (2026)
Where is the Mind? Persona Vectors and LLM Individuation
von: Beckmann, Pierre, et al.
Veröffentlicht: (2026)
von: Beckmann, Pierre, et al.
Veröffentlicht: (2026)
NewsBench: A Systematic Evaluation Framework for Assessing Editorial Capabilities of Large Language Models in Chinese Journalism
von: Li, Miao, et al.
Veröffentlicht: (2024)
von: Li, Miao, et al.
Veröffentlicht: (2024)
From Persona to Personalization: A Survey on Role-Playing Language Agents
von: Chen, Jiangjie, et al.
Veröffentlicht: (2024)
von: Chen, Jiangjie, et al.
Veröffentlicht: (2024)
Adaptive Activation Steering: A Tuning-Free LLM Truthfulness Improvement Method for Diverse Hallucinations Categories
von: Wang, Tianlong, et al.
Veröffentlicht: (2024)
von: Wang, Tianlong, et al.
Veröffentlicht: (2024)
Persona-judge: Personalized Alignment of Large Language Models via Token-level Self-judgment
von: Zhang, Xiaotian, et al.
Veröffentlicht: (2025)
von: Zhang, Xiaotian, et al.
Veröffentlicht: (2025)
Polypersona: Persona-Grounded LLM for Synthetic Survey Responses
von: Dash, Tejaswani, et al.
Veröffentlicht: (2025)
von: Dash, Tejaswani, et al.
Veröffentlicht: (2025)
Bullying the Machine: How Personas Increase LLM Vulnerability
von: Xu, Ziwei, et al.
Veröffentlicht: (2025)
von: Xu, Ziwei, et al.
Veröffentlicht: (2025)
Effectively Steer LLM To Follow Preference via Building Confident Directions
von: Song, Bingqing, et al.
Veröffentlicht: (2025)
von: Song, Bingqing, et al.
Veröffentlicht: (2025)
WorkForceAgent-R1: Incentivizing Reasoning Capability in LLM-based Web Agents via Reinforcement Learning
von: Zhuang, Yuchen, et al.
Veröffentlicht: (2025)
von: Zhuang, Yuchen, et al.
Veröffentlicht: (2025)
CoSteer: Collaborative Decoding-Time Personalization via Local Delta Steering
von: Lv, Hang, et al.
Veröffentlicht: (2025)
von: Lv, Hang, et al.
Veröffentlicht: (2025)
Steer Like the LLM: Activation Steering that Mimics Prompting
von: Heyman, Geert, et al.
Veröffentlicht: (2026)
von: Heyman, Geert, et al.
Veröffentlicht: (2026)
CoSER: A Comprehensive Literary Dataset and Framework for Training and Evaluating LLM Role-Playing and Persona Simulation
von: Wang, Xintao, et al.
Veröffentlicht: (2025)
von: Wang, Xintao, et al.
Veröffentlicht: (2025)
Judge's Verdict: A Comprehensive Analysis of LLM Judge Capability Through Human Agreement
von: Han, Steve, et al.
Veröffentlicht: (2025)
von: Han, Steve, et al.
Veröffentlicht: (2025)
TalkDep: Clinically Grounded LLM Personas for Conversation-Centric Depression Screening
von: Wang, Xi, et al.
Veröffentlicht: (2025)
von: Wang, Xi, et al.
Veröffentlicht: (2025)
Population-Aligned Persona Generation for LLM-based Social Simulation
von: Hu, Zhengyu, et al.
Veröffentlicht: (2025)
von: Hu, Zhengyu, et al.
Veröffentlicht: (2025)
Concept-Level Explainability for Auditing & Steering LLM Responses
von: Amara, Kenza, et al.
Veröffentlicht: (2025)
von: Amara, Kenza, et al.
Veröffentlicht: (2025)
Steer LLM Latents for Hallucination Detection
von: Park, Seongheon, et al.
Veröffentlicht: (2025)
von: Park, Seongheon, et al.
Veröffentlicht: (2025)
Algorithmic Fragility and Persona Bias in LLM-Generated Autistic Communication
von: Rizvi, Naba, et al.
Veröffentlicht: (2026)
von: Rizvi, Naba, et al.
Veröffentlicht: (2026)
SensorPersona: An LLM-Empowered System for Continual Persona Extraction from Longitudinal Mobile Sensor Streams
von: Yang, Bufang, et al.
Veröffentlicht: (2026)
von: Yang, Bufang, et al.
Veröffentlicht: (2026)
PlanGenLLMs: A Modern Survey of LLM Planning Capabilities
von: Wei, Hui, et al.
Veröffentlicht: (2025)
von: Wei, Hui, et al.
Veröffentlicht: (2025)
The Impact of Persona-based Political Perspectives on Hateful Content Detection
von: Civelli, Stefano, et al.
Veröffentlicht: (2025)
von: Civelli, Stefano, et al.
Veröffentlicht: (2025)
TReB: A Comprehensive Benchmark for Evaluating Table Reasoning Capabilities of Large Language Models
von: Li, Ce, et al.
Veröffentlicht: (2025)
von: Li, Ce, et al.
Veröffentlicht: (2025)
TEaR: Improving LLM-based Machine Translation with Systematic Self-Refinement
von: Feng, Zhaopeng, et al.
Veröffentlicht: (2024)
von: Feng, Zhaopeng, et al.
Veröffentlicht: (2024)
German General Social Survey Personas: A Survey-Derived Persona Prompt Collection for Population-Aligned LLM Studies
von: Rupprecht, Jens, et al.
Veröffentlicht: (2025)
von: Rupprecht, Jens, et al.
Veröffentlicht: (2025)
When Wording Steers the Evaluation: Framing Bias in LLM judges
von: Hwang, Yerin, et al.
Veröffentlicht: (2026)
von: Hwang, Yerin, et al.
Veröffentlicht: (2026)
Critical-Questions-of-Thought: Steering LLM reasoning with Argumentative Querying
von: Castagna, Federico, et al.
Veröffentlicht: (2024)
von: Castagna, Federico, et al.
Veröffentlicht: (2024)
Unleashing LLM Reasoning Capability via Scalable Question Synthesis from Scratch
von: Ding, Yuyang, et al.
Veröffentlicht: (2024)
von: Ding, Yuyang, et al.
Veröffentlicht: (2024)
Breaking the Capability Ceiling of LLM Post-Training by Reintroducing Markov States
von: Yuan, Yurun, et al.
Veröffentlicht: (2026)
von: Yuan, Yurun, et al.
Veröffentlicht: (2026)
Pay What LLM Wants: Can LLM Simulate Economics Experiment with 522 Real-human Persona?
von: Choi, Junhyuk, et al.
Veröffentlicht: (2025)
von: Choi, Junhyuk, et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
The Impact of Steering Large Language Models with Persona Vectors in Educational Applications
von: Wu, Yongchao, et al.
Veröffentlicht: (2026) -
Controllable LLM Reasoning via Sparse Autoencoder-Based Steering
von: Fang, Yi, et al.
Veröffentlicht: (2026) -
Steering at the Source: Style Modulation Heads for Robust Persona Control
von: Izawa, Yoshihiro, et al.
Veröffentlicht: (2026) -
BILLY: Steering Large Language Models via Merging Persona Vectors for Creative Generation
von: Pai, Tsung-Min, et al.
Veröffentlicht: (2025) -
SteeringSafety: A Systematic Safety Evaluation Framework of Representation Steering in LLMs
von: Siu, Vincent, et al.
Veröffentlicht: (2025)