How Emotion Shapes the Behavior of LLMs and Agents: A Mechanistic Study
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Sun, Moran, Li, Tianlin, Zheng, Yuwei, Zhou, Zhenhong, Liu, Aishan, Liu, Xianglong, Liu, Yang |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2026
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
LLMCBench: Benchmarking Large Language Model Compression for Efficient Deployment
von: Yang, Ge, et al.
Veröffentlicht: (2024)
von: Yang, Ge, et al.
Veröffentlicht: (2024)
CORBA: Contagious Recursive Blocking Attacks on Multi-Agent Systems Based on Large Language Models
von: Zhou, Zhenhong, et al.
Veröffentlicht: (2025)
von: Zhou, Zhenhong, et al.
Veröffentlicht: (2025)
Towards Understanding the Safety Boundaries of DeepSeek Models: Evaluation and Findings
von: Ying, Zonghao, et al.
Veröffentlicht: (2025)
von: Ying, Zonghao, et al.
Veröffentlicht: (2025)
When Can Large Reasoning Models Save Thinking? Mechanistic Analysis of Behavioral Divergence in Reasoning
von: Zhu, Rongzhi, et al.
Veröffentlicht: (2025)
von: Zhu, Rongzhi, et al.
Veröffentlicht: (2025)
RoboSafe: Safeguarding Embodied Agents via Executable Safety Logic
von: Wang, Le, et al.
Veröffentlicht: (2025)
von: Wang, Le, et al.
Veröffentlicht: (2025)
From Context to Intent: Reasoning-Guided Function-Level Code Completion
von: Li, Yanzhou, et al.
Veröffentlicht: (2025)
von: Li, Yanzhou, et al.
Veröffentlicht: (2025)
DiffuGuard: How Intrinsic Safety is Lost and Found in Diffusion Large Language Models
von: Li, Zherui, et al.
Veröffentlicht: (2025)
von: Li, Zherui, et al.
Veröffentlicht: (2025)
LLMs for Relational Reasoning: How Far are We?
von: Li, Zhiming, et al.
Veröffentlicht: (2024)
von: Li, Zhiming, et al.
Veröffentlicht: (2024)
KernelSkill: A Multi-Agent Framework for GPU Kernel Optimization
von: Sun, Qitong, et al.
Veröffentlicht: (2026)
von: Sun, Qitong, et al.
Veröffentlicht: (2026)
Minimal and Mechanistic Conditions for Behavioral Self-Awareness in LLMs
von: Bozoukov, Matthew, et al.
Veröffentlicht: (2025)
von: Bozoukov, Matthew, et al.
Veröffentlicht: (2025)
Alignment-Enhanced Decoding:Defending via Token-Level Adaptive Refining of Probability Distributions
von: Liu, Quan, et al.
Veröffentlicht: (2024)
von: Liu, Quan, et al.
Veröffentlicht: (2024)
Semantic Mirror Jailbreak: Genetic Algorithm Based Jailbreak Prompts Against Open-source LLMs
von: Li, Xiaoxia, et al.
Veröffentlicht: (2024)
von: Li, Xiaoxia, et al.
Veröffentlicht: (2024)
Investigating Training Data Detection in AI Coders
von: Li, Tianlin, et al.
Veröffentlicht: (2025)
von: Li, Tianlin, et al.
Veröffentlicht: (2025)
Valence-Arousal Subspace in LLMs: Circular Emotion Geometry and Multi-Behavioral Control
von: Sun, Lihao, et al.
Veröffentlicht: (2026)
von: Sun, Lihao, et al.
Veröffentlicht: (2026)
Speak Out of Turn: Safety Vulnerability of Large Language Models in Multi-turn Dialogue
von: Zhou, Zhenhong, et al.
Veröffentlicht: (2024)
von: Zhou, Zhenhong, et al.
Veröffentlicht: (2024)
GSM-Infinite: How Do Your LLMs Behave over Infinitely Increasing Context Length and Reasoning Complexity?
von: Zhou, Yang, et al.
Veröffentlicht: (2025)
von: Zhou, Yang, et al.
Veröffentlicht: (2025)
How Post-Training Reshapes LLMs: A Mechanistic View on Knowledge, Truthfulness, Refusal, and Confidence
von: Du, Hongzhe, et al.
Veröffentlicht: (2025)
von: Du, Hongzhe, et al.
Veröffentlicht: (2025)
Reasoning-Augmented Conversation for Multi-Turn Jailbreak Attacks on Large Language Models
von: Ying, Zonghao, et al.
Veröffentlicht: (2025)
von: Ying, Zonghao, et al.
Veröffentlicht: (2025)
From Helpfulness to Toxic Proactivity: Diagnosing Behavioral Misalignment in LLM Agents
von: Wang, Xinyue, et al.
Veröffentlicht: (2026)
von: Wang, Xinyue, et al.
Veröffentlicht: (2026)
Mechanistic Interpretability of Emotion Inference in Large Language Models
von: Tak, Ala N., et al.
Veröffentlicht: (2025)
von: Tak, Ala N., et al.
Veröffentlicht: (2025)
Mechanistic Behavior Editing of Language Models
von: Singh, Joykirat, et al.
Veröffentlicht: (2024)
von: Singh, Joykirat, et al.
Veröffentlicht: (2024)
Under the Shadow of Babel: How Language Shapes Reasoning in LLMs
von: Wang, Chenxi, et al.
Veröffentlicht: (2025)
von: Wang, Chenxi, et al.
Veröffentlicht: (2025)
BiLLM: Pushing the Limit of Post-Training Quantization for LLMs
von: Huang, Wei, et al.
Veröffentlicht: (2024)
von: Huang, Wei, et al.
Veröffentlicht: (2024)
Software Development Life Cycle Perspective: A Survey of Benchmarks for Code Large Language Models and Agents
von: Wang, Kaixin, et al.
Veröffentlicht: (2025)
von: Wang, Kaixin, et al.
Veröffentlicht: (2025)
Measuring Reasoning Quality in LLMs: A Multi-Dimensional Behavioral Framework
von: Şenol, Ali, et al.
Veröffentlicht: (2026)
von: Şenol, Ali, et al.
Veröffentlicht: (2026)
DeepPlanner: Scaling Planning Capability for Deep Research Agents via Advantage Shaping
von: Fan, Wei, et al.
Veröffentlicht: (2025)
von: Fan, Wei, et al.
Veröffentlicht: (2025)
MetaAligner: Towards Generalizable Multi-Objective Alignment of Language Models
von: Yang, Kailai, et al.
Veröffentlicht: (2024)
von: Yang, Kailai, et al.
Veröffentlicht: (2024)
Personality-affected Emotion Generation in Dialog Systems
von: Wen, Zhiyuan, et al.
Veröffentlicht: (2024)
von: Wen, Zhiyuan, et al.
Veröffentlicht: (2024)
How Jailbreak Defenses Work and Ensemble? A Mechanistic Investigation
von: Long, Zhuohang, et al.
Veröffentlicht: (2025)
von: Long, Zhuohang, et al.
Veröffentlicht: (2025)
Are LLMs Ready for Neural-integrated Mechanistic Modeling? A Benchmark and Agentic Framework
von: Guan, Zihan, et al.
Veröffentlicht: (2026)
von: Guan, Zihan, et al.
Veröffentlicht: (2026)
Aligning Machiavellian Agents: Behavior Steering via Test-Time Policy Shaping
von: Mujtaba, Dena, et al.
Veröffentlicht: (2025)
von: Mujtaba, Dena, et al.
Veröffentlicht: (2025)
OAgents: An Empirical Study of Building Effective Agents
von: Zhu, He, et al.
Veröffentlicht: (2025)
von: Zhu, He, et al.
Veröffentlicht: (2025)
Black-Box Adversarial Attack on Vision Language Models for Autonomous Driving
von: Wang, Lu, et al.
Veröffentlicht: (2025)
von: Wang, Lu, et al.
Veröffentlicht: (2025)
Centering Emotion Hotspots: Multimodal Local-Global Fusion and Cross-Modal Alignment for Emotion Recognition in Conversations
von: Liu, Yu, et al.
Veröffentlicht: (2025)
von: Liu, Yu, et al.
Veröffentlicht: (2025)
DB-LLM: Accurate Dual-Binarization for Efficient LLMs
von: Chen, Hong, et al.
Veröffentlicht: (2024)
von: Chen, Hong, et al.
Veröffentlicht: (2024)
Dissecting Bias in LLMs: A Mechanistic Interpretability Perspective
von: Chandna, Bhavik, et al.
Veröffentlicht: (2025)
von: Chandna, Bhavik, et al.
Veröffentlicht: (2025)
MASteer: Multi-Agent Adaptive Steer Strategy for End-to-End LLM Trustworthiness Repair
von: Li, Changqing, et al.
Veröffentlicht: (2025)
von: Li, Changqing, et al.
Veröffentlicht: (2025)
TheraAgent: Self-Improving Therapeutic Agent for Precise and Comprehensive Treatment Planning
von: Li, Junkai, et al.
Veröffentlicht: (2026)
von: Li, Junkai, et al.
Veröffentlicht: (2026)
Densing Law of LLMs
von: Xiao, Chaojun, et al.
Veröffentlicht: (2024)
von: Xiao, Chaojun, et al.
Veröffentlicht: (2024)
How Context Shapes Truth: Geometric Transformations of Statement-level Truth Representations in LLMs
von: Adarsh, Shivam, et al.
Veröffentlicht: (2026)
von: Adarsh, Shivam, et al.
Veröffentlicht: (2026)
Ähnliche Einträge
-
LLMCBench: Benchmarking Large Language Model Compression for Efficient Deployment
von: Yang, Ge, et al.
Veröffentlicht: (2024) -
CORBA: Contagious Recursive Blocking Attacks on Multi-Agent Systems Based on Large Language Models
von: Zhou, Zhenhong, et al.
Veröffentlicht: (2025) -
Towards Understanding the Safety Boundaries of DeepSeek Models: Evaluation and Findings
von: Ying, Zonghao, et al.
Veröffentlicht: (2025) -
When Can Large Reasoning Models Save Thinking? Mechanistic Analysis of Behavioral Divergence in Reasoning
von: Zhu, Rongzhi, et al.
Veröffentlicht: (2025) -
RoboSafe: Safeguarding Embodied Agents via Executable Safety Logic
von: Wang, Le, et al.
Veröffentlicht: (2025)