Do LLMs Experience an Internal Polylogue? Investigating Reasoning through the Lens of Personas
Fuente:
arXiv
Guardado en:
| Autores principales: | Herrmann, Nils A., Girrbach, Leander, Bykov, Kirill, Akata, Zeynep |
|---|---|
| Formato: | Preprint |
| Publicado: |
2026
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
Are Reasoning LLMs Robust to Interventions on Their Chain-of-Thought?
por: von Recum, Alexander, et al.
Publicado: (2026)
por: von Recum, Alexander, et al.
Publicado: (2026)
SUB: Benchmarking CBM Generalization via Synthetic Attribute Substitutions
por: Bader, Jessica, et al.
Publicado: (2025)
por: Bader, Jessica, et al.
Publicado: (2025)
Sparse Autoencoders are Topic Models
por: Girrbach, Leander, et al.
Publicado: (2025)
por: Girrbach, Leander, et al.
Publicado: (2025)
APM: Evaluating Style Personalization in LLMs with Arbitrary Preference Mappings
por: Spohn, Philipp, et al.
Publicado: (2026)
por: Spohn, Philipp, et al.
Publicado: (2026)
Reference-Free Rating of LLM Responses via Latent Information
por: Girrbach, Leander, et al.
Publicado: (2025)
por: Girrbach, Leander, et al.
Publicado: (2025)
DeLoRA: Decoupling Angles and Strength in Low-rank Adaptation
por: Bini, Massimo, et al.
Publicado: (2025)
por: Bini, Massimo, et al.
Publicado: (2025)
A Systematic Study of In-the-Wild Model Merging for Large Language Models
por: Hitit, Oğuz Kağan, et al.
Publicado: (2025)
por: Hitit, Oğuz Kağan, et al.
Publicado: (2025)
A Large Scale Analysis of Gender Biases in Text-to-Image Generative Models
por: Girrbach, Leander, et al.
Publicado: (2025)
por: Girrbach, Leander, et al.
Publicado: (2025)
Align-then-Unlearn: Embedding Alignment for LLM Unlearning
por: Spohn, Philipp, et al.
Publicado: (2025)
por: Spohn, Philipp, et al.
Publicado: (2025)
Person-Centric Annotations of LAION-400M: Auditing Bias and Its Transfer to Models
por: Girrbach, Leander, et al.
Publicado: (2025)
por: Girrbach, Leander, et al.
Publicado: (2025)
Revealing and Reducing Gender Biases in Vision and Language Assistants (VLAs)
por: Girrbach, Leander, et al.
Publicado: (2024)
por: Girrbach, Leander, et al.
Publicado: (2024)
Do Persona-Infused LLMs Affect Performance in a Strategic Reasoning Game?
por: Licato, John, et al.
Publicado: (2025)
por: Licato, John, et al.
Publicado: (2025)
Do LLMs Know When to Flip a Coin? Strategic Randomization through Reasoning and Experience
por: Yang, Lingyu
Publicado: (2025)
por: Yang, Lingyu
Publicado: (2025)
SemioLLM: Evaluating Large Language Models for Diagnostic Reasoning from Unstructured Clinical Narratives in Epilepsy
por: Dani, Meghal, et al.
Publicado: (2024)
por: Dani, Meghal, et al.
Publicado: (2024)
Feasibility with Language Models for Open-World Compositional Zero-Shot Learning
por: Kim, Jae Myung, et al.
Publicado: (2025)
por: Kim, Jae Myung, et al.
Publicado: (2025)
From Drop-off to Recovery: A Mechanistic Analysis of Segmentation in MLLMs
por: Wu, Boyong, et al.
Publicado: (2026)
por: Wu, Boyong, et al.
Publicado: (2026)
Discovering Chunks in Neural Embeddings for Interpretability
por: Wu, Shuchen, et al.
Publicado: (2025)
por: Wu, Shuchen, et al.
Publicado: (2025)
SOTAlign: Semi-Supervised Alignment of Unimodal Vision and Language Models via Optimal Transport
por: Roschmann, Simon, et al.
Publicado: (2026)
por: Roschmann, Simon, et al.
Publicado: (2026)
An Investigation of Neuron Activation as a Unified Lens to Explain Chain-of-Thought Eliciting Arithmetic Reasoning of LLMs
por: Rai, Daking, et al.
Publicado: (2024)
por: Rai, Daking, et al.
Publicado: (2024)
Persona Non Grata: Single-Method Safety Evaluation Is Incomplete for Persona-Imbued LLMs
por: Li, Wenkai, et al.
Publicado: (2026)
por: Li, Wenkai, et al.
Publicado: (2026)
FINER: MLLMs Hallucinate under Fine-grained Negative Queries
por: Xiao, Rui, et al.
Publicado: (2026)
por: Xiao, Rui, et al.
Publicado: (2026)
Building, Reusing, and Generalizing Abstract Representations from Concrete Sequences
por: Wu, Shuchen, et al.
Publicado: (2024)
por: Wu, Shuchen, et al.
Publicado: (2024)
Standards for Belief Representations in LLMs
por: Herrmann, Daniel A., et al.
Publicado: (2024)
por: Herrmann, Daniel A., et al.
Publicado: (2024)
Improving Intervention Efficacy via Concept Realignment in Concept Bottleneck Models
por: Singhi, Nishad, et al.
Publicado: (2024)
por: Singhi, Nishad, et al.
Publicado: (2024)
Styles + Persona-plug = Customized LLMs
por: Song, Yutong, et al.
Publicado: (2026)
por: Song, Yutong, et al.
Publicado: (2026)
PEAM: Parametric Embodied Agent Memory through Contrastive Internalization of Experience in Minecraft
por: Guo, Yuchen, et al.
Publicado: (2026)
por: Guo, Yuchen, et al.
Publicado: (2026)
Can LLMs Reason Structurally? Benchmarking via the Lens of Data Structures
por: He, Yu, et al.
Publicado: (2025)
por: He, Yu, et al.
Publicado: (2025)
Rethinking the Text-Vision Reasoning Imbalance in MLLMs through the Lens of Training Recipes
por: Yao, Guanyu, et al.
Publicado: (2025)
por: Yao, Guanyu, et al.
Publicado: (2025)
FLAIR: VLM with Fine-grained Language-informed Image Representations
por: Xiao, Rui, et al.
Publicado: (2024)
por: Xiao, Rui, et al.
Publicado: (2024)
Stitch: Training-Free Position Control in Multimodal Diffusion Transformers
por: Bader, Jessica, et al.
Publicado: (2025)
por: Bader, Jessica, et al.
Publicado: (2025)
Do Linear Probes Generalize Better in Persona Coordinates?
por: Mahadik, Prasad, et al.
Publicado: (2026)
por: Mahadik, Prasad, et al.
Publicado: (2026)
PersonaGym: Evaluating Persona Agents and LLMs
por: Samuel, Vinay, et al.
Publicado: (2024)
por: Samuel, Vinay, et al.
Publicado: (2024)
Localizing Persona Representations in LLMs
por: Cintas, Celia, et al.
Publicado: (2025)
por: Cintas, Celia, et al.
Publicado: (2025)
The Latent Color Subspace: Emergent Order in High-Dimensional Chaos
por: Pach, Mateusz, et al.
Publicado: (2026)
por: Pach, Mateusz, et al.
Publicado: (2026)
Time Series Representations for Classification Lie Hidden in Pretrained Vision Transformers
por: Roschmann, Simon, et al.
Publicado: (2025)
por: Roschmann, Simon, et al.
Publicado: (2025)
Sparse Autoencoders Learn Monosemantic Features in Vision-Language Models
por: Pach, Mateusz, et al.
Publicado: (2025)
por: Pach, Mateusz, et al.
Publicado: (2025)
Internalizing Agency from Reflective Experience
por: Ge, Rui, et al.
Publicado: (2026)
por: Ge, Rui, et al.
Publicado: (2026)
Beyond Words: A Latent Memory Approach to Internal Reasoning in LLMs
por: Orlicki, José I.
Publicado: (2025)
por: Orlicki, José I.
Publicado: (2025)
Concise Reasoning in the Lens of Lagrangian Optimization
por: Gao, Chengqian, et al.
Publicado: (2025)
por: Gao, Chengqian, et al.
Publicado: (2025)
Safety Compliance: Rethinking LLM Safety Reasoning through the Lens of Compliance
por: Hu, Wenbin, et al.
Publicado: (2025)
por: Hu, Wenbin, et al.
Publicado: (2025)
Ejemplares similares
-
Are Reasoning LLMs Robust to Interventions on Their Chain-of-Thought?
por: von Recum, Alexander, et al.
Publicado: (2026) -
SUB: Benchmarking CBM Generalization via Synthetic Attribute Substitutions
por: Bader, Jessica, et al.
Publicado: (2025) -
Sparse Autoencoders are Topic Models
por: Girrbach, Leander, et al.
Publicado: (2025) -
APM: Evaluating Style Personalization in LLMs with Arbitrary Preference Mappings
por: Spohn, Philipp, et al.
Publicado: (2026) -
Reference-Free Rating of LLM Responses via Latent Information
por: Girrbach, Leander, et al.
Publicado: (2025)