Latent Debate: A Surrogate Framework for Interpreting LLM Thinking
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Chen, Lihu, Yin, Xiang, Toni, Francesca |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2025
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Identifying Query-Relevant Neurons in Large Language Models for Long-Form Texts
von: Chen, Lihu, et al.
Veröffentlicht: (2024)
von: Chen, Lihu, et al.
Veröffentlicht: (2024)
What is the Role of Small Models in the LLM Era: A Survey
von: Chen, Lihu, et al.
Veröffentlicht: (2024)
von: Chen, Lihu, et al.
Veröffentlicht: (2024)
Pub-Guard-LLM: Detecting Retracted Biomedical Articles with Reliable Explanations
von: Chen, Lihu, et al.
Veröffentlicht: (2025)
von: Chen, Lihu, et al.
Veröffentlicht: (2025)
Evaluating Uncertainty Quantification Methods in Argumentative Large Language Models
von: Zhou, Kevin, et al.
Veröffentlicht: (2025)
von: Zhou, Kevin, et al.
Veröffentlicht: (2025)
Towards a Framework for Evaluating Explanations in Automated Fact Verification
von: Kotonya, Neema, et al.
Veröffentlicht: (2024)
von: Kotonya, Neema, et al.
Veröffentlicht: (2024)
Think Silently, Think Fast: Dynamic Latent Compression of LLM Reasoning Chains
von: Tan, Wenhui, et al.
Veröffentlicht: (2025)
von: Tan, Wenhui, et al.
Veröffentlicht: (2025)
Selective Latent Thinking: Adaptive Compression of LLM Reasoning Chains
von: Xie, Hui, et al.
Veröffentlicht: (2026)
von: Xie, Hui, et al.
Veröffentlicht: (2026)
ArgLLM-App: An Interactive System for Argumentative Reasoning with Large Language Models
von: Dejl, Adam, et al.
Veröffentlicht: (2026)
von: Dejl, Adam, et al.
Veröffentlicht: (2026)
Learning High-Quality and General-Purpose Phrase Representations
von: Chen, Lihu, et al.
Veröffentlicht: (2024)
von: Chen, Lihu, et al.
Veröffentlicht: (2024)
ThinkRouter: Efficient Reasoning via Routing Thinking between Latent and Discrete Spaces
von: Xu, Xin, et al.
Veröffentlicht: (2026)
von: Xu, Xin, et al.
Veröffentlicht: (2026)
Debate, Deliberate, Decide (D3): A Cost-Aware Adversarial Framework for Reliable and Interpretable LLM Evaluation
von: Harrasse, Abir, et al.
Veröffentlicht: (2024)
von: Harrasse, Abir, et al.
Veröffentlicht: (2024)
Single LLM Debate, MoLaCE: Mixture of Latent Concept Experts Against Confirmation Bias
von: Kim, Hazel, et al.
Veröffentlicht: (2025)
von: Kim, Hazel, et al.
Veröffentlicht: (2025)
Exploring the Potential for Large Language Models to Demonstrate Rational Probabilistic Beliefs
von: Freedman, Gabriel, et al.
Veröffentlicht: (2025)
von: Freedman, Gabriel, et al.
Veröffentlicht: (2025)
LLM-POTUS Score: A Framework of Analyzing Presidential Debates with Large Language Models
von: Liu, Zhengliang, et al.
Veröffentlicht: (2024)
von: Liu, Zhengliang, et al.
Veröffentlicht: (2024)
Debatable Intelligence: Benchmarking LLM Judges via Debate Speech Evaluation
von: Sternlicht, Noy, et al.
Veröffentlicht: (2025)
von: Sternlicht, Noy, et al.
Veröffentlicht: (2025)
Query-Level Uncertainty in Large Language Models
von: Chen, Lihu, et al.
Veröffentlicht: (2025)
von: Chen, Lihu, et al.
Veröffentlicht: (2025)
Reconfidencing LLMs from the Grouping Loss Perspective
von: Chen, Lihu, et al.
Veröffentlicht: (2024)
von: Chen, Lihu, et al.
Veröffentlicht: (2024)
Tree-of-Debate: Multi-Persona Debate Trees Elicit Critical Thinking for Scientific Comparative Analysis
von: Kargupta, Priyanka, et al.
Veröffentlicht: (2025)
von: Kargupta, Priyanka, et al.
Veröffentlicht: (2025)
Encouraging Divergent Thinking in Large Language Models through Multi-Agent Debate
von: Liang, Tian, et al.
Veröffentlicht: (2023)
von: Liang, Tian, et al.
Veröffentlicht: (2023)
MechELK: A Mechanistic Interpretability Framework for Eliciting Latent Knowledge in Large Language Models
von: Park, Ji-jun, et al.
Veröffentlicht: (2026)
von: Park, Ji-jun, et al.
Veröffentlicht: (2026)
A Debate-Driven Experiment on LLM Hallucinations and Accuracy
von: Li, Ray, et al.
Veröffentlicht: (2024)
von: Li, Ray, et al.
Veröffentlicht: (2024)
Latent Reasoning with Supervised Thinking States
von: Amos, Ido, et al.
Veröffentlicht: (2026)
von: Amos, Ido, et al.
Veröffentlicht: (2026)
Argumentative Large Language Models for Explainable and Contestable Claim Verification
von: Freedman, Gabriel, et al.
Veröffentlicht: (2024)
von: Freedman, Gabriel, et al.
Veröffentlicht: (2024)
Think Deep, Not Just Long: Measuring LLM Reasoning Effort via Deep-Thinking Tokens
von: Chen, Wei-Lin, et al.
Veröffentlicht: (2026)
von: Chen, Wei-Lin, et al.
Veröffentlicht: (2026)
Latent Thinking Optimization: Your Latent Reasoning Language Model Secretly Encodes Reward Signals in Its Latent Thoughts
von: Du, Hanwen, et al.
Veröffentlicht: (2025)
von: Du, Hanwen, et al.
Veröffentlicht: (2025)
Can LLMs Beat Humans in Debating? A Dynamic Multi-agent Framework for Competitive Debate
von: Zhang, Yiqun, et al.
Veröffentlicht: (2024)
von: Zhang, Yiqun, et al.
Veröffentlicht: (2024)
Think Twice: Enhancing LLM Reasoning by Scaling Multi-round Test-time Thinking
von: Tian, Xiaoyu, et al.
Veröffentlicht: (2025)
von: Tian, Xiaoyu, et al.
Veröffentlicht: (2025)
AgentCoMa: A Compositional Benchmark Mixing Commonsense and Mathematical Reasoning in Real-World Scenarios
von: Alazraki, Lisa, et al.
Veröffentlicht: (2025)
von: Alazraki, Lisa, et al.
Veröffentlicht: (2025)
When to Think, When to Speak: Learning Disclosure Policies for LLM Reasoning
von: Wei, Jiaqi, et al.
Veröffentlicht: (2026)
von: Wei, Jiaqi, et al.
Veröffentlicht: (2026)
Debate-to-Write: A Persona-Driven Multi-Agent Framework for Diverse Argument Generation
von: Hu, Zhe, et al.
Veröffentlicht: (2024)
von: Hu, Zhe, et al.
Veröffentlicht: (2024)
Argumentation for Explainable and Globally Contestable Decision Support with LLMs
von: Dejl, Adam, et al.
Veröffentlicht: (2026)
von: Dejl, Adam, et al.
Veröffentlicht: (2026)
Thinking into the Future: Latent Lookahead Training for Transformers
von: Noci, Lorenzo, et al.
Veröffentlicht: (2026)
von: Noci, Lorenzo, et al.
Veröffentlicht: (2026)
Learning to Explain: Prototype-Based Surrogate Models for LLM Classification
von: Wei, Bowen, et al.
Veröffentlicht: (2025)
von: Wei, Bowen, et al.
Veröffentlicht: (2025)
Interpretable Depression Detection from Social Media Text Using LLM-Derived Embeddings
von: Kim, Samuel, et al.
Veröffentlicht: (2025)
von: Kim, Samuel, et al.
Veröffentlicht: (2025)
Thinking Without Words: Efficient Latent Reasoning with Abstract Chain-of-Thought
von: Ramji, Keshav, et al.
Veröffentlicht: (2026)
von: Ramji, Keshav, et al.
Veröffentlicht: (2026)
Systematic Biases in LLM Simulations of Debates
von: Taubenfeld, Amir, et al.
Veröffentlicht: (2024)
von: Taubenfeld, Amir, et al.
Veröffentlicht: (2024)
Conveying Imagistic Thinking in Traditional Chinese Medicine Translation: A Prompt Engineering and LLM-Based Evaluation Framework
von: Han, Jiatong
Veröffentlicht: (2025)
von: Han, Jiatong
Veröffentlicht: (2025)
Semantic-Augmented Latent Topic Modeling with LLM-in-the-Loop
von: Hong, Mengze, et al.
Veröffentlicht: (2025)
von: Hong, Mengze, et al.
Veröffentlicht: (2025)
LatentChem: From Textual CoT to Latent Thinking in Chemical Reasoning
von: Ye, Xinwu, et al.
Veröffentlicht: (2026)
von: Ye, Xinwu, et al.
Veröffentlicht: (2026)
DiffusionDialog: A Diffusion Model for Diverse Dialog Generation with Latent Space
von: Xiang, Jianxiang, et al.
Veröffentlicht: (2024)
von: Xiang, Jianxiang, et al.
Veröffentlicht: (2024)
Ähnliche Einträge
-
Identifying Query-Relevant Neurons in Large Language Models for Long-Form Texts
von: Chen, Lihu, et al.
Veröffentlicht: (2024) -
What is the Role of Small Models in the LLM Era: A Survey
von: Chen, Lihu, et al.
Veröffentlicht: (2024) -
Pub-Guard-LLM: Detecting Retracted Biomedical Articles with Reliable Explanations
von: Chen, Lihu, et al.
Veröffentlicht: (2025) -
Evaluating Uncertainty Quantification Methods in Argumentative Large Language Models
von: Zhou, Kevin, et al.
Veröffentlicht: (2025) -
Towards a Framework for Evaluating Explanations in Automated Fact Verification
von: Kotonya, Neema, et al.
Veröffentlicht: (2024)