Exploring the generalization of LLM truth directions on conversational formats
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Ichmoukhamedov, Timour, Martens, David |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2025
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
How good is my story? Towards quantitative metrics for evaluating LLM-generated XAI narratives
von: Ichmoukhamedov, Timour, et al.
Veröffentlicht: (2024)
von: Ichmoukhamedov, Timour, et al.
Veröffentlicht: (2024)
Cash or Comfort? How LLMs Value Your Inconvenience
von: Cedro, Mateusz, et al.
Veröffentlicht: (2025)
von: Cedro, Mateusz, et al.
Veröffentlicht: (2025)
The Human and the Mechanical: logos, truthfulness, and ChatGPT
von: Giannakidou, Anastasia, et al.
Veröffentlicht: (2024)
von: Giannakidou, Anastasia, et al.
Veröffentlicht: (2024)
On the Importance and Evaluation of Narrativity in Natural Language AI Explanations
von: Cedro, Mateusz, et al.
Veröffentlicht: (2026)
von: Cedro, Mateusz, et al.
Veröffentlicht: (2026)
Dynamic benchmarking framework for LLM-based conversational data capture
von: Aluffi, Pietro Alessandro, et al.
Veröffentlicht: (2025)
von: Aluffi, Pietro Alessandro, et al.
Veröffentlicht: (2025)
Efficient RL for optimizing conversation level outcomes with an LLM-based tutor
von: Nam, Hyunji, et al.
Veröffentlicht: (2025)
von: Nam, Hyunji, et al.
Veröffentlicht: (2025)
Reinforcement learning for path integrals in quantum statistical physics
von: Ichmoukhamedov, Timour, et al.
Veröffentlicht: (2026)
von: Ichmoukhamedov, Timour, et al.
Veröffentlicht: (2026)
Would a Large Language Model Pay Extra for a View? Inferring Willingness to Pay from Subjective Choices
von: Reusens, Manon, et al.
Veröffentlicht: (2026)
von: Reusens, Manon, et al.
Veröffentlicht: (2026)
Tell me the truth: A system to measure the trustworthiness of Large Language Models
von: Lipizzi, Carlo
Veröffentlicht: (2024)
von: Lipizzi, Carlo
Veröffentlicht: (2024)
Amphista: Bi-directional Multi-head Decoding for Accelerating LLM Inference
von: Li, Zeping, et al.
Veröffentlicht: (2024)
von: Li, Zeping, et al.
Veröffentlicht: (2024)
Experiments with truth using Machine Learning: Spectral analysis and explainable classification of synthetic, false, and genuine information
von: Pendyala, Vishnu S., et al.
Veröffentlicht: (2024)
von: Pendyala, Vishnu S., et al.
Veröffentlicht: (2024)
GuideLLM: Exploring LLM-Guided Conversation with Applications in Autobiography Interviewing
von: Duan, Jinhao, et al.
Veröffentlicht: (2025)
von: Duan, Jinhao, et al.
Veröffentlicht: (2025)
Exploring and Controlling Diversity in LLM-Agent Conversation
von: Chu, KuanChao, et al.
Veröffentlicht: (2024)
von: Chu, KuanChao, et al.
Veröffentlicht: (2024)
Mitigating LLM biases toward spurious social contexts using direct preference optimization
von: Nam, Hyunji, et al.
Veröffentlicht: (2026)
von: Nam, Hyunji, et al.
Veröffentlicht: (2026)
Table as Thought: Exploring Structured Thoughts in LLM Reasoning
von: Sun, Zhenjie, et al.
Veröffentlicht: (2025)
von: Sun, Zhenjie, et al.
Veröffentlicht: (2025)
Exploring the Necessity of Reasoning in LLM-based Agent Scenarios
von: Zhou, Xueyang, et al.
Veröffentlicht: (2025)
von: Zhou, Xueyang, et al.
Veröffentlicht: (2025)
TravelBench : Exploring LLM Performance in Low-Resource Domains
von: Billa, Srinivas, et al.
Veröffentlicht: (2025)
von: Billa, Srinivas, et al.
Veröffentlicht: (2025)
Exploring Knowledge Conflicts for Faithful LLM Reasoning: Benchmark and Method
von: Zhao, Tianzhe, et al.
Veröffentlicht: (2026)
von: Zhao, Tianzhe, et al.
Veröffentlicht: (2026)
Exploring the LLM Journey from Cognition to Expression with Linear Representations
von: Yan, Yuzi, et al.
Veröffentlicht: (2024)
von: Yan, Yuzi, et al.
Veröffentlicht: (2024)
Attention Mechanisms Perspective: Exploring LLM Processing of Graph-Structured Data
von: Guan, Zhong, et al.
Veröffentlicht: (2025)
von: Guan, Zhong, et al.
Veröffentlicht: (2025)
Developing an AI framework to automatically detect shared decision-making in patient-doctor conversations
von: Ponce-Ponte, Oscar J., et al.
Veröffentlicht: (2025)
von: Ponce-Ponte, Oscar J., et al.
Veröffentlicht: (2025)
Survive at All Costs: Exploring LLM's Risky Behaviors under Survival Pressure
von: Lu, Yida, et al.
Veröffentlicht: (2026)
von: Lu, Yida, et al.
Veröffentlicht: (2026)
Quantization Meets Reasoning: Exploring LLM Low-Bit Quantization Degradation for Mathematical Reasoning
von: Li, Zhen, et al.
Veröffentlicht: (2025)
von: Li, Zhen, et al.
Veröffentlicht: (2025)
Exploring Spatial Representations in the Historical Lake District Texts with LLM-based Relation Extraction
von: Haris, Erum, et al.
Veröffentlicht: (2024)
von: Haris, Erum, et al.
Veröffentlicht: (2024)
English is Not All You Need: Systematically Exploring the Role of Multilinguality in LLM Post-Training
von: Dhaliwal, Mehak, et al.
Veröffentlicht: (2026)
von: Dhaliwal, Mehak, et al.
Veröffentlicht: (2026)
Exploring Persona-dependent LLM Alignment for the Moral Machine Experiment
von: Kim, Jiseon, et al.
Veröffentlicht: (2025)
von: Kim, Jiseon, et al.
Veröffentlicht: (2025)
Generative Explore-Exploit: Training-free Optimization of Generative Recommender Systems using LLM Optimizers
von: Senel, Lütfi Kerem, et al.
Veröffentlicht: (2024)
von: Senel, Lütfi Kerem, et al.
Veröffentlicht: (2024)
Simul-LLM: A Framework for Exploring High-Quality Simultaneous Translation with Large Language Models
von: Agostinelli, Victor, et al.
Veröffentlicht: (2023)
von: Agostinelli, Victor, et al.
Veröffentlicht: (2023)
Towards Transparency: Exploring LLM Trainings Datasets through Visual Topic Modeling and Semantic Frame
von: de Dampierre, Charles, et al.
Veröffentlicht: (2024)
von: de Dampierre, Charles, et al.
Veröffentlicht: (2024)
Automated test generation to evaluate tool-augmented LLMs as conversational AI agents
von: Arcadinho, Samuel, et al.
Veröffentlicht: (2024)
von: Arcadinho, Samuel, et al.
Veröffentlicht: (2024)
Do explanations generalize across large reasoning models?
von: Pal, Koyena, et al.
Veröffentlicht: (2026)
von: Pal, Koyena, et al.
Veröffentlicht: (2026)
MobileLLM-R1: Exploring the Limits of Sub-Billion Language Model Reasoners with Open Training Recipes
von: Zhao, Changsheng, et al.
Veröffentlicht: (2025)
von: Zhao, Changsheng, et al.
Veröffentlicht: (2025)
An automated method for the ontological representation of security directives
von: Bella, Giampaolo, et al.
Veröffentlicht: (2023)
von: Bella, Giampaolo, et al.
Veröffentlicht: (2023)
Every time I fire a conversational designer, the performance of the dialog system goes down
von: Xompero, Giancarlo A., et al.
Veröffentlicht: (2021)
von: Xompero, Giancarlo A., et al.
Veröffentlicht: (2021)
Tell me what I need to know: Exploring LLM-based (Personalized) Abstractive Multi-Source Meeting Summarization
von: Kirstein, Frederic, et al.
Veröffentlicht: (2024)
von: Kirstein, Frederic, et al.
Veröffentlicht: (2024)
Patch-Effect Graph Kernels for LLM Interpretability
von: Fernandez-Boullon, Ruben, et al.
Veröffentlicht: (2026)
von: Fernandez-Boullon, Ruben, et al.
Veröffentlicht: (2026)
Exploring LLM Reasoning Through Controlled Prompt Variations
von: Chatziveroglou, Giannis, et al.
Veröffentlicht: (2025)
von: Chatziveroglou, Giannis, et al.
Veröffentlicht: (2025)
Exploring LLM biases to manipulate AI search overview
von: Smirnov, Roman
Veröffentlicht: (2026)
von: Smirnov, Roman
Veröffentlicht: (2026)
The Illusion of Latent Generalization: Bi-directionality and the Reversal Curse
von: Coda-Forno, Julian, et al.
Veröffentlicht: (2026)
von: Coda-Forno, Julian, et al.
Veröffentlicht: (2026)
Planning to Explore: Curiosity-Driven Planning for LLM Test Generation
von: Amayuelas, Alfonso, et al.
Veröffentlicht: (2026)
von: Amayuelas, Alfonso, et al.
Veröffentlicht: (2026)
Ähnliche Einträge
-
How good is my story? Towards quantitative metrics for evaluating LLM-generated XAI narratives
von: Ichmoukhamedov, Timour, et al.
Veröffentlicht: (2024) -
Cash or Comfort? How LLMs Value Your Inconvenience
von: Cedro, Mateusz, et al.
Veröffentlicht: (2025) -
The Human and the Mechanical: logos, truthfulness, and ChatGPT
von: Giannakidou, Anastasia, et al.
Veröffentlicht: (2024) -
On the Importance and Evaluation of Narrativity in Natural Language AI Explanations
von: Cedro, Mateusz, et al.
Veröffentlicht: (2026) -
Dynamic benchmarking framework for LLM-based conversational data capture
von: Aluffi, Pietro Alessandro, et al.
Veröffentlicht: (2025)