Can LLMs Lie? Investigation beyond Hallucination
Fuente:
arXiv
Guardado en:
| Autores principales: | Huan, Haoran, Prabhudesai, Mihir, Wu, Mengning, Jaiswal, Shantanu, Pathak, Deepak |
|---|---|
| Formato: | Preprint |
| Publicado: |
2025
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
Diffusion Beats Autoregressive in Data-Constrained Settings
por: Prabhudesai, Mihir, et al.
Publicado: (2025)
por: Prabhudesai, Mihir, et al.
Publicado: (2025)
Iterative Refinement Improves Compositional Image Generation
por: Jaiswal, Shantanu, et al.
Publicado: (2026)
por: Jaiswal, Shantanu, et al.
Publicado: (2026)
Aligning Text-to-Image Diffusion Models with Reward Backpropagation
por: Prabhudesai, Mihir, et al.
Publicado: (2023)
por: Prabhudesai, Mihir, et al.
Publicado: (2023)
Self-Questioning Language Models
por: Chen, Lili, et al.
Publicado: (2025)
por: Chen, Lili, et al.
Publicado: (2025)
Unified Multimodal Discrete Diffusion
por: Swerdlow, Alexander, et al.
Publicado: (2025)
por: Swerdlow, Alexander, et al.
Publicado: (2025)
Video Diffusion Alignment via Reward Gradients
por: Prabhudesai, Mihir, et al.
Publicado: (2024)
por: Prabhudesai, Mihir, et al.
Publicado: (2024)
Maximizing Confidence Alone Improves Reasoning
por: Prabhudesai, Mihir, et al.
Publicado: (2025)
por: Prabhudesai, Mihir, et al.
Publicado: (2025)
Can Knowledge Graphs Reduce Hallucinations in LLMs? : A Survey
por: Agrawal, Garima, et al.
Publicado: (2023)
por: Agrawal, Garima, et al.
Publicado: (2023)
Solving Physics Olympiad via Reinforcement Learning on Physics Simulators
por: Prabhudesai, Mihir, et al.
Publicado: (2026)
por: Prabhudesai, Mihir, et al.
Publicado: (2026)
Measuring Visual Understanding in Telecom domain: Performance Metrics for Image-to-UML conversion using VLMs
por: Ranjani, HG, et al.
Publicado: (2025)
por: Ranjani, HG, et al.
Publicado: (2025)
Can LLMs be Fooled? Investigating Vulnerabilities in LLMs
por: Abdali, Sara, et al.
Publicado: (2024)
por: Abdali, Sara, et al.
Publicado: (2024)
Zero-Shot Visual Reasoning by Vision-Language Models: Benchmarking and Analysis
por: Nagar, Aishik, et al.
Publicado: (2024)
por: Nagar, Aishik, et al.
Publicado: (2024)
Learning to Reason Iteratively and Parallelly for Complex Visual Reasoning Scenarios
por: Jaiswal, Shantanu, et al.
Publicado: (2024)
por: Jaiswal, Shantanu, et al.
Publicado: (2024)
Teaming LLMs to Detect and Mitigate Hallucinations
por: Till, Demian, et al.
Publicado: (2025)
por: Till, Demian, et al.
Publicado: (2025)
Which LLMs are Difficult to Detect? A Detailed Analysis of Potential Factors Contributing to Difficulties in LLM Text Detection
por: Thorat, Shantanu, et al.
Publicado: (2024)
por: Thorat, Shantanu, et al.
Publicado: (2024)
LLMs Will Always Hallucinate, and We Need to Live With This
por: Banerjee, Sourav, et al.
Publicado: (2024)
por: Banerjee, Sourav, et al.
Publicado: (2024)
Can Multimodal LLMs Perform Time Series Anomaly Detection?
por: Xu, Xiongxiao, et al.
Publicado: (2025)
por: Xu, Xiongxiao, et al.
Publicado: (2025)
Sharpness-Aware Minimization Can Hallucinate Minimizers
por: Park, Chanwoong, et al.
Publicado: (2025)
por: Park, Chanwoong, et al.
Publicado: (2025)
Robust Hallucination Detection in LLMs via Adaptive Token Selection
por: Niu, Mengjia, et al.
Publicado: (2025)
por: Niu, Mengjia, et al.
Publicado: (2025)
Intrinsic Explainability of Multimodal Learning for Crop Yield Prediction
por: Najjar, Hiba, et al.
Publicado: (2025)
por: Najjar, Hiba, et al.
Publicado: (2025)
Beyond Prompt-Induced Lies: Investigating LLM Deception on Benign Prompts
por: Wu, Zhaomin, et al.
Publicado: (2025)
por: Wu, Zhaomin, et al.
Publicado: (2025)
Semantic Energy: Detecting LLM Hallucination Beyond Entropy
por: Ma, Huan, et al.
Publicado: (2025)
por: Ma, Huan, et al.
Publicado: (2025)
A Concise Review of Hallucinations in LLMs and their Mitigation
por: Pulkundwar, Parth, et al.
Publicado: (2025)
por: Pulkundwar, Parth, et al.
Publicado: (2025)
HalluGuard: Demystifying Data-Driven and Reasoning-Driven Hallucinations in LLMs
por: Zeng, Xinyue, et al.
Publicado: (2026)
por: Zeng, Xinyue, et al.
Publicado: (2026)
Marginals Before Conditionals
por: Sahasrabudhe, Mihir
Publicado: (2026)
por: Sahasrabudhe, Mihir
Publicado: (2026)
Meta-Evolve: Continuous Robot Evolution for One-to-many Policy Transfer
por: Liu, Xingyu, et al.
Publicado: (2024)
por: Liu, Xingyu, et al.
Publicado: (2024)
Generation Constraint Scaling Can Mitigate Hallucination
por: Kollias, Georgios, et al.
Publicado: (2024)
por: Kollias, Georgios, et al.
Publicado: (2024)
Evolutionary Policy Optimization
por: Wang, Jianren, et al.
Publicado: (2025)
por: Wang, Jianren, et al.
Publicado: (2025)
Cost-Effective Hallucination Detection for LLMs
por: Valentin, Simon, et al.
Publicado: (2024)
por: Valentin, Simon, et al.
Publicado: (2024)
Is Implicit Knowledge Enough for LLMs? A RAG Approach for Tree-based Structures
por: Gupte, Mihir, et al.
Publicado: (2025)
por: Gupte, Mihir, et al.
Publicado: (2025)
Failure Modes of LLMs for Causal Reasoning on Narratives
por: Yamin, Khurram, et al.
Publicado: (2024)
por: Yamin, Khurram, et al.
Publicado: (2024)
Unlearners Can Lie: Evaluating and Improving Honesty in LLM Unlearning
por: Gu, Renjie, et al.
Publicado: (2026)
por: Gu, Renjie, et al.
Publicado: (2026)
Controlled Causal Hallucinations Can Estimate Phantom Nodes in Multiexpert Mixtures of Fuzzy Cognitive Maps
por: Panda, Akash Kumar, et al.
Publicado: (2024)
por: Panda, Akash Kumar, et al.
Publicado: (2024)
SparseSSM: Efficient Selective Structured State Space Models Can Be Pruned in One-Shot
por: Tuo, Kaiwen, et al.
Publicado: (2025)
por: Tuo, Kaiwen, et al.
Publicado: (2025)
Benchmarking Reliability of Deep Learning Models for Pathological Gait Classification
por: Jaiswal, Abhishek, et al.
Publicado: (2024)
por: Jaiswal, Abhishek, et al.
Publicado: (2024)
Can LLMs Understand Time Series Anomalies?
por: Zhou, Zihao, et al.
Publicado: (2024)
por: Zhou, Zihao, et al.
Publicado: (2024)
Efficient Distributed Training through Gradient Compression with Sparsification and Quantization Techniques
por: Singh, Shruti, et al.
Publicado: (2024)
por: Singh, Shruti, et al.
Publicado: (2024)
Agent Factories for High Level Synthesis: How Far Can General-Purpose Coding Agents Go in Hardware Optimization?
por: Bhandwaldar, Abhishek, et al.
Publicado: (2026)
por: Bhandwaldar, Abhishek, et al.
Publicado: (2026)
Can LLMs subtract numbers?
por: Jobanputra, Mayank, et al.
Publicado: (2025)
por: Jobanputra, Mayank, et al.
Publicado: (2025)
MobileKernelBench: Can LLMs Write Efficient Kernels for Mobile Devices?
por: Zou, Xingze, et al.
Publicado: (2026)
por: Zou, Xingze, et al.
Publicado: (2026)
Ejemplares similares
-
Diffusion Beats Autoregressive in Data-Constrained Settings
por: Prabhudesai, Mihir, et al.
Publicado: (2025) -
Iterative Refinement Improves Compositional Image Generation
por: Jaiswal, Shantanu, et al.
Publicado: (2026) -
Aligning Text-to-Image Diffusion Models with Reward Backpropagation
por: Prabhudesai, Mihir, et al.
Publicado: (2023) -
Self-Questioning Language Models
por: Chen, Lili, et al.
Publicado: (2025) -
Unified Multimodal Discrete Diffusion
por: Swerdlow, Alexander, et al.
Publicado: (2025)