Can LLMs Lie? Investigation beyond Hallucination
Fuente:
arXiv
Salvato in:
| Autori principali: | Huan, Haoran, Prabhudesai, Mihir, Wu, Mengning, Jaiswal, Shantanu, Pathak, Deepak |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2025
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
Diffusion Beats Autoregressive in Data-Constrained Settings
di: Prabhudesai, Mihir, et al.
Pubblicazione: (2025)
di: Prabhudesai, Mihir, et al.
Pubblicazione: (2025)
Iterative Refinement Improves Compositional Image Generation
di: Jaiswal, Shantanu, et al.
Pubblicazione: (2026)
di: Jaiswal, Shantanu, et al.
Pubblicazione: (2026)
Aligning Text-to-Image Diffusion Models with Reward Backpropagation
di: Prabhudesai, Mihir, et al.
Pubblicazione: (2023)
di: Prabhudesai, Mihir, et al.
Pubblicazione: (2023)
Self-Questioning Language Models
di: Chen, Lili, et al.
Pubblicazione: (2025)
di: Chen, Lili, et al.
Pubblicazione: (2025)
Unified Multimodal Discrete Diffusion
di: Swerdlow, Alexander, et al.
Pubblicazione: (2025)
di: Swerdlow, Alexander, et al.
Pubblicazione: (2025)
Video Diffusion Alignment via Reward Gradients
di: Prabhudesai, Mihir, et al.
Pubblicazione: (2024)
di: Prabhudesai, Mihir, et al.
Pubblicazione: (2024)
Maximizing Confidence Alone Improves Reasoning
di: Prabhudesai, Mihir, et al.
Pubblicazione: (2025)
di: Prabhudesai, Mihir, et al.
Pubblicazione: (2025)
Can Knowledge Graphs Reduce Hallucinations in LLMs? : A Survey
di: Agrawal, Garima, et al.
Pubblicazione: (2023)
di: Agrawal, Garima, et al.
Pubblicazione: (2023)
Solving Physics Olympiad via Reinforcement Learning on Physics Simulators
di: Prabhudesai, Mihir, et al.
Pubblicazione: (2026)
di: Prabhudesai, Mihir, et al.
Pubblicazione: (2026)
Measuring Visual Understanding in Telecom domain: Performance Metrics for Image-to-UML conversion using VLMs
di: Ranjani, HG, et al.
Pubblicazione: (2025)
di: Ranjani, HG, et al.
Pubblicazione: (2025)
Can LLMs be Fooled? Investigating Vulnerabilities in LLMs
di: Abdali, Sara, et al.
Pubblicazione: (2024)
di: Abdali, Sara, et al.
Pubblicazione: (2024)
Zero-Shot Visual Reasoning by Vision-Language Models: Benchmarking and Analysis
di: Nagar, Aishik, et al.
Pubblicazione: (2024)
di: Nagar, Aishik, et al.
Pubblicazione: (2024)
Learning to Reason Iteratively and Parallelly for Complex Visual Reasoning Scenarios
di: Jaiswal, Shantanu, et al.
Pubblicazione: (2024)
di: Jaiswal, Shantanu, et al.
Pubblicazione: (2024)
Teaming LLMs to Detect and Mitigate Hallucinations
di: Till, Demian, et al.
Pubblicazione: (2025)
di: Till, Demian, et al.
Pubblicazione: (2025)
Which LLMs are Difficult to Detect? A Detailed Analysis of Potential Factors Contributing to Difficulties in LLM Text Detection
di: Thorat, Shantanu, et al.
Pubblicazione: (2024)
di: Thorat, Shantanu, et al.
Pubblicazione: (2024)
LLMs Will Always Hallucinate, and We Need to Live With This
di: Banerjee, Sourav, et al.
Pubblicazione: (2024)
di: Banerjee, Sourav, et al.
Pubblicazione: (2024)
Can Multimodal LLMs Perform Time Series Anomaly Detection?
di: Xu, Xiongxiao, et al.
Pubblicazione: (2025)
di: Xu, Xiongxiao, et al.
Pubblicazione: (2025)
Sharpness-Aware Minimization Can Hallucinate Minimizers
di: Park, Chanwoong, et al.
Pubblicazione: (2025)
di: Park, Chanwoong, et al.
Pubblicazione: (2025)
Robust Hallucination Detection in LLMs via Adaptive Token Selection
di: Niu, Mengjia, et al.
Pubblicazione: (2025)
di: Niu, Mengjia, et al.
Pubblicazione: (2025)
Intrinsic Explainability of Multimodal Learning for Crop Yield Prediction
di: Najjar, Hiba, et al.
Pubblicazione: (2025)
di: Najjar, Hiba, et al.
Pubblicazione: (2025)
Beyond Prompt-Induced Lies: Investigating LLM Deception on Benign Prompts
di: Wu, Zhaomin, et al.
Pubblicazione: (2025)
di: Wu, Zhaomin, et al.
Pubblicazione: (2025)
Semantic Energy: Detecting LLM Hallucination Beyond Entropy
di: Ma, Huan, et al.
Pubblicazione: (2025)
di: Ma, Huan, et al.
Pubblicazione: (2025)
A Concise Review of Hallucinations in LLMs and their Mitigation
di: Pulkundwar, Parth, et al.
Pubblicazione: (2025)
di: Pulkundwar, Parth, et al.
Pubblicazione: (2025)
HalluGuard: Demystifying Data-Driven and Reasoning-Driven Hallucinations in LLMs
di: Zeng, Xinyue, et al.
Pubblicazione: (2026)
di: Zeng, Xinyue, et al.
Pubblicazione: (2026)
Marginals Before Conditionals
di: Sahasrabudhe, Mihir
Pubblicazione: (2026)
di: Sahasrabudhe, Mihir
Pubblicazione: (2026)
Meta-Evolve: Continuous Robot Evolution for One-to-many Policy Transfer
di: Liu, Xingyu, et al.
Pubblicazione: (2024)
di: Liu, Xingyu, et al.
Pubblicazione: (2024)
Generation Constraint Scaling Can Mitigate Hallucination
di: Kollias, Georgios, et al.
Pubblicazione: (2024)
di: Kollias, Georgios, et al.
Pubblicazione: (2024)
Evolutionary Policy Optimization
di: Wang, Jianren, et al.
Pubblicazione: (2025)
di: Wang, Jianren, et al.
Pubblicazione: (2025)
Cost-Effective Hallucination Detection for LLMs
di: Valentin, Simon, et al.
Pubblicazione: (2024)
di: Valentin, Simon, et al.
Pubblicazione: (2024)
Is Implicit Knowledge Enough for LLMs? A RAG Approach for Tree-based Structures
di: Gupte, Mihir, et al.
Pubblicazione: (2025)
di: Gupte, Mihir, et al.
Pubblicazione: (2025)
Failure Modes of LLMs for Causal Reasoning on Narratives
di: Yamin, Khurram, et al.
Pubblicazione: (2024)
di: Yamin, Khurram, et al.
Pubblicazione: (2024)
Unlearners Can Lie: Evaluating and Improving Honesty in LLM Unlearning
di: Gu, Renjie, et al.
Pubblicazione: (2026)
di: Gu, Renjie, et al.
Pubblicazione: (2026)
Controlled Causal Hallucinations Can Estimate Phantom Nodes in Multiexpert Mixtures of Fuzzy Cognitive Maps
di: Panda, Akash Kumar, et al.
Pubblicazione: (2024)
di: Panda, Akash Kumar, et al.
Pubblicazione: (2024)
SparseSSM: Efficient Selective Structured State Space Models Can Be Pruned in One-Shot
di: Tuo, Kaiwen, et al.
Pubblicazione: (2025)
di: Tuo, Kaiwen, et al.
Pubblicazione: (2025)
Benchmarking Reliability of Deep Learning Models for Pathological Gait Classification
di: Jaiswal, Abhishek, et al.
Pubblicazione: (2024)
di: Jaiswal, Abhishek, et al.
Pubblicazione: (2024)
Can LLMs Understand Time Series Anomalies?
di: Zhou, Zihao, et al.
Pubblicazione: (2024)
di: Zhou, Zihao, et al.
Pubblicazione: (2024)
Efficient Distributed Training through Gradient Compression with Sparsification and Quantization Techniques
di: Singh, Shruti, et al.
Pubblicazione: (2024)
di: Singh, Shruti, et al.
Pubblicazione: (2024)
Agent Factories for High Level Synthesis: How Far Can General-Purpose Coding Agents Go in Hardware Optimization?
di: Bhandwaldar, Abhishek, et al.
Pubblicazione: (2026)
di: Bhandwaldar, Abhishek, et al.
Pubblicazione: (2026)
Can LLMs subtract numbers?
di: Jobanputra, Mayank, et al.
Pubblicazione: (2025)
di: Jobanputra, Mayank, et al.
Pubblicazione: (2025)
MobileKernelBench: Can LLMs Write Efficient Kernels for Mobile Devices?
di: Zou, Xingze, et al.
Pubblicazione: (2026)
di: Zou, Xingze, et al.
Pubblicazione: (2026)
Documenti analoghi
-
Diffusion Beats Autoregressive in Data-Constrained Settings
di: Prabhudesai, Mihir, et al.
Pubblicazione: (2025) -
Iterative Refinement Improves Compositional Image Generation
di: Jaiswal, Shantanu, et al.
Pubblicazione: (2026) -
Aligning Text-to-Image Diffusion Models with Reward Backpropagation
di: Prabhudesai, Mihir, et al.
Pubblicazione: (2023) -
Self-Questioning Language Models
di: Chen, Lili, et al.
Pubblicazione: (2025) -
Unified Multimodal Discrete Diffusion
di: Swerdlow, Alexander, et al.
Pubblicazione: (2025)