First Hallucination Tokens Are Different from Conditional Ones
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Snel, Jakob, Oh, Seong Joon |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2025
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
DISCO: Diversifying Sample Condensation for Efficient Model Evaluation
von: Rubinstein, Alexander, et al.
Veröffentlicht: (2025)
von: Rubinstein, Alexander, et al.
Veröffentlicht: (2025)
Do Deep Neural Network Solutions Form a Star Domain?
von: Sonthalia, Ankit, et al.
Veröffentlicht: (2024)
von: Sonthalia, Ankit, et al.
Veröffentlicht: (2024)
Dynamics Reveals Structure: Challenging the Linear Propagation Assumption
von: Chang, Hoyeon, et al.
Veröffentlicht: (2026)
von: Chang, Hoyeon, et al.
Veröffentlicht: (2026)
CPR: Mitigating Large Language Model Hallucinations with Curative Prompt Refinement
von: Shim, Jung-Woo, et al.
Veröffentlicht: (2025)
von: Shim, Jung-Woo, et al.
Veröffentlicht: (2025)
Multi-stage Prompt Refinement for Mitigating Hallucinations in Large Language Models
von: Shim, Jung-Woo, et al.
Veröffentlicht: (2025)
von: Shim, Jung-Woo, et al.
Veröffentlicht: (2025)
Scaling Up Membership Inference: When and How Attacks Succeed on Large Language Models
von: Puerto, Haritz, et al.
Veröffentlicht: (2024)
von: Puerto, Haritz, et al.
Veröffentlicht: (2024)
Universal Algorithm-Implicit Learning
von: Woerner, Stefano, et al.
Veröffentlicht: (2026)
von: Woerner, Stefano, et al.
Veröffentlicht: (2026)
Squish and Release: Exposing Hidden Hallucinations by Making Them Surface as Safety Signals
von: Oh, Nathaniel, et al.
Veröffentlicht: (2026)
von: Oh, Nathaniel, et al.
Veröffentlicht: (2026)
LLM generation novelty through the lens of semantic similarity
von: Davydov, Philipp, et al.
Veröffentlicht: (2025)
von: Davydov, Philipp, et al.
Veröffentlicht: (2025)
Dr.LLM: Dynamic Layer Routing in LLMs
von: Heakl, Ahmed, et al.
Veröffentlicht: (2025)
von: Heakl, Ahmed, et al.
Veröffentlicht: (2025)
Calibrating Large Language Models Using Their Generations Only
von: Ulmer, Dennis, et al.
Veröffentlicht: (2024)
von: Ulmer, Dennis, et al.
Veröffentlicht: (2024)
OrthoRank: Token Selection via Sink Token Orthogonality for Efficient LLM inference
von: Shin, Seungjun, et al.
Veröffentlicht: (2025)
von: Shin, Seungjun, et al.
Veröffentlicht: (2025)
Are We Done with Object-Centric Learning?
von: Rubinstein, Alexander, et al.
Veröffentlicht: (2025)
von: Rubinstein, Alexander, et al.
Veröffentlicht: (2025)
Scalable Ensemble Diversification for OOD Generalization and Detection
von: Rubinstein, Alexander, et al.
Veröffentlicht: (2024)
von: Rubinstein, Alexander, et al.
Veröffentlicht: (2024)
Towards Dynamic Trend Filtering through Trend Point Detection with Reinforcement Learning
von: Seong, Jihyeon, et al.
Veröffentlicht: (2024)
von: Seong, Jihyeon, et al.
Veröffentlicht: (2024)
From Generator to Embedder: Harnessing Innate Abilities of Multimodal LLMs via Building Zero-Shot Discriminative Embedding Model
von: Ju, Yeong-Joon, et al.
Veröffentlicht: (2025)
von: Ju, Yeong-Joon, et al.
Veröffentlicht: (2025)
Energy-Efficient Wireless LLM Inference via Uncertainty and Importance-Aware Speculative Decoding
von: Park, Jihoon, et al.
Veröffentlicht: (2025)
von: Park, Jihoon, et al.
Veröffentlicht: (2025)
MASEval: Extending Multi-Agent Evaluation from Models to Systems
von: Emde, Cornelius, et al.
Veröffentlicht: (2026)
von: Emde, Cornelius, et al.
Veröffentlicht: (2026)
Two Heads Are Better than One: Simulating Large Transformers with Small Ones
von: Yu, Hantao, et al.
Veröffentlicht: (2025)
von: Yu, Hantao, et al.
Veröffentlicht: (2025)
What Makes Looped Transformers Perform Better Than Non-Recursive Ones
von: Gong, Zixuan, et al.
Veröffentlicht: (2025)
von: Gong, Zixuan, et al.
Veröffentlicht: (2025)
Towards User-Focused Research in Training Data Attribution for Human-Centered Explainable AI
von: Nguyen, Elisa, et al.
Veröffentlicht: (2024)
von: Nguyen, Elisa, et al.
Veröffentlicht: (2024)
Scalable Token-Level Hallucination Detection in Large Language Models
von: Min, Rui, et al.
Veröffentlicht: (2026)
von: Min, Rui, et al.
Veröffentlicht: (2026)
Performance Control in Early Exiting to Deploy Large Models at the Same Cost of Smaller Ones
von: Mofakhami, Mehrnaz, et al.
Veröffentlicht: (2024)
von: Mofakhami, Mehrnaz, et al.
Veröffentlicht: (2024)
TRAP: Targeted Random Adversarial Prompt Honeypot for Black-Box Identification
von: Gubri, Martin, et al.
Veröffentlicht: (2024)
von: Gubri, Martin, et al.
Veröffentlicht: (2024)
Mitigating Shortcut Learning with Diffusion Counterfactuals and Diverse Ensembles
von: Scimeca, Luca, et al.
Veröffentlicht: (2023)
von: Scimeca, Luca, et al.
Veröffentlicht: (2023)
Qubit-centric Transformer for Surface Code Decoding
von: Park, Seong-Joon, et al.
Veröffentlicht: (2025)
von: Park, Seong-Joon, et al.
Veröffentlicht: (2025)
SelfReflect: Can LLMs Communicate Their Internal Answer Distribution?
von: Kirchhof, Michael, et al.
Veröffentlicht: (2025)
von: Kirchhof, Michael, et al.
Veröffentlicht: (2025)
Scratching Visual Transformer's Back with Uniform Attention
von: Hyeon-Woo, Nam, et al.
Veröffentlicht: (2022)
von: Hyeon-Woo, Nam, et al.
Veröffentlicht: (2022)
Abstraction for Offline Goal-Conditioned Reinforcement Learning
von: Wibault, Clarisse, et al.
Veröffentlicht: (2026)
von: Wibault, Clarisse, et al.
Veröffentlicht: (2026)
NBDI: A Simple and Effective Termination Condition for Skill Extraction from Task-Agnostic Demonstrations
von: Kim, Myunsoo, et al.
Veröffentlicht: (2025)
von: Kim, Myunsoo, et al.
Veröffentlicht: (2025)
Design Conditions for Intra-Group Learning of Sequence-Level Rewards: Token Gradient Cancellation
von: Ding, Fei, et al.
Veröffentlicht: (2026)
von: Ding, Fei, et al.
Veröffentlicht: (2026)
Goal-Conditioned Agents that Learn Everything All at Once
von: Matthews, Michael, et al.
Veröffentlicht: (2026)
von: Matthews, Michael, et al.
Veröffentlicht: (2026)
It Takes Two: Complementary Self-Distillation for Contextual Integrity in LLMs
von: Park, Sangwoo, et al.
Veröffentlicht: (2026)
von: Park, Sangwoo, et al.
Veröffentlicht: (2026)
Explicit Diversity Conditions for Effective Question Answer Generation with Large Language Models
von: Yadav, Vikas, et al.
Veröffentlicht: (2024)
von: Yadav, Vikas, et al.
Veröffentlicht: (2024)
Unsupervised Feature Selection to Identify Important ICD-10 Codes for Machine Learning: A Case Study on a Coronary Artery Disease Patient Cohort
von: Ghasemi, Peyman, et al.
Veröffentlicht: (2023)
von: Ghasemi, Peyman, et al.
Veröffentlicht: (2023)
Efficient Contrastive Decoding with Probabilistic Hallucination Detection - Mitigating Hallucinations in Large Vision Language Models -
von: Fieback, Laura, et al.
Veröffentlicht: (2025)
von: Fieback, Laura, et al.
Veröffentlicht: (2025)
Let Me Think! A Long Chain-of-Thought Can Be Worth Exponentially Many Short Ones
von: Mirtaheri, Parsa, et al.
Veröffentlicht: (2025)
von: Mirtaheri, Parsa, et al.
Veröffentlicht: (2025)
ERBench: An Entity-Relationship based Automatically Verifiable Hallucination Benchmark for Large Language Models
von: Oh, Jio, et al.
Veröffentlicht: (2024)
von: Oh, Jio, et al.
Veröffentlicht: (2024)
Rethinking Hallucinations: Correctness, Consistency, and Prompt Multiplicity
von: Ganesh, Prakhar, et al.
Veröffentlicht: (2026)
von: Ganesh, Prakhar, et al.
Veröffentlicht: (2026)
Role-Aware Conditional Inference for Spatiotemporal Ecosystem Carbon Flux Prediction
von: Sun, Yiming, et al.
Veröffentlicht: (2026)
von: Sun, Yiming, et al.
Veröffentlicht: (2026)
Ähnliche Einträge
-
DISCO: Diversifying Sample Condensation for Efficient Model Evaluation
von: Rubinstein, Alexander, et al.
Veröffentlicht: (2025) -
Do Deep Neural Network Solutions Form a Star Domain?
von: Sonthalia, Ankit, et al.
Veröffentlicht: (2024) -
Dynamics Reveals Structure: Challenging the Linear Propagation Assumption
von: Chang, Hoyeon, et al.
Veröffentlicht: (2026) -
CPR: Mitigating Large Language Model Hallucinations with Curative Prompt Refinement
von: Shim, Jung-Woo, et al.
Veröffentlicht: (2025) -
Multi-stage Prompt Refinement for Mitigating Hallucinations in Large Language Models
von: Shim, Jung-Woo, et al.
Veröffentlicht: (2025)