Adaptive Activation Cancellation for Hallucination Mitigation in Large Language Models
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Yocam, Eric, Vaidyan, Varghese, Comert, Gurcan, Kalathas, Paris, Wang, Yong, Mwakalonge, Judith L. |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2026
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Computational Economics in Large Language Models: Exploring Model Behavior and Incentive Design under Resource Constraints
von: Reddy, Sandeep, et al.
Veröffentlicht: (2025)
von: Reddy, Sandeep, et al.
Veröffentlicht: (2025)
A Prescriptive Framework for Determining Optimal Days for Short-Term Traffic Counts
von: Mukwaya, Arthur, et al.
Veröffentlicht: (2025)
von: Mukwaya, Arthur, et al.
Veröffentlicht: (2025)
Listwise Direct Preference Optimization with Multi-Dimensional Preference Mixing
von: Sun, Yuhui, et al.
Veröffentlicht: (2025)
von: Sun, Yuhui, et al.
Veröffentlicht: (2025)
Cost-Aware Model Selection for Text Classification: Multi-Objective Trade-offs Between Fine-Tuned Encoders and LLM Prompting in Production
von: Gonzalez, Alberto Andres Valdes
Veröffentlicht: (2026)
von: Gonzalez, Alberto Andres Valdes
Veröffentlicht: (2026)
Efficient Strategy for Improving Large Language Model (LLM) Capabilities
von: Gutiérrez, Julián Camilo Velandia
Veröffentlicht: (2025)
von: Gutiérrez, Julián Camilo Velandia
Veröffentlicht: (2025)
Beyond Subtokens: A Rich Character Embedding for Low-resource and Morphologically Complex Languages
von: Schneider, Felix, et al.
Veröffentlicht: (2026)
von: Schneider, Felix, et al.
Veröffentlicht: (2026)
Entropy-Based Measurement of Value Drift and Alignment Work in Large Language Models
von: Fadli, Samih
Veröffentlicht: (2025)
von: Fadli, Samih
Veröffentlicht: (2025)
KSHSeek: Data-Driven Approaches to Mitigating and Detecting Knowledge-Shortcut Hallucinations in Generative Models
von: Liu, Zhongxin, et al.
Veröffentlicht: (2025)
von: Liu, Zhongxin, et al.
Veröffentlicht: (2025)
When is dataset cartography ineffective? Using training dynamics does not improve robustness against Adversarial SQuAD
von: Mandal, Paul K.
Veröffentlicht: (2025)
von: Mandal, Paul K.
Veröffentlicht: (2025)
Less is More: Learning Graph Tasks with Just LLMs
von: Shirai, Sola, et al.
Veröffentlicht: (2025)
von: Shirai, Sola, et al.
Veröffentlicht: (2025)
LLM Performance Predictors: Learning When to Escalate in Hybrid Human-AI Moderation Systems
von: Bachar, Or, et al.
Veröffentlicht: (2026)
von: Bachar, Or, et al.
Veröffentlicht: (2026)
Merge-Bench: Resolve Merge Conflicts with Large Language Models
von: Schesch, Benedikt, et al.
Veröffentlicht: (2026)
von: Schesch, Benedikt, et al.
Veröffentlicht: (2026)
Automated Bug Triaging using Instruction-Tuned Large Language Models
von: Kiashemshaki, Kiana, et al.
Veröffentlicht: (2025)
von: Kiashemshaki, Kiana, et al.
Veröffentlicht: (2025)
Adversarially Probing Cross-Family Sound Symbolism in 27 Languages
von: Sharma, Anika, et al.
Veröffentlicht: (2025)
von: Sharma, Anika, et al.
Veröffentlicht: (2025)
Evaluating the Systematic Reasoning Abilities of Large Language Models through Graph Coloring
von: Heyman, Alex, et al.
Veröffentlicht: (2025)
von: Heyman, Alex, et al.
Veröffentlicht: (2025)
Reasoning Large Language Model Errors Arise from Hallucinating Critical Problem Features
von: Heyman, Alex, et al.
Veröffentlicht: (2025)
von: Heyman, Alex, et al.
Veröffentlicht: (2025)
Towards Alignment-Centric Paradigm: A Survey of Instruction Tuning in Large Language Models
von: Han, Xudong, et al.
Veröffentlicht: (2025)
von: Han, Xudong, et al.
Veröffentlicht: (2025)
Mitigating Cross-Lingual Cultural Inconsistencies in LLMs via Consensus-Driven Preference Optimisation
von: Resck, Lucas, et al.
Veröffentlicht: (2026)
von: Resck, Lucas, et al.
Veröffentlicht: (2026)
MCP: A Control-Theoretic Orchestration Framework for Synergistic Efficiency and Interpretability in Multimodal Large Language Models
von: Zhang, Luyan
Veröffentlicht: (2025)
von: Zhang, Luyan
Veröffentlicht: (2025)
ACE: Exploring Activation Cosine Similarity and Variance for Accurate and Calibration-Efficient LLM Pruning
von: Mi, Zhendong, et al.
Veröffentlicht: (2025)
von: Mi, Zhendong, et al.
Veröffentlicht: (2025)
Automated CAD Modeling Sequence Generation from Text Descriptions via Transformer-Based Large Language Models
von: Liao, Jianxing, et al.
Veröffentlicht: (2025)
von: Liao, Jianxing, et al.
Veröffentlicht: (2025)
ADALog: Adaptive Unsupervised Anomaly detection in Logs with Self-attention Masked Language Model
von: Pospieszny, Przemek, et al.
Veröffentlicht: (2025)
von: Pospieszny, Przemek, et al.
Veröffentlicht: (2025)
Latent Instruction Representation Alignment: defending against jailbreaks, backdoors and undesired knowledge in LLMs
von: Easley, Eric, et al.
Veröffentlicht: (2026)
von: Easley, Eric, et al.
Veröffentlicht: (2026)
A Comparative Study of Feature Selection in Tsetlin Machines
von: Halenka, Vojtech, et al.
Veröffentlicht: (2025)
von: Halenka, Vojtech, et al.
Veröffentlicht: (2025)
Cytoarchitecture in Words: Weakly Supervised Vision-Language Modeling for Human Brain Microscopy
von: Sutton, Matthew, et al.
Veröffentlicht: (2026)
von: Sutton, Matthew, et al.
Veröffentlicht: (2026)
Exploring Model Invariance with Discrete Search for Ultra-Low-Bit Quantization
von: Wen, Yuqiao, et al.
Veröffentlicht: (2025)
von: Wen, Yuqiao, et al.
Veröffentlicht: (2025)
EBBS: An Ensemble with Bi-Level Beam Search for Zero-Shot Machine Translation
von: Wen, Yuqiao, et al.
Veröffentlicht: (2024)
von: Wen, Yuqiao, et al.
Veröffentlicht: (2024)
OntoLogX: Ontology-Guided Knowledge Graph Extraction from Cybersecurity Logs with Large Language Models
von: Cotti, Luca, et al.
Veröffentlicht: (2025)
von: Cotti, Luca, et al.
Veröffentlicht: (2025)
Predictive Analytics for Collaborators Answers, Code Quality, and Dropout on Stack Overflow
von: Zolduoarrati, Elijah, et al.
Veröffentlicht: (2025)
von: Zolduoarrati, Elijah, et al.
Veröffentlicht: (2025)
Survey Transfer Learning: Recycling Data with Silicon Responses
von: Amini, Ali
Veröffentlicht: (2025)
von: Amini, Ali
Veröffentlicht: (2025)
Neural Attention: A Novel Mechanism for Enhanced Expressive Power in Transformer Models
von: DiGiugno, Andrew, et al.
Veröffentlicht: (2025)
von: DiGiugno, Andrew, et al.
Veröffentlicht: (2025)
D-COT: Disciplined Chain-of-Thought Learning for Efficient Reasoning in Small Language Models
von: Ubukata, Shunsuke
Veröffentlicht: (2026)
von: Ubukata, Shunsuke
Veröffentlicht: (2026)
Emergent Lexical Semantics in Neural Language Models: Testing Martin's Law on LLM-Generated Text
von: Kugler, Kai
Veröffentlicht: (2025)
von: Kugler, Kai
Veröffentlicht: (2025)
Transformer Scalability Crisis: The First Comprehensive Empirical Analysis of Performance Walls in Modern Language Models
von: Moghadasi, Mahdi Naser, et al.
Veröffentlicht: (2026)
von: Moghadasi, Mahdi Naser, et al.
Veröffentlicht: (2026)
Bridging the Gap: An Intermediate Language for Enhanced and Cost-Effective Grapheme-to-Phoneme Conversion with Homographs with Multiple Pronunciations Disambiguation
von: Bertina, Abbas, et al.
Veröffentlicht: (2025)
von: Bertina, Abbas, et al.
Veröffentlicht: (2025)
Let Your Graph Do the Talking: Encoding Structured Data for LLMs
von: Perozzi, Bryan, et al.
Veröffentlicht: (2024)
von: Perozzi, Bryan, et al.
Veröffentlicht: (2024)
LLM-Rubric: A Multidimensional, Calibrated Approach to Automated Evaluation of Natural Language Texts
von: Hashemi, Helia, et al.
Veröffentlicht: (2024)
von: Hashemi, Helia, et al.
Veröffentlicht: (2024)
Neural Activation Patterns Across Language Model Architectures: A Comprehensive Analysis of Cognitive Task Performance
von: Naser-Moghadasi, Mahdi, et al.
Veröffentlicht: (2026)
von: Naser-Moghadasi, Mahdi, et al.
Veröffentlicht: (2026)
Layer-Aware Embedding Fusion for LLMs in Text Classifications
von: Gwak, Jiho, et al.
Veröffentlicht: (2025)
von: Gwak, Jiho, et al.
Veröffentlicht: (2025)
On the Influence of Discourse Relations in Persuasive Texts
von: Turk, Nawar, et al.
Veröffentlicht: (2025)
von: Turk, Nawar, et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
Computational Economics in Large Language Models: Exploring Model Behavior and Incentive Design under Resource Constraints
von: Reddy, Sandeep, et al.
Veröffentlicht: (2025) -
A Prescriptive Framework for Determining Optimal Days for Short-Term Traffic Counts
von: Mukwaya, Arthur, et al.
Veröffentlicht: (2025) -
Listwise Direct Preference Optimization with Multi-Dimensional Preference Mixing
von: Sun, Yuhui, et al.
Veröffentlicht: (2025) -
Cost-Aware Model Selection for Text Classification: Multi-Objective Trade-offs Between Fine-Tuned Encoders and LLM Prompting in Production
von: Gonzalez, Alberto Andres Valdes
Veröffentlicht: (2026) -
Efficient Strategy for Improving Large Language Model (LLM) Capabilities
von: Gutiérrez, Julián Camilo Velandia
Veröffentlicht: (2025)