Expanding before Inferring: Enhancing Factuality in Large Language Models through Premature Layers Interpolation
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Chen, Dingwei, Liu, Ziqiang, Fang, Feiteng, Leong, Chak Tou, Ni, Shiwen, Argha, Ahmadreza, Alinejad-Rokny, Hamid, Yang, Min, Li, Chengming |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2025
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Lower Layers Matter: Alleviating Hallucination via Multi-Layer Fusion Contrastive Decoding with Truthfulness Refocused
von: Chen, Dingwei, et al.
Veröffentlicht: (2024)
von: Chen, Dingwei, et al.
Veröffentlicht: (2024)
ETAGE: Enhanced Test Time Adaptation with Integrated Entropy and Gradient Norms for Robust Model Performance
von: Shamsi, Afshar, et al.
Veröffentlicht: (2024)
von: Shamsi, Afshar, et al.
Veröffentlicht: (2024)
Interpretable graph-based models on multimodal biomedical data integration: A technical review and benchmarking
von: Sadeghi, Alireza, et al.
Veröffentlicht: (2025)
von: Sadeghi, Alireza, et al.
Veröffentlicht: (2025)
xJailbreak: Representation Space Guided Reinforcement Learning for Interpretable LLM Jailbreaking
von: Lee, Sunbowen, et al.
Veröffentlicht: (2025)
von: Lee, Sunbowen, et al.
Veröffentlicht: (2025)
RxSafeBench: Identifying Medication Safety Issues of Large Language Models in Simulated Consultation
von: Zhao, Jiahao, et al.
Veröffentlicht: (2025)
von: Zhao, Jiahao, et al.
Veröffentlicht: (2025)
CLinNET: An Interpretable and Uncertainty‐Aware Deep Learning Framework for Multi‐Modal Clinical Genomics
von: Ivan Bakhshayeshi, et al.
Veröffentlicht: (2026)
von: Ivan Bakhshayeshi, et al.
Veröffentlicht: (2026)
Forgetting before Learning: Utilizing Parametric Arithmetic for Knowledge Updating in Large Language Models
von: Ni, Shiwen, et al.
Veröffentlicht: (2023)
von: Ni, Shiwen, et al.
Veröffentlicht: (2023)
SemanticST: Spatially Informed Semantic Graph Learning for Clustering, Integration, and Scalable Analysis of Spatial Transcriptomics
von: Zahedi, Roxana, et al.
Veröffentlicht: (2025)
von: Zahedi, Roxana, et al.
Veröffentlicht: (2025)
Small Language Model as Data Prospector for Large Language Model
von: Ni, Shiwen, et al.
Veröffentlicht: (2024)
von: Ni, Shiwen, et al.
Veröffentlicht: (2024)
AgentCourt: Simulating Court with Adversarial Evolvable Lawyer Agents
von: Chen, Guhong, et al.
Veröffentlicht: (2024)
von: Chen, Guhong, et al.
Veröffentlicht: (2024)
PersonaMath: Boosting Mathematical Reasoning via Persona-Driven Data Augmentation
von: Luo, Jing, et al.
Veröffentlicht: (2024)
von: Luo, Jing, et al.
Veröffentlicht: (2024)
STORYTELLER: An Enhanced Plot-Planning Framework for Coherent and Cohesive Story Generation
von: Li, Jiaming, et al.
Veröffentlicht: (2025)
von: Li, Jiaming, et al.
Veröffentlicht: (2025)
PLOT: Enhancing Preference Learning via Optimal Transport
von: Zhu, Liang, et al.
Veröffentlicht: (2026)
von: Zhu, Liang, et al.
Veröffentlicht: (2026)
Act-Adaptive Margin: Dynamically Calibrating Reward Models for Subjective Ambiguity
von: Fang, Feiteng, et al.
Veröffentlicht: (2025)
von: Fang, Feiteng, et al.
Veröffentlicht: (2025)
Beyond Quantity: Trajectory Diversity Scaling for Code Agents
von: Chen, Guhong, et al.
Veröffentlicht: (2026)
von: Chen, Guhong, et al.
Veröffentlicht: (2026)
Automatic Paper Reviewing with Heterogeneous Graph Reasoning over LLM-Simulated Reviewer-Author Debates
von: Li, Shuaimin, et al.
Veröffentlicht: (2025)
von: Li, Shuaimin, et al.
Veröffentlicht: (2025)
CLaSp: In-Context Layer Skip for Self-Speculative Decoding
von: Chen, Longze, et al.
Veröffentlicht: (2025)
von: Chen, Longze, et al.
Veröffentlicht: (2025)
Enhancing Noise Robustness of Retrieval-Augmented Language Models with Adaptive Adversarial Training
von: Fang, Feiteng, et al.
Veröffentlicht: (2024)
von: Fang, Feiteng, et al.
Veröffentlicht: (2024)
Transcriptomic Models for Immunotherapy Response Prediction Show Limited Cross-cohort Generalisability
von: Liang, Yuheng, et al.
Veröffentlicht: (2026)
von: Liang, Yuheng, et al.
Veröffentlicht: (2026)
Enhancing Monte Carlo Dropout Performance for Uncertainty Quantification
von: Asgharnezhad, Hamzeh, et al.
Veröffentlicht: (2025)
von: Asgharnezhad, Hamzeh, et al.
Veröffentlicht: (2025)
Probing the Difficulty Perception Mechanism of Large Language Models
von: Lee, Sunbowen, et al.
Veröffentlicht: (2025)
von: Lee, Sunbowen, et al.
Veröffentlicht: (2025)
CollectiveSFT: Scaling Large Language Models for Chinese Medical Benchmark with Collective Instructions in Healthcare
von: Zhu, Jingwei, et al.
Veröffentlicht: (2024)
von: Zhu, Jingwei, et al.
Veröffentlicht: (2024)
How chromatin interactions shed light on interpreting non-coding genomic variants: opportunities and future direc-tions
von: Liang, Yuheng, et al.
Veröffentlicht: (2024)
von: Liang, Yuheng, et al.
Veröffentlicht: (2024)
Layer-wise Regularized Dropout for Neural Language Models
von: Ni, Shiwen, et al.
Veröffentlicht: (2024)
von: Ni, Shiwen, et al.
Veröffentlicht: (2024)
Finding RELIEF: Shaping Reasoning Behavior without Reasoning Supervision via Belief Engineering
von: Leong, Chak Tou, et al.
Veröffentlicht: (2026)
von: Leong, Chak Tou, et al.
Veröffentlicht: (2026)
Structuring Reasoning for Complex Rules Beyond Flat Representations
von: Yang, Zhihao, et al.
Veröffentlicht: (2025)
von: Yang, Zhihao, et al.
Veröffentlicht: (2025)
AutoPatent: A Multi-Agent Framework for Automatic Patent Generation
von: Wang, Qiyao, et al.
Veröffentlicht: (2024)
von: Wang, Qiyao, et al.
Veröffentlicht: (2024)
Seeing Isn't Believing: Mitigating Belief Inertia via Active Intervention in Embodied Agents
von: Wang, Hanlin, et al.
Veröffentlicht: (2026)
von: Wang, Hanlin, et al.
Veröffentlicht: (2026)
STeCa: Step-level Trajectory Calibration for LLM Agent Learning
von: Wang, Hanlin, et al.
Veröffentlicht: (2025)
von: Wang, Hanlin, et al.
Veröffentlicht: (2025)
Why Safeguarded Ships Run Aground? Aligned Large Language Models' Safety Mechanisms Tend to Be Anchored in The Template Region
von: Leong, Chak Tou, et al.
Veröffentlicht: (2025)
von: Leong, Chak Tou, et al.
Veröffentlicht: (2025)
E2CL: Exploration-based Error Correction Learning for Embodied Agents
von: Wang, Hanlin, et al.
Veröffentlicht: (2024)
von: Wang, Hanlin, et al.
Veröffentlicht: (2024)
DeliLaw: A Chinese Legal Counselling System Based on a Large Language Model
von: Xie, Nan, et al.
Veröffentlicht: (2024)
von: Xie, Nan, et al.
Veröffentlicht: (2024)
CoTJudger: A Graph-Driven Framework for Automatic Evaluation of Chain-of-Thought Efficiency and Redundancy in LRMs
von: Li, Siyi, et al.
Veröffentlicht: (2026)
von: Li, Siyi, et al.
Veröffentlicht: (2026)
InteractWeb-Bench: Can Multimodal Agent Escape Blind Execution in Interactive Website Generation?
von: Wang, Qiyao, et al.
Veröffentlicht: (2026)
von: Wang, Qiyao, et al.
Veröffentlicht: (2026)
PatRe: A Full-Stage Office Action and Rebuttal Generation Benchmark for Patent Examination
von: Wang, Qiyao, et al.
Veröffentlicht: (2026)
von: Wang, Qiyao, et al.
Veröffentlicht: (2026)
KNN-SSD: Enabling Dynamic Self-Speculative Decoding via Nearest Neighbor Layer Set Optimization
von: Song, Mingbo, et al.
Veröffentlicht: (2025)
von: Song, Mingbo, et al.
Veröffentlicht: (2025)
Scaling over Scaling: Exploring Test-Time Scaling Plateau in Large Reasoning Models
von: Wang, Jian, et al.
Veröffentlicht: (2025)
von: Wang, Jian, et al.
Veröffentlicht: (2025)
TokenSkip: Controllable Chain-of-Thought Compression in LLMs
von: Xia, Heming, et al.
Veröffentlicht: (2025)
von: Xia, Heming, et al.
Veröffentlicht: (2025)
Induction Head Toxicity Mechanistically Explains Repetition Curse in Large Language Models
von: Wang, Shuxun, et al.
Veröffentlicht: (2025)
von: Wang, Shuxun, et al.
Veröffentlicht: (2025)
SPA-RL: Reinforcing LLM Agents via Stepwise Progress Attribution
von: Wang, Hanlin, et al.
Veröffentlicht: (2025)
von: Wang, Hanlin, et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
Lower Layers Matter: Alleviating Hallucination via Multi-Layer Fusion Contrastive Decoding with Truthfulness Refocused
von: Chen, Dingwei, et al.
Veröffentlicht: (2024) -
ETAGE: Enhanced Test Time Adaptation with Integrated Entropy and Gradient Norms for Robust Model Performance
von: Shamsi, Afshar, et al.
Veröffentlicht: (2024) -
Interpretable graph-based models on multimodal biomedical data integration: A technical review and benchmarking
von: Sadeghi, Alireza, et al.
Veröffentlicht: (2025) -
xJailbreak: Representation Space Guided Reinforcement Learning for Interpretable LLM Jailbreaking
von: Lee, Sunbowen, et al.
Veröffentlicht: (2025) -
RxSafeBench: Identifying Medication Safety Issues of Large Language Models in Simulated Consultation
von: Zhao, Jiahao, et al.
Veröffentlicht: (2025)