Salvato in:
| Autori principali: | Bao, Yuntai, Zhang, Xuhong, Chen, Jintao, Su, Ge, Cai, Yuxiang, Peng, Hao, Sun, Bing, Weng, Haiqin, Yan, Liu, Yin, Jianwei |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2026
|
| Soggetti: | |
| Accesso online: | https://arxiv.org/abs/2602.05234 |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
Towards Steering without Sacrifice: Principled Training of Steering Vectors for Prompt-only Interventions
di: Bao, Yuntai, et al.
Pubblicazione: (2026)
di: Bao, Yuntai, et al.
Pubblicazione: (2026)
Probing the Geometry of Truth: Consistency and Generalization of Truth Directions in LLMs Across Logical Transformations and Question Answering Tasks
di: Bao, Yuntai, et al.
Pubblicazione: (2025)
di: Bao, Yuntai, et al.
Pubblicazione: (2025)
Scalable Multi-Stage Influence Function for Large Language Models via Eigenvalue-Corrected Kronecker-Factored Parameterization
di: Bao, Yuntai, et al.
Pubblicazione: (2025)
di: Bao, Yuntai, et al.
Pubblicazione: (2025)
PragLocker: Protecting Agent Intellectual Property in Untrusted Deployments via Non-Portable Prompts
di: Li, Qinfeng, et al.
Pubblicazione: (2026)
di: Li, Qinfeng, et al.
Pubblicazione: (2026)
Ground What You See: Hallucination-Resistant MLLMs via Caption Feedback, Diversity-Aware Sampling, and Conflict Regularization
di: Pan, Miao, et al.
Pubblicazione: (2026)
di: Pan, Miao, et al.
Pubblicazione: (2026)
Neural Quantum States in Variational Monte Carlo Method: A Brief Summary
di: Song, Yuntai
Pubblicazione: (2024)
di: Song, Yuntai
Pubblicazione: (2024)
VITAL: Visual-Semantic Dual Supervision for Enhanced and Interpretable Latent Reasoning in Medical MLLMs
di: Li, Qiaoru, et al.
Pubblicazione: (2026)
di: Li, Qiaoru, et al.
Pubblicazione: (2026)
Steer2Edit: From Activation Steering to Component-Level Editing
di: Sun, Chung-En, et al.
Pubblicazione: (2026)
di: Sun, Chung-En, et al.
Pubblicazione: (2026)
Do Not Merge My Model! Safeguarding Open-Source LLMs Against Unauthorized Model Merging
di: Li, Qinfeng, et al.
Pubblicazione: (2025)
di: Li, Qinfeng, et al.
Pubblicazione: (2025)
Faithful and Stable Neuron Explanations for Trustworthy Mechanistic Interpretability
di: Yan, Ge, et al.
Pubblicazione: (2025)
di: Yan, Ge, et al.
Pubblicazione: (2025)
ToolGate: Contract-Grounded and Verified Tool Execution for LLMs
di: Liu, Yanming, et al.
Pubblicazione: (2026)
di: Liu, Yanming, et al.
Pubblicazione: (2026)
RAGFort: Dual-Path Defense Against Proprietary Knowledge Base Extraction in Retrieval-Augmented Generation
di: Li, Qinfeng, et al.
Pubblicazione: (2025)
di: Li, Qinfeng, et al.
Pubblicazione: (2025)
IBISAgent: Reinforcing Pixel-Level Visual Reasoning in MLLMs for Universal Biomedical Object Referring and Segmentation
di: Jiang, Yankai, et al.
Pubblicazione: (2026)
di: Jiang, Yankai, et al.
Pubblicazione: (2026)
GFT: From Imitation to Reward Fine-Tuning with Unbiased Group Advantages and Dynamic Coefficient Rectification
di: Gan, Wangjie, et al.
Pubblicazione: (2026)
di: Gan, Wangjie, et al.
Pubblicazione: (2026)
ReFIne: A Framework for Trustworthy Large Reasoning Models with Reliability, Faithfulness, and Interpretability
di: Sun, Chung-En, et al.
Pubblicazione: (2025)
di: Sun, Chung-En, et al.
Pubblicazione: (2025)
Emerging Trends and Research Hotspots of Remote Ischemic Preconditioning in Cardiac Surgery: A Bibliometric Analysis
di: Linlin Chen, et al.
Pubblicazione: (2025)
di: Linlin Chen, et al.
Pubblicazione: (2025)
TK-Mamba: Marrying KAN With Mamba for Text-Driven 3D Medical Image Segmentation
di: Yang, Haoyu, et al.
Pubblicazione: (2025)
di: Yang, Haoyu, et al.
Pubblicazione: (2025)
Graph Out-of-Distribution Generalization via Causal Intervention
di: Wu, Qitian, et al.
Pubblicazione: (2024)
di: Wu, Qitian, et al.
Pubblicazione: (2024)
Interchange.
Pubblicazione: (1982)
Pubblicazione: (1982)
Interchange.
Pubblicazione: (1983)
Pubblicazione: (1983)
Interchange.
Pubblicazione: (1982)
Pubblicazione: (1982)
Interchange.
Pubblicazione: (1980)
Pubblicazione: (1980)
ERA-CoT: Improving Chain-of-Thought through Entity Relationship Analysis
di: Liu, Yanming, et al.
Pubblicazione: (2024)
di: Liu, Yanming, et al.
Pubblicazione: (2024)
Distributed fault estimation observer design for nonlinear multi‐agent systems with directed graphs
di: Yuhui Weng, et al.
Pubblicazione: (2024)
di: Yuhui Weng, et al.
Pubblicazione: (2024)
Data-Locality-Aware Task Assignment and Scheduling for Distributed Job Executions
di: Zhao, Hailiang, et al.
Pubblicazione: (2024)
di: Zhao, Hailiang, et al.
Pubblicazione: (2024)
Disentangling Long-Short Term State Under Unknown Interventions for Online Time Series Forecasting
di: Cai, Ruichu, et al.
Pubblicazione: (2025)
di: Cai, Ruichu, et al.
Pubblicazione: (2025)
Influence of Macrocracks Extending along the Propagation Direction of Detonation Wave on the Detonation Velocity of Explosives
di: Jun Liu, et al.
Pubblicazione: (2025)
di: Jun Liu, et al.
Pubblicazione: (2025)
Learning Few-Step Diffusion Models by Trajectory Distribution Matching
di: Luo, Yihong, et al.
Pubblicazione: (2025)
di: Luo, Yihong, et al.
Pubblicazione: (2025)
First Extraction of Transverse Momentum Dependent Helicity Distributions
di: Yang, Ke, et al.
Pubblicazione: (2024)
di: Yang, Ke, et al.
Pubblicazione: (2024)
RA-ISF: Learning to Answer and Understand from Retrieval Augmentation via Iterative Self-Feedback
di: Liu, Yanming, et al.
Pubblicazione: (2024)
di: Liu, Yanming, et al.
Pubblicazione: (2024)
Nonstationary Time Series Forecasting via Unknown Distribution Adaptation
di: Li, Zijian, et al.
Pubblicazione: (2024)
di: Li, Zijian, et al.
Pubblicazione: (2024)
A Wideband Distributed Massive MIMO Channel Sounder for Communication and Sensing
di: Sandra, Michiel, et al.
Pubblicazione: (2024)
di: Sandra, Michiel, et al.
Pubblicazione: (2024)
Exploring ChatGPT's Capabilities on Vulnerability Management
di: Liu, Peiyu, et al.
Pubblicazione: (2023)
di: Liu, Peiyu, et al.
Pubblicazione: (2023)
AhaRobot: A Low-Cost Open-Source Bimanual Mobile Manipulator for Embodied AI
di: Cui, Haiqin, et al.
Pubblicazione: (2025)
di: Cui, Haiqin, et al.
Pubblicazione: (2025)
SteerFlow: Steering Rectified Flows for Faithful Inversion-Based Image Editing
di: Dao, Thinh, et al.
Pubblicazione: (2026)
di: Dao, Thinh, et al.
Pubblicazione: (2026)
GeoSteer: Faithful Chain-of-Thought Steering via Latent Manifold Gradients
di: Kazama, Kentaro, et al.
Pubblicazione: (2026)
di: Kazama, Kentaro, et al.
Pubblicazione: (2026)
Evaluating Human Alignment and Model Faithfulness of LLM Rationale
di: Fayyaz, Mohsen, et al.
Pubblicazione: (2024)
di: Fayyaz, Mohsen, et al.
Pubblicazione: (2024)
Der Mediendiskurs zu Referenzgesellschaften und PISA: Ein Vergleich zwischen Deutschland und Festlandchina aus einer postkolonialen Perspektive
di: Ning, Haiqin
Pubblicazione: (2024)
di: Ning, Haiqin
Pubblicazione: (2024)
CIRR: Causal-Invariant Retrieval-Augmented Recommendation with Faithful Explanations under Distribution Shift
di: Sun, Sebastian
Pubblicazione: (2025)
di: Sun, Sebastian
Pubblicazione: (2025)
LANCET: Neural Intervention via Structural Entropy for Mitigating Faithfulness Hallucinations in LLMs
di: Wang, Chenxu, et al.
Pubblicazione: (2026)
di: Wang, Chenxu, et al.
Pubblicazione: (2026)
Documenti analoghi
-
Towards Steering without Sacrifice: Principled Training of Steering Vectors for Prompt-only Interventions
di: Bao, Yuntai, et al.
Pubblicazione: (2026) -
Probing the Geometry of Truth: Consistency and Generalization of Truth Directions in LLMs Across Logical Transformations and Question Answering Tasks
di: Bao, Yuntai, et al.
Pubblicazione: (2025) -
Scalable Multi-Stage Influence Function for Large Language Models via Eigenvalue-Corrected Kronecker-Factored Parameterization
di: Bao, Yuntai, et al.
Pubblicazione: (2025) -
PragLocker: Protecting Agent Intellectual Property in Untrusted Deployments via Non-Portable Prompts
di: Li, Qinfeng, et al.
Pubblicazione: (2026) -
Ground What You See: Hallucination-Resistant MLLMs via Caption Feedback, Diversity-Aware Sampling, and Conflict Regularization
di: Pan, Miao, et al.
Pubblicazione: (2026)