Advancing the Robustness of Large Language Models through Self-Denoised Smoothing
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Ji, Jiabao, Hou, Bairu, Zhang, Zhen, Zhang, Guanhua, Fan, Wenqi, Li, Qing, Zhang, Yang, Liu, Gaowen, Liu, Sijia, Chang, Shiyu |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2024
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Defending Large Language Models against Jailbreak Attacks via Semantic Smoothing
von: Ji, Jiabao, et al.
Veröffentlicht: (2024)
von: Ji, Jiabao, et al.
Veröffentlicht: (2024)
Decomposing Uncertainty for Large Language Models through Input Clarification Ensembling
von: Hou, Bairu, et al.
Veröffentlicht: (2023)
von: Hou, Bairu, et al.
Veröffentlicht: (2023)
ThinkPrune: Pruning Long Chain-of-Thought of LLMs via Reinforcement Learning
von: Hou, Bairu, et al.
Veröffentlicht: (2025)
von: Hou, Bairu, et al.
Veröffentlicht: (2025)
Reversing the Forget-Retain Objectives: An Efficient LLM Unlearning Framework from Logit Difference
von: Ji, Jiabao, et al.
Veröffentlicht: (2024)
von: Ji, Jiabao, et al.
Veröffentlicht: (2024)
A Probabilistic Framework for LLM Hallucination Detection via Belief Tree Propagation
von: Hou, Bairu, et al.
Veröffentlicht: (2024)
von: Hou, Bairu, et al.
Veröffentlicht: (2024)
KVLink: Accelerating Large Language Models via Efficient KV Cache Reuse
von: Yang, Jingbo, et al.
Veröffentlicht: (2025)
von: Yang, Jingbo, et al.
Veröffentlicht: (2025)
How Well Do Agentic Skills Work in the Wild: Benchmarking LLM Skill Usage in Realistic Settings
von: Liu, Yujian, et al.
Veröffentlicht: (2026)
von: Liu, Yujian, et al.
Veröffentlicht: (2026)
Instruction-Following Pruning for Large Language Models
von: Hou, Bairu, et al.
Veröffentlicht: (2025)
von: Hou, Bairu, et al.
Veröffentlicht: (2025)
HarnessLLM: Automatic Testing Harness Generation via Reinforcement Learning
von: Liu, Yujian, et al.
Veröffentlicht: (2025)
von: Liu, Yujian, et al.
Veröffentlicht: (2025)
Advancing Large Language Model Attribution through Self-Improving
von: Huang, Lei, et al.
Veröffentlicht: (2024)
von: Huang, Lei, et al.
Veröffentlicht: (2024)
Navigating the Clutter: Waypoint-Based Bi-Level Planning for Multi-Robot Systems
von: Ji, Jiabao, et al.
Veröffentlicht: (2026)
von: Ji, Jiabao, et al.
Veröffentlicht: (2026)
Collision- and Reachability-Aware Multi-Robot Control with Grounded LLM Planners
von: Ji, Jiabao, et al.
Veröffentlicht: (2025)
von: Ji, Jiabao, et al.
Veröffentlicht: (2025)
CompliBench: Benchmarking LLM Judges for Compliance Violation Detection in Dialogue Systems
von: Yang, Jingbo, et al.
Veröffentlicht: (2026)
von: Yang, Jingbo, et al.
Veröffentlicht: (2026)
SEUF: Is Unlearning One Expert Enough for Mixture-of-Experts LLMs?
von: Zhuang, Haomin, et al.
Veröffentlicht: (2024)
von: Zhuang, Haomin, et al.
Veröffentlicht: (2024)
Augment before You Try: Knowledge-Enhanced Table Question Answering via Table Expansion
von: Liu, Yujian, et al.
Veröffentlicht: (2024)
von: Liu, Yujian, et al.
Veröffentlicht: (2024)
DynaThink: Fast or Slow? A Dynamic Decision-Making Framework for Large Language Models
von: Pan, Jiabao, et al.
Veröffentlicht: (2024)
von: Pan, Jiabao, et al.
Veröffentlicht: (2024)
Large Language Models are In-Context Molecule Learners
von: Li, Jiatong, et al.
Veröffentlicht: (2024)
von: Li, Jiatong, et al.
Veröffentlicht: (2024)
DEEPAMBIGQA: Ambiguous Multi-hop Questions for Benchmarking LLM Answer Completeness
von: Ji, Jiabao, et al.
Veröffentlicht: (2025)
von: Ji, Jiabao, et al.
Veröffentlicht: (2025)
Seer Self-Consistency: Advance Budget Estimation for Adaptive Test-Time Scaling
von: Ji, Shiyu, et al.
Veröffentlicht: (2025)
von: Ji, Shiyu, et al.
Veröffentlicht: (2025)
Advancing Process Verification for Large Language Models via Tree-Based Preference Learning
von: He, Mingqian, et al.
Veröffentlicht: (2024)
von: He, Mingqian, et al.
Veröffentlicht: (2024)
Advancing Tool-Augmented Large Language Models: Integrating Insights from Errors in Inference Trees
von: Chen, Sijia, et al.
Veröffentlicht: (2024)
von: Chen, Sijia, et al.
Veröffentlicht: (2024)
Enhancing Multilingual Capabilities of Large Language Models through Self-Distillation from Resource-Rich Languages
von: Zhang, Yuanchi, et al.
Veröffentlicht: (2024)
von: Zhang, Yuanchi, et al.
Veröffentlicht: (2024)
Continual Learning Using Only Large Language Model Prompting
von: Qiu, Jiabao, et al.
Veröffentlicht: (2024)
von: Qiu, Jiabao, et al.
Veröffentlicht: (2024)
Fictitious Synthetic Data Can Improve LLM Factuality via Prerequisite Learning
von: Liu, Yujian, et al.
Veröffentlicht: (2024)
von: Liu, Yujian, et al.
Veröffentlicht: (2024)
Revisiting Who's Harry Potter: Towards Targeted Unlearning from a Causal Intervention Perspective
von: Liu, Yujian, et al.
Veröffentlicht: (2024)
von: Liu, Yujian, et al.
Veröffentlicht: (2024)
TimeToM: Temporal Space is the Key to Unlocking the Door of Large Language Models' Theory-of-Mind
von: Hou, Guiyang, et al.
Veröffentlicht: (2024)
von: Hou, Guiyang, et al.
Veröffentlicht: (2024)
Self-Guided Function Calling in Large Language Models via Stepwise Experience Recall
von: Cui, Sijia, et al.
Veröffentlicht: (2025)
von: Cui, Sijia, et al.
Veröffentlicht: (2025)
Attention Reveals More Than Tokens: Training-Free Long-Context Reasoning with Attention-guided Retrieval
von: Zhang, Yuwei, et al.
Veröffentlicht: (2025)
von: Zhang, Yuwei, et al.
Veröffentlicht: (2025)
DRPruning: Efficient Large Language Model Pruning through Distributionally Robust Optimization
von: Deng, Hexuan, et al.
Veröffentlicht: (2024)
von: Deng, Hexuan, et al.
Veröffentlicht: (2024)
PID Control-Based Self-Healing to Improve the Robustness of Large Language Models
von: Chen, Zhuotong, et al.
Veröffentlicht: (2024)
von: Chen, Zhuotong, et al.
Veröffentlicht: (2024)
Evaluating Robustness of Large Audio Language Models to Audio Injection: An Empirical Study
von: Hou, Guanyu, et al.
Veröffentlicht: (2025)
von: Hou, Guanyu, et al.
Veröffentlicht: (2025)
Empowering Molecule Discovery for Molecule-Caption Translation with Large Language Models: A ChatGPT Perspective
von: Li, Jiatong, et al.
Veröffentlicht: (2023)
von: Li, Jiatong, et al.
Veröffentlicht: (2023)
Better Language Model-Based Judging Reward Modeling through Scaling Comprehension Boundaries
von: Ning, Meiling, et al.
Veröffentlicht: (2025)
von: Ning, Meiling, et al.
Veröffentlicht: (2025)
Towards Lightweight, Adaptive and Attribute-Aware Multi-Aspect Controllable Text Generation with Large Language Models
von: Zhu, Chenyu, et al.
Veröffentlicht: (2025)
von: Zhu, Chenyu, et al.
Veröffentlicht: (2025)
Self-Evolutionary Large Language Models through Uncertainty-Enhanced Preference Optimization
von: Wang, Jianing, et al.
Veröffentlicht: (2024)
von: Wang, Jianing, et al.
Veröffentlicht: (2024)
DiffAgent: Fast and Accurate Text-to-Image API Selection with Large Language Model
von: Zhao, Lirui, et al.
Veröffentlicht: (2024)
von: Zhao, Lirui, et al.
Veröffentlicht: (2024)
TasTe: Teaching Large Language Models to Translate through Self-Reflection
von: Wang, Yutong, et al.
Veröffentlicht: (2024)
von: Wang, Yutong, et al.
Veröffentlicht: (2024)
WebSeer: Training Deeper Search Agents through Reinforcement Learning with Self-Reflection
von: He, Guanzhong, et al.
Veröffentlicht: (2025)
von: He, Guanzhong, et al.
Veröffentlicht: (2025)
ChartAdapter: Large Vision-Language Model for Chart Summarization
von: Xu, Peixin, et al.
Veröffentlicht: (2024)
von: Xu, Peixin, et al.
Veröffentlicht: (2024)
SCAN: Self-Denoising Monte Carlo Annotation for Robust Process Reward Learning
von: Ding, Yuyang, et al.
Veröffentlicht: (2025)
von: Ding, Yuyang, et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
Defending Large Language Models against Jailbreak Attacks via Semantic Smoothing
von: Ji, Jiabao, et al.
Veröffentlicht: (2024) -
Decomposing Uncertainty for Large Language Models through Input Clarification Ensembling
von: Hou, Bairu, et al.
Veröffentlicht: (2023) -
ThinkPrune: Pruning Long Chain-of-Thought of LLMs via Reinforcement Learning
von: Hou, Bairu, et al.
Veröffentlicht: (2025) -
Reversing the Forget-Retain Objectives: An Efficient LLM Unlearning Framework from Logit Difference
von: Ji, Jiabao, et al.
Veröffentlicht: (2024) -
A Probabilistic Framework for LLM Hallucination Detection via Belief Tree Propagation
von: Hou, Bairu, et al.
Veröffentlicht: (2024)