Enhancing LLM Robustness to Perturbed Instructions: An Empirical Study
Fuente:
arXiv
Salvato in:
| Autori principali: | Agrawal, Aryan, Alazraki, Lisa, Honarvar, Shahin, Rei, Marek |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2025
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
Meta-Reasoning Improves Tool Use in Large Language Models
di: Alazraki, Lisa, et al.
Pubblicazione: (2024)
di: Alazraki, Lisa, et al.
Pubblicazione: (2024)
AgentCoMa: A Compositional Benchmark Mixing Commonsense and Mathematical Reasoning in Real-World Scenarios
di: Alazraki, Lisa, et al.
Pubblicazione: (2025)
di: Alazraki, Lisa, et al.
Pubblicazione: (2025)
Improving the OOD Performance of Closed-Source LLMs on NLI Through Strategic Data Selection
di: Stacey, Joe, et al.
Pubblicazione: (2025)
di: Stacey, Joe, et al.
Pubblicazione: (2025)
Reverse Engineering Human Preferences with Reinforcement Learning
di: Alazraki, Lisa, et al.
Pubblicazione: (2025)
di: Alazraki, Lisa, et al.
Pubblicazione: (2025)
No Need for Explanations: LLMs can implicitly learn from mistakes in-context
di: Alazraki, Lisa, et al.
Pubblicazione: (2025)
di: Alazraki, Lisa, et al.
Pubblicazione: (2025)
StateAct: Enhancing LLM Base Agents via Self-prompting and State-tracking
di: Rozanov, Nikolai, et al.
Pubblicazione: (2024)
di: Rozanov, Nikolai, et al.
Pubblicazione: (2024)
Culturally Grounded Physical Commonsense Reasoning in Italian and English: A Submission to the MRL 2025 Shared Task
di: De Santis, Marco, et al.
Pubblicazione: (2025)
di: De Santis, Marco, et al.
Pubblicazione: (2025)
Distilling Robustness into Natural Language Inference Models with Domain-Targeted Augmentation
di: Stacey, Joe, et al.
Pubblicazione: (2023)
di: Stacey, Joe, et al.
Pubblicazione: (2023)
Fine-tuning with RAG for Improving LLM Learning of New Skills
di: Ibrahim, Humaid, et al.
Pubblicazione: (2025)
di: Ibrahim, Humaid, et al.
Pubblicazione: (2025)
Illuminate: A novel approach for depression detection with explainable analysis and proactive therapy using prompt engineering
di: Agrawal, Aryan
Pubblicazione: (2024)
di: Agrawal, Aryan
Pubblicazione: (2024)
DiffuseDef: Improved Robustness to Adversarial Attacks via Iterative Denoising
di: Li, Zhenhao, et al.
Pubblicazione: (2024)
di: Li, Zhenhao, et al.
Pubblicazione: (2024)
Is Preference Alignment Always the Best Option to Enhance LLM-Based Translation? An Empirical Analysis
di: Gisserot-Boukhlef, Hippolyte, et al.
Pubblicazione: (2024)
di: Gisserot-Boukhlef, Hippolyte, et al.
Pubblicazione: (2024)
Scaling Small Agents Through Strategy Auctions
di: Alazraki, Lisa, et al.
Pubblicazione: (2026)
di: Alazraki, Lisa, et al.
Pubblicazione: (2026)
Improving LLM Unlearning Robustness via Random Perturbations
di: Huu-Tien, Dang, et al.
Pubblicazione: (2025)
di: Huu-Tien, Dang, et al.
Pubblicazione: (2025)
More Than a Quick Glance: Overcoming the Greedy Bias in KV-Cache Compression
di: Sood, Aryan, et al.
Pubblicazione: (2026)
di: Sood, Aryan, et al.
Pubblicazione: (2026)
An Empirical Study of Conformal Prediction in LLM with ASP Scaffolds for Robust Reasoning
di: Kaur, Navdeep, et al.
Pubblicazione: (2025)
di: Kaur, Navdeep, et al.
Pubblicazione: (2025)
An Empathetic AI Coach for Self-Attachment Therapy
di: Alazraki, Lisa, et al.
Pubblicazione: (2022)
di: Alazraki, Lisa, et al.
Pubblicazione: (2022)
Can Automatic Metrics Assess High-Quality Translations?
di: Agrawal, Sweta, et al.
Pubblicazione: (2024)
di: Agrawal, Sweta, et al.
Pubblicazione: (2024)
Resource-Aware Arabic LLM Creation: Model Adaptation, Integration, and Multi-Domain Testing
di: Aryan, Prakash
Pubblicazione: (2024)
di: Aryan, Prakash
Pubblicazione: (2024)
Teach Me Sign: Stepwise Prompting LLM for Sign Language Production
di: An, Zhaoyi, et al.
Pubblicazione: (2025)
di: An, Zhaoyi, et al.
Pubblicazione: (2025)
DFKI-NLP at SemEval-2024 Task 2: Towards Robust LLMs Using Data Perturbations and MinMax Training
di: Verma, Bhuvanesh, et al.
Pubblicazione: (2024)
di: Verma, Bhuvanesh, et al.
Pubblicazione: (2024)
Towards Understanding the Robustness of LLM-based Evaluations under Perturbations
di: Chaudhary, Manav, et al.
Pubblicazione: (2024)
di: Chaudhary, Manav, et al.
Pubblicazione: (2024)
Parameter Efficient Instruction Tuning: An Empirical Study
di: He, Pengfei
Pubblicazione: (2024)
di: He, Pengfei
Pubblicazione: (2024)
Is Context Helpful for Chat Translation Evaluation?
di: Agrawal, Sweta, et al.
Pubblicazione: (2024)
di: Agrawal, Sweta, et al.
Pubblicazione: (2024)
EASYTOOL: Enhancing LLM-based Agents with Concise Tool Instruction
di: Yuan, Siyu, et al.
Pubblicazione: (2024)
di: Yuan, Siyu, et al.
Pubblicazione: (2024)
Can LLM-Generated Textual Explanations Enhance Model Classification Performance? An Empirical Study
di: Dhaini, Mahdi, et al.
Pubblicazione: (2025)
di: Dhaini, Mahdi, et al.
Pubblicazione: (2025)
An Empirical Study on the Robustness of Massively Multilingual Neural Machine Translation
di: Supryadi, et al.
Pubblicazione: (2024)
di: Supryadi, et al.
Pubblicazione: (2024)
Structure First, Reason Next: Enhancing a Large Language Model using Knowledge Graph for Numerical Reasoning in Financial Documents
di: Mishra, Aryan, et al.
Pubblicazione: (2026)
di: Mishra, Aryan, et al.
Pubblicazione: (2026)
Enhancing LLM Reasoning with Iterative DPO: A Comprehensive Empirical Investigation
di: Tu, Songjun, et al.
Pubblicazione: (2025)
di: Tu, Songjun, et al.
Pubblicazione: (2025)
LLMSR@XLLM25: An Empirical Study of LLM for Structural Reasoning
di: Li, Xinye, et al.
Pubblicazione: (2025)
di: Li, Xinye, et al.
Pubblicazione: (2025)
LLMRank: Understanding LLM Strengths for Model Routing
di: Agrawal, Shubham, et al.
Pubblicazione: (2025)
di: Agrawal, Shubham, et al.
Pubblicazione: (2025)
Robust Biomedical Publication Type and Study Design Classification with Knowledge-Guided Perturbations
di: Ming, Shufan, et al.
Pubblicazione: (2026)
di: Ming, Shufan, et al.
Pubblicazione: (2026)
xTower: A Multilingual LLM for Explaining and Correcting Translation Errors
di: Treviso, Marcos, et al.
Pubblicazione: (2024)
di: Treviso, Marcos, et al.
Pubblicazione: (2024)
Effective Context Selection in LLM-based Leaderboard Generation: An Empirical Study
di: Kabongo, Salomon, et al.
Pubblicazione: (2024)
di: Kabongo, Salomon, et al.
Pubblicazione: (2024)
Instructional Segment Embedding: Improving LLM Safety with Instruction Hierarchy
di: Wu, Tong, et al.
Pubblicazione: (2024)
di: Wu, Tong, et al.
Pubblicazione: (2024)
Evaluating Robustness of Large Audio Language Models to Audio Injection: An Empirical Study
di: Hou, Guanyu, et al.
Pubblicazione: (2025)
di: Hou, Guanyu, et al.
Pubblicazione: (2025)
Instruction Finetuning for Leaderboard Generation from Empirical AI Research
di: Kabongo, Salomon, et al.
Pubblicazione: (2024)
di: Kabongo, Salomon, et al.
Pubblicazione: (2024)
Self-Review Framework for Enhancing Instruction Following Capability of LLM
di: Park, Sihyun
Pubblicazione: (2025)
di: Park, Sihyun
Pubblicazione: (2025)
Assessing Adversarial Robustness of Large Language Models: An Empirical Study
di: Yang, Zeyu, et al.
Pubblicazione: (2024)
di: Yang, Zeyu, et al.
Pubblicazione: (2024)
Taming Sensitive Weights : Noise Perturbation Fine-tuning for Robust LLM Quantization
di: Wang, Dongwei, et al.
Pubblicazione: (2024)
di: Wang, Dongwei, et al.
Pubblicazione: (2024)
Documenti analoghi
-
Meta-Reasoning Improves Tool Use in Large Language Models
di: Alazraki, Lisa, et al.
Pubblicazione: (2024) -
AgentCoMa: A Compositional Benchmark Mixing Commonsense and Mathematical Reasoning in Real-World Scenarios
di: Alazraki, Lisa, et al.
Pubblicazione: (2025) -
Improving the OOD Performance of Closed-Source LLMs on NLI Through Strategic Data Selection
di: Stacey, Joe, et al.
Pubblicazione: (2025) -
Reverse Engineering Human Preferences with Reinforcement Learning
di: Alazraki, Lisa, et al.
Pubblicazione: (2025) -
No Need for Explanations: LLMs can implicitly learn from mistakes in-context
di: Alazraki, Lisa, et al.
Pubblicazione: (2025)