SelfPrompt: Autonomously Evaluating LLM Robustness via Domain-Constrained Knowledge Guidelines and Refined Adversarial Prompts
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Pei, Aihua, Yang, Zehua, Zhu, Shunan, Cheng, Ruoxi, Jia, Ju |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2024
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
KGPA: Robustness Evaluation for Large Language Models via Cross-Domain Knowledge Graphs
von: Pei, Aihua, et al.
Veröffentlicht: (2024)
von: Pei, Aihua, et al.
Veröffentlicht: (2024)
PromptRobust: Towards Evaluating the Robustness of Large Language Models on Adversarial Prompts
von: Zhu, Kaijie, et al.
Veröffentlicht: (2023)
von: Zhu, Kaijie, et al.
Veröffentlicht: (2023)
SelfPrompt: Confidence-Aware Semi-Supervised Tuning for Robust Vision-Language Model Adaptation
von: Roy, Shuvendu, et al.
Veröffentlicht: (2025)
von: Roy, Shuvendu, et al.
Veröffentlicht: (2025)
Prompt Engineering: How Prompt Vocabulary affects Domain Knowledge
von: Schreiter, Dimitri
Veröffentlicht: (2025)
von: Schreiter, Dimitri
Veröffentlicht: (2025)
HPSS: Heuristic Prompting Strategy Search for LLM Evaluators
von: Wen, Bosi, et al.
Veröffentlicht: (2025)
von: Wen, Bosi, et al.
Veröffentlicht: (2025)
Beyond Prompt Content: Enhancing LLM Performance via Content-Format Integrated Prompt Optimization
von: Liu, Yuanye, et al.
Veröffentlicht: (2025)
von: Liu, Yuanye, et al.
Veröffentlicht: (2025)
Enhancing Medical Dialogue Generation through Knowledge Refinement and Dynamic Prompt Adjustment
von: Sun, Hongda, et al.
Veröffentlicht: (2025)
von: Sun, Hongda, et al.
Veröffentlicht: (2025)
Prompt Chaining or Stepwise Prompt? Refinement in Text Summarization
von: Sun, Shichao, et al.
Veröffentlicht: (2024)
von: Sun, Shichao, et al.
Veröffentlicht: (2024)
A Better LLM Evaluator for Text Generation: The Impact of Prompt Output Sequencing and Optimization
von: Chu, KuanChao, et al.
Veröffentlicht: (2024)
von: Chu, KuanChao, et al.
Veröffentlicht: (2024)
Understanding Learner-LLM Chatbot Interactions and the Impact of Prompting Guidelines
von: Koyuturk, Cansu, et al.
Veröffentlicht: (2025)
von: Koyuturk, Cansu, et al.
Veröffentlicht: (2025)
SafeReview: Defending LLM-based Review Systems Against Adversarial Hidden Prompts
von: Xin, Yuan, et al.
Veröffentlicht: (2026)
von: Xin, Yuan, et al.
Veröffentlicht: (2026)
Towards Self-Robust LLMs: Intrinsic Prompt Noise Resistance via CoIPO
von: Yang, Xin, et al.
Veröffentlicht: (2026)
von: Yang, Xin, et al.
Veröffentlicht: (2026)
PromptFix: Few-shot Backdoor Removal via Adversarial Prompt Tuning
von: Zhang, Tianrong, et al.
Veröffentlicht: (2024)
von: Zhang, Tianrong, et al.
Veröffentlicht: (2024)
RulePrompt: Weakly Supervised Text Classification with Prompting PLMs and Self-Iterative Logical Rules
von: Li, Miaomiao, et al.
Veröffentlicht: (2024)
von: Li, Miaomiao, et al.
Veröffentlicht: (2024)
Certifying LLM Safety against Adversarial Prompting
von: Kumar, Aounon, et al.
Veröffentlicht: (2023)
von: Kumar, Aounon, et al.
Veröffentlicht: (2023)
Prompt-Reverse Inconsistency: LLM Self-Inconsistency Beyond Generative Randomness and Prompt Paraphrasing
von: Ahn, Jihyun Janice, et al.
Veröffentlicht: (2025)
von: Ahn, Jihyun Janice, et al.
Veröffentlicht: (2025)
Self-Prompt Tuning: Enable Autonomous Role-Playing in LLMs
von: Kong, Aobo, et al.
Veröffentlicht: (2024)
von: Kong, Aobo, et al.
Veröffentlicht: (2024)
Prompt Optimization via Adversarial In-Context Learning
von: Do, Xuan Long, et al.
Veröffentlicht: (2023)
von: Do, Xuan Long, et al.
Veröffentlicht: (2023)
SciPrompt: Knowledge-augmented Prompting for Fine-grained Categorization of Scientific Topics
von: You, Zhiwen, et al.
Veröffentlicht: (2024)
von: You, Zhiwen, et al.
Veröffentlicht: (2024)
Can Prompts Rewind Time for LLMs? Evaluating the Effectiveness of Prompted Knowledge Cutoffs
von: Gao, Xin, et al.
Veröffentlicht: (2025)
von: Gao, Xin, et al.
Veröffentlicht: (2025)
LLM Prompt Evaluation for Educational Applications
von: Holmes, Langdon, et al.
Veröffentlicht: (2026)
von: Holmes, Langdon, et al.
Veröffentlicht: (2026)
RASPRef: Retrieval-Augmented Self-Supervised Prompt Refinement for Large Reasoning Models
von: Soni, Rahul
Veröffentlicht: (2026)
von: Soni, Rahul
Veröffentlicht: (2026)
Refining and Reusing Annotation Guidelines for LLM Annotation
von: Kim, Kon Woo, et al.
Veröffentlicht: (2026)
von: Kim, Kon Woo, et al.
Veröffentlicht: (2026)
Evaluating Knowledge Generation and Self-Refinement Strategies for LLM-based Column Type Annotation
von: Korini, Keti, et al.
Veröffentlicht: (2025)
von: Korini, Keti, et al.
Veröffentlicht: (2025)
GuideBench: Benchmarking Domain-Oriented Guideline Following for LLM Agents
von: Diao, Lingxiao, et al.
Veröffentlicht: (2025)
von: Diao, Lingxiao, et al.
Veröffentlicht: (2025)
PromptRad: Knowledge-Enhanced Multi-Label Prompt-Tuning for Low-Resource Radiology Report Labeling
von: Lin, Ying-Jia, et al.
Veröffentlicht: (2026)
von: Lin, Ying-Jia, et al.
Veröffentlicht: (2026)
CPR: Mitigating Large Language Model Hallucinations with Curative Prompt Refinement
von: Shim, Jung-Woo, et al.
Veröffentlicht: (2025)
von: Shim, Jung-Woo, et al.
Veröffentlicht: (2025)
Multi-stage Prompt Refinement for Mitigating Hallucinations in Large Language Models
von: Shim, Jung-Woo, et al.
Veröffentlicht: (2025)
von: Shim, Jung-Woo, et al.
Veröffentlicht: (2025)
Knowledge Restoration-driven Prompt Optimization: Unlocking LLM Potential for Open-Domain Relational Triplet Extraction
von: Jing, Xiaonan, et al.
Veröffentlicht: (2026)
von: Jing, Xiaonan, et al.
Veröffentlicht: (2026)
PromptKD: Distilling Student-Friendly Knowledge for Generative Language Models via Prompt Tuning
von: Kim, Gyeongman, et al.
Veröffentlicht: (2024)
von: Kim, Gyeongman, et al.
Veröffentlicht: (2024)
PromptOptMe: Error-Aware Prompt Compression for LLM-based MT Evaluation Metrics
von: Larionov, Daniil, et al.
Veröffentlicht: (2024)
von: Larionov, Daniil, et al.
Veröffentlicht: (2024)
A Taxonomy of Prompt Defects in LLM Systems
von: Tian, Haoye, et al.
Veröffentlicht: (2025)
von: Tian, Haoye, et al.
Veröffentlicht: (2025)
CPJ: Explainable Agricultural Pest Diagnosis via Caption-Prompt-Judge with LLM-Judged Refinement
von: Zhang, Wentao, et al.
Veröffentlicht: (2025)
von: Zhang, Wentao, et al.
Veröffentlicht: (2025)
Efficient and Stealthy Jailbreak Attacks via Adversarial Prompt Distillation from LLMs to SLMs
von: Li, Xiang, et al.
Veröffentlicht: (2025)
von: Li, Xiang, et al.
Veröffentlicht: (2025)
LLARS: Enabling Domain Expert & Developer Collaboration for LLM Prompting, Generation and Evaluation
von: Steigerwald, Philipp, et al.
Veröffentlicht: (2026)
von: Steigerwald, Philipp, et al.
Veröffentlicht: (2026)
Enhancing LLM Factual Accuracy with RAG to Counter Hallucinations: A Case Study on Domain-Specific Queries in Private Knowledge-Bases
von: Li, Jiarui, et al.
Veröffentlicht: (2024)
von: Li, Jiarui, et al.
Veröffentlicht: (2024)
Are All Prompt Components Value-Neutral? Understanding the Heterogeneous Adversarial Robustness of Dissected Prompt in Large Language Models
von: Zheng, Yujia, et al.
Veröffentlicht: (2025)
von: Zheng, Yujia, et al.
Veröffentlicht: (2025)
GRL-Prompt: Towards Knowledge Graph based Prompt Optimization via Reinforcement Learning
von: Liu, Yuze, et al.
Veröffentlicht: (2024)
von: Liu, Yuze, et al.
Veröffentlicht: (2024)
PISanitizer: Preventing Prompt Injection to Long-Context LLMs via Prompt Sanitization
von: Geng, Runpeng, et al.
Veröffentlicht: (2025)
von: Geng, Runpeng, et al.
Veröffentlicht: (2025)
An Enhanced Prompt-Based LLM Reasoning Scheme via Knowledge Graph-Integrated Collaboration
von: Li, Yihao, et al.
Veröffentlicht: (2024)
von: Li, Yihao, et al.
Veröffentlicht: (2024)
Ähnliche Einträge
-
KGPA: Robustness Evaluation for Large Language Models via Cross-Domain Knowledge Graphs
von: Pei, Aihua, et al.
Veröffentlicht: (2024) -
PromptRobust: Towards Evaluating the Robustness of Large Language Models on Adversarial Prompts
von: Zhu, Kaijie, et al.
Veröffentlicht: (2023) -
SelfPrompt: Confidence-Aware Semi-Supervised Tuning for Robust Vision-Language Model Adaptation
von: Roy, Shuvendu, et al.
Veröffentlicht: (2025) -
Prompt Engineering: How Prompt Vocabulary affects Domain Knowledge
von: Schreiter, Dimitri
Veröffentlicht: (2025) -
HPSS: Heuristic Prompting Strategy Search for LLM Evaluators
von: Wen, Bosi, et al.
Veröffentlicht: (2025)