Gespeichert in:
| Hauptverfasser: | Li, Zongjie, Wang, Chaozheng, Ma, Pingchuan, Wu, Daoyuan, Wang, Shuai, Gao, Cuiyun, Liu, Yang |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2023
|
| Schlagworte: | |
| Online-Zugang: | https://arxiv.org/abs/2310.01432 |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Testing and Understanding Erroneous Planning in LLM Agents through Synthesized User Inputs
von: Ji, Zhenlan, et al.
Veröffentlicht: (2024)
von: Ji, Zhenlan, et al.
Veröffentlicht: (2024)
STShield: Single-Token Sentinel for Real-Time Jailbreak Detection in Large Language Models
von: Wang, Xunguang, et al.
Veröffentlicht: (2025)
von: Wang, Xunguang, et al.
Veröffentlicht: (2025)
WARBENCH: A Comprehensive Benchmark for Evaluating LLMs in Military Decision-Making
von: Li, Zongjie, et al.
Veröffentlicht: (2026)
von: Li, Zongjie, et al.
Veröffentlicht: (2026)
Taxonomy, Evaluation and Exploitation of IPI-Centric LLM Agent Defense Frameworks
von: Ji, Zimo, et al.
Veröffentlicht: (2025)
von: Ji, Zimo, et al.
Veröffentlicht: (2025)
Measuring and Augmenting Large Language Models for Solving Capture-the-Flag Challenges
von: Ji, Zimo, et al.
Veröffentlicht: (2025)
von: Ji, Zimo, et al.
Veröffentlicht: (2025)
IP Leakage Attacks Targeting LLM-Based Multi-Agent Systems
von: Wang, Liwen, et al.
Veröffentlicht: (2025)
von: Wang, Liwen, et al.
Veröffentlicht: (2025)
The Prompt Alchemist: Automated LLM-Tailored Prompt Optimization for Test Case Generation
von: Gao, Shuzheng, et al.
Veröffentlicht: (2025)
von: Gao, Shuzheng, et al.
Veröffentlicht: (2025)
Digging Into the Internal: Causality-Based Analysis of LLM Function Calling
von: Ji, Zhenlan, et al.
Veröffentlicht: (2025)
von: Ji, Zhenlan, et al.
Veröffentlicht: (2025)
An Empirical Study on Large Language Models in Accuracy and Robustness under Chinese Industrial Scenarios
von: Li, Zongjie, et al.
Veröffentlicht: (2024)
von: Li, Zongjie, et al.
Veröffentlicht: (2024)
SelfDefend: LLMs Can Defend Themselves against Jailbreaking in a Practical Manner
von: Wang, Xunguang, et al.
Veröffentlicht: (2024)
von: Wang, Xunguang, et al.
Veröffentlicht: (2024)
Differentiation-Based Extraction of Proprietary Data from Fine-Tuned LLMs
von: Li, Zongjie, et al.
Veröffentlicht: (2025)
von: Li, Zongjie, et al.
Veröffentlicht: (2025)
API-guided Dataset Synthesis to Finetune Large Code Models
von: Li, Zongjie, et al.
Veröffentlicht: (2024)
von: Li, Zongjie, et al.
Veröffentlicht: (2024)
SoK: Evaluating Jailbreak Guardrails for Large Language Models
von: Wang, Xunguang, et al.
Veröffentlicht: (2025)
von: Wang, Xunguang, et al.
Veröffentlicht: (2025)
GuidedBench: Measuring and Mitigating the Evaluation Discrepancies of In-the-wild LLM Jailbreak Methods
von: Huang, Ruixuan, et al.
Veröffentlicht: (2025)
von: Huang, Ruixuan, et al.
Veröffentlicht: (2025)
CodeVisionary: An Agent-based Framework for Evaluating Large Language Models in Code Generation
von: Wang, Xinchen, et al.
Veröffentlicht: (2025)
von: Wang, Xinchen, et al.
Veröffentlicht: (2025)
Optimal Brain Iterative Merging: Mitigating Interference in LLM Merging
von: Wang, Zhixiang, et al.
Veröffentlicht: (2025)
von: Wang, Zhixiang, et al.
Veröffentlicht: (2025)
Beyond Content Safety: Real-Time Monitoring for Reasoning Vulnerabilities in Large Language Models
von: Wang, Xunguang, et al.
Veröffentlicht: (2026)
von: Wang, Xunguang, et al.
Veröffentlicht: (2026)
Extrapolation Merging: Keep Improving With Extrapolation and Merging
von: Lin, Yiguan, et al.
Veröffentlicht: (2025)
von: Lin, Yiguan, et al.
Veröffentlicht: (2025)
Empirical Study of Code Large Language Models for Binary Security Patch Detection
von: Li, Qingyuan, et al.
Veröffentlicht: (2025)
von: Li, Qingyuan, et al.
Veröffentlicht: (2025)
Reasoning as a Resource: Optimizing Fast and Slow Thinking in Code Generation Models
von: Li, Zongjie, et al.
Veröffentlicht: (2025)
von: Li, Zongjie, et al.
Veröffentlicht: (2025)
Probing Association Biases in LLM Moderation Over-Sensitivity
von: Wang, Yuxin, et al.
Veröffentlicht: (2025)
von: Wang, Yuxin, et al.
Veröffentlicht: (2025)
HALF: Harm-Aware LLM Fairness Evaluation Aligned with Deployment
von: Mekky, Ali, et al.
Veröffentlicht: (2025)
von: Mekky, Ali, et al.
Veröffentlicht: (2025)
Gender and Positional Biases in LLM-Based Hiring Decisions: Evidence from Comparative CV/Résumé Evaluations
von: Rozado, David
Veröffentlicht: (2025)
von: Rozado, David
Veröffentlicht: (2025)
Justice or Prejudice? Quantifying Biases in LLM-as-a-Judge
von: Ye, Jiayi, et al.
Veröffentlicht: (2024)
von: Ye, Jiayi, et al.
Veröffentlicht: (2024)
SPVR: syntax-to-prompt vulnerability repair based on large language models
von: Wang, Ruoke, et al.
Veröffentlicht: (2024)
von: Wang, Ruoke, et al.
Veröffentlicht: (2024)
Grounded in Reality: Learning and Deploying Proactive LLM from Offline Logs
von: Wei, Fei, et al.
Veröffentlicht: (2025)
von: Wei, Fei, et al.
Veröffentlicht: (2025)
Taming Various Privilege Escalation in LLM-Based Agent Systems: A Mandatory Access Control Framework
von: Ji, Zimo, et al.
Veröffentlicht: (2026)
von: Ji, Zimo, et al.
Veröffentlicht: (2026)
Knowledge-Infused Legal Wisdom: Navigating LLM Consultation through the Lens of Diagnostics and Positive-Unlabeled Reinforcement Learning
von: Wu, Yang, et al.
Veröffentlicht: (2024)
von: Wu, Yang, et al.
Veröffentlicht: (2024)
From Biased Chatbots to Biased Agents: Examining Role Assignment Effects on LLM Agent Robustness
von: Cao, Linbo, et al.
Veröffentlicht: (2026)
von: Cao, Linbo, et al.
Veröffentlicht: (2026)
Reason-Align-Respond: Aligning LLM Reasoning with Knowledge Graphs for KGQA
von: Shen, Xiangqing, et al.
Veröffentlicht: (2025)
von: Shen, Xiangqing, et al.
Veröffentlicht: (2025)
BoRP: Bootstrapped Regression Probing for Scalable and Human-Aligned LLM Evaluation
von: Sun, Peng, et al.
Veröffentlicht: (2026)
von: Sun, Peng, et al.
Veröffentlicht: (2026)
Training-free LLM Merging for Multi-task Learning
von: Fu, Zichuan, et al.
Veröffentlicht: (2025)
von: Fu, Zichuan, et al.
Veröffentlicht: (2025)
LLMs Can Defend Themselves Against Jailbreaking in a Practical Manner: A Vision Paper
von: Wu, Daoyuan, et al.
Veröffentlicht: (2024)
von: Wu, Daoyuan, et al.
Veröffentlicht: (2024)
LPFQA: A Long-Tail Professional Forum-based Benchmark for LLM Evaluation
von: Zhu, Liya, et al.
Veröffentlicht: (2025)
von: Zhu, Liya, et al.
Veröffentlicht: (2025)
MRScore: Evaluating Radiology Report Generation with LLM-based Reward System
von: Liu, Yunyi, et al.
Veröffentlicht: (2024)
von: Liu, Yunyi, et al.
Veröffentlicht: (2024)
LLMs Judge Themselves: A Game-Theoretic Framework for Human-Aligned Evaluation
von: Yang, Gao, et al.
Veröffentlicht: (2025)
von: Yang, Gao, et al.
Veröffentlicht: (2025)
Citation-Enhanced Generation for LLM-based Chatbots
von: Li, Weitao, et al.
Veröffentlicht: (2024)
von: Li, Weitao, et al.
Veröffentlicht: (2024)
Is Chain-of-Thought Reasoning of LLMs a Mirage? A Data Distribution Lens
von: Zhao, Chengshuai, et al.
Veröffentlicht: (2025)
von: Zhao, Chengshuai, et al.
Veröffentlicht: (2025)
Pruning via Merging: Compressing LLMs via Manifold Alignment Based Layer Merging
von: Liu, Deyuan, et al.
Veröffentlicht: (2024)
von: Liu, Deyuan, et al.
Veröffentlicht: (2024)
Search-Based LLMs for Code Optimization
von: Gao, Shuzheng, et al.
Veröffentlicht: (2024)
von: Gao, Shuzheng, et al.
Veröffentlicht: (2024)
Ähnliche Einträge
-
Testing and Understanding Erroneous Planning in LLM Agents through Synthesized User Inputs
von: Ji, Zhenlan, et al.
Veröffentlicht: (2024) -
STShield: Single-Token Sentinel for Real-Time Jailbreak Detection in Large Language Models
von: Wang, Xunguang, et al.
Veröffentlicht: (2025) -
WARBENCH: A Comprehensive Benchmark for Evaluating LLMs in Military Decision-Making
von: Li, Zongjie, et al.
Veröffentlicht: (2026) -
Taxonomy, Evaluation and Exploitation of IPI-Centric LLM Agent Defense Frameworks
von: Ji, Zimo, et al.
Veröffentlicht: (2025) -
Measuring and Augmenting Large Language Models for Solving Capture-the-Flag Challenges
von: Ji, Zimo, et al.
Veröffentlicht: (2025)