Guardado en:
| Autores principales: | Ji, Zhenlan, Wu, Daoyuan, Ma, Pingchuan, Li, Zongjie, Wang, Shuai |
|---|---|
| Formato: | Preprint |
| Publicado: |
2024
|
| Materias: | |
| Acceso en línea: | https://arxiv.org/abs/2404.17833 |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
STShield: Single-Token Sentinel for Real-Time Jailbreak Detection in Large Language Models
por: Wang, Xunguang, et al.
Publicado: (2025)
por: Wang, Xunguang, et al.
Publicado: (2025)
IP Leakage Attacks Targeting LLM-Based Multi-Agent Systems
por: Wang, Liwen, et al.
Publicado: (2025)
por: Wang, Liwen, et al.
Publicado: (2025)
Split and Merge: Aligning Position Biases in LLM-based Evaluators
por: Li, Zongjie, et al.
Publicado: (2023)
por: Li, Zongjie, et al.
Publicado: (2023)
Digging Into the Internal: Causality-Based Analysis of LLM Function Calling
por: Ji, Zhenlan, et al.
Publicado: (2025)
por: Ji, Zhenlan, et al.
Publicado: (2025)
Measuring and Augmenting Large Language Models for Solving Capture-the-Flag Challenges
por: Ji, Zimo, et al.
Publicado: (2025)
por: Ji, Zimo, et al.
Publicado: (2025)
SoK: Evaluating Jailbreak Guardrails for Large Language Models
por: Wang, Xunguang, et al.
Publicado: (2025)
por: Wang, Xunguang, et al.
Publicado: (2025)
SelfDefend: LLMs Can Defend Themselves against Jailbreaking in a Practical Manner
por: Wang, Xunguang, et al.
Publicado: (2024)
por: Wang, Xunguang, et al.
Publicado: (2024)
Taxonomy, Evaluation and Exploitation of IPI-Centric LLM Agent Defense Frameworks
por: Ji, Zimo, et al.
Publicado: (2025)
por: Ji, Zimo, et al.
Publicado: (2025)
Beyond Content Safety: Real-Time Monitoring for Reasoning Vulnerabilities in Large Language Models
por: Wang, Xunguang, et al.
Publicado: (2026)
por: Wang, Xunguang, et al.
Publicado: (2026)
An Empirical Study on Large Language Models in Accuracy and Robustness under Chinese Industrial Scenarios
por: Li, Zongjie, et al.
Publicado: (2024)
por: Li, Zongjie, et al.
Publicado: (2024)
Evaluating LLMs on Sequential API Call Through Automated Test Generation
por: Huang, Yuheng, et al.
Publicado: (2025)
por: Huang, Yuheng, et al.
Publicado: (2025)
API-guided Dataset Synthesis to Finetune Large Code Models
por: Li, Zongjie, et al.
Publicado: (2024)
por: Li, Zongjie, et al.
Publicado: (2024)
Differentiation-Based Extraction of Proprietary Data from Fine-Tuned LLMs
por: Li, Zongjie, et al.
Publicado: (2025)
por: Li, Zongjie, et al.
Publicado: (2025)
Understanding and Bridging the Planner-Coder Gap: A Systematic Study on the Robustness of Multi-Agent Systems for Code Generation
por: Lyu, Zongyi, et al.
Publicado: (2025)
por: Lyu, Zongyi, et al.
Publicado: (2025)
Low-Cost and Comprehensive Non-textual Input Fuzzing with LLM-Synthesized Input Generators
por: Zhang, Kunpeng, et al.
Publicado: (2025)
por: Zhang, Kunpeng, et al.
Publicado: (2025)
EAMET: Robust Massive Model Editing via Embedding Alignment Optimization
por: Dai, Yanbo, et al.
Publicado: (2025)
por: Dai, Yanbo, et al.
Publicado: (2025)
InstructTA: Instruction-Tuned Targeted Attack for Large Vision-Language Models
por: Wang, Xunguang, et al.
Publicado: (2023)
por: Wang, Xunguang, et al.
Publicado: (2023)
WARBENCH: A Comprehensive Benchmark for Evaluating LLMs in Military Decision-Making
por: Li, Zongjie, et al.
Publicado: (2026)
por: Li, Zongjie, et al.
Publicado: (2026)
Disabling Self-Correction in Retrieval-Augmented Generation via Stealthy Retriever Poisoning
por: Dai, Yanbo, et al.
Publicado: (2025)
por: Dai, Yanbo, et al.
Publicado: (2025)
GuidedBench: Measuring and Mitigating the Evaluation Discrepancies of In-the-wild LLM Jailbreak Methods
por: Huang, Ruixuan, et al.
Publicado: (2025)
por: Huang, Ruixuan, et al.
Publicado: (2025)
Once4All: Skeleton-Guided SMT Solver Fuzzing with LLM-Synthesized Generators
por: Sun, Maolin, et al.
Publicado: (2025)
por: Sun, Maolin, et al.
Publicado: (2025)
Provable Coordination for LLM Agents via Message Sequence Charts
por: Bollig, Benedikt, et al.
Publicado: (2026)
por: Bollig, Benedikt, et al.
Publicado: (2026)
Taming Various Privilege Escalation in LLM-Based Agent Systems: A Mandatory Access Control Framework
por: Ji, Zimo, et al.
Publicado: (2026)
por: Ji, Zimo, et al.
Publicado: (2026)
CodeARC: Benchmarking Reasoning Capabilities of LLM Agents for Inductive Program Synthesis
por: Wei, Anjiang, et al.
Publicado: (2025)
por: Wei, Anjiang, et al.
Publicado: (2025)
An LLM-powered Natural-to-Robotic Language Translation Framework with Correctness Guarantees
por: Chen, ZhenDong, et al.
Publicado: (2025)
por: Chen, ZhenDong, et al.
Publicado: (2025)
Benchmark Test-Time Scaling of General LLM Agents
por: Li, Xiaochuan, et al.
Publicado: (2026)
por: Li, Xiaochuan, et al.
Publicado: (2026)
OBsmith: LLM-Powered JavaScript Obfuscator Testing
por: Jiang, Shan, et al.
Publicado: (2025)
por: Jiang, Shan, et al.
Publicado: (2025)
AIOS Compiler: LLM as Interpreter for Natural Language Programming and Flow Programming of AI Agents
por: Xu, Shuyuan, et al.
Publicado: (2024)
por: Xu, Shuyuan, et al.
Publicado: (2024)
Generating Pragmatic Examples to Train Neural Program Synthesizers
por: Vaduguru, Saujas, et al.
Publicado: (2023)
por: Vaduguru, Saujas, et al.
Publicado: (2023)
AutoPDL: Automatic Prompt Optimization for LLM Agents
por: Spiess, Claudio, et al.
Publicado: (2025)
por: Spiess, Claudio, et al.
Publicado: (2025)
SEAL: Subspace-Anchored Watermarks for LLM Ownership
por: Dai, Yanbo, et al.
Publicado: (2025)
por: Dai, Yanbo, et al.
Publicado: (2025)
LACUNA: Safe Agents as Recursive Program Holes
por: Zhao, Yaoyu, et al.
Publicado: (2026)
por: Zhao, Yaoyu, et al.
Publicado: (2026)
BuildBench: Benchmarking LLM Agents on Compiling Real-World Open-Source Software
por: Zhang, Zehua, et al.
Publicado: (2025)
por: Zhang, Zehua, et al.
Publicado: (2025)
Enhancing Dialogue State Tracking Models through LLM-backed User-Agents Simulation
por: Niu, Cheng, et al.
Publicado: (2024)
por: Niu, Cheng, et al.
Publicado: (2024)
Inference Plans for Hybrid Particle Filtering
por: Cheng, Ellie Y., et al.
Publicado: (2024)
por: Cheng, Ellie Y., et al.
Publicado: (2024)
A Declarative Language for Building And Orchestrating LLM-Powered Agent Workflows
por: Daunis, Ivan
Publicado: (2025)
por: Daunis, Ivan
Publicado: (2025)
Synthesizing Post-Training Data for LLMs through Multi-Agent Simulation
por: Tang, Shuo, et al.
Publicado: (2024)
por: Tang, Shuo, et al.
Publicado: (2024)
Synthesizing Programmatic Reinforcement Learning Policies with Large Language Model Guided Search
por: Liu, Max, et al.
Publicado: (2024)
por: Liu, Max, et al.
Publicado: (2024)
CodeV: Empowering LLMs with HDL Generation through Multi-Level Summarization
por: Zhao, Yang, et al.
Publicado: (2024)
por: Zhao, Yang, et al.
Publicado: (2024)
Enforcing Temporal Constraints for LLM Agents
por: Kamath, Adharsh, et al.
Publicado: (2025)
por: Kamath, Adharsh, et al.
Publicado: (2025)
Ejemplares similares
-
STShield: Single-Token Sentinel for Real-Time Jailbreak Detection in Large Language Models
por: Wang, Xunguang, et al.
Publicado: (2025) -
IP Leakage Attacks Targeting LLM-Based Multi-Agent Systems
por: Wang, Liwen, et al.
Publicado: (2025) -
Split and Merge: Aligning Position Biases in LLM-based Evaluators
por: Li, Zongjie, et al.
Publicado: (2023) -
Digging Into the Internal: Causality-Based Analysis of LLM Function Calling
por: Ji, Zhenlan, et al.
Publicado: (2025) -
Measuring and Augmenting Large Language Models for Solving Capture-the-Flag Challenges
por: Ji, Zimo, et al.
Publicado: (2025)