Language-Based Agent Control
Fuente:
arXiv
Saved in:
| Main Authors: | Zhou, Timothy, D'Antoni, Loris, Polikarpova, Nadia |
|---|---|
| Format: | Preprint |
| Published: |
2026
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
ChopChop: a Programmable Framework for Semantically Constraining the Output of Language Models
by: Nagy, Shaan, et al.
Published: (2025)
by: Nagy, Shaan, et al.
Published: (2025)
AC4A: Access Control for Agents
by: Sharma, Reshabh K, et al.
Published: (2026)
by: Sharma, Reshabh K, et al.
Published: (2026)
PAuth - Precise Task-Scoped Authorization For Agents
by: Sharma, Reshabh K, et al.
Published: (2026)
by: Sharma, Reshabh K, et al.
Published: (2026)
Semia: Auditing Agent Skills via Constraint-Guided Representation Synthesis
by: Wen, Hongbo, et al.
Published: (2026)
by: Wen, Hongbo, et al.
Published: (2026)
How secure is AI-generated Code: A Large-Scale Comparison of Large Language Models
by: Tihanyi, Norbert, et al.
Published: (2024)
by: Tihanyi, Norbert, et al.
Published: (2024)
A Fast, Reliable, and Secure Programming Language for LLM Agents with Code Actions
by: Mell, Stephen, et al.
Published: (2025)
by: Mell, Stephen, et al.
Published: (2025)
Grammar-Aligned Decoding
by: Park, Kanghee, et al.
Published: (2024)
by: Park, Kanghee, et al.
Published: (2024)
CrypTorch: PyTorch-based Auto-tuning Compiler for Machine Learning with Multi-party Computation
by: Liu, Jinyu, et al.
Published: (2025)
by: Liu, Jinyu, et al.
Published: (2025)
C2RUST-BENCH: A Minimized, Representative Dataset for C-to-Rust Transpilation Evaluation
by: Sirlanci, Melih, et al.
Published: (2025)
by: Sirlanci, Melih, et al.
Published: (2025)
Arbiter: Detecting Interference in LLM Agent System Prompts
by: Mason, Tony
Published: (2026)
by: Mason, Tony
Published: (2026)
Watch Out for Your Agents! Investigating Backdoor Threats to LLM-Based Agents
by: Yang, Wenkai, et al.
Published: (2024)
by: Yang, Wenkai, et al.
Published: (2024)
PECAN: A Deterministic Certified Defense Against Backdoor Attacks
by: Zhang, Yuhao, et al.
Published: (2023)
by: Zhang, Yuhao, et al.
Published: (2023)
Fuzzing: Randomness? Reasoning! Efficient Directed Fuzzing via Large Language Models
by: Feng, Xiaotao, et al.
Published: (2025)
by: Feng, Xiaotao, et al.
Published: (2025)
Flexible and Efficient Grammar-Constrained Decoding
by: Park, Kanghee, et al.
Published: (2025)
by: Park, Kanghee, et al.
Published: (2025)
R1-Fuzz: Specializing Language Models for Textual Fuzzing via Reinforcement Learning
by: Lin, Jiayi, et al.
Published: (2025)
by: Lin, Jiayi, et al.
Published: (2025)
GradingAttack: Exposing Security Vulnerabilities in LLM Based Educational Grading Agents
by: Li, Xueyi, et al.
Published: (2026)
by: Li, Xueyi, et al.
Published: (2026)
RedAgent: Red Teaming Large Language Models with Context-aware Autonomous Language Agent
by: Xu, Huiyu, et al.
Published: (2024)
by: Xu, Huiyu, et al.
Published: (2024)
AgentAlign: Navigating Safety Alignment in the Shift from Informative to Agentic Large Language Models
by: Zhang, Jinchuan, et al.
Published: (2025)
by: Zhang, Jinchuan, et al.
Published: (2025)
IPIGuard: A Novel Tool Dependency Graph-Based Defense Against Indirect Prompt Injection in LLM Agents
by: An, Hengyu, et al.
Published: (2025)
by: An, Hengyu, et al.
Published: (2025)
Large Language Model Sentinel: LLM Agent for Adversarial Purification
by: Lin, Guang, et al.
Published: (2024)
by: Lin, Guang, et al.
Published: (2024)
ai.txt: A Domain-Specific Language for Guiding AI Interactions with the Internet
by: Li, Yuekang, et al.
Published: (2025)
by: Li, Yuekang, et al.
Published: (2025)
Defenses & Enablers For Skill Injection Attacks on Terminal Based Agents
by: Fujinuma, Yoshinari, et al.
Published: (2026)
by: Fujinuma, Yoshinari, et al.
Published: (2026)
IP Leakage Attacks Targeting LLM-Based Multi-Agent Systems
by: Wang, Liwen, et al.
Published: (2025)
by: Wang, Liwen, et al.
Published: (2025)
SafeSearch: Automated Red-Teaming of LLM-Based Search Agents
by: Dong, Jianshuo, et al.
Published: (2025)
by: Dong, Jianshuo, et al.
Published: (2025)
Cognitive Control Architecture (CCA): A Lifecycle Supervision Framework for Robustly Aligned AI Agents
by: Liang, Zhibo, et al.
Published: (2025)
by: Liang, Zhibo, et al.
Published: (2025)
When Agents "Misremember" Collectively: Exploring the Mandela Effect in LLM-based Multi-Agent Systems
by: Xu, Naen, et al.
Published: (2026)
by: Xu, Naen, et al.
Published: (2026)
Hound: Relation-First Knowledge Graphs for Complex-System Reasoning in Security Audits
by: Mueller, Bernhard
Published: (2025)
by: Mueller, Bernhard
Published: (2025)
BaxBench: Can LLMs Generate Correct and Secure Backends?
by: Vero, Mark, et al.
Published: (2025)
by: Vero, Mark, et al.
Published: (2025)
Agentic Specification Generator for Move Programs
by: Fu, Yu-Fu, et al.
Published: (2025)
by: Fu, Yu-Fu, et al.
Published: (2025)
AutoBaxBuilder: Bootstrapping Code Security Benchmarking
by: von Arx, Tobias, et al.
Published: (2025)
by: von Arx, Tobias, et al.
Published: (2025)
Agent Tools Orchestration Leaks More: Dataset, Benchmark, and Mitigation
by: Qiao, Yuxuan, et al.
Published: (2025)
by: Qiao, Yuxuan, et al.
Published: (2025)
ACIArena: Toward Unified Evaluation for Agent Cascading Injection
by: An, Hengyu, et al.
Published: (2026)
by: An, Hengyu, et al.
Published: (2026)
SIRAJ: Diverse and Efficient Red-Teaming for LLM Agents via Distilled Structured Reasoning
by: Zhou, Kaiwen, et al.
Published: (2025)
by: Zhou, Kaiwen, et al.
Published: (2025)
Tempest: Autonomous Multi-Turn Jailbreaking of Large Language Models with Tree Search
by: Zhou, Andy, et al.
Published: (2025)
by: Zhou, Andy, et al.
Published: (2025)
CCJA: Context-Coherent Jailbreak Attack for Aligned Large Language Models
by: Zhou, Guanghao, et al.
Published: (2025)
by: Zhou, Guanghao, et al.
Published: (2025)
MRJ-Agent: An Effective Jailbreak Agent for Multi-Round Dialogue
by: Wang, Fengxiang, et al.
Published: (2024)
by: Wang, Fengxiang, et al.
Published: (2024)
A Content-Based Framework for Cybersecurity Refusal Decisions in Large Language Models
by: Linder, Noa, et al.
Published: (2026)
by: Linder, Noa, et al.
Published: (2026)
Protecting Users From Themselves: Safeguarding Contextual Privacy in Interactions with Conversational Agents
by: Ngong, Ivoline, et al.
Published: (2025)
by: Ngong, Ivoline, et al.
Published: (2025)
NSmark: Null Space Based Black-box Watermarking Defense Framework for Language Models
by: Zhao, Haodong, et al.
Published: (2024)
by: Zhao, Haodong, et al.
Published: (2024)
Your Agent, Their Asset: A Real-World Safety Analysis of OpenClaw
by: Wang, Zijun, et al.
Published: (2026)
by: Wang, Zijun, et al.
Published: (2026)
Similar Items
-
ChopChop: a Programmable Framework for Semantically Constraining the Output of Language Models
by: Nagy, Shaan, et al.
Published: (2025) -
AC4A: Access Control for Agents
by: Sharma, Reshabh K, et al.
Published: (2026) -
PAuth - Precise Task-Scoped Authorization For Agents
by: Sharma, Reshabh K, et al.
Published: (2026) -
Semia: Auditing Agent Skills via Constraint-Guided Representation Synthesis
by: Wen, Hongbo, et al.
Published: (2026) -
How secure is AI-generated Code: A Large-Scale Comparison of Large Language Models
by: Tihanyi, Norbert, et al.
Published: (2024)