How does information access affect LLM monitors' ability to detect sabotage?
Fuente:
arXiv
Saved in:
| Main Authors: | Arike, Rauno, Moreno, Raja Mehta, Subramani, Rohan, Biswas, Shubhorup, Ward, Francis Rhys |
|---|---|
| Format: | Preprint |
| Published: |
2026
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Consistency Amplifies: How Behavioral Variance Shapes Agent Accuracy
by: Mehta, Aman
Published: (2026)
by: Mehta, Aman
Published: (2026)
C2-Faith: Benchmarking LLM Judges for Causal and Coverage Faithfulness in Chain-of-Thought Reasoning
by: Mittal, Avni, et al.
Published: (2026)
by: Mittal, Avni, et al.
Published: (2026)
Green My LLM: Studying the key factors affecting the energy consumption of code assistants
by: Coignion, Tristan, et al.
Published: (2024)
by: Coignion, Tristan, et al.
Published: (2024)
LLM For Loop Invariant Generation and Fixing: How Far Are We?
by: Akhond, Mostafijur Rahman, et al.
Published: (2025)
by: Akhond, Mostafijur Rahman, et al.
Published: (2025)
CASET: Complexity Analysis using Simple Execution Traces for CS* submissions
by: Mehta, Aaryen, et al.
Published: (2024)
by: Mehta, Aaryen, et al.
Published: (2024)
Five Fatal Assumptions: Why T-Shirt Sizing Systematically Fails for AI Projects
by: Soundaramourty, Raja, et al.
Published: (2026)
by: Soundaramourty, Raja, et al.
Published: (2026)
An Empirical Framework for Evaluating Semantic Preservation Using Hugging Face
by: Jia, Nan, et al.
Published: (2025)
by: Jia, Nan, et al.
Published: (2025)
The Hidden Cost of Readability: How Code Formatting Silently Consumes Your LLM Budget
by: Pan, Dangfeng, et al.
Published: (2025)
by: Pan, Dangfeng, et al.
Published: (2025)
Less Is More: Measuring How LLM Involvement affects Chatbot Accuracy in Static Analysis
by: Narasimhan, Krishna
Published: (2026)
by: Narasimhan, Krishna
Published: (2026)
How to Trick Your AI TA: A Systematic Study of Academic Jailbreaking in LLM Code Evaluation
by: Sahoo, Devanshu, et al.
Published: (2025)
by: Sahoo, Devanshu, et al.
Published: (2025)
What Should Frontier AI Developers Disclose About Internal Deployments?
by: Charnock, Jacob, et al.
Published: (2026)
by: Charnock, Jacob, et al.
Published: (2026)
How Many Tries Does It Take? Iterative Self-Repair in LLM Code Generation Across Model Scales and Benchmarks
by: Arimbur, Johin Johny
Published: (2026)
by: Arimbur, Johin Johny
Published: (2026)
Higher-Order Belief in Incomplete Information MAIDs
by: Foxabbott, Jack, et al.
Published: (2025)
by: Foxabbott, Jack, et al.
Published: (2025)
Automated detection of atomicity violations in large-scale systems
by: He, Hang, et al.
Published: (2025)
by: He, Hang, et al.
Published: (2025)
How do annotations affect Java code readability?
by: Guerra, Eduardo, et al.
Published: (2024)
by: Guerra, Eduardo, et al.
Published: (2024)
Automated Creation and Enrichment Framework for Improved Invocation of Enterprise APIs as Tools
by: Agarwal, Prerna, et al.
Published: (2025)
by: Agarwal, Prerna, et al.
Published: (2025)
Password-Activated Shutdown Protocols for Misaligned Frontier Agents
by: Williams, Kai, et al.
Published: (2025)
by: Williams, Kai, et al.
Published: (2025)
MLScent A tool for Anti-pattern detection in ML projects
by: Shivashankar, Karthik, et al.
Published: (2025)
by: Shivashankar, Karthik, et al.
Published: (2025)
AdversariaLLM: A Unified and Modular Toolbox for LLM Robustness Research
by: Beyer, Tim, et al.
Published: (2025)
by: Beyer, Tim, et al.
Published: (2025)
LLM-Explorer: Towards Efficient and Affordable LLM-based Exploration for Mobile Apps
by: Zhao, Shanhui, et al.
Published: (2025)
by: Zhao, Shanhui, et al.
Published: (2025)
Quality Assessment of Tabular Data using Large Language Models and Code Generation
by: Akella, Ashlesha, et al.
Published: (2025)
by: Akella, Ashlesha, et al.
Published: (2025)
How does Simulation-based Testing for Self-driving Cars match Human Perception?
by: Birchler, Christian, et al.
Published: (2024)
by: Birchler, Christian, et al.
Published: (2024)
Artificial intelligence for context-aware visual change detection in software test automation
by: Moradi, Milad, et al.
Published: (2024)
by: Moradi, Milad, et al.
Published: (2024)
How Efficient is LLM-Generated Code? A Rigorous & High-Standard Benchmark
by: Qiu, Ruizhong, et al.
Published: (2024)
by: Qiu, Ruizhong, et al.
Published: (2024)
LLM-Rosetta: A Hub-and-Spoke Intermediate Representation for Cross-Provider LLM API Translation
by: Ding, Peng
Published: (2026)
by: Ding, Peng
Published: (2026)
Specification and Detection of LLM Code Smells
by: Mahmoudi, Brahim, et al.
Published: (2025)
by: Mahmoudi, Brahim, et al.
Published: (2025)
Evaluating the effectiveness of LLM-based interoperability
by: Falcão, Rodrigo, et al.
Published: (2025)
by: Falcão, Rodrigo, et al.
Published: (2025)
Investigating The Smells of LLM Generated Code
by: Paul, Debalina Ghosh, et al.
Published: (2025)
by: Paul, Debalina Ghosh, et al.
Published: (2025)
Uncertainty Propagation in LLM-Based Systems
by: Xia, Boming, et al.
Published: (2026)
by: Xia, Boming, et al.
Published: (2026)
Breaking the Illusion of Identity in LLM Tooling
by: Miller, Marek
Published: (2026)
by: Miller, Marek
Published: (2026)
Enhancing Code Translation in Language Models with Few-Shot Learning via Retrieval-Augmented Generation
by: Bhattarai, Manish, et al.
Published: (2024)
by: Bhattarai, Manish, et al.
Published: (2024)
How Robust are LLM-Generated Library Imports? An Empirical Study using Stack Overflow
by: Latendresse, Jasmine, et al.
Published: (2025)
by: Latendresse, Jasmine, et al.
Published: (2025)
Retrieval-Augmented Test Generation: How Far Are We?
by: Shin, Jiho, et al.
Published: (2024)
by: Shin, Jiho, et al.
Published: (2024)
Reducing Cost of LLM Agents with Trajectory Reduction
by: Xiao, Yuan-An, et al.
Published: (2025)
by: Xiao, Yuan-An, et al.
Published: (2025)
LLM Collaboration With Multi-Agent Reinforcement Learning
by: Liu, Shuo, et al.
Published: (2025)
by: Liu, Shuo, et al.
Published: (2025)
Uncertainty Quantification for LLM-based Code Generation
by: Xu, Senrong, et al.
Published: (2026)
by: Xu, Senrong, et al.
Published: (2026)
Performance Review on LLM for solving leetcode problems
by: Wang, Lun, et al.
Published: (2025)
by: Wang, Lun, et al.
Published: (2025)
LLM-based Iterative Approach to Metamodeling in Automotive
by: Petrovic, Nenad, et al.
Published: (2025)
by: Petrovic, Nenad, et al.
Published: (2025)
Impact of Comments on LLM Comprehension of Legacy Code
by: Sabetto, Rock, et al.
Published: (2025)
by: Sabetto, Rock, et al.
Published: (2025)
Rover: Context-aware Conflict Resolution with LLM
by: Zhang, Qingyu, et al.
Published: (2026)
by: Zhang, Qingyu, et al.
Published: (2026)
Similar Items
-
Consistency Amplifies: How Behavioral Variance Shapes Agent Accuracy
by: Mehta, Aman
Published: (2026) -
C2-Faith: Benchmarking LLM Judges for Causal and Coverage Faithfulness in Chain-of-Thought Reasoning
by: Mittal, Avni, et al.
Published: (2026) -
Green My LLM: Studying the key factors affecting the energy consumption of code assistants
by: Coignion, Tristan, et al.
Published: (2024) -
LLM For Loop Invariant Generation and Fixing: How Far Are We?
by: Akhond, Mostafijur Rahman, et al.
Published: (2025) -
CASET: Complexity Analysis using Simple Execution Traces for CS* submissions
by: Mehta, Aaryen, et al.
Published: (2024)