Saved in:
| Main Authors: | Marwah, Manish, Narayanan, Asad, Jou, Stephan, Arlitt, Martin, Pospelova, Maria |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | https://arxiv.org/abs/2407.14664 |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
PhreshPhish: A Real-World, High-Quality, Large-Scale Phishing Website Dataset and Benchmark
by: Dalton, Thomas, et al.
Published: (2025)
by: Dalton, Thomas, et al.
Published: (2025)
Beyond Self-Reports: Multi-Observer Agents for Personality Assessment in Large Language Models
by: Huang, Yin Jou, et al.
Published: (2025)
by: Huang, Yin Jou, et al.
Published: (2025)
Semantic F1 Scores: Fair Evaluation Under Fuzzy Class Boundaries
by: Chochlakis, Georgios, et al.
Published: (2025)
by: Chochlakis, Georgios, et al.
Published: (2025)
Quantifying Laziness, Decoding Suboptimality, and Context Degradation in Large Language Models
by: Ma, Yiqing, et al.
Published: (2025)
by: Ma, Yiqing, et al.
Published: (2025)
Chimera: State Space Models Beyond Sequences
by: Lahoti, Aakash, et al.
Published: (2025)
by: Lahoti, Aakash, et al.
Published: (2025)
Effective Integration of Weighted Cost-to-go and Conflict Heuristic within Suboptimal CBS
by: Veerapaneni, Rishi, et al.
Published: (2022)
by: Veerapaneni, Rishi, et al.
Published: (2022)
How Personality Traits Influence Negotiation Outcomes? A Simulation based on Large Language Models
by: Huang, Yin Jou, et al.
Published: (2024)
by: Huang, Yin Jou, et al.
Published: (2024)
On the Benefits of Memory for Modeling Time-Dependent PDEs
by: Ruiz, Ricardo Buitrago, et al.
Published: (2024)
by: Ruiz, Ricardo Buitrago, et al.
Published: (2024)
Dissecting Discrete Soft Actor-Critic: Limitations and Principled Alternatives
by: Asad, Reza, et al.
Published: (2025)
by: Asad, Reza, et al.
Published: (2025)
Suboptimal Shapley Value Explanations
by: Lu, Xiaolei
Published: (2025)
by: Lu, Xiaolei
Published: (2025)
Collaborative Intelligence: Topic Modelling of Large Language Model use in Live Cybersecurity Operations
by: Lochner, Martin, et al.
Published: (2025)
by: Lochner, Martin, et al.
Published: (2025)
Dynamic Risk Assessments for Offensive Cybersecurity Agents
by: Wei, Boyi, et al.
Published: (2025)
by: Wei, Boyi, et al.
Published: (2025)
Investigating Cost-Efficiency of LLM-Generated Training Data for Conversational Semantic Frame Analysis
by: Matta, Shiho, et al.
Published: (2024)
by: Matta, Shiho, et al.
Published: (2024)
Cost-Aware Model Orchestration for LLM-based Systems
by: Smirnova, Daria, et al.
Published: (2025)
by: Smirnova, Daria, et al.
Published: (2025)
Trust, or Don't Predict: Introducing the CWSA Family for Confidence-Aware Model Evaluation
by: Shahnazari, Kourosh, et al.
Published: (2025)
by: Shahnazari, Kourosh, et al.
Published: (2025)
PAS : Prelim Attention Score for Detecting Object Hallucinations in Large Vision--Language Models
by: Hoang-Xuan, Nhat, et al.
Published: (2025)
by: Hoang-Xuan, Nhat, et al.
Published: (2025)
E-Scores for (In)Correctness Assessment of Generative Model Outputs
by: Dhillon, Guneet S., et al.
Published: (2025)
by: Dhillon, Guneet S., et al.
Published: (2025)
Bidirectional Bounded-Suboptimal Heuristic Search with Consistent Heuristics
by: Shperberg, Shahaf S., et al.
Published: (2025)
by: Shperberg, Shahaf S., et al.
Published: (2025)
Adapting Language Models via Token Translation
by: Feng, Zhili, et al.
Published: (2024)
by: Feng, Zhili, et al.
Published: (2024)
Improving ARDS Diagnosis Through Context-Aware Concept Bottleneck Models
by: Narain, Anish, et al.
Published: (2025)
by: Narain, Anish, et al.
Published: (2025)
A Survey of Large Language Models in Cybersecurity
by: da Silva, Gabriel de Jesus Coelho, et al.
Published: (2024)
by: da Silva, Gabriel de Jesus Coelho, et al.
Published: (2024)
Weaponizing Language Models for Cybersecurity Offensive Operations: Automating Vulnerability Assessment Report Validation; A Review Paper
by: Almuhaidib, Abdulrahman S, et al.
Published: (2025)
by: Almuhaidib, Abdulrahman S, et al.
Published: (2025)
Semantic Labeling for Third-Party Cybersecurity Risk Assessment: A Semi-Supervised Approach to Intent-Aware Question Retrieval
by: Eldin, Ali Nour, et al.
Published: (2026)
by: Eldin, Ali Nour, et al.
Published: (2026)
SV-TrustEval-C: Evaluating Structure and Semantic Reasoning in Large Language Models for Source Code Vulnerability Analysis
by: Li, Yansong, et al.
Published: (2025)
by: Li, Yansong, et al.
Published: (2025)
FinOps Agent -- A Use-Case for IT Infrastructure and Cost Optimization
by: Vo, Ngoc Phuoc An, et al.
Published: (2025)
by: Vo, Ngoc Phuoc An, et al.
Published: (2025)
Ontology-Aware RAG for Improved Question-Answering in Cybersecurity Education
by: Zhao, Chengshuai, et al.
Published: (2024)
by: Zhao, Chengshuai, et al.
Published: (2024)
Retaining Suboptimal Actions to Follow Shifting Optima in Multi-Agent Reinforcement Learning
by: Jo, Yonghyeon, et al.
Published: (2026)
by: Jo, Yonghyeon, et al.
Published: (2026)
New Mechanisms in Flex Distribution for Bounded Suboptimal Multi-Agent Path Finding
by: Chan, Shao-Hung, et al.
Published: (2025)
by: Chan, Shao-Hung, et al.
Published: (2025)
Toward Cybersecurity-Expert Small Language Models
by: Levi, Matan, et al.
Published: (2025)
by: Levi, Matan, et al.
Published: (2025)
Predictive Coding and Information Bottleneck for Hallucination Detection in Large Language Models
by: Bhatt, Manish
Published: (2026)
by: Bhatt, Manish
Published: (2026)
Evaluating Large Language Models Using Contrast Sets: An Experimental Approach
by: Sanwal, Manish
Published: (2024)
by: Sanwal, Manish
Published: (2024)
Leveraging Large Language Models for Cybersecurity Risk Assessment -- A Case from Forestry Cyber-Physical Systems
by: Gultekin, Fikret Mert, et al.
Published: (2025)
by: Gultekin, Fikret Mert, et al.
Published: (2025)
GPT-Enabled Cybersecurity Training: A Tailored Approach for Effective Awareness
by: Al-Dhamari, Nabil, et al.
Published: (2024)
by: Al-Dhamari, Nabil, et al.
Published: (2024)
CALM : A Multi-task Benchmark for Comprehensive Assessment of Language Model Bias
by: Gupta, Vipul, et al.
Published: (2023)
by: Gupta, Vipul, et al.
Published: (2023)
From Texts to Shields: Convergence of Large Language Models and Cybersecurity
by: Li, Tao, et al.
Published: (2025)
by: Li, Tao, et al.
Published: (2025)
Token-based Decision Criteria Are Suboptimal in In-context Learning
by: Cho, Hakaze, et al.
Published: (2024)
by: Cho, Hakaze, et al.
Published: (2024)
Breakthrough the Suboptimal Stable Point in Value-Factorization-Based Multi-Agent Reinforcement Learning
by: Tao, Lesong, et al.
Published: (2026)
by: Tao, Lesong, et al.
Published: (2026)
A survey on Concept-based Approaches For Model Improvement
by: Gupta, Avani, et al.
Published: (2024)
by: Gupta, Avani, et al.
Published: (2024)
Integrating Large Language Models For Monte Carlo Simulation of Chemical Reaction Networks
by: Gyawali, Sadikshya, et al.
Published: (2025)
by: Gyawali, Sadikshya, et al.
Published: (2025)
SECURE: Benchmarking Large Language Models for Cybersecurity
by: Bhusal, Dipkamal, et al.
Published: (2024)
by: Bhusal, Dipkamal, et al.
Published: (2024)
Similar Items
-
PhreshPhish: A Real-World, High-Quality, Large-Scale Phishing Website Dataset and Benchmark
by: Dalton, Thomas, et al.
Published: (2025) -
Beyond Self-Reports: Multi-Observer Agents for Personality Assessment in Large Language Models
by: Huang, Yin Jou, et al.
Published: (2025) -
Semantic F1 Scores: Fair Evaluation Under Fuzzy Class Boundaries
by: Chochlakis, Georgios, et al.
Published: (2025) -
Quantifying Laziness, Decoding Suboptimality, and Context Degradation in Large Language Models
by: Ma, Yiqing, et al.
Published: (2025) -
Chimera: State Space Models Beyond Sequences
by: Lahoti, Aakash, et al.
Published: (2025)