Safety, Security, and Cognitive Risks in World Models
Fuente:
arXiv
Salvato in:
| Autore principale: | Parmar, Manoj |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2026
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
"Abuse Risks are Often Inherent to Product Features": Exploring AI Vendors' Bug Bounty and Responsible Disclosure Policies
di: Piao, Yangheran, et al.
Pubblicazione: (2025)
di: Piao, Yangheran, et al.
Pubblicazione: (2025)
PACT: Reducing Alert Fatigue in Low-Prevalence SOC Streams with Triggered Active Learning
di: Ndichu, Samuel, et al.
Pubblicazione: (2026)
di: Ndichu, Samuel, et al.
Pubblicazione: (2026)
Explainable Attention-Based LSTM Framework for Early Detection of AI-Assisted Ransomware via File System Behavioral Analysis
di: Nayak, Prabhudarshi, et al.
Pubblicazione: (2026)
di: Nayak, Prabhudarshi, et al.
Pubblicazione: (2026)
Operationalizing Cybersecurity Governance for Mitigation Planning with Attack-Path Modeling and Reinforcement Learning
di: Huff, Philip, et al.
Pubblicazione: (2026)
di: Huff, Philip, et al.
Pubblicazione: (2026)
$δ$-STEAL: LLM Stealing Attack with Local Differential Privacy
di: Dang, Kieu, et al.
Pubblicazione: (2025)
di: Dang, Kieu, et al.
Pubblicazione: (2025)
Cross-Domain Malware Detection via Probability-Level Fusion of Lightweight Gradient Boosting Models
di: Mohamed, Omar Khalid Ali
Pubblicazione: (2025)
di: Mohamed, Omar Khalid Ali
Pubblicazione: (2025)
BreakFun: Jailbreaking LLMs via Schema Exploitation
di: Oskooei, Amirkia Rafiei, et al.
Pubblicazione: (2025)
di: Oskooei, Amirkia Rafiei, et al.
Pubblicazione: (2025)
Exploratory Analysis of Cyberattack Patterns on E-Commerce Platforms Using Statistical Methods
di: Adeniya, Fatimo Adenike
Pubblicazione: (2025)
di: Adeniya, Fatimo Adenike
Pubblicazione: (2025)
Semantically Guided Adversarial Testing of Vision Models Using Language Models
di: Filus, Katarzyna, et al.
Pubblicazione: (2025)
di: Filus, Katarzyna, et al.
Pubblicazione: (2025)
Breaking to Build: A Threat Model of Prompt-Based Attacks for Securing LLMs
di: Hill, Brennen, et al.
Pubblicazione: (2025)
di: Hill, Brennen, et al.
Pubblicazione: (2025)
An Improved Adaptive PID Optimizer with Enhanced Convergence and Stability for Deep Learning
di: Saini, Saurabh, et al.
Pubblicazione: (2026)
di: Saini, Saurabh, et al.
Pubblicazione: (2026)
DRIFT: Drift-Resilient Invariant-Feature Transformer for DGA Detection
di: Lee, Chaeyoung, et al.
Pubblicazione: (2026)
di: Lee, Chaeyoung, et al.
Pubblicazione: (2026)
Quantum Machine Learning for Cyber-Physical Anomaly Detection in Unmanned Aerial Vehicles: A Leakage-Free Evaluation with Proxy-Audited Feature Sets
di: Paredes, Carlos A. Durán, et al.
Pubblicazione: (2026)
di: Paredes, Carlos A. Durán, et al.
Pubblicazione: (2026)
VOLTRON: Detecting Unknown Malware Using Graph-Based Zero-Shot Learning
di: Akdeniz, M. Tahir, et al.
Pubblicazione: (2025)
di: Akdeniz, M. Tahir, et al.
Pubblicazione: (2025)
Feature Selection Based on Reinforcement Learning and Hazard State Classification for Magnetic Adhesion Wall-Climbing Robots
di: Ma, Zhen, et al.
Pubblicazione: (2025)
di: Ma, Zhen, et al.
Pubblicazione: (2025)
Agentic Artificial Intelligence for Ethical Cybersecurity in Uganda: A Reinforcement Learning Framework for Threat Detection in Resource-Constrained Environments
di: Adabara, Ibrahim, et al.
Pubblicazione: (2025)
di: Adabara, Ibrahim, et al.
Pubblicazione: (2025)
Connectivity-Aware Representations for Constrained Motion Planning via Multi-Scale Contrastive Learning
di: Jeon, Suhyun, et al.
Pubblicazione: (2026)
di: Jeon, Suhyun, et al.
Pubblicazione: (2026)
Vis-CoT: A Human-in-the-Loop Framework for Interactive Visualization and Intervention in LLM Chain-of-Thought Reasoning
di: Pather, Kaviraj, et al.
Pubblicazione: (2025)
di: Pather, Kaviraj, et al.
Pubblicazione: (2025)
ConvXformer: Differentially Private Hybrid ConvNeXt-Transformer for Inertial Navigation
di: Tariq, Omer, et al.
Pubblicazione: (2025)
di: Tariq, Omer, et al.
Pubblicazione: (2025)
RAR: Setting Knowledge Tripwires for Retrieval Augmented Rejection
di: Buonocore, Tommaso Mario, et al.
Pubblicazione: (2025)
di: Buonocore, Tommaso Mario, et al.
Pubblicazione: (2025)
Improved ICNN-LSTM Model Classification Based on Attitude Sensor Data for Hazardous State Assessment of Magnetic Adhesion Climbing Wall Robots
di: Ma, Zhen, et al.
Pubblicazione: (2024)
di: Ma, Zhen, et al.
Pubblicazione: (2024)
A Comparison Between Decision Transformers and Traditional Offline Reinforcement Learning Algorithms
di: Caunhye, Ali Murtaza, et al.
Pubblicazione: (2025)
di: Caunhye, Ali Murtaza, et al.
Pubblicazione: (2025)
Evaluating Model-Agnostic Meta-Learning on MetaWorld ML10 Benchmark: Fast Adaptation in Robotic Manipulation Tasks
di: Atamuradov, Sanjar
Pubblicazione: (2025)
di: Atamuradov, Sanjar
Pubblicazione: (2025)
CellARC: Measuring Intelligence with Cellular Automata
di: Lžičař, Miroslav
Pubblicazione: (2025)
di: Lžičař, Miroslav
Pubblicazione: (2025)
Swarm Intelligence-Driven Client Selection for Federated Learning in Cybersecurity applications
di: Khan, Koffka, et al.
Pubblicazione: (2024)
di: Khan, Koffka, et al.
Pubblicazione: (2024)
Emotion-Inspired Learning Signals (EILS): A Homeostatic Framework for Adaptive Autonomous Agents
di: Tiwari, Dhruv
Pubblicazione: (2025)
di: Tiwari, Dhruv
Pubblicazione: (2025)
A Dual-Path Generative Framework for Zero-Day Fraud Detection in Banking Systems
di: Ismail, Nasim Abdirahman, et al.
Pubblicazione: (2026)
di: Ismail, Nasim Abdirahman, et al.
Pubblicazione: (2026)
Breaking Boundaries: Balancing Performance and Robustness in Deep Wireless Traffic Forecasting
di: Ilbert, Romain, et al.
Pubblicazione: (2023)
di: Ilbert, Romain, et al.
Pubblicazione: (2023)
Evaluation of Differential Privacy Mechanisms on Federated Learning
di: Varsani, Tejash
Pubblicazione: (2025)
di: Varsani, Tejash
Pubblicazione: (2025)
Multiple data-driven missing imputation
di: Kavun, Sergii
Pubblicazione: (2025)
di: Kavun, Sergii
Pubblicazione: (2025)
OpCode-Based Malware Classification Using Machine Learning and Deep Learning Techniques
di: Saini, Varij, et al.
Pubblicazione: (2025)
di: Saini, Varij, et al.
Pubblicazione: (2025)
Adaptive Negative Scheduling for Graph Contrastive Learning
di: Ali, Adnan, et al.
Pubblicazione: (2026)
di: Ali, Adnan, et al.
Pubblicazione: (2026)
The Impact of Data Characteristics on GNN Evaluation for Detecting Fake News
di: Karn, Isha, et al.
Pubblicazione: (2025)
di: Karn, Isha, et al.
Pubblicazione: (2025)
Generative AI and the Transformation of Software Development Practices
di: Acharya, Vivek
Pubblicazione: (2025)
di: Acharya, Vivek
Pubblicazione: (2025)
Federated Learning in Adversarial Environments: Testbed Design and Poisoning Resilience in Cybersecurity
di: Huang, Hao Jian, et al.
Pubblicazione: (2024)
di: Huang, Hao Jian, et al.
Pubblicazione: (2024)
Normalisation and Initialisation Strategies for Graph Neural Networks in Blockchain Anomaly Detection
di: Duy, Dang Sy, et al.
Pubblicazione: (2026)
di: Duy, Dang Sy, et al.
Pubblicazione: (2026)
The Age of Sensorial Zero Trust: Why We Can No Longer Trust Our Senses
di: Xavier, Fabio Correa
Pubblicazione: (2025)
di: Xavier, Fabio Correa
Pubblicazione: (2025)
From Imitation to Interaction: Mastering Game of Schnapsen with Shallow Reinforcement Learning
di: Klačan, Ján, et al.
Pubblicazione: (2026)
di: Klačan, Ján, et al.
Pubblicazione: (2026)
ElasticFlow: One-Step Physics-Consistent Policy with Elastic Time Horizons for Language-Guided Manipulation
di: Chen, Kewei, et al.
Pubblicazione: (2026)
di: Chen, Kewei, et al.
Pubblicazione: (2026)
AI Agents: Evolution, Architecture, and Real-World Applications
di: Krishnan, Naveen
Pubblicazione: (2025)
di: Krishnan, Naveen
Pubblicazione: (2025)
Documenti analoghi
-
"Abuse Risks are Often Inherent to Product Features": Exploring AI Vendors' Bug Bounty and Responsible Disclosure Policies
di: Piao, Yangheran, et al.
Pubblicazione: (2025) -
PACT: Reducing Alert Fatigue in Low-Prevalence SOC Streams with Triggered Active Learning
di: Ndichu, Samuel, et al.
Pubblicazione: (2026) -
Explainable Attention-Based LSTM Framework for Early Detection of AI-Assisted Ransomware via File System Behavioral Analysis
di: Nayak, Prabhudarshi, et al.
Pubblicazione: (2026) -
Operationalizing Cybersecurity Governance for Mitigation Planning with Attack-Path Modeling and Reinforcement Learning
di: Huff, Philip, et al.
Pubblicazione: (2026) -
$δ$-STEAL: LLM Stealing Attack with Local Differential Privacy
di: Dang, Kieu, et al.
Pubblicazione: (2025)