Saved in:
| Main Author: | Vassilev, Apostol |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | https://arxiv.org/abs/2512.10100 |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Arrhythmia Classification from 12-Lead ECG Signals Using Convolutional and Transformer-Based Deep Learning Models
by: Apostol, Andrei, et al.
Published: (2025)
by: Apostol, Andrei, et al.
Published: (2025)
Advancements in eHealth Data Analytics through Natural Language Processing and Deep Learning
by: Apostol, Elena-Simona, et al.
Published: (2024)
by: Apostol, Elena-Simona, et al.
Published: (2024)
CONTAIN: A Community-based Algorithm for Network Immunization
by: Apostol, Elena-Simona, et al.
Published: (2023)
by: Apostol, Elena-Simona, et al.
Published: (2023)
A Novel Reinforcement Learning Model for Post-Incident Malware Investigations
by: Dunsin, Dipo, et al.
Published: (2024)
by: Dunsin, Dipo, et al.
Published: (2024)
Clone-Robust AI Alignment
by: Procaccia, Ariel D., et al.
Published: (2025)
by: Procaccia, Ariel D., et al.
Published: (2025)
GETAE: Graph information Enhanced deep neural NeTwork ensemble ArchitecturE for fake news detection
by: Truică, Ciprian-Octavian, et al.
Published: (2024)
by: Truică, Ciprian-Octavian, et al.
Published: (2024)
Security-First AI: Foundations for Robust and Trustworthy Systems
by: Tallam, Krti
Published: (2025)
by: Tallam, Krti
Published: (2025)
The Adaptive Arms Race: Redefining Robustness in AI Security
by: Tsingenopoulos, Ilias, et al.
Published: (2023)
by: Tsingenopoulos, Ilias, et al.
Published: (2023)
A Comprehensive Analysis of the Role of Artificial Intelligence and Machine Learning in Modern Digital Forensics and Incident Response
by: Dunsin, Dipo, et al.
Published: (2023)
by: Dunsin, Dipo, et al.
Published: (2023)
Reinforcement Learning for an Efficient and Effective Malware Investigation during Cyber Incident Response
by: Dunsin, Dipo, et al.
Published: (2024)
by: Dunsin, Dipo, et al.
Published: (2024)
Towards Robust and Secure Embodied AI: A Survey on Vulnerabilities and Attacks
by: Xing, Wenpeng, et al.
Published: (2025)
by: Xing, Wenpeng, et al.
Published: (2025)
AI Alignment: A Comprehensive Survey
by: Ji, Jiaming, et al.
Published: (2023)
by: Ji, Jiaming, et al.
Published: (2023)
Hermes Seal: Zero-Knowledge Assurance for Autonomous Vehicle Communications
by: Hasan, Munawar, et al.
Published: (2026)
by: Hasan, Munawar, et al.
Published: (2026)
Beyond Preferences in AI Alignment
by: Zhi-Xuan, Tan, et al.
Published: (2024)
by: Zhi-Xuan, Tan, et al.
Published: (2024)
Explainable and Robust Millimeter Wave Beam Alignment for AI-Native 6G Networks
by: Khan, Nasir, et al.
Published: (2025)
by: Khan, Nasir, et al.
Published: (2025)
The AI Alignment Paradox
by: West, Robert, et al.
Published: (2024)
by: West, Robert, et al.
Published: (2024)
Measuring AI Alignment with Human Flourishing
by: Hilliard, Elizabeth, et al.
Published: (2025)
by: Hilliard, Elizabeth, et al.
Published: (2025)
Measuring Safety Alignment Effects in Autonomous Security Agents
by: David, Isaac, et al.
Published: (2026)
by: David, Isaac, et al.
Published: (2026)
Security of AI Agents
by: He, Yifeng, et al.
Published: (2024)
by: He, Yifeng, et al.
Published: (2024)
AI Security Map: Holistic Organization of AI Security Technologies and Impacts on Stakeholders
by: Kato, Hiroya, et al.
Published: (2025)
by: Kato, Hiroya, et al.
Published: (2025)
Improved Bounds for Private and Robust Alignment
by: Weng, Wenqian, et al.
Published: (2025)
by: Weng, Wenqian, et al.
Published: (2025)
Rethinking AI Cultural Alignment
by: Bravansky, Michal, et al.
Published: (2025)
by: Bravansky, Michal, et al.
Published: (2025)
Identifying and Mitigating the Security Risks of Generative AI
by: Barrett, Clark, et al.
Published: (2023)
by: Barrett, Clark, et al.
Published: (2023)
Interactive AI Alignment: Specification, Process, and Evaluation Alignment
by: Terry, Michael, et al.
Published: (2023)
by: Terry, Michael, et al.
Published: (2023)
Conformal Feedback Alignment: Quantifying Answer-Level Reliability for Robust LLM Alignment
by: Chen, Tiejin, et al.
Published: (2026)
by: Chen, Tiejin, et al.
Published: (2026)
Security of and by Generative AI platforms
by: Hayagreevan, Hari, et al.
Published: (2024)
by: Hayagreevan, Hari, et al.
Published: (2024)
The AI Security Pyramid of Pain
by: Ward, Chris M., et al.
Published: (2024)
by: Ward, Chris M., et al.
Published: (2024)
Secure Multiparty Generative AI
by: Shrestha, Manil, et al.
Published: (2024)
by: Shrestha, Manil, et al.
Published: (2024)
Strong Preferences Affect the Robustness of Preference Models and Value Alignment
by: Xu, Ziwei, et al.
Published: (2024)
by: Xu, Ziwei, et al.
Published: (2024)
Maia-2: A Unified Model for Human-AI Alignment in Chess
by: Tang, Zhenwei, et al.
Published: (2024)
by: Tang, Zhenwei, et al.
Published: (2024)
Trustworthy AI-Generative Content for Intelligent Network Service: Robustness, Security, and Fairness
by: Li, Siyuan, et al.
Published: (2024)
by: Li, Siyuan, et al.
Published: (2024)
Robust LLM Alignment via Distributionally Robust Direct Preference Optimization
by: Xu, Zaiyan, et al.
Published: (2025)
by: Xu, Zaiyan, et al.
Published: (2025)
Resource Rational Contractualism Should Guide AI Alignment
by: Levine, Sydney, et al.
Published: (2025)
by: Levine, Sydney, et al.
Published: (2025)
Evaluating Cognitive Age Alignment in Interactive AI Agents
by: Shen, Yifan, et al.
Published: (2026)
by: Shen, Yifan, et al.
Published: (2026)
Co-Alignment: Rethinking Alignment as Bidirectional Human-AI Cognitive Adaptation
by: Li, Yubo, et al.
Published: (2025)
by: Li, Yubo, et al.
Published: (2025)
MCWDST: a Minimum-Cost Weighted Directed Spanning Tree Algorithm for Real-Time Fake News Mitigation in Social Media
by: Truică, Ciprian-Octavian, et al.
Published: (2023)
by: Truică, Ciprian-Octavian, et al.
Published: (2023)
Adversarial Preference Learning for Robust LLM Alignment
by: Wang, Yuanfu, et al.
Published: (2025)
by: Wang, Yuanfu, et al.
Published: (2025)
Manifold Approximation leads to Robust Kernel Alignment
by: Islam, Mohammad Tariqul, et al.
Published: (2025)
by: Islam, Mohammad Tariqul, et al.
Published: (2025)
Doubly Robust Alignment for Large Language Models
by: Xu, Erhan, et al.
Published: (2025)
by: Xu, Erhan, et al.
Published: (2025)
Dialogical Reasoning Across AI Architectures: A Multi-Model Framework for Testing AI Alignment Strategies
by: Cox, Gray
Published: (2026)
by: Cox, Gray
Published: (2026)
Similar Items
-
Arrhythmia Classification from 12-Lead ECG Signals Using Convolutional and Transformer-Based Deep Learning Models
by: Apostol, Andrei, et al.
Published: (2025) -
Advancements in eHealth Data Analytics through Natural Language Processing and Deep Learning
by: Apostol, Elena-Simona, et al.
Published: (2024) -
CONTAIN: A Community-based Algorithm for Network Immunization
by: Apostol, Elena-Simona, et al.
Published: (2023) -
A Novel Reinforcement Learning Model for Post-Incident Malware Investigations
by: Dunsin, Dipo, et al.
Published: (2024) -
Clone-Robust AI Alignment
by: Procaccia, Ariel D., et al.
Published: (2025)