Supervised and Unsupervised Alignments for Spoofing Behavioral Biometrics
Fuente:
arXiv
Guardado en:
| Autores principales: | Thebaud, Thomas, Lan, Gaël Le, Larcher, Anthony |
|---|---|
| Formato: | Preprint |
| Publicado: |
2024
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
Enhancing LLM Watermark Resilience Against Both Scrubbing and Spoofing Attacks
por: Shen, Huanming, et al.
Publicado: (2025)
por: Shen, Huanming, et al.
Publicado: (2025)
A Survey of Threats Against Voice Authentication and Anti-Spoofing Systems
por: Kamel, Kamel, et al.
Publicado: (2025)
por: Kamel, Kamel, et al.
Publicado: (2025)
DITTO: A Spoofing Attack Framework on Watermarked LLMs via Knowledge Distillation
por: An, Hyeseon, et al.
Publicado: (2025)
por: An, Hyeseon, et al.
Publicado: (2025)
Experimental Validation of Sensor Fusion-based GNSS Spoofing Attack Detection Framework for Autonomous Vehicles
por: Dasgupta, Sagar, et al.
Publicado: (2024)
por: Dasgupta, Sagar, et al.
Publicado: (2024)
Discovering Spoofing Attempts on Language Model Watermarks
por: Gloaguen, Thibaud, et al.
Publicado: (2024)
por: Gloaguen, Thibaud, et al.
Publicado: (2024)
GPS Spoofing Attack Detection in Autonomous Vehicles Using Adaptive DBSCAN
por: Mohammadi, Ahmad, et al.
Publicado: (2025)
por: Mohammadi, Ahmad, et al.
Publicado: (2025)
Entropy-Synchronized Neural Hashing for Unsupervised Ransomware Detection
por: Idliman, Peter, et al.
Publicado: (2025)
por: Idliman, Peter, et al.
Publicado: (2025)
Reimagining Safety Alignment with An Image
por: Xia, Yifan, et al.
Publicado: (2025)
por: Xia, Yifan, et al.
Publicado: (2025)
Towards Unsupervised Adversarial Document Detection in Retrieval Augmented Generation Systems
por: Levi, Patrick
Publicado: (2026)
por: Levi, Patrick
Publicado: (2026)
Unsupervised Threat Hunting using Continuous Bag-of-Terms-and-Time (CBoTT)
por: Kayhan, Varol, et al.
Publicado: (2024)
por: Kayhan, Varol, et al.
Publicado: (2024)
Agent Safety Alignment via Reinforcement Learning
por: Sha, Zeyang, et al.
Publicado: (2025)
por: Sha, Zeyang, et al.
Publicado: (2025)
EVA: Editing for Versatile Alignment against Jailbreaks
por: Wang, Yi, et al.
Publicado: (2026)
por: Wang, Yi, et al.
Publicado: (2026)
UK AISI Alignment Evaluation Case-Study
por: Souly, Alexandra, et al.
Publicado: (2026)
por: Souly, Alexandra, et al.
Publicado: (2026)
Adversarial Evasion in Non-Stationary Malware Detection: Minimizing Drift Signals through Similarity-Constrained Perturbations
por: Acharya, Pawan, et al.
Publicado: (2026)
por: Acharya, Pawan, et al.
Publicado: (2026)
The Cognitive Firewall:Securing Browser Based AI Agents Against Indirect Prompt Injection Via Hybrid Edge Cloud Defense
por: Lan, Qianlong, et al.
Publicado: (2026)
por: Lan, Qianlong, et al.
Publicado: (2026)
Targeting Alignment: Extracting Safety Classifiers of Aligned LLMs
por: Ferrand, Jean-Charles Noirot, et al.
Publicado: (2025)
por: Ferrand, Jean-Charles Noirot, et al.
Publicado: (2025)
Measuring Safety Alignment Effects in Autonomous Security Agents
por: David, Isaac, et al.
Publicado: (2026)
por: David, Isaac, et al.
Publicado: (2026)
VisuoAlign: Safety Alignment of LVLMs with Multimodal Tree Search
por: Li, MingSheng, et al.
Publicado: (2025)
por: Li, MingSheng, et al.
Publicado: (2025)
FreakOut-LLM: The Effect of Emotional Stimuli on Safety Alignment
por: Kuznetsov, Daniel, et al.
Publicado: (2026)
por: Kuznetsov, Daniel, et al.
Publicado: (2026)
LLM-Driven Feature-Level Adversarial Attacks on Android Malware Detectors
por: Lan, Tianwei, et al.
Publicado: (2025)
por: Lan, Tianwei, et al.
Publicado: (2025)
Sequential Behavioral Watermarking for LLM Agents
por: An, Hyeseon, et al.
Publicado: (2026)
por: An, Hyeseon, et al.
Publicado: (2026)
Identification of Malicious Posts on the Dark Web Using Supervised Machine Learning
por: Filho, Sebastião Alves de Jesus, et al.
Publicado: (2025)
por: Filho, Sebastião Alves de Jesus, et al.
Publicado: (2025)
SpoofTrackBench: Interpretable AI for Spoof-Aware UAV Tracking and Benchmarking
por: Le, Van, et al.
Publicado: (2025)
por: Le, Van, et al.
Publicado: (2025)
Medical Multimodal Model Stealing Attacks via Adversarial Domain Alignment
por: Shen, Yaling, et al.
Publicado: (2025)
por: Shen, Yaling, et al.
Publicado: (2025)
Ablating Safety: Mechanisms for Removing Alignment in Language Models for Security Applications
por: David, Isaac, et al.
Publicado: (2026)
por: David, Isaac, et al.
Publicado: (2026)
Matching Ranks Over Probability Yields Truly Deep Safety Alignment
por: Vega, Jason, et al.
Publicado: (2025)
por: Vega, Jason, et al.
Publicado: (2025)
Defensive Refusal Bias: How Safety Alignment Fails Cyber Defenders
por: Campbell, David, et al.
Publicado: (2026)
por: Campbell, David, et al.
Publicado: (2026)
PRISM: Robust VLM Alignment with Principled Reasoning for Integrated Safety in Multimodality
por: Li, Nanxi, et al.
Publicado: (2025)
por: Li, Nanxi, et al.
Publicado: (2025)
Co-Evolutionary Multi-Modal Alignment via Structured Adversarial Evolution
por: Shi, Guoxin, et al.
Publicado: (2026)
por: Shi, Guoxin, et al.
Publicado: (2026)
When Alignment Isn't Enough: Response-Path Attacks on LLM Agents
por: Luo, Mingyu, et al.
Publicado: (2026)
por: Luo, Mingyu, et al.
Publicado: (2026)
Biometrics in Extended Reality: A Review
por: Agarwal, Ayush, et al.
Publicado: (2024)
por: Agarwal, Ayush, et al.
Publicado: (2024)
Refusal Falls off a Cliff: How Safety Alignment Fails in Reasoning?
por: Yin, Qingyu, et al.
Publicado: (2025)
por: Yin, Qingyu, et al.
Publicado: (2025)
On the Impossibility of Separating Intelligence from Judgment: The Computational Intractability of Filtering for AI Alignment
por: Ball, Sarah, et al.
Publicado: (2025)
por: Ball, Sarah, et al.
Publicado: (2025)
SafeThinker: Reasoning about Risk to Deepen Safety Beyond Shallow Alignment
por: Fang, Xianya, et al.
Publicado: (2026)
por: Fang, Xianya, et al.
Publicado: (2026)
Generalizing Speaker Verification for Spoof Awareness in the Embedding Space
por: Liu, Xuechen, et al.
Publicado: (2024)
por: Liu, Xuechen, et al.
Publicado: (2024)
AgentMark: Utility-Preserving Behavioral Watermarking for Agents
por: Huang, Kaibo, et al.
Publicado: (2026)
por: Huang, Kaibo, et al.
Publicado: (2026)
An Investigation into the Performance of Non-Contrastive Self-Supervised Learning Methods for Network Intrusion Detection
por: Fard, Hamed, et al.
Publicado: (2025)
por: Fard, Hamed, et al.
Publicado: (2025)
Safety Alignment Should Be Made More Than Just a Few Tokens Deep
por: Qi, Xiangyu, et al.
Publicado: (2024)
por: Qi, Xiangyu, et al.
Publicado: (2024)
Medoid Prototype Alignment for Cross-Plant Unknown Attack Detection in Industrial Control Systems
por: Wang, Luyao
Publicado: (2026)
por: Wang, Luyao
Publicado: (2026)
MTSA: Multi-turn Safety Alignment for LLMs through Multi-round Red-teaming
por: Guo, Weiyang, et al.
Publicado: (2025)
por: Guo, Weiyang, et al.
Publicado: (2025)
Ejemplares similares
-
Enhancing LLM Watermark Resilience Against Both Scrubbing and Spoofing Attacks
por: Shen, Huanming, et al.
Publicado: (2025) -
A Survey of Threats Against Voice Authentication and Anti-Spoofing Systems
por: Kamel, Kamel, et al.
Publicado: (2025) -
DITTO: A Spoofing Attack Framework on Watermarked LLMs via Knowledge Distillation
por: An, Hyeseon, et al.
Publicado: (2025) -
Experimental Validation of Sensor Fusion-based GNSS Spoofing Attack Detection Framework for Autonomous Vehicles
por: Dasgupta, Sagar, et al.
Publicado: (2024) -
Discovering Spoofing Attempts on Language Model Watermarks
por: Gloaguen, Thibaud, et al.
Publicado: (2024)