Are Robust LLM Fingerprints Adversarially Robust?
Fuente:
arXiv
Saved in:
| Main Authors: | Nasery, Anshul, Contente, Edoardo, Kaz, Alkin, Viswanath, Pramod, Oh, Sewoong |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Scalable Fingerprinting of Large Language Models
by: Nasery, Anshul, et al.
Published: (2025)
by: Nasery, Anshul, et al.
Published: (2025)
OML: A Primitive for Reconciling Open Access with Owner Control in AI Model Distribution
by: Cheng, Zerui, et al.
Published: (2024)
by: Cheng, Zerui, et al.
Published: (2024)
Training AI to be Loyal
by: Oh, Sewoong, et al.
Published: (2025)
by: Oh, Sewoong, et al.
Published: (2025)
SELF: A Robust Singular Value and Eigenvalue Approach for LLM Fingerprinting
by: Zhang, Hanxiu, et al.
Published: (2025)
by: Zhang, Hanxiu, et al.
Published: (2025)
Investigating the Impact of Quantization on Adversarial Robustness
by: Li, Qun, et al.
Published: (2024)
by: Li, Qun, et al.
Published: (2024)
On the Robustness of Bayesian Neural Networks to Adversarial Attacks
by: Bortolussi, Luca, et al.
Published: (2022)
by: Bortolussi, Luca, et al.
Published: (2022)
Variational Randomized Smoothing for Sample-Wise Adversarial Robustness
by: Hase, Ryo, et al.
Published: (2024)
by: Hase, Ryo, et al.
Published: (2024)
Robust and Reliable Early-Stage Website Fingerprinting Attacks via Spatial-Temporal Distribution Analysis
by: Deng, Xinhao, et al.
Published: (2024)
by: Deng, Xinhao, et al.
Published: (2024)
Attacks and Defenses Against LLM Fingerprinting
by: Kurian, Kevin, et al.
Published: (2025)
by: Kurian, Kevin, et al.
Published: (2025)
Characterizing the Training Dynamics of Private Fine-tuning with Langevin diffusion
by: Ke, Shuqi, et al.
Published: (2024)
by: Ke, Shuqi, et al.
Published: (2024)
Explainable Transformer-Based Email Phishing Classification with Adversarial Robustness
by: P, Sajad U
Published: (2025)
by: P, Sajad U
Published: (2025)
Empirical Analysis of Adversarial Robustness and Explainability Drift in Cybersecurity Classifiers
by: Rajhans, Mona, et al.
Published: (2026)
by: Rajhans, Mona, et al.
Published: (2026)
Complexity Matters: Effective Dimensionality as a Measure for Adversarial Robustness
by: Khachaturov, David, et al.
Published: (2024)
by: Khachaturov, David, et al.
Published: (2024)
Adversarial Robustness in Financial Machine Learning: Defenses, Economic Impact, and Governance Evidence
by: Baviskar, Samruddhi
Published: (2025)
by: Baviskar, Samruddhi
Published: (2025)
VISAT: Benchmarking Adversarial and Distribution Shift Robustness in Traffic Sign Recognition with Visual Attributes
by: Yu, Simon, et al.
Published: (2025)
by: Yu, Simon, et al.
Published: (2025)
Enhance DNN Adversarial Robustness and Efficiency via Injecting Noise to Non-Essential Neurons
by: Liu, Zhenyu, et al.
Published: (2024)
by: Liu, Zhenyu, et al.
Published: (2024)
A Survey on Adversarial Robustness of LiDAR-based Machine Learning Perception in Autonomous Vehicles
by: Kim, Junae, et al.
Published: (2024)
by: Kim, Junae, et al.
Published: (2024)
AEGIS: Adversarial Target-Guided Retention-Data-Free Robust Concept Erasure from Diffusion Models
by: Li, Fengpeng, et al.
Published: (2026)
by: Li, Fengpeng, et al.
Published: (2026)
Robustness of LLM-enabled vehicle trajectory prediction under data security threats
by: Wang, Feilong, et al.
Published: (2025)
by: Wang, Feilong, et al.
Published: (2025)
Towards Robust Stability Prediction in Smart Grids: GAN-based Approach under Data Constraints and Adversarial Challenges
by: Efatinasab, Emad, et al.
Published: (2025)
by: Efatinasab, Emad, et al.
Published: (2025)
On Adversarial Robustness of Language Models in Transfer Learning
by: Turbal, Bohdan, et al.
Published: (2024)
by: Turbal, Bohdan, et al.
Published: (2024)
Robust Privacy: Inference-Time Privacy through Certified Robustness
by: Jin, Jiankai, et al.
Published: (2026)
by: Jin, Jiankai, et al.
Published: (2026)
Towards Robust Knowledge Unlearning: An Adversarial Framework for Assessing and Improving Unlearning Robustness in Large Language Models
by: Yuan, Hongbang, et al.
Published: (2024)
by: Yuan, Hongbang, et al.
Published: (2024)
TAO: Tolerance-Aware Optimistic Verification for Floating-Point Neural Networks
by: Yao, Jianzhu, et al.
Published: (2025)
by: Yao, Jianzhu, et al.
Published: (2025)
Prompt2Fingerprint: Plug-and-Play LLM Fingerprinting via Text-to-Weight Generation
by: Chen, Sixu, et al.
Published: (2026)
by: Chen, Sixu, et al.
Published: (2026)
Adversarial Augmentation and Active Sampling for Robust Cyber Anomaly Detection
by: Benabderrahmane, Sidahmed, et al.
Published: (2025)
by: Benabderrahmane, Sidahmed, et al.
Published: (2025)
Wolfpack Adversarial Attack for Robust Multi-Agent Reinforcement Learning
by: Lee, Sunwoo, et al.
Published: (2025)
by: Lee, Sunwoo, et al.
Published: (2025)
Enhancing Cloud Network Resilience via a Robust LLM-Empowered Multi-Agent Reinforcement Learning Framework
by: Peng, Yixiao, et al.
Published: (2026)
by: Peng, Yixiao, et al.
Published: (2026)
Mixture of Robust Experts (MoRE):A Robust Denoising Method towards multiple perturbations
by: Cheng, Hao, et al.
Published: (2021)
by: Cheng, Hao, et al.
Published: (2021)
FPEdit: Robust LLM Fingerprinting through Localized Parameter Editing
by: Wang, Shida, et al.
Published: (2025)
by: Wang, Shida, et al.
Published: (2025)
Robustness of Agentic AI Systems via Adversarially-Aligned Jacobian Regularization
by: Mumcu, Furkan, et al.
Published: (2026)
by: Mumcu, Furkan, et al.
Published: (2026)
Probing Latent Subspaces in LLM for AI Security: Identifying and Manipulating Adversarial States
by: Chia, Xin Wei, et al.
Published: (2025)
by: Chia, Xin Wei, et al.
Published: (2025)
Watch your steps: Dormant Adversarial Behaviors that Activate upon LLM Finetuning
by: Gloaguen, Thibaud, et al.
Published: (2025)
by: Gloaguen, Thibaud, et al.
Published: (2025)
Attacking Byzantine Robust Aggregation in High Dimensions
by: Choudhary, Sarthak, et al.
Published: (2023)
by: Choudhary, Sarthak, et al.
Published: (2023)
Bridging Privacy and Robustness for Trustworthy Machine Learning
by: Zhang, Xiaojin, et al.
Published: (2024)
by: Zhang, Xiaojin, et al.
Published: (2024)
Soft-Label Integration for Robust Toxicity Classification
by: Cheng, Zelei, et al.
Published: (2024)
by: Cheng, Zelei, et al.
Published: (2024)
Enhancing Variational Autoencoders with Smooth Robust Latent Encoding
by: Lee, Hyomin, et al.
Published: (2025)
by: Lee, Hyomin, et al.
Published: (2025)
Linearizing Models for Efficient yet Robust Private Inference
by: Sarkar, Sreetama, et al.
Published: (2024)
by: Sarkar, Sreetama, et al.
Published: (2024)
Unforgeable Watermarks for Language Models via Robust Signatures
by: Lin, Huijia, et al.
Published: (2026)
by: Lin, Huijia, et al.
Published: (2026)
FLARE: A Wireless Side-Channel Fingerprinting Attack on Federated Learning
by: Shuvo, Md Nahid Hasan, et al.
Published: (2025)
by: Shuvo, Md Nahid Hasan, et al.
Published: (2025)
Similar Items
-
Scalable Fingerprinting of Large Language Models
by: Nasery, Anshul, et al.
Published: (2025) -
OML: A Primitive for Reconciling Open Access with Owner Control in AI Model Distribution
by: Cheng, Zerui, et al.
Published: (2024) -
Training AI to be Loyal
by: Oh, Sewoong, et al.
Published: (2025) -
SELF: A Robust Singular Value and Eigenvalue Approach for LLM Fingerprinting
by: Zhang, Hanxiu, et al.
Published: (2025) -
Investigating the Impact of Quantization on Adversarial Robustness
by: Li, Qun, et al.
Published: (2024)