Making Models Unmergeable via Scaling-Sensitive Loss Landscape
Fuente:
arXiv
Saved in:
| Main Authors: | Jang, Minwoo, Kim, Hoyoung, Koo, Jabin, Ok, Jungseul |
|---|---|
| Format: | Preprint |
| Published: |
2026
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Federated Variational Preference Alignment with Gumbel-Softmax Prior for Personalized User Preferences
by: Koo, Jabin, et al.
Published: (2026)
by: Koo, Jabin, et al.
Published: (2026)
Towards Robust and Efficient Federated Low-Rank Adaptation with Heterogeneous Clients
by: Koo, Jabin, et al.
Published: (2024)
by: Koo, Jabin, et al.
Published: (2024)
LLM Watermark Evasion via Bias Inversion
by: Hwang, Jeongyeon, et al.
Published: (2025)
by: Hwang, Jeongyeon, et al.
Published: (2025)
ChimeraLoRA: Multi-Head LoRA-Guided Synthetic Datasets
by: Kim, Hoyoung, et al.
Published: (2026)
by: Kim, Hoyoung, et al.
Published: (2026)
Parallel Test-Time Scaling with Multi-Sequence Verifiers
by: Kim, Yegon, et al.
Published: (2026)
by: Kim, Yegon, et al.
Published: (2026)
The Attack and Defense Landscape of Agentic AI: A Comprehensive Survey
by: Kim, Juhee, et al.
Published: (2026)
by: Kim, Juhee, et al.
Published: (2026)
Lessons from Penetration Tests on Large-Scale Agent Systems
by: Eykholt, Kevin, et al.
Published: (2026)
by: Eykholt, Kevin, et al.
Published: (2026)
Quantifying Loss Aversion in Cyber Adversaries via LLM Analysis
by: Hans, Soham, et al.
Published: (2025)
by: Hans, Soham, et al.
Published: (2025)
Model Context Protocol (MCP): Landscape, Security Threats, and Future Research Directions
by: Hou, Xinyi, et al.
Published: (2025)
by: Hou, Xinyi, et al.
Published: (2025)
Making Them Ask and Answer: Jailbreaking Large Language Models in Few Queries via Disguise and Reconstruction
by: Liu, Tong, et al.
Published: (2024)
by: Liu, Tong, et al.
Published: (2024)
Conflicts Make Large Reasoning Models Vulnerable to Attacks
by: Liu, Honghao, et al.
Published: (2026)
by: Liu, Honghao, et al.
Published: (2026)
An Engorgio Prompt Makes Large Language Model Babble on
by: Dong, Jianshuo, et al.
Published: (2024)
by: Dong, Jianshuo, et al.
Published: (2024)
Rethinking Model Inversion Attacks With Patch-Wise Reconstruction
by: Jang, Jonggyu, et al.
Published: (2023)
by: Jang, Jonggyu, et al.
Published: (2023)
ATLANTIS: AI-driven Threat Localization, Analysis, and Triage Intelligence System
by: Kim, Taesoo, et al.
Published: (2025)
by: Kim, Taesoo, et al.
Published: (2025)
The Weight of a Bit: EMFI Sensitivity Analysis of Embedded Deep Learning Models
by: Breier, Jakub, et al.
Published: (2026)
by: Breier, Jakub, et al.
Published: (2026)
ThreatModeling-LLM: Automating Threat Modeling using Large Language Models for Banking System
by: Wu, Tingmin, et al.
Published: (2024)
by: Wu, Tingmin, et al.
Published: (2024)
Mapping LLM Security Landscapes: A Comprehensive Stakeholder Risk Assessment Proposal
by: Pankajakshan, Rahul, et al.
Published: (2024)
by: Pankajakshan, Rahul, et al.
Published: (2024)
SecInfer: Preventing Prompt Injection via Inference-time Scaling
by: Liu, Yupei, et al.
Published: (2025)
by: Liu, Yupei, et al.
Published: (2025)
Gradient Cuff: Detecting Jailbreak Attacks on Large Language Models by Exploring Refusal Loss Landscapes
by: Hu, Xiaomeng, et al.
Published: (2024)
by: Hu, Xiaomeng, et al.
Published: (2024)
SIExVulTS: Sensitive Information Exposure Vulnerability Detection System using Transformer Models and Static Analysis
by: Katz, Kyler, et al.
Published: (2025)
by: Katz, Kyler, et al.
Published: (2025)
Surveying the Operational Cybersecurity and Supply Chain Threat Landscape when Developing and Deploying AI Systems
by: Smith, Michael R, et al.
Published: (2025)
by: Smith, Michael R, et al.
Published: (2025)
CSF: Black-box Fingerprinting via Compositional Semantics for Text-to-Image Models
by: Lee, Junhoo, et al.
Published: (2026)
by: Lee, Junhoo, et al.
Published: (2026)
Can LLMs Make (Personalized) Access Control Decisions?
by: Groschupp, Friederike, et al.
Published: (2025)
by: Groschupp, Friederike, et al.
Published: (2025)
Scrub It Out! Erasing Sensitive Memorization in Code Language Models via Machine Unlearning
by: Chu, Zhaoyang, et al.
Published: (2025)
by: Chu, Zhaoyang, et al.
Published: (2025)
Detection and Analysis of Sensitive and Illegal Content on the Ethereum Blockchain Using Machine Learning Techniques
by: Feng, Xingyu
Published: (2025)
by: Feng, Xingyu
Published: (2025)
Make Split, not Hijack: Preventing Feature-Space Hijacking Attacks in Split Learning
by: Khan, Tanveer, et al.
Published: (2024)
by: Khan, Tanveer, et al.
Published: (2024)
GhostCite: A Large-Scale Analysis of Citation Validity in the Age of Large Language Models
by: Xu, Zuyao, et al.
Published: (2026)
by: Xu, Zuyao, et al.
Published: (2026)
The Mirror Design Pattern: Strict Data Geometry over Model Scale for Prompt Injection Detection
by: Corll, J Alex
Published: (2026)
by: Corll, J Alex
Published: (2026)
Make Your Home Safe: Time-aware Unsupervised User Behavior Anomaly Detection in Smart Homes via Loss-guided Mask
by: Xiao, Jingyu, et al.
Published: (2024)
by: Xiao, Jingyu, et al.
Published: (2024)
From Beats to Breaches:How Offensive AI Infers Sensitive User Information from Playlists
by: Cecconello, Stefano, et al.
Published: (2026)
by: Cecconello, Stefano, et al.
Published: (2026)
Trivial Trojans: How Minimal MCP Servers Enable Cross-Tool Exfiltration of Sensitive Data
by: Croce, Nicola, et al.
Published: (2025)
by: Croce, Nicola, et al.
Published: (2025)
Aero-LLM: A Distributed Framework for Secure UAV Communication and Intelligent Decision-Making
by: Dharmalingam, Balakrishnan, et al.
Published: (2025)
by: Dharmalingam, Balakrishnan, et al.
Published: (2025)
Hallucination-Resistant Security Planning with a Large Language Model
by: Hammar, Kim, et al.
Published: (2026)
by: Hammar, Kim, et al.
Published: (2026)
Payload-Aware Intrusion Detection with CMAE and Large Language Models
by: Kim, Yongcheol, et al.
Published: (2025)
by: Kim, Yongcheol, et al.
Published: (2025)
Scaling Homomorphic Applications in Deployment
by: Marinelli, Ryan, et al.
Published: (2025)
by: Marinelli, Ryan, et al.
Published: (2025)
Defending MoE LLMs against Harmful Fine-Tuning via Safety Routing Alignment
by: Kim, Jaehan, et al.
Published: (2025)
by: Kim, Jaehan, et al.
Published: (2025)
Argus: A Multi-Agent Sensitive Information Leakage Detection Framework Based on Hierarchical Reference Relationships
by: Wang, Bin, et al.
Published: (2025)
by: Wang, Bin, et al.
Published: (2025)
Enhancing Decision-Making in Windows PE Malware Classification During Dataset Shifts with Uncertainty Estimation
by: Yumlembam, Rahul, et al.
Published: (2025)
by: Yumlembam, Rahul, et al.
Published: (2025)
In-Context Autonomous Network Incident Response: An End-to-End Large Language Model Agent Approach
by: Gao, Yiran, et al.
Published: (2026)
by: Gao, Yiran, et al.
Published: (2026)
Silent Egress: When Implicit Prompt Injection Makes LLM Agents Leak Without a Trace
by: Lan, Qianlong, et al.
Published: (2026)
by: Lan, Qianlong, et al.
Published: (2026)
Similar Items
-
Federated Variational Preference Alignment with Gumbel-Softmax Prior for Personalized User Preferences
by: Koo, Jabin, et al.
Published: (2026) -
Towards Robust and Efficient Federated Low-Rank Adaptation with Heterogeneous Clients
by: Koo, Jabin, et al.
Published: (2024) -
LLM Watermark Evasion via Bias Inversion
by: Hwang, Jeongyeon, et al.
Published: (2025) -
ChimeraLoRA: Multi-Head LoRA-Guided Synthetic Datasets
by: Kim, Hoyoung, et al.
Published: (2026) -
Parallel Test-Time Scaling with Multi-Sequence Verifiers
by: Kim, Yegon, et al.
Published: (2026)