Identity Lock: Locking API Fine-tuned LLMs With Identity-based Wake Words
Fuente:
arXiv
Saved in:
| Main Authors: | Su, Hongyu, Gao, Yifeng, Ding, Yifan, Ma, Xingjun |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
ADMIT: Few-shot Knowledge Poisoning Attacks on RAG-based Fact Checking
by: Wu, Yutao, et al.
Published: (2025)
by: Wu, Yutao, et al.
Published: (2025)
Mind the (Belief) Gap: Group Identity in the World of LLMs
by: Borah, Angana, et al.
Published: (2025)
by: Borah, Angana, et al.
Published: (2025)
Fine-tuning and Utilization Methods of Domain-specific LLMs
by: Jeong, Cheonsu
Published: (2024)
by: Jeong, Cheonsu
Published: (2024)
API-BLEND: A Comprehensive Corpora for Training and Benchmarking API LLMs
by: Basu, Kinjal, et al.
Published: (2024)
by: Basu, Kinjal, et al.
Published: (2024)
Aloe: A Family of Fine-tuned Open Healthcare LLMs
by: Gururajan, Ashwin Kumar, et al.
Published: (2024)
by: Gururajan, Ashwin Kumar, et al.
Published: (2024)
Towards Context-Invariant Safety Alignment for Large Language Models
by: Wang, Yixu, et al.
Published: (2026)
by: Wang, Yixu, et al.
Published: (2026)
The Lock-in Hypothesis: Stagnation by Algorithm
by: Qiu, Tianyi Alex, et al.
Published: (2025)
by: Qiu, Tianyi Alex, et al.
Published: (2025)
ITERTL: An Iterative Framework for Fine-tuning LLMs for RTL Code Generation
by: Wu, Peiyang, et al.
Published: (2024)
by: Wu, Peiyang, et al.
Published: (2024)
PaperAsk: A Benchmark for Reliability Evaluation of LLMs in Paper Search and Reading
by: Wu, Yutao, et al.
Published: (2025)
by: Wu, Yutao, et al.
Published: (2025)
Affordably Fine-tuned LLMs Provide Better Answers to Course-specific MCQs
by: Raimondi, Bianca, et al.
Published: (2025)
by: Raimondi, Bianca, et al.
Published: (2025)
RepCali: High Efficient Fine-tuning Via Representation Calibration in Latent Space for Pre-trained Language Models
by: Zhang, Fujun, et al.
Published: (2025)
by: Zhang, Fujun, et al.
Published: (2025)
BackdoorAgent: A Unified Framework for Backdoor Attacks on LLM-based Agents
by: Feng, Yunhao, et al.
Published: (2026)
by: Feng, Yunhao, et al.
Published: (2026)
Fine-tuned Large Language Models (LLMs): Improved Prompt Injection Attacks Detection
by: Rahman, Md Abdur, et al.
Published: (2024)
by: Rahman, Md Abdur, et al.
Published: (2024)
Principled Fine-tuning of LLMs from User-Edits: A Medley of Preference, Supervision, and Reward
by: Misra, Dipendra, et al.
Published: (2026)
by: Misra, Dipendra, et al.
Published: (2026)
TREX: Automating LLM Fine-tuning via Agent-Driven Tree-based Exploration
by: Ma, Zerun, et al.
Published: (2026)
by: Ma, Zerun, et al.
Published: (2026)
Disentangling Hate Across Target Identities
by: Jin, Yiping, et al.
Published: (2024)
by: Jin, Yiping, et al.
Published: (2024)
Maastricht University at AMIYA: Adapting LLMs for Dialectal Arabic using Fine-tuning and MBR Decoding
by: Alali, Abdulhai, et al.
Published: (2026)
by: Alali, Abdulhai, et al.
Published: (2026)
When MOE Meets LLMs: Parameter Efficient Fine-tuning for Multi-task Medical Applications
by: Liu, Qidong, et al.
Published: (2023)
by: Liu, Qidong, et al.
Published: (2023)
"OK Aura, Be Fair With Me": Demographics-Agnostic Training for Bias Mitigation in Wake-up Word Detection
by: López, Fernando, et al.
Published: (2026)
by: López, Fernando, et al.
Published: (2026)
Progtuning: Progressive Fine-tuning Framework for Transformer-based Language Models
by: Ji, Xiaoshuang, et al.
Published: (2025)
by: Ji, Xiaoshuang, et al.
Published: (2025)
Fine-tuning Transformer-based Encoder for Turkish Language Understanding Tasks
by: Yildirim, Savas
Published: (2024)
by: Yildirim, Savas
Published: (2024)
Small Updates, Big Doubts: Does Parameter-Efficient Fine-tuning Enhance Hallucination Detection ?
by: Hu, Xu, et al.
Published: (2026)
by: Hu, Xu, et al.
Published: (2026)
FineSurE: Fine-grained Summarization Evaluation using LLMs
by: Song, Hwanjun, et al.
Published: (2024)
by: Song, Hwanjun, et al.
Published: (2024)
MTUncertainty: Assessing the Need for Post-editing of Machine Translation Outputs by Fine-tuning OpenAI LLMs
by: Gladkoff, Serge, et al.
Published: (2023)
by: Gladkoff, Serge, et al.
Published: (2023)
NESTFUL: A Benchmark for Evaluating LLMs on Nested Sequences of API Calls
by: Basu, Kinjal, et al.
Published: (2024)
by: Basu, Kinjal, et al.
Published: (2024)
Word Form Matters: LLMs' Semantic Reconstruction under Typoglycemia
by: Wang, Chenxi, et al.
Published: (2025)
by: Wang, Chenxi, et al.
Published: (2025)
AccLock: Unlocking Identity with Heartbeat Using In-Ear Accelerometers
by: Wang, Lei, et al.
Published: (2026)
by: Wang, Lei, et al.
Published: (2026)
Multitask Mayhem: Unveiling and Mitigating Safety Gaps in LLMs Fine-tuning
by: Jan, Essa, et al.
Published: (2024)
by: Jan, Essa, et al.
Published: (2024)
RLSF: Fine-tuning LLMs via Symbolic Feedback
by: Jha, Piyush, et al.
Published: (2024)
by: Jha, Piyush, et al.
Published: (2024)
Locking Down the Finetuned LLMs Safety
by: Zhu, Minjun, et al.
Published: (2024)
by: Zhu, Minjun, et al.
Published: (2024)
Evaluating LLMs on Sequential API Call Through Automated Test Generation
by: Huang, Yuheng, et al.
Published: (2025)
by: Huang, Yuheng, et al.
Published: (2025)
Schema Key Wording as an Instruction Channel in Structured Generation under Constrained Decoding
by: Le, Yifan
Published: (2026)
by: Le, Yifan
Published: (2026)
KInIT at SemEval-2024 Task 8: Fine-tuned LLMs for Multilingual Machine-Generated Text Detection
by: Spiegel, Michal, et al.
Published: (2024)
by: Spiegel, Michal, et al.
Published: (2024)
Retrieval-Augmented Generation Systems for Intellectual Property via Synthetic Multi-Angle Fine-tuning
by: Ren, Runtao, et al.
Published: (2025)
by: Ren, Runtao, et al.
Published: (2025)
RobustFT: Robust Supervised Fine-tuning for Large Language Models under Noisy Response
by: Luo, Junyu, et al.
Published: (2024)
by: Luo, Junyu, et al.
Published: (2024)
Whose Journey Matters? Investigating Identity Biases in Large Language Models (LLMs) for Travel Planning Assistance
by: Ren, Ruiping, et al.
Published: (2024)
by: Ren, Ruiping, et al.
Published: (2024)
ChatKBQA: A Generate-then-Retrieve Framework for Knowledge Base Question Answering with Fine-tuned Large Language Models
by: Luo, Haoran, et al.
Published: (2023)
by: Luo, Haoran, et al.
Published: (2023)
Outlier-weighed Layerwise Sampling for LLM Fine-tuning
by: Li, Pengxiang, et al.
Published: (2024)
by: Li, Pengxiang, et al.
Published: (2024)
Keeping LLMs Aligned After Fine-tuning: The Crucial Role of Prompt Templates
by: Lyu, Kaifeng, et al.
Published: (2024)
by: Lyu, Kaifeng, et al.
Published: (2024)
Polysemanticity or Polysemy? Lexical Identity Confounds Superposition Metrics
by: Hou, Iyad Ait, et al.
Published: (2026)
by: Hou, Iyad Ait, et al.
Published: (2026)
Similar Items
-
ADMIT: Few-shot Knowledge Poisoning Attacks on RAG-based Fact Checking
by: Wu, Yutao, et al.
Published: (2025) -
Mind the (Belief) Gap: Group Identity in the World of LLMs
by: Borah, Angana, et al.
Published: (2025) -
Fine-tuning and Utilization Methods of Domain-specific LLMs
by: Jeong, Cheonsu
Published: (2024) -
API-BLEND: A Comprehensive Corpora for Training and Benchmarking API LLMs
by: Basu, Kinjal, et al.
Published: (2024) -
Aloe: A Family of Fine-tuned Open Healthcare LLMs
by: Gururajan, Ashwin Kumar, et al.
Published: (2024)