Consecutive Batch Model Editing with HooK Layers
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Li, Shuaiyi, Deng, Yang, Cai, Deng, Lu, Hongyuan, Chen, Liang, Lam, Wai |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2024
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
DepWiGNN: A Depth-wise Graph Neural Network for Multi-hop Spatial Reasoning in Text
von: Li, Shuaiyi, et al.
Veröffentlicht: (2023)
von: Li, Shuaiyi, et al.
Veröffentlicht: (2023)
InComeS: Integrating Compression and Selection Mechanisms into LLMs for Efficient Model Editing
von: Li, Shuaiyi, et al.
Veröffentlicht: (2025)
von: Li, Shuaiyi, et al.
Veröffentlicht: (2025)
Dictionary Insertion Prompting for Multilingual Reasoning on Multilingual Large Language Models
von: Lu, Hongyuan, et al.
Veröffentlicht: (2024)
von: Lu, Hongyuan, et al.
Veröffentlicht: (2024)
Clean Evaluations on Contaminated Visual Language Models
von: Lu, Hongyuan, et al.
Veröffentlicht: (2024)
von: Lu, Hongyuan, et al.
Veröffentlicht: (2024)
Towards Generalization of Block Attention via Automatic Segmentation and Block Distillation
von: Li, Shuaiyi, et al.
Veröffentlicht: (2026)
von: Li, Shuaiyi, et al.
Veröffentlicht: (2026)
SLoW: Select Low-frequency Words! Automatic Dictionary Selection for Translation on Large Language Models
von: Lu, Hongyuan, et al.
Veröffentlicht: (2025)
von: Lu, Hongyuan, et al.
Veröffentlicht: (2025)
WatME: Towards Lossless Watermarking Through Lexical Redundancy
von: Chen, Liang, et al.
Veröffentlicht: (2023)
von: Chen, Liang, et al.
Veröffentlicht: (2023)
InfiniteICL: Breaking the Limit of Context Window Size via Long Short-term Memory Transformation
von: Cao, Bowen, et al.
Veröffentlicht: (2025)
von: Cao, Bowen, et al.
Veröffentlicht: (2025)
Reasons to Reject? Aligning Language Models with Judgments
von: Xu, Weiwen, et al.
Veröffentlicht: (2023)
von: Xu, Weiwen, et al.
Veröffentlicht: (2023)
Chain-of-Dictionary Prompting Elicits Translation in Large Language Models
von: Lu, Hongyuan, et al.
Veröffentlicht: (2023)
von: Lu, Hongyuan, et al.
Veröffentlicht: (2023)
StrategyLLM: Large Language Models as Strategy Generators, Executors, Optimizers, and Evaluators for Problem Solving
von: Gao, Chang, et al.
Veröffentlicht: (2023)
von: Gao, Chang, et al.
Veröffentlicht: (2023)
Unveiling the Generalization Power of Fine-Tuned Large Language Models
von: Yang, Haoran, et al.
Veröffentlicht: (2024)
von: Yang, Haoran, et al.
Veröffentlicht: (2024)
On the Worst Prompt Performance of Large Language Models
von: Cao, Bowen, et al.
Veröffentlicht: (2024)
von: Cao, Bowen, et al.
Veröffentlicht: (2024)
A Frustratingly Simple Decoding Method for Neural Text Generation
von: Yang, Haoran, et al.
Veröffentlicht: (2023)
von: Yang, Haoran, et al.
Veröffentlicht: (2023)
LNE-Blocking: An Efficient Framework for Contamination Mitigation Evaluation on Large Language Models
von: Hou, Ruijie, et al.
Veröffentlicht: (2025)
von: Hou, Ruijie, et al.
Veröffentlicht: (2025)
Chain-of-Symbol Prompting Elicits Planning in Large Langauge Models
von: Hu, Hanxu, et al.
Veröffentlicht: (2023)
von: Hu, Hanxu, et al.
Veröffentlicht: (2023)
Not All Preference Pairs Are Created Equal: A Recipe for Annotation-Efficient Iterative Preference Learning
von: Yang, Sen, et al.
Veröffentlicht: (2024)
von: Yang, Sen, et al.
Veröffentlicht: (2024)
A Thorough Examination of Decoding Methods in the Era of LLMs
von: Shi, Chufan, et al.
Veröffentlicht: (2024)
von: Shi, Chufan, et al.
Veröffentlicht: (2024)
Knowledge Boundary of Large Language Models: A Survey
von: Li, Moxin, et al.
Veröffentlicht: (2024)
von: Li, Moxin, et al.
Veröffentlicht: (2024)
Plug-and-Play Policy Planner for Large Language Model Powered Dialogue Agents
von: Deng, Yang, et al.
Veröffentlicht: (2023)
von: Deng, Yang, et al.
Veröffentlicht: (2023)
Adam's Law: Textual Frequency Law on Large Language Models
von: Lu, Hongyuan Adam, et al.
Veröffentlicht: (2026)
von: Lu, Hongyuan Adam, et al.
Veröffentlicht: (2026)
Stable Knowledge Editing in Large Language Models
von: Wei, Zihao, et al.
Veröffentlicht: (2024)
von: Wei, Zihao, et al.
Veröffentlicht: (2024)
Attention Entropy is a Key Factor: An Analysis of Parallel Context Encoding with Full-attention-based Pre-trained Language Models
von: Zhang, Zhisong, et al.
Veröffentlicht: (2024)
von: Zhang, Zhisong, et al.
Veröffentlicht: (2024)
A Silver Bullet or a Compromise for Full Attention? A Comprehensive Study of Gist Token-based Context Compression
von: Deng, Chenlong, et al.
Veröffentlicht: (2024)
von: Deng, Chenlong, et al.
Veröffentlicht: (2024)
MLaKE: Multilingual Knowledge Editing Benchmark for Large Language Models
von: Wei, Zihao, et al.
Veröffentlicht: (2024)
von: Wei, Zihao, et al.
Veröffentlicht: (2024)
Stephanie: Step-by-Step Dialogues for Mimicking Human Interactions in Social Conversations
von: Yang, Hao, et al.
Veröffentlicht: (2024)
von: Yang, Hao, et al.
Veröffentlicht: (2024)
Everything is Editable: Extend Knowledge Editing to Unstructured Data in Large Language Models
von: Deng, Jingcheng, et al.
Veröffentlicht: (2024)
von: Deng, Jingcheng, et al.
Veröffentlicht: (2024)
O-Edit: Orthogonal Subspace Editing for Language Model Sequential Editing
von: Cai, Yuchen, et al.
Veröffentlicht: (2024)
von: Cai, Yuchen, et al.
Veröffentlicht: (2024)
UniGist: Towards General and Hardware-aligned Sequence-level Long Context Compression
von: Deng, Chenlong, et al.
Veröffentlicht: (2025)
von: Deng, Chenlong, et al.
Veröffentlicht: (2025)
PFME: A Modular Approach for Fine-grained Hallucination Detection and Editing of Large Language Models
von: Deng, Kunquan, et al.
Veröffentlicht: (2024)
von: Deng, Kunquan, et al.
Veröffentlicht: (2024)
Stephanie2: Thinking, Waiting, and Making Decisions Like Humans in Step-by-Step AI Social Chat
von: Yang, Hao, et al.
Veröffentlicht: (2026)
von: Yang, Hao, et al.
Veröffentlicht: (2026)
Consistency-Aware Parameter-Preserving Knowledge Editing Framework for Multi-Hop Question Answering
von: Deng, Lingwen, et al.
Veröffentlicht: (2025)
von: Deng, Lingwen, et al.
Veröffentlicht: (2025)
GLBench: A Comprehensive Benchmark for Graph with Large Language Models
von: Li, Yuhan, et al.
Veröffentlicht: (2024)
von: Li, Yuhan, et al.
Veröffentlicht: (2024)
Multi-LLM Collaborative Search for Complex Problem Solving
von: Yang, Sen, et al.
Veröffentlicht: (2025)
von: Yang, Sen, et al.
Veröffentlicht: (2025)
LiFi: Lightweight Controlled Text Generation with Fine-Grained Control Codes
von: Shi, Chufan, et al.
Veröffentlicht: (2024)
von: Shi, Chufan, et al.
Veröffentlicht: (2024)
Retrieval Backward Attention without Additional Training: Enhance Embeddings of Large Language Models via Repetition
von: Duan, Yifei, et al.
Veröffentlicht: (2025)
von: Duan, Yifei, et al.
Veröffentlicht: (2025)
JsonTuning: Towards Generalizable, Robust, and Controllable Instruction Tuning
von: Gao, Chang, et al.
Veröffentlicht: (2023)
von: Gao, Chang, et al.
Veröffentlicht: (2023)
RePO: Replay-Enhanced Policy Optimization
von: Li, Siheng, et al.
Veröffentlicht: (2025)
von: Li, Siheng, et al.
Veröffentlicht: (2025)
LLM2: Let Large Language Models Harness System 2 Reasoning
von: Yang, Cheng, et al.
Veröffentlicht: (2024)
von: Yang, Cheng, et al.
Veröffentlicht: (2024)
On the Transformations across Reward Model, Parameter Update, and In-Context Prompt
von: Cai, Deng, et al.
Veröffentlicht: (2024)
von: Cai, Deng, et al.
Veröffentlicht: (2024)
Ähnliche Einträge
-
DepWiGNN: A Depth-wise Graph Neural Network for Multi-hop Spatial Reasoning in Text
von: Li, Shuaiyi, et al.
Veröffentlicht: (2023) -
InComeS: Integrating Compression and Selection Mechanisms into LLMs for Efficient Model Editing
von: Li, Shuaiyi, et al.
Veröffentlicht: (2025) -
Dictionary Insertion Prompting for Multilingual Reasoning on Multilingual Large Language Models
von: Lu, Hongyuan, et al.
Veröffentlicht: (2024) -
Clean Evaluations on Contaminated Visual Language Models
von: Lu, Hongyuan, et al.
Veröffentlicht: (2024) -
Towards Generalization of Block Attention via Automatic Segmentation and Block Distillation
von: Li, Shuaiyi, et al.
Veröffentlicht: (2026)