WARP: Guaranteed Inner-Layer Repair of NLP Transformers
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Hsu, Hsin-Ling, Chen, Min-Yu, Chen, Nai-Chia, Chen, Yan-Ru, Chang, Yi-Ling, Yu, Fang |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2026
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
WARP: On the Benefits of Weight Averaged Rewarded Policies
von: Ramé, Alexandre, et al.
Veröffentlicht: (2024)
von: Ramé, Alexandre, et al.
Veröffentlicht: (2024)
Sparse Adapter Fusion for Continual Learning in NLP
von: Zeng, Min, et al.
Veröffentlicht: (2026)
von: Zeng, Min, et al.
Veröffentlicht: (2026)
DECRL: A Deep Evolutionary Clustering Jointed Temporal Knowledge Graph Representation Learning Approach
von: Chen, Qian, et al.
Veröffentlicht: (2024)
von: Chen, Qian, et al.
Veröffentlicht: (2024)
Jailbreaking with Universal Multi-Prompts
von: Hsu, Yu-Ling, et al.
Veröffentlicht: (2025)
von: Hsu, Yu-Ling, et al.
Veröffentlicht: (2025)
WARP: Weight Teleportation for Attack-Resilient Unlearning Protocols
von: Maheri, Mohammad M, et al.
Veröffentlicht: (2025)
von: Maheri, Mohammad M, et al.
Veröffentlicht: (2025)
Outlier-Efficient Hopfield Layers for Large Transformer-Based Models
von: Hu, Jerry Yao-Chieh, et al.
Veröffentlicht: (2024)
von: Hu, Jerry Yao-Chieh, et al.
Veröffentlicht: (2024)
Regret-Guided Search Control for Efficient Learning in AlphaZero
von: Tsai, Yun-Jui, et al.
Veröffentlicht: (2026)
von: Tsai, Yun-Jui, et al.
Veröffentlicht: (2026)
DAQ: Density-Aware Post-Training Weight-Only Quantization For LLMs
von: Luo, Yingsong, et al.
Veröffentlicht: (2024)
von: Luo, Yingsong, et al.
Veröffentlicht: (2024)
Inner-Instance Normalization for Time Series Forecasting
von: Jibao, Zipo, et al.
Veröffentlicht: (2025)
von: Jibao, Zipo, et al.
Veröffentlicht: (2025)
Beyond In-Domain Detection: SpikeScore for Cross-Domain Hallucination Detection
von: Deng, Yongxin, et al.
Veröffentlicht: (2026)
von: Deng, Yongxin, et al.
Veröffentlicht: (2026)
A Survey of Data Synthesis Approaches
von: Chang, Hsin-Yu, et al.
Veröffentlicht: (2024)
von: Chang, Hsin-Yu, et al.
Veröffentlicht: (2024)
AdaMixup: A Dynamic Defense Framework for Membership Inference Attack Mitigation
von: Chen, Ying, et al.
Veröffentlicht: (2025)
von: Chen, Ying, et al.
Veröffentlicht: (2025)
MIM-Reasoner: Learning with Theoretical Guarantees for Multiplex Influence Maximization
von: Do, Nguyen, et al.
Veröffentlicht: (2024)
von: Do, Nguyen, et al.
Veröffentlicht: (2024)
TAGA: Text-Attributed Graph Self-Supervised Learning by Synergizing Graph and Text Mutual Transformations
von: Zhang, Zheng, et al.
Veröffentlicht: (2024)
von: Zhang, Zheng, et al.
Veröffentlicht: (2024)
Demystifying MuZero Planning: Interpreting the Learned Model
von: Guei, Hung, et al.
Veröffentlicht: (2024)
von: Guei, Hung, et al.
Veröffentlicht: (2024)
Climate Downscaling: A Deep-Learning Based Super-resolution Model of Precipitation Data with Attention Block and Skip Connections
von: Chiang, Chia-Hao, et al.
Veröffentlicht: (2024)
von: Chiang, Chia-Hao, et al.
Veröffentlicht: (2024)
On-device AI: Quantization-aware Training of Transformers in Time-Series
von: Ling, Tianheng, et al.
Veröffentlicht: (2024)
von: Ling, Tianheng, et al.
Veröffentlicht: (2024)
Forecasting Unseen Points of Interest Visits Using Context and Proximity Priors
von: Li, Ziyao, et al.
Veröffentlicht: (2024)
von: Li, Ziyao, et al.
Veröffentlicht: (2024)
BitPipe: Bidirectional Interleaved Pipeline Parallelism for Accelerating Large Models Training
von: Wu, Houming, et al.
Veröffentlicht: (2024)
von: Wu, Houming, et al.
Veröffentlicht: (2024)
LoopQ: Quantization for Recursive Transformers
von: Fang, Rui, et al.
Veröffentlicht: (2026)
von: Fang, Rui, et al.
Veröffentlicht: (2026)
Fate: Fast Edge Inference of Mixture-of-Experts Models via Cross-Layer Gate
von: Fang, Zhiyuan, et al.
Veröffentlicht: (2025)
von: Fang, Zhiyuan, et al.
Veröffentlicht: (2025)
Explainable LLM Unlearning Through Reasoning
von: Liao, Junfeng, et al.
Veröffentlicht: (2026)
von: Liao, Junfeng, et al.
Veröffentlicht: (2026)
Data-Driven Lipschitz Continuity: A Cost-Effective Approach to Improve Adversarial Robustness
von: Chen, Erh-Chung, et al.
Veröffentlicht: (2024)
von: Chen, Erh-Chung, et al.
Veröffentlicht: (2024)
Generalized Phase Pressure Control Enhanced Reinforcement Learning for Traffic Signal Control
von: Liao, Xiao-Cheng, et al.
Veröffentlicht: (2025)
von: Liao, Xiao-Cheng, et al.
Veröffentlicht: (2025)
Theoretical Insights in Model Inversion Robustness and Conditional Entropy Maximization for Collaborative Inference Systems
von: Xia, Song, et al.
Veröffentlicht: (2025)
von: Xia, Song, et al.
Veröffentlicht: (2025)
Curriculum Learning with Quality-Driven Data Selection
von: Wu, Biao, et al.
Veröffentlicht: (2024)
von: Wu, Biao, et al.
Veröffentlicht: (2024)
Diagonal Adaptive Non-local Observables on Quantum Neural Networks
von: Tseng, Huan-Hsin, et al.
Veröffentlicht: (2026)
von: Tseng, Huan-Hsin, et al.
Veröffentlicht: (2026)
POIFormer: A Transformer-Based Framework for Accurate and Scalable Point-of-Interest Attribution
von: Saxena, Nripsuta Ani, et al.
Veröffentlicht: (2025)
von: Saxena, Nripsuta Ani, et al.
Veröffentlicht: (2025)
Open-Vocabulary Panoptic Segmentation Using BERT Pre-Training of Vision-Language Multiway Transformer Model
von: Chen, Yi-Chia, et al.
Veröffentlicht: (2024)
von: Chen, Yi-Chia, et al.
Veröffentlicht: (2024)
Sample-based Dynamic Hierarchical Transformer with Layer and Head Flexibility via Contextual Bandit
von: Meng, Fanfei, et al.
Veröffentlicht: (2023)
von: Meng, Fanfei, et al.
Veröffentlicht: (2023)
Material Property Prediction with Element Attribute Knowledge Graphs and Multimodal Representation Learning
von: Huang, Chao, et al.
Veröffentlicht: (2024)
von: Huang, Chao, et al.
Veröffentlicht: (2024)
HIT-ROCKET: Hadamard-vector Inner-product Transformer for ROCKET
von: Hao, Wang, et al.
Veröffentlicht: (2025)
von: Hao, Wang, et al.
Veröffentlicht: (2025)
From Pruning to Grafting: Dynamic Knowledge Redistribution via Learnable Layer Fusion
von: Pei, Zehua, et al.
Veröffentlicht: (2024)
von: Pei, Zehua, et al.
Veröffentlicht: (2024)
MOSS: Efficient and Accurate FP8 LLM Training with Microscaling and Automatic Scaling
von: Zhang, Yu, et al.
Veröffentlicht: (2025)
von: Zhang, Yu, et al.
Veröffentlicht: (2025)
Meta-DiffuB: A Contextualized Sequence-to-Sequence Text Diffusion Model with Meta-Exploration
von: Chuang, Yun-Yen, et al.
Veröffentlicht: (2024)
von: Chuang, Yun-Yen, et al.
Veröffentlicht: (2024)
Integer-only Quantized Transformers for Embedded FPGA-based Time-series Forecasting in AIoT
von: Ling, Tianheng, et al.
Veröffentlicht: (2024)
von: Ling, Tianheng, et al.
Veröffentlicht: (2024)
Spatial-temporal Graph Convolutional Networks with Diversified Transformation for Dynamic Graph Representation Learning
von: Wang, Ling, et al.
Veröffentlicht: (2024)
von: Wang, Ling, et al.
Veröffentlicht: (2024)
ST-Hyper: Learning High-Order Dependencies Across Multiple Spatial-Temporal Scales for Multivariate Time Series Forecasting
von: Wu, Binqing, et al.
Veröffentlicht: (2025)
von: Wu, Binqing, et al.
Veröffentlicht: (2025)
AnomalyGFM: Graph Foundation Model for Zero/Few-shot Anomaly Detection
von: Qiao, Hezhe, et al.
Veröffentlicht: (2025)
von: Qiao, Hezhe, et al.
Veröffentlicht: (2025)
MillGNN: Learning Multi-Scale Lead-Lag Dependencies for Multi-Variate Time Series Forecasting
von: Wu, Binqing, et al.
Veröffentlicht: (2025)
von: Wu, Binqing, et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
WARP: On the Benefits of Weight Averaged Rewarded Policies
von: Ramé, Alexandre, et al.
Veröffentlicht: (2024) -
Sparse Adapter Fusion for Continual Learning in NLP
von: Zeng, Min, et al.
Veröffentlicht: (2026) -
DECRL: A Deep Evolutionary Clustering Jointed Temporal Knowledge Graph Representation Learning Approach
von: Chen, Qian, et al.
Veröffentlicht: (2024) -
Jailbreaking with Universal Multi-Prompts
von: Hsu, Yu-Ling, et al.
Veröffentlicht: (2025) -
WARP: Weight Teleportation for Attack-Resilient Unlearning Protocols
von: Maheri, Mohammad M, et al.
Veröffentlicht: (2025)