Generalization or Memorization: Dynamic Decoding for Mode Steering
Fuente:
arXiv
Gespeichert in:
| 1. Verfasser: | Zhang, Xuanming |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2025
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Mitigating Memorization in LLMs using Activation Steering
von: Suri, Manan, et al.
Veröffentlicht: (2025)
von: Suri, Manan, et al.
Veröffentlicht: (2025)
Be like a Goldfish, Don't Memorize! Mitigating Memorization in Generative LLMs
von: Hans, Abhimanyu, et al.
Veröffentlicht: (2024)
von: Hans, Abhimanyu, et al.
Veröffentlicht: (2024)
Memorization Dynamics in Knowledge Distillation for Language Models
von: Borkar, Jaydeep, et al.
Veröffentlicht: (2026)
von: Borkar, Jaydeep, et al.
Veröffentlicht: (2026)
MatheMagic: Generating Dynamic Mathematics Benchmarks Robust to Memorization
von: O'Brien, Dayyán, et al.
Veröffentlicht: (2025)
von: O'Brien, Dayyán, et al.
Veröffentlicht: (2025)
FairSteer: Inference Time Debiasing for LLMs with Dynamic Activation Steering
von: Li, Yichen, et al.
Veröffentlicht: (2025)
von: Li, Yichen, et al.
Veröffentlicht: (2025)
ProLex: A Benchmark for Language Proficiency-oriented Lexical Substitution
von: Zhang, Xuanming, et al.
Veröffentlicht: (2024)
von: Zhang, Xuanming, et al.
Veröffentlicht: (2024)
CoSteer: Collaborative Decoding-Time Personalization via Local Delta Steering
von: Lv, Hang, et al.
Veröffentlicht: (2025)
von: Lv, Hang, et al.
Veröffentlicht: (2025)
Long-Term Ad Memorability: Understanding & Generating Memorable Ads
von: SI, Harini, et al.
Veröffentlicht: (2023)
von: SI, Harini, et al.
Veröffentlicht: (2023)
Towards Exception Safety Code Generation with Intermediate Representation Agents Framework
von: Zhang, Xuanming, et al.
Veröffentlicht: (2024)
von: Zhang, Xuanming, et al.
Veröffentlicht: (2024)
Prototype-Based Dynamic Steering for Large Language Models
von: Kayan, Ceyhun Efe, et al.
Veröffentlicht: (2025)
von: Kayan, Ceyhun Efe, et al.
Veröffentlicht: (2025)
Steering LLMs for Culturally Localized Generation
von: Khanuja, Simran, et al.
Veröffentlicht: (2026)
von: Khanuja, Simran, et al.
Veröffentlicht: (2026)
Gradient-Controlled Decoding: A Safety Guardrail for LLMs with Dual-Anchor Steering
von: Chiniya, Purva, et al.
Veröffentlicht: (2026)
von: Chiniya, Purva, et al.
Veröffentlicht: (2026)
Cognition-of-Thought Elicits Social-Aligned Reasoning in Large Language Models
von: Zhang, Xuanming, et al.
Veröffentlicht: (2025)
von: Zhang, Xuanming, et al.
Veröffentlicht: (2025)
MetaMind: Modeling Human Social Thoughts with Metacognitive Multi-Agent Systems
von: Zhang, Xuanming, et al.
Veröffentlicht: (2025)
von: Zhang, Xuanming, et al.
Veröffentlicht: (2025)
Memorization Dynamics of Fill-in-the-Middle Pretraining
von: von Arx, Tobias, et al.
Veröffentlicht: (2026)
von: von Arx, Tobias, et al.
Veröffentlicht: (2026)
Neuron-Level Differentiation of Memorization and Generalization in Large Language Models
von: Huang, Ko-Wei, et al.
Veröffentlicht: (2024)
von: Huang, Ko-Wei, et al.
Veröffentlicht: (2024)
VarBench: Robust Language Model Benchmarking Through Dynamic Variable Perturbation
von: Qian, Kun, et al.
Veröffentlicht: (2024)
von: Qian, Kun, et al.
Veröffentlicht: (2024)
Context Memorization for Efficient Long Context Generation
von: Okoshi, Yasuyuki, et al.
Veröffentlicht: (2026)
von: Okoshi, Yasuyuki, et al.
Veröffentlicht: (2026)
Steer Model beyond Assistant: Controlling System Prompt Strength via Contrastive Decoding
von: Dong, Yijiang River, et al.
Veröffentlicht: (2026)
von: Dong, Yijiang River, et al.
Veröffentlicht: (2026)
Personalized Text Generation with Contrastive Activation Steering
von: Zhang, Jinghao, et al.
Veröffentlicht: (2025)
von: Zhang, Jinghao, et al.
Veröffentlicht: (2025)
Steering Multimodal Large Language Models Decoding for Context-Aware Safety
von: Liu, Zheyuan, et al.
Veröffentlicht: (2025)
von: Liu, Zheyuan, et al.
Veröffentlicht: (2025)
Rote Learning Considered Useful: Generalizing over Memorized Data in LLMs
von: Wu, Qinyuan, et al.
Veröffentlicht: (2025)
von: Wu, Qinyuan, et al.
Veröffentlicht: (2025)
Characterizing Memorization in Diffusion Language Models: Generalized Extraction and Sampling Effects
von: Luo, Xiaoyu, et al.
Veröffentlicht: (2026)
von: Luo, Xiaoyu, et al.
Veröffentlicht: (2026)
Seeker: Towards Exception Safety Code Generation with Intermediate Language Agents Framework
von: Zhang, Xuanming, et al.
Veröffentlicht: (2024)
von: Zhang, Xuanming, et al.
Veröffentlicht: (2024)
Towards Better Generalization in Open-Domain Question Answering by Mitigating Context Memorization
von: Zhang, Zixuan, et al.
Veröffentlicht: (2024)
von: Zhang, Zixuan, et al.
Veröffentlicht: (2024)
SteerX: Disentangled Steering for LLM Personalization
von: Zhao, Xiaoyan, et al.
Veröffentlicht: (2025)
von: Zhao, Xiaoyan, et al.
Veröffentlicht: (2025)
On Memorization of Large Language Models in Logical Reasoning
von: Xie, Chulin, et al.
Veröffentlicht: (2024)
von: Xie, Chulin, et al.
Veröffentlicht: (2024)
Can LLM Graph Reasoning Generalize beyond Pattern Memorization?
von: Zhang, Yizhuo, et al.
Veröffentlicht: (2024)
von: Zhang, Yizhuo, et al.
Veröffentlicht: (2024)
DSVD: Dynamic Self-Verify Decoding for Faithful Generation in Large Language Models
von: Guo, YiQiu, et al.
Veröffentlicht: (2025)
von: Guo, YiQiu, et al.
Veröffentlicht: (2025)
Memorization or Interpolation ? Detecting LLM Memorization through Input Perturbation Analysis
von: Djiré, Albérick Euraste, et al.
Veröffentlicht: (2025)
von: Djiré, Albérick Euraste, et al.
Veröffentlicht: (2025)
Memorization and Knowledge Injection in Gated LLMs
von: Pan, Xu, et al.
Veröffentlicht: (2025)
von: Pan, Xu, et al.
Veröffentlicht: (2025)
Continual Memorization of Factoids in Language Models
von: Chen, Howard, et al.
Veröffentlicht: (2024)
von: Chen, Howard, et al.
Veröffentlicht: (2024)
Memory Dial: A Training Framework for Controllable Memorization in Language Models
von: Zhang, Xiangbo, et al.
Veröffentlicht: (2026)
von: Zhang, Xiangbo, et al.
Veröffentlicht: (2026)
Generalization or Memorization? Brittleness Testing for Chess-Trained Language Models
von: Tang, Ethan
Veröffentlicht: (2026)
von: Tang, Ethan
Veröffentlicht: (2026)
EvoCodeBench: An Evolving Code Generation Benchmark Aligned with Real-World Code Repositories
von: Li, Jia, et al.
Veröffentlicht: (2024)
von: Li, Jia, et al.
Veröffentlicht: (2024)
Data Compressibility Quantifies LLM Memorization
von: Huang, Yizhan, et al.
Veröffentlicht: (2025)
von: Huang, Yizhan, et al.
Veröffentlicht: (2025)
Fine-Grained Activation Steering: Steering Less, Achieving More
von: Feng, Zijian, et al.
Veröffentlicht: (2026)
von: Feng, Zijian, et al.
Veröffentlicht: (2026)
Memorization-Compression Cycles Improve Generalization
von: Yu, Fangyuan
Veröffentlicht: (2025)
von: Yu, Fangyuan
Veröffentlicht: (2025)
Memorization or Reasoning? Exploring the Idiom Understanding of LLMs
von: Kim, Jisu, et al.
Veröffentlicht: (2025)
von: Kim, Jisu, et al.
Veröffentlicht: (2025)
Steered Generation via Gradient Descent on Sparse Features
von: Bhattacharyya, Sumanta, et al.
Veröffentlicht: (2025)
von: Bhattacharyya, Sumanta, et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
Mitigating Memorization in LLMs using Activation Steering
von: Suri, Manan, et al.
Veröffentlicht: (2025) -
Be like a Goldfish, Don't Memorize! Mitigating Memorization in Generative LLMs
von: Hans, Abhimanyu, et al.
Veröffentlicht: (2024) -
Memorization Dynamics in Knowledge Distillation for Language Models
von: Borkar, Jaydeep, et al.
Veröffentlicht: (2026) -
MatheMagic: Generating Dynamic Mathematics Benchmarks Robust to Memorization
von: O'Brien, Dayyán, et al.
Veröffentlicht: (2025) -
FairSteer: Inference Time Debiasing for LLMs with Dynamic Activation Steering
von: Li, Yichen, et al.
Veröffentlicht: (2025)