TransformerFAM: Feedback attention is working memory
Fuente:
arXiv
Enregistré dans:
| Auteurs principaux: | Hwang, Dongseong, Wang, Weiran, Huo, Zhuoyuan, Sim, Khe Chai, Mengibar, Pedro Moreno |
|---|---|
| Format: | Preprint |
| Publié: |
2024
|
| Sujets: | |
| Accès en ligne: | |
| Tags: |
Ajouter un tag
Pas de tags, Soyez le premier à ajouter un tag!
|
Documents similaires
Aligner-Encoders: Self-Attention Transformers Can Be Self-Transducers
par: Stooke, Adam, et autres
Publié: (2025)
par: Stooke, Adam, et autres
Publié: (2025)
Hierarchical Recurrent Adapters for Efficient Multi-Task Adaptation of Large Speech Models
par: Munkhdalai, Tsendsuren, et autres
Publié: (2024)
par: Munkhdalai, Tsendsuren, et autres
Publié: (2024)
FAdam: Adam is a natural gradient optimizer using diagonal empirical Fisher information
par: Hwang, Dongseong
Publié: (2024)
par: Hwang, Dongseong
Publié: (2024)
HARP: Hesitation-Aware Reframing in Transformer Inference Pass
par: Storaï, Romain, et autres
Publié: (2024)
par: Storaï, Romain, et autres
Publié: (2024)
Critique of Impure Reason: Unveiling the reasoning behaviour of medical Large Language Models
par: Sim, Shamus, et autres
Publié: (2024)
par: Sim, Shamus, et autres
Publié: (2024)
Leave No Context Behind: Efficient Infinite Context Transformers with Infini-attention
par: Munkhdalai, Tsendsuren, et autres
Publié: (2024)
par: Munkhdalai, Tsendsuren, et autres
Publié: (2024)
Geometry-Lite: Interpretable Safety Probing via Layer-Wise Margin Geometry
par: Sim, Woo Seob, et autres
Publié: (2026)
par: Sim, Woo Seob, et autres
Publié: (2026)
RLBFF: Binary Flexible Feedback to bridge between Human Feedback & Verifiable Rewards
par: Wang, Zhilin, et autres
Publié: (2025)
par: Wang, Zhilin, et autres
Publié: (2025)
What are you sinking? A geometric approach on attention sink
par: Ruscio, Valeria, et autres
Publié: (2025)
par: Ruscio, Valeria, et autres
Publié: (2025)
SELF: Self-Evolution with Language Feedback
par: Lu, Jianqiao, et autres
Publié: (2023)
par: Lu, Jianqiao, et autres
Publié: (2023)
Constructing Synthetic Instruction Datasets for Improving Reasoning in Domain-Specific LLMs: A Case Study in the Japanese Financial Domain
par: Okochi, Yuma, et autres
Publié: (2026)
par: Okochi, Yuma, et autres
Publié: (2026)
UltraFeedback: Boosting Language Models with Scaled AI Feedback
par: Cui, Ganqu, et autres
Publié: (2023)
par: Cui, Ganqu, et autres
Publié: (2023)
Learning to Reason from Feedback at Test-Time
par: Li, Yanyang, et autres
Publié: (2025)
par: Li, Yanyang, et autres
Publié: (2025)
Knowledge Editing on Black-box Large Language Models
par: Song, Xiaoshuai, et autres
Publié: (2024)
par: Song, Xiaoshuai, et autres
Publié: (2024)
Why mask diffusion does not work
par: Sun, Haocheng, et autres
Publié: (2025)
par: Sun, Haocheng, et autres
Publié: (2025)
RLAIF vs. RLHF: Scaling Reinforcement Learning from Human Feedback with AI Feedback
par: Lee, Harrison, et autres
Publié: (2023)
par: Lee, Harrison, et autres
Publié: (2023)
The Information of Large Language Model Geometry
par: Tan, Zhiquan, et autres
Publié: (2024)
par: Tan, Zhiquan, et autres
Publié: (2024)
Reinforcement Learning with Backtracking Feedback
par: Sel, Bilgehan, et autres
Publié: (2026)
par: Sel, Bilgehan, et autres
Publié: (2026)
MINT: Evaluating LLMs in Multi-turn Interaction with Tools and Language Feedback
par: Wang, Xingyao, et autres
Publié: (2023)
par: Wang, Xingyao, et autres
Publié: (2023)
Editing Arbitrary Propositions in LLMs without Subject Labels
par: Feigenbaum, Itai, et autres
Publié: (2024)
par: Feigenbaum, Itai, et autres
Publié: (2024)
Pretraining with hierarchical memories: separating long-tail and common knowledge
par: Pouransari, Hadi, et autres
Publié: (2025)
par: Pouransari, Hadi, et autres
Publié: (2025)
SIPDO: Closed-Loop Prompt Optimization via Synthetic Data Feedback
par: Yu, Yaoning, et autres
Publié: (2025)
par: Yu, Yaoning, et autres
Publié: (2025)
Hindsight-Anchored Policy Optimization: Turning Failure into Feedback in Sparse Reward Settings
par: Wu, Yuning, et autres
Publié: (2026)
par: Wu, Yuning, et autres
Publié: (2026)
Closing the Loop: Learning to Generate Writing Feedback via Language Model Simulated Student Revisions
par: Nair, Inderjeet, et autres
Publié: (2024)
par: Nair, Inderjeet, et autres
Publié: (2024)
Understanding Subword Compositionality of Large Language Models
par: Peng, Qiwei, et autres
Publié: (2025)
par: Peng, Qiwei, et autres
Publié: (2025)
LILO: Bayesian Optimization with Natural Language Feedback
par: Kobalczyk, Katarzyna, et autres
Publié: (2025)
par: Kobalczyk, Katarzyna, et autres
Publié: (2025)
Provable Interactive Learning with Hindsight Instruction Feedback
par: Misra, Dipendra, et autres
Publié: (2024)
par: Misra, Dipendra, et autres
Publié: (2024)
RLTHF: Targeted Human Feedback for LLM Alignment
par: Xu, Yifei, et autres
Publié: (2025)
par: Xu, Yifei, et autres
Publié: (2025)
Learning Personalized Agents from Human Feedback
par: Liang, Kaiqu, et autres
Publié: (2026)
par: Liang, Kaiqu, et autres
Publié: (2026)
Distributionally Robust Reinforcement Learning with Human Feedback
par: Mandal, Debmalya, et autres
Publié: (2025)
par: Mandal, Debmalya, et autres
Publié: (2025)
Training Language Models with Language Feedback at Scale
par: Scheurer, Jérémy, et autres
Publié: (2023)
par: Scheurer, Jérémy, et autres
Publié: (2023)
Policy Improvement using Language Feedback Models
par: Zhong, Victor, et autres
Publié: (2024)
par: Zhong, Victor, et autres
Publié: (2024)
Towards Aligning Language Models with Textual Feedback
par: Lloret, Saüc Abadal, et autres
Publié: (2024)
par: Lloret, Saüc Abadal, et autres
Publié: (2024)
Reward Modeling with Ordinal Feedback: Wisdom of the Crowd
par: Liu, Shang, et autres
Publié: (2024)
par: Liu, Shang, et autres
Publié: (2024)
Mitigating Tail Narrowing in LLM Self-Improvement via Socratic-Guided Sampling
par: Ding, Yiwen, et autres
Publié: (2024)
par: Ding, Yiwen, et autres
Publié: (2024)
RLPF: Reinforcement Learning from Prediction Feedback for User Summarization with LLMs
par: Wu, Jiaxing, et autres
Publié: (2024)
par: Wu, Jiaxing, et autres
Publié: (2024)
Diff-eRank: A Novel Rank-Based Metric for Evaluating Large Language Models
par: Wei, Lai, et autres
Publié: (2024)
par: Wei, Lai, et autres
Publié: (2024)
Fresh in memory: Training-order recency is linearly encoded in language model activations
par: Krasheninnikov, Dmitrii, et autres
Publié: (2025)
par: Krasheninnikov, Dmitrii, et autres
Publié: (2025)
Parameter Efficient Reinforcement Learning from Human Feedback
par: Sidahmed, Hakim, et autres
Publié: (2024)
par: Sidahmed, Hakim, et autres
Publié: (2024)
Joint Learning of Context and Feedback Embeddings in Spoken Dialogue
par: Qian, Livia, et autres
Publié: (2024)
par: Qian, Livia, et autres
Publié: (2024)
Documents similaires
-
Aligner-Encoders: Self-Attention Transformers Can Be Self-Transducers
par: Stooke, Adam, et autres
Publié: (2025) -
Hierarchical Recurrent Adapters for Efficient Multi-Task Adaptation of Large Speech Models
par: Munkhdalai, Tsendsuren, et autres
Publié: (2024) -
FAdam: Adam is a natural gradient optimizer using diagonal empirical Fisher information
par: Hwang, Dongseong
Publié: (2024) -
HARP: Hesitation-Aware Reframing in Transformer Inference Pass
par: Storaï, Romain, et autres
Publié: (2024) -
Critique of Impure Reason: Unveiling the reasoning behaviour of medical Large Language Models
par: Sim, Shamus, et autres
Publié: (2024)