Gespeichert in:
| Hauptverfasser: | Yang, Liu, Lin, Ziqian, Lee, Kangwook, Papailiopoulos, Dimitris, Nowak, Robert |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2025
|
| Schlagworte: | |
| Online-Zugang: | https://arxiv.org/abs/2501.09240 |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Looped Transformers are Better at Learning Learning Algorithms
von: Yang, Liu, et al.
Veröffentlicht: (2023)
von: Yang, Liu, et al.
Veröffentlicht: (2023)
Dual Operating Modes of In-Context Learning
von: Lin, Ziqian, et al.
Veröffentlicht: (2024)
von: Lin, Ziqian, et al.
Veröffentlicht: (2024)
Can Mamba Learn How to Learn? A Comparative Study on In-Context Learning Tasks
von: Park, Jongho, et al.
Veröffentlicht: (2024)
von: Park, Jongho, et al.
Veröffentlicht: (2024)
From Artificial Needles to Real Haystacks: Improving Retrieval Capabilities in LLMs by Finetuning on Synthetic Data
von: Xiong, Zheyang, et al.
Veröffentlicht: (2024)
von: Xiong, Zheyang, et al.
Veröffentlicht: (2024)
Predictive Pipelined Decoding: A Compute-Latency Trade-off for Exact LLM Decoding
von: Yang, Seongjun, et al.
Veröffentlicht: (2023)
von: Yang, Seongjun, et al.
Veröffentlicht: (2023)
In-Context Learning with Hypothesis-Class Guidance
von: Lin, Ziqian, et al.
Veröffentlicht: (2025)
von: Lin, Ziqian, et al.
Veröffentlicht: (2025)
Self-Improving Transformers Overcome Easy-to-Hard and Length Generalization Challenges
von: Lee, Nayoung, et al.
Veröffentlicht: (2025)
von: Lee, Nayoung, et al.
Veröffentlicht: (2025)
Everything Everywhere All at Once: LLMs can In-Context Learn Multiple Tasks in Superposition
von: Xiong, Zheyang, et al.
Veröffentlicht: (2024)
von: Xiong, Zheyang, et al.
Veröffentlicht: (2024)
Variation Spaces for Multi-Output Neural Networks: Insights on Multi-Task Learning and Network Compression
von: Shenouda, Joseph, et al.
Veröffentlicht: (2023)
von: Shenouda, Joseph, et al.
Veröffentlicht: (2023)
ReJump: A Tree-Jump Representation for Analyzing and Improving LLM Reasoning
von: Zeng, Yuchen, et al.
Veröffentlicht: (2025)
von: Zeng, Yuchen, et al.
Veröffentlicht: (2025)
Understanding Task Vectors in In-Context Learning: Emergence, Functionality, and Limitations
von: Dong, Yuxin, et al.
Veröffentlicht: (2025)
von: Dong, Yuxin, et al.
Veröffentlicht: (2025)
How Well Can Transformers Emulate In-context Newton's Method?
von: Giannou, Angeliki, et al.
Veröffentlicht: (2024)
von: Giannou, Angeliki, et al.
Veröffentlicht: (2024)
Lexico: Extreme KV Cache Compression via Sparse Coding over Universal Dictionaries
von: Kim, Junhyuck, et al.
Veröffentlicht: (2024)
von: Kim, Junhyuck, et al.
Veröffentlicht: (2024)
Fine-Tuning Without Forgetting In-Context Learning: A Theoretical Analysis of Linear Attention Models
von: Lee, Chungpa, et al.
Veröffentlicht: (2026)
von: Lee, Chungpa, et al.
Veröffentlicht: (2026)
Emergence and Effectiveness of Task Vectors in In-Context Learning: An Encoder Decoder Perspective
von: Han, Seungwook, et al.
Veröffentlicht: (2024)
von: Han, Seungwook, et al.
Veröffentlicht: (2024)
Endless Terminals: Scaling RL Environments for Terminal Agents
von: Gandhi, Kanishk, et al.
Veröffentlicht: (2026)
von: Gandhi, Kanishk, et al.
Veröffentlicht: (2026)
Not All Bits Are Equal: Scale-Dependent Memory Optimization Strategies for Reasoning Models
von: Kim, Junhyuck, et al.
Veröffentlicht: (2025)
von: Kim, Junhyuck, et al.
Veröffentlicht: (2025)
Can MLLMs Perform Text-to-Image In-Context Learning?
von: Zeng, Yuchen, et al.
Veröffentlicht: (2024)
von: Zeng, Yuchen, et al.
Veröffentlicht: (2024)
Transformers in the Dark: Navigating Unknown Search Spaces via Bandit Feedback
von: Kim, Jungtaek, et al.
Veröffentlicht: (2026)
von: Kim, Jungtaek, et al.
Veröffentlicht: (2026)
Optimizing DDPM Sampling with Shortcut Fine-Tuning
von: Fan, Ying, et al.
Veröffentlicht: (2023)
von: Fan, Ying, et al.
Veröffentlicht: (2023)
Wait, Wait, Wait... Why Do Reasoning Models Loop?
von: Pipis, Charilaos, et al.
Veröffentlicht: (2025)
von: Pipis, Charilaos, et al.
Veröffentlicht: (2025)
MEMENTO: Teaching LLMs to Manage Their Own Context
von: Kontonis, Vasilis, et al.
Veröffentlicht: (2026)
von: Kontonis, Vasilis, et al.
Veröffentlicht: (2026)
Sample More to Think Less: Group Filtered Policy Optimization for Concise Reasoning
von: Shrivastava, Vaishnavi, et al.
Veröffentlicht: (2025)
von: Shrivastava, Vaishnavi, et al.
Veröffentlicht: (2025)
VersaPRM: Multi-Domain Process Reward Model via Synthetic Reasoning Data
von: Zeng, Thomas, et al.
Veröffentlicht: (2025)
von: Zeng, Thomas, et al.
Veröffentlicht: (2025)
The Expressive Power of Low-Rank Adaptation
von: Zeng, Yuchen, et al.
Veröffentlicht: (2023)
von: Zeng, Yuchen, et al.
Veröffentlicht: (2023)
The Effects of Multi-Task Learning on ReLU Neural Network Functions
von: Nakhleh, Julia, et al.
Veröffentlicht: (2024)
von: Nakhleh, Julia, et al.
Veröffentlicht: (2024)
CHAI: Clustered Head Attention for Efficient LLM Inference
von: Agarwal, Saurabh, et al.
Veröffentlicht: (2024)
von: Agarwal, Saurabh, et al.
Veröffentlicht: (2024)
Looped Transformers for Length Generalization
von: Fan, Ying, et al.
Veröffentlicht: (2024)
von: Fan, Ying, et al.
Veröffentlicht: (2024)
Memorization Capacity for Additive Fine-Tuning with Small ReLU Networks
von: Sohn, Jy-yong, et al.
Veröffentlicht: (2024)
von: Sohn, Jy-yong, et al.
Veröffentlicht: (2024)
Pretraining Decision Transformers with Reward Prediction for In-Context Multi-task Structured Bandit Learning
von: Mukherjee, Subhojyoti, et al.
Veröffentlicht: (2024)
von: Mukherjee, Subhojyoti, et al.
Veröffentlicht: (2024)
Context and Diversity Matter: The Emergence of In-Context Learning in World Models
von: Wang, Fan, et al.
Veröffentlicht: (2025)
von: Wang, Fan, et al.
Veröffentlicht: (2025)
Multi-Task Corrupted Prediction for Learning Robust Audio-Visual Speech Representation
von: Kim, Sungnyun, et al.
Veröffentlicht: (2025)
von: Kim, Sungnyun, et al.
Veröffentlicht: (2025)
Convergence and Emergence of In-Context Reinforcement Learning with Chain of Thought
von: Xie, Zixuan, et al.
Veröffentlicht: (2026)
von: Xie, Zixuan, et al.
Veröffentlicht: (2026)
Efficient Active Learning with Abstention
von: Zhu, Yinglun, et al.
Veröffentlicht: (2022)
von: Zhu, Yinglun, et al.
Veröffentlicht: (2022)
Vocabulary In-Context Learning in Transformers: Benefits of Positional Encoding
von: Ma, Qian, et al.
Veröffentlicht: (2025)
von: Ma, Qian, et al.
Veröffentlicht: (2025)
Provable In-Context Vector Arithmetic via Retrieving Task Concepts
von: Bu, Dake, et al.
Veröffentlicht: (2025)
von: Bu, Dake, et al.
Veröffentlicht: (2025)
Active Learning with Neural Networks: Insights from Nonparametric Statistics
von: Zhu, Yinglun, et al.
Veröffentlicht: (2022)
von: Zhu, Yinglun, et al.
Veröffentlicht: (2022)
Towards Provable Emergence of In-Context Reinforcement Learning
von: Wang, Jiuqi, et al.
Veröffentlicht: (2025)
von: Wang, Jiuqi, et al.
Veröffentlicht: (2025)
Infected Smallville: How Disease Threat Shapes Sociality in LLM Agents
von: Choi, Soyeon, et al.
Veröffentlicht: (2025)
von: Choi, Soyeon, et al.
Veröffentlicht: (2025)
Learning Task Representations from In-Context Learning
von: Saglam, Baturay, et al.
Veröffentlicht: (2025)
von: Saglam, Baturay, et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
Looped Transformers are Better at Learning Learning Algorithms
von: Yang, Liu, et al.
Veröffentlicht: (2023) -
Dual Operating Modes of In-Context Learning
von: Lin, Ziqian, et al.
Veröffentlicht: (2024) -
Can Mamba Learn How to Learn? A Comparative Study on In-Context Learning Tasks
von: Park, Jongho, et al.
Veröffentlicht: (2024) -
From Artificial Needles to Real Haystacks: Improving Retrieval Capabilities in LLMs by Finetuning on Synthetic Data
von: Xiong, Zheyang, et al.
Veröffentlicht: (2024) -
Predictive Pipelined Decoding: A Compute-Latency Trade-off for Exact LLM Decoding
von: Yang, Seongjun, et al.
Veröffentlicht: (2023)