In-Context Learning with Representations: Contextual Generalization of Trained Transformers
Fuente:
arXiv
Salvato in:
| Autori principali: | Yang, Tong, Huang, Yu, Liang, Yingbin, Chi, Yuejie |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2024
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
Multi-head Transformers Provably Learn Symbolic Multi-step Reasoning via Gradient Descent
di: Yang, Tong, et al.
Pubblicazione: (2025)
di: Yang, Tong, et al.
Pubblicazione: (2025)
Agentic Transformers Provably Learn to Search via Reinforcement Learning
di: Yang, Tong, et al.
Pubblicazione: (2026)
di: Yang, Tong, et al.
Pubblicazione: (2026)
A Theoretical Analysis of Self-Supervised Learning for Vision Transformers
di: Huang, Yu, et al.
Pubblicazione: (2024)
di: Huang, Yu, et al.
Pubblicazione: (2024)
Breaking the Sample Size Barrier in Model-Based Reinforcement Learning with a Generative Model
di: Li, Gen, et al.
Pubblicazione: (2020)
di: Li, Gen, et al.
Pubblicazione: (2020)
Is Q-Learning Minimax Optimal? A Tight Sample Complexity Analysis
di: Li, Gen, et al.
Pubblicazione: (2021)
di: Li, Gen, et al.
Pubblicazione: (2021)
Accelerating Convergence of Score-Based Diffusion Models, Provably
di: Li, Gen, et al.
Pubblicazione: (2024)
di: Li, Gen, et al.
Pubblicazione: (2024)
High-probability sample complexities for policy evaluation with linear function approximation
di: Li, Gen, et al.
Pubblicazione: (2023)
di: Li, Gen, et al.
Pubblicazione: (2023)
Achieving Logarithmic Regret in KL-Regularized Zero-Sum Markov Games
di: Nayak, Anupam, et al.
Pubblicazione: (2025)
di: Nayak, Anupam, et al.
Pubblicazione: (2025)
Fast Computation of Optimal Transport via Entropy-Regularized Extragradient Methods
di: Li, Gen, et al.
Pubblicazione: (2023)
di: Li, Gen, et al.
Pubblicazione: (2023)
Incentivize without Bonus: Provably Efficient Model-based Online Multi-agent RL for Markov Games
di: Yang, Tong, et al.
Pubblicazione: (2025)
di: Yang, Tong, et al.
Pubblicazione: (2025)
The Implicit Curriculum: Learning Dynamics in RL with Verifiable Rewards
di: Huang, Yu, et al.
Pubblicazione: (2026)
di: Huang, Yu, et al.
Pubblicazione: (2026)
Transformers Provably Learn Chain-of-Thought Reasoning with Length Generalization
di: Huang, Yu, et al.
Pubblicazione: (2025)
di: Huang, Yu, et al.
Pubblicazione: (2025)
Statistical and Algorithmic Foundations of Reinforcement Learning
di: Chi, Yuejie, et al.
Pubblicazione: (2025)
di: Chi, Yuejie, et al.
Pubblicazione: (2025)
The Sample-Communication Complexity Trade-off in Federated Q-Learning
di: Salgia, Sudeep, et al.
Pubblicazione: (2024)
di: Salgia, Sudeep, et al.
Pubblicazione: (2024)
Exploration from a Primal-Dual Lens: Value-Incentivized Actor-Critic Methods for Sample-Efficient Online RL
di: Yang, Tong, et al.
Pubblicazione: (2025)
di: Yang, Tong, et al.
Pubblicazione: (2025)
Adversarial Water-Filling: Theory, Algorithms and Foundation Model
di: Tong, Xindi, et al.
Pubblicazione: (2026)
di: Tong, Xindi, et al.
Pubblicazione: (2026)
Unveiling Induction Heads: Provable Training Dynamics and Feature Learning in Transformers
di: Chen, Siyu, et al.
Pubblicazione: (2024)
di: Chen, Siyu, et al.
Pubblicazione: (2024)
Stochastic Zeroth-Order Optimization under Strongly Convexity and Lipschitz Hessian: Minimax Sample Complexity
di: Yu, Qian, et al.
Pubblicazione: (2024)
di: Yu, Qian, et al.
Pubblicazione: (2024)
On the Convergence Analysis of Muon
di: Shen, Wei, et al.
Pubblicazione: (2025)
di: Shen, Wei, et al.
Pubblicazione: (2025)
Generalized Orthogonal Procrustes Problem under Arbitrary Adversaries
di: Ling, Shuyang
Pubblicazione: (2021)
di: Ling, Shuyang
Pubblicazione: (2021)
Span-Based Optimal Sample Complexity for Weakly Communicating and General Average Reward MDPs
di: Zurek, Matthew, et al.
Pubblicazione: (2024)
di: Zurek, Matthew, et al.
Pubblicazione: (2024)
Stochastic Smoothed Gradient Descent Ascent for Federated Minimax Optimization
di: Shen, Wei, et al.
Pubblicazione: (2023)
di: Shen, Wei, et al.
Pubblicazione: (2023)
A Single-Loop First-Order Algorithm for Linearly Constrained Bilevel Optimization
di: Shen, Wei, et al.
Pubblicazione: (2025)
di: Shen, Wei, et al.
Pubblicazione: (2025)
Vertical Federated Learning with Missing Features During Training and Inference
di: Valdeira, Pedro, et al.
Pubblicazione: (2024)
di: Valdeira, Pedro, et al.
Pubblicazione: (2024)
On the Robustness of Cross-Concentrated Sampling for Matrix Completion
di: Cai, HanQin, et al.
Pubblicazione: (2024)
di: Cai, HanQin, et al.
Pubblicazione: (2024)
Learning How to Strategically Disclose Information
di: Velicheti, Raj Kiriti, et al.
Pubblicazione: (2024)
di: Velicheti, Raj Kiriti, et al.
Pubblicazione: (2024)
Communication-Efficient Federated Optimization over Semi-Decentralized Networks
di: Wang, He, et al.
Pubblicazione: (2023)
di: Wang, He, et al.
Pubblicazione: (2023)
MLorc: Momentum Low-rank Compression for Memory Efficient Large Language Model Adaptation
di: Shen, Wei, et al.
Pubblicazione: (2025)
di: Shen, Wei, et al.
Pubblicazione: (2025)
Submodular Information Selection for Hypothesis Testing with Misclassification Penalties
di: Bhargav, Jayanth, et al.
Pubblicazione: (2024)
di: Bhargav, Jayanth, et al.
Pubblicazione: (2024)
Ensemble-Conditional Gaussian Processes (Ens-CGP): Representation, Geometry, and Inference
di: Ravela, Sai, et al.
Pubblicazione: (2026)
di: Ravela, Sai, et al.
Pubblicazione: (2026)
The Plug-in Approach for Average-Reward and Discounted MDPs: Optimal Sample Complexity Analysis
di: Zurek, Matthew, et al.
Pubblicazione: (2024)
di: Zurek, Matthew, et al.
Pubblicazione: (2024)
Variational Inference on the Boolean Hypercube with the Quantum Entropy
di: Beyler, Eliot, et al.
Pubblicazione: (2024)
di: Beyler, Eliot, et al.
Pubblicazione: (2024)
Structured Sampling for Robust Euclidean Distance Geometry
di: Kundu, Chandra, et al.
Pubblicazione: (2024)
di: Kundu, Chandra, et al.
Pubblicazione: (2024)
Group Projected Subspace Pursuit for Block Sparse Signal Reconstruction: Convergence Analysis and Applications
di: He, Roy Y., et al.
Pubblicazione: (2024)
di: He, Roy Y., et al.
Pubblicazione: (2024)
Convexity in Disguise: A Theoretical Framework for Nonconvex Low-Rank Matrix Estimation
di: Cui, Chengyu, et al.
Pubblicazione: (2026)
di: Cui, Chengyu, et al.
Pubblicazione: (2026)
Linear regression with overparameterized linear neural networks: Tight upper and lower bounds for implicit $\ell^1$-regularization
di: Matt, Hannes, et al.
Pubblicazione: (2025)
di: Matt, Hannes, et al.
Pubblicazione: (2025)
Recovering Simultaneously Structured Data via Non-Convex Iteratively Reweighted Least Squares
di: Kümmerle, Christian, et al.
Pubblicazione: (2023)
di: Kümmerle, Christian, et al.
Pubblicazione: (2023)
Finite-Time Minimax Bounds and an Optimal Lyapunov Policy in Queueing Control
di: Liu, Yujie, et al.
Pubblicazione: (2025)
di: Liu, Yujie, et al.
Pubblicazione: (2025)
Optimal Variance-Dependent Regret Bounds for Infinite-Horizon MDPs
di: Zamir, Guy, et al.
Pubblicazione: (2026)
di: Zamir, Guy, et al.
Pubblicazione: (2026)
A Neural Network Algorithm for KL Divergence Estimation with Quantitative Error Bounds
di: Foss, Mikil, et al.
Pubblicazione: (2025)
di: Foss, Mikil, et al.
Pubblicazione: (2025)
Documenti analoghi
-
Multi-head Transformers Provably Learn Symbolic Multi-step Reasoning via Gradient Descent
di: Yang, Tong, et al.
Pubblicazione: (2025) -
Agentic Transformers Provably Learn to Search via Reinforcement Learning
di: Yang, Tong, et al.
Pubblicazione: (2026) -
A Theoretical Analysis of Self-Supervised Learning for Vision Transformers
di: Huang, Yu, et al.
Pubblicazione: (2024) -
Breaking the Sample Size Barrier in Model-Based Reinforcement Learning with a Generative Model
di: Li, Gen, et al.
Pubblicazione: (2020) -
Is Q-Learning Minimax Optimal? A Tight Sample Complexity Analysis
di: Li, Gen, et al.
Pubblicazione: (2021)