Iterate to Accelerate: A Unified Framework for Iterative Reasoning and Feedback Convergence
Fuente:
arXiv
Guardado en:
| Autor principal: | Fein-Ashley, Jacob |
|---|---|
| Formato: | Preprint |
| Publicado: |
2025
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
Contextual Feedback Loops: Amplifying Deep Reasoning with Iterative Top-Down Feedback
por: Fein-Ashley, Jacob, et al.
Publicado: (2024)
por: Fein-Ashley, Jacob, et al.
Publicado: (2024)
Linear Diffusion Networks
por: Fein-Ashley, Jacob
Publicado: (2025)
por: Fein-Ashley, Jacob
Publicado: (2025)
Flowing Through Layers: A Continuous Dynamical Systems Perspective on Transformers
por: Fein-Ashley, Jacob
Publicado: (2025)
por: Fein-Ashley, Jacob
Publicado: (2025)
A Comparison of Traditional and Deep Learning Methods for Parameter Estimation of the Ornstein-Uhlenbeck Process
por: Fein-Ashley, Jacob
Publicado: (2024)
por: Fein-Ashley, Jacob
Publicado: (2024)
Solve the Loop: Attractor Models for Language and Reasoning
por: Fein-Ashley, Jacob, et al.
Publicado: (2026)
por: Fein-Ashley, Jacob, et al.
Publicado: (2026)
On Value Iteration Convergence in Connected MDPs
por: Mustafin, Arsenii, et al.
Publicado: (2024)
por: Mustafin, Arsenii, et al.
Publicado: (2024)
Non-Asymptotic Convergence of Stochastic Iterative Algorithms: A Lyapunov Framework
por: Chen, Zaiwei, et al.
Publicado: (2026)
por: Chen, Zaiwei, et al.
Publicado: (2026)
The Harder Path: Last Iterate Convergence for Uncoupled Learning in Zero-Sum Games with Bandit Feedback
por: Fiegel, Côme, et al.
Publicado: (2026)
por: Fiegel, Côme, et al.
Publicado: (2026)
Near-Optimal Last-Iterate Convergence for Zero-Sum Games with Bandit Feedback and Opponent Actions
por: Hait, Soumita, et al.
Publicado: (2026)
por: Hait, Soumita, et al.
Publicado: (2026)
Studying the Effects of Self-Attention on SAR Automatic Target Recognition
por: Fein-Ashley, Jacob, et al.
Publicado: (2024)
por: Fein-Ashley, Jacob, et al.
Publicado: (2024)
Extragradient Preference Optimization (EGPO): Beyond Last-Iterate Convergence for Nash Learning from Human Feedback
por: Zhou, Runlong, et al.
Publicado: (2025)
por: Zhou, Runlong, et al.
Publicado: (2025)
SPECTRE: An FFT-Based Efficient Drop-In Replacement to Self-Attention for Long Contexts
por: Fein-Ashley, Jacob, et al.
Publicado: (2025)
por: Fein-Ashley, Jacob, et al.
Publicado: (2025)
Mixture of Thoughts: Learning to Aggregate What Experts Think, Not Just What They Say
por: Fein-Ashley, Jacob, et al.
Publicado: (2025)
por: Fein-Ashley, Jacob, et al.
Publicado: (2025)
On Separation Between Best-Iterate, Random-Iterate, and Last-Iterate Convergence of Learning in Games
por: Cai, Yang, et al.
Publicado: (2025)
por: Cai, Yang, et al.
Publicado: (2025)
Exponential Convergence Guarantees for Iterative Markovian Fitting
por: Silveri, Marta Gentiloni, et al.
Publicado: (2025)
por: Silveri, Marta Gentiloni, et al.
Publicado: (2025)
Vanishing Contributions: A Unified Framework for Smooth and Iterative Model Compression
por: Nikiforos, Lorenzo, et al.
Publicado: (2025)
por: Nikiforos, Lorenzo, et al.
Publicado: (2025)
On the Last-Iterate Convergence of Shuffling Gradient Methods
por: Liu, Zijian, et al.
Publicado: (2024)
por: Liu, Zijian, et al.
Publicado: (2024)
LocalKMeans: Convergence of Lloyd's Algorithm with Distributed Local Iterations
por: Vardhan, Harsh, et al.
Publicado: (2025)
por: Vardhan, Harsh, et al.
Publicado: (2025)
Augmented Lagrangian Method for Last-Iterate Convergence for Constrained MDPs
por: Lu, Michael, et al.
Publicado: (2026)
por: Lu, Michael, et al.
Publicado: (2026)
From Average-Iterate to Last-Iterate Convergence in Games: A Reduction and Its Applications
por: Cai, Yang, et al.
Publicado: (2025)
por: Cai, Yang, et al.
Publicado: (2025)
Robust Sublinear Convergence Rates for Iterative Bregman Projections
por: Peyré, Gabriel
Publicado: (2026)
por: Peyré, Gabriel
Publicado: (2026)
Convergence Rate of the Last Iterate of Stochastic Proximal Algorithms
por: Vaidyan, Kevin Kurian Thomas, et al.
Publicado: (2026)
por: Vaidyan, Kevin Kurian Thomas, et al.
Publicado: (2026)
Revisiting the Last-Iterate Convergence of Stochastic Gradient Methods
por: Liu, Zijian, et al.
Publicado: (2023)
por: Liu, Zijian, et al.
Publicado: (2023)
Accelerating Sinkhorn Algorithm with Sparse Newton Iterations
por: Tang, Xun, et al.
Publicado: (2024)
por: Tang, Xun, et al.
Publicado: (2024)
Stationary Reweighting Yields Local Convergence of Soft Fitted Q-Iteration
por: van der Laan, Lars, et al.
Publicado: (2025)
por: van der Laan, Lars, et al.
Publicado: (2025)
Last-Iterate Global Convergence of Policy Gradients for Constrained Reinforcement Learning
por: Montenegro, Alessandro, et al.
Publicado: (2024)
por: Montenegro, Alessandro, et al.
Publicado: (2024)
Machine Learning-Augmented Acceleration of Iterative Ptychographic Reconstruction
por: Zheng, Bowen, et al.
Publicado: (2026)
por: Zheng, Bowen, et al.
Publicado: (2026)
An Iterative Approach to Topic Modelling
por: Wong, Albert, et al.
Publicado: (2024)
por: Wong, Albert, et al.
Publicado: (2024)
NOWS: Neural Operator Warm Starts for Accelerating Iterative Solvers
por: Eshaghi, Mohammad Sadegh, et al.
Publicado: (2025)
por: Eshaghi, Mohammad Sadegh, et al.
Publicado: (2025)
Fast Last-Iterate Convergence of SGD in the Smooth Interpolation Regime
por: Attia, Amit, et al.
Publicado: (2025)
por: Attia, Amit, et al.
Publicado: (2025)
Last-Iterate Convergence of General Parameterized Policies in Constrained MDPs
por: Mondal, Washim Uddin, et al.
Publicado: (2024)
por: Mondal, Washim Uddin, et al.
Publicado: (2024)
Last Iterate Convergence of Incremental Methods and Applications in Continual Learning
por: Cai, Xufeng, et al.
Publicado: (2024)
por: Cai, Xufeng, et al.
Publicado: (2024)
iFlip: Iterative Feedback-driven Counterfactual Example Refinement
por: Wang, Yilong, et al.
Publicado: (2026)
por: Wang, Yilong, et al.
Publicado: (2026)
Resource-Efficient Iterative LLM-Based NAS with Feedback Memory
por: Gu, Xiaojie, et al.
Publicado: (2026)
por: Gu, Xiaojie, et al.
Publicado: (2026)
Learning Iterative Reasoning through Energy Diffusion
por: Du, Yilun, et al.
Publicado: (2024)
por: Du, Yilun, et al.
Publicado: (2024)
Linear Feedback Control Systems for Iterative Prompt Optimization in Large Language Models
por: Karn, Rupesh Raj
Publicado: (2025)
por: Karn, Rupesh Raj
Publicado: (2025)
Online Iterative Reinforcement Learning from Human Feedback with General Preference Model
por: Ye, Chenlu, et al.
Publicado: (2024)
por: Ye, Chenlu, et al.
Publicado: (2024)
Accelerated Training through Iterative Gradient Propagation Along the Residual Path
por: Fagnou, Erwan, et al.
Publicado: (2025)
por: Fagnou, Erwan, et al.
Publicado: (2025)
Faster WIND: Accelerating Iterative Best-of-$N$ Distillation for LLM Alignment
por: Yang, Tong, et al.
Publicado: (2024)
por: Yang, Tong, et al.
Publicado: (2024)
Revisiting Value Iteration: Unified Analysis of Discounted and Average-Reward Cases
por: Mustafin, Arsenii, et al.
Publicado: (2025)
por: Mustafin, Arsenii, et al.
Publicado: (2025)
Ejemplares similares
-
Contextual Feedback Loops: Amplifying Deep Reasoning with Iterative Top-Down Feedback
por: Fein-Ashley, Jacob, et al.
Publicado: (2024) -
Linear Diffusion Networks
por: Fein-Ashley, Jacob
Publicado: (2025) -
Flowing Through Layers: A Continuous Dynamical Systems Perspective on Transformers
por: Fein-Ashley, Jacob
Publicado: (2025) -
A Comparison of Traditional and Deep Learning Methods for Parameter Estimation of the Ornstein-Uhlenbeck Process
por: Fein-Ashley, Jacob
Publicado: (2024) -
Solve the Loop: Attractor Models for Language and Reasoning
por: Fein-Ashley, Jacob, et al.
Publicado: (2026)