Covariance-Aware Transformers for Quadratic Programming and Decision Making
Fuente:
arXiv
Salvato in:
| Autori principali: | Tire, Kutay, Zhang, Yufan, Taga, Ege Onur, Oymak, Samet |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2026
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
Retrieval Augmented Time Series Forecasting
di: Tire, Kutay, et al.
Pubblicazione: (2024)
di: Tire, Kutay, et al.
Pubblicazione: (2024)
Learning to Bet for Horizon-Aware Anytime-Valid Testing
di: Taga, Ege Onur, et al.
Pubblicazione: (2026)
di: Taga, Ege Onur, et al.
Pubblicazione: (2026)
TimePFN: Effective Multivariate Time Series Forecasting with Synthetic Data
di: Taga, Ege Onur, et al.
Pubblicazione: (2025)
di: Taga, Ege Onur, et al.
Pubblicazione: (2025)
Evolutionary Multi-Task Optimization for LLM-Guided Program Discovery
di: Gozeten, Halil Alperen, et al.
Pubblicazione: (2026)
di: Gozeten, Halil Alperen, et al.
Pubblicazione: (2026)
Learning to Correct: Calibrated Reinforcement Learning for Multi-Attempt Chain-of-Thought
di: Ildiz, Muhammed Emrullah, et al.
Pubblicazione: (2026)
di: Ildiz, Muhammed Emrullah, et al.
Pubblicazione: (2026)
High-dimensional Analysis of Knowledge Distillation: Weak-to-Strong Generalization and Scaling Laws
di: Ildiz, M. Emrullah, et al.
Pubblicazione: (2024)
di: Ildiz, M. Emrullah, et al.
Pubblicazione: (2024)
Efficient Contextual LLM Cascades through Budget-Constrained Policy Learning
di: Zhang, Xuechen, et al.
Pubblicazione: (2024)
di: Zhang, Xuechen, et al.
Pubblicazione: (2024)
Latent Chain-of-Thought Improves Structured-Data Transformers
di: Dudley, Carson, et al.
Pubblicazione: (2026)
di: Dudley, Carson, et al.
Pubblicazione: (2026)
On the Power of Convolution Augmented Transformer
di: Li, Mingchen, et al.
Pubblicazione: (2024)
di: Li, Mingchen, et al.
Pubblicazione: (2024)
Making Small Language Models Efficient Reasoners: Intervention, Supervision, Reinforcement
di: Zhang, Xuechen, et al.
Pubblicazione: (2025)
di: Zhang, Xuechen, et al.
Pubblicazione: (2025)
Can Transformers Learn Optimal Filtering for Unknown Systems?
di: Balim, Haldun, et al.
Pubblicazione: (2023)
di: Balim, Haldun, et al.
Pubblicazione: (2023)
Transformers as Support Vector Machines
di: Tarzanagh, Davoud Ataee, et al.
Pubblicazione: (2023)
di: Tarzanagh, Davoud Ataee, et al.
Pubblicazione: (2023)
SmartChunk Retrieval: Query-Aware Chunk Compression with Planning for Efficient Document RAG
di: Zhang, Xuechen, et al.
Pubblicazione: (2025)
di: Zhang, Xuechen, et al.
Pubblicazione: (2025)
Test-Time Training Provably Improves Transformers as In-context Learners
di: Gozeten, Halil Alperen, et al.
Pubblicazione: (2025)
di: Gozeten, Halil Alperen, et al.
Pubblicazione: (2025)
Selective Attention: Enhancing Transformer through Principled Context Control
di: Zhang, Xuechen, et al.
Pubblicazione: (2024)
di: Zhang, Xuechen, et al.
Pubblicazione: (2024)
In-Context Learning Under Regime Change
di: Dudley, Carson, et al.
Pubblicazione: (2026)
di: Dudley, Carson, et al.
Pubblicazione: (2026)
Attention with Trained Embeddings Provably Selects Important Tokens
di: Wu, Diyuan, et al.
Pubblicazione: (2025)
di: Wu, Diyuan, et al.
Pubblicazione: (2025)
Fine-grained Analysis of In-context Linear Estimation: Data, Architecture, and Beyond
di: Li, Yingcong, et al.
Pubblicazione: (2024)
di: Li, Yingcong, et al.
Pubblicazione: (2024)
Class-attribute Priors: Adapting Optimization to Heterogeneity and Fairness Objective
di: Zhang, Xuechen, et al.
Pubblicazione: (2024)
di: Zhang, Xuechen, et al.
Pubblicazione: (2024)
From Self-Attention to Markov Models: Unveiling the Dynamics of Generative Transformers
di: Ildiz, M. Emrullah, et al.
Pubblicazione: (2024)
di: Ildiz, M. Emrullah, et al.
Pubblicazione: (2024)
BREAD: Branched Rollouts from Expert Anchors Bridge SFT & RL for Reasoning
di: Zhang, Xuechen, et al.
Pubblicazione: (2025)
di: Zhang, Xuechen, et al.
Pubblicazione: (2025)
VSPO: Vector-Steered Policy Optimization for Behavioral Control
di: Zhang, Xuechen, et al.
Pubblicazione: (2026)
di: Zhang, Xuechen, et al.
Pubblicazione: (2026)
Input-Label Correlation Governs a Linear-to-Nonlinear Transition in Random Features under Spiked Covariance
di: Demir, Samet, et al.
Pubblicazione: (2024)
di: Demir, Samet, et al.
Pubblicazione: (2024)
SPARKLE: A Nonparametric Approach for Online Decision-Making with High-Dimensional Covariates
di: Wang, Wenjia, et al.
Pubblicazione: (2025)
di: Wang, Wenjia, et al.
Pubblicazione: (2025)
Plug-and-Play Transformer Modules for Test-Time Adaptation
di: Chang, Xiangyu, et al.
Pubblicazione: (2024)
di: Chang, Xiangyu, et al.
Pubblicazione: (2024)
Continuous Chain of Thought Enables Parallel Exploration and Reasoning
di: Gozeten, Halil Alperen, et al.
Pubblicazione: (2025)
di: Gozeten, Halil Alperen, et al.
Pubblicazione: (2025)
Regret Minimization and Statistical Inference in Online Decision Making with High-dimensional Covariates
di: Duan, Congyuan, et al.
Pubblicazione: (2024)
di: Duan, Congyuan, et al.
Pubblicazione: (2024)
Evaluating the Impact of Data Cleaning on the Quality of Generated Pull Request Descriptions
di: Tire, Kutay, et al.
Pubblicazione: (2025)
di: Tire, Kutay, et al.
Pubblicazione: (2025)
RECOVAR: Representation Covariances on Deep Latent Spaces for Seismic Event Detection
di: Efe, Onur, et al.
Pubblicazione: (2024)
di: Efe, Onur, et al.
Pubblicazione: (2024)
Data-driven Piecewise Affine Decision Rules for Stochastic Programming with Covariate Information
di: Zhang, Yiyang, et al.
Pubblicazione: (2023)
di: Zhang, Yiyang, et al.
Pubblicazione: (2023)
Language Agents for Hypothesis-driven Clinical Decision Making with Reinforcement Learning
di: Bani-Harouni, David, et al.
Pubblicazione: (2025)
di: Bani-Harouni, David, et al.
Pubblicazione: (2025)
GUIDE-VAE: Advancing Data Generation with User Information and Pattern Dictionaries
di: Bölat, Kutay, et al.
Pubblicazione: (2024)
di: Bölat, Kutay, et al.
Pubblicazione: (2024)
Asymptotic Study of In-context Learning with Random Transformers through Equivalent Models
di: Demir, Samet, et al.
Pubblicazione: (2025)
di: Demir, Samet, et al.
Pubblicazione: (2025)
How Data Mixing Shapes In-Context Learning: Asymptotic Equivalence for Transformers with MLPs
di: Demir, Samet, et al.
Pubblicazione: (2025)
di: Demir, Samet, et al.
Pubblicazione: (2025)
Can Mamba Learn How to Learn? A Comparative Study on In-Context Learning Tasks
di: Park, Jongho, et al.
Pubblicazione: (2024)
di: Park, Jongho, et al.
Pubblicazione: (2024)
CONTRAST: Continual Multi-source Adaptation to Dynamic Distributions
di: Ahmed, Sk Miraj, et al.
Pubblicazione: (2024)
di: Ahmed, Sk Miraj, et al.
Pubblicazione: (2024)
Identification and Adaptive Control of Markov Jump Systems: Sample Complexity and Regret Bounds
di: Sattar, Yahya, et al.
Pubblicazione: (2021)
di: Sattar, Yahya, et al.
Pubblicazione: (2021)
Gating is Weighting: Understanding Gated Linear Attention through In-context Learning
di: Li, Yingcong, et al.
Pubblicazione: (2025)
di: Li, Yingcong, et al.
Pubblicazione: (2025)
Mechanics of Next Token Prediction with Self-Attention
di: Li, Yingcong, et al.
Pubblicazione: (2024)
di: Li, Yingcong, et al.
Pubblicazione: (2024)
Beyond the Known: Decision Making with Counterfactual Reasoning Decision Transformer
di: Nguyen, Minh Hoang, et al.
Pubblicazione: (2025)
di: Nguyen, Minh Hoang, et al.
Pubblicazione: (2025)
Documenti analoghi
-
Retrieval Augmented Time Series Forecasting
di: Tire, Kutay, et al.
Pubblicazione: (2024) -
Learning to Bet for Horizon-Aware Anytime-Valid Testing
di: Taga, Ege Onur, et al.
Pubblicazione: (2026) -
TimePFN: Effective Multivariate Time Series Forecasting with Synthetic Data
di: Taga, Ege Onur, et al.
Pubblicazione: (2025) -
Evolutionary Multi-Task Optimization for LLM-Guided Program Discovery
di: Gozeten, Halil Alperen, et al.
Pubblicazione: (2026) -
Learning to Correct: Calibrated Reinforcement Learning for Multi-Attempt Chain-of-Thought
di: Ildiz, Muhammed Emrullah, et al.
Pubblicazione: (2026)