Transformers Don't In-Context Learn Least Squares Regression
Fuente:
arXiv
Salvato in:
| Autori principali: | Hill, Joshua, Eyre, Benjamin, Creager, Elliot |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2025
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
Ordinary Least Squares is a Special Case of Transformer
di: Tan, Xiaojun, et al.
Pubblicazione: (2026)
di: Tan, Xiaojun, et al.
Pubblicazione: (2026)
Don't Forget the Critic: Value-Based Data Rehearsal for Multi-Cyclic Continual Reinforcement Learning
di: Poole, Benjamin, et al.
Pubblicazione: (2026)
di: Poole, Benjamin, et al.
Pubblicazione: (2026)
Attention Mechanisms Don't Learn Additive Models: Rethinking Feature Importance for Transformers
di: Leemann, Tobias, et al.
Pubblicazione: (2024)
di: Leemann, Tobias, et al.
Pubblicazione: (2024)
Do's and Don'ts: Learning Desirable Skills with Instruction Videos
di: Kim, Hyunseung, et al.
Pubblicazione: (2024)
di: Kim, Hyunseung, et al.
Pubblicazione: (2024)
Out of the Ordinary: Spectrally Adapting Regression for Covariate Shift
di: Eyre, Benjamin, et al.
Pubblicazione: (2023)
di: Eyre, Benjamin, et al.
Pubblicazione: (2023)
Don't Freeze, Don't Crash: Extending the Safe Operating Range of Neural Navigation in Dense Crowds
di: Zhang, Jiefu, et al.
Pubblicazione: (2026)
di: Zhang, Jiefu, et al.
Pubblicazione: (2026)
APFL: Analytic Personalized Federated Learning via Dual-Stream Least Squares
di: Fan, Kejia, et al.
Pubblicazione: (2025)
di: Fan, Kejia, et al.
Pubblicazione: (2025)
Concurrent Learning with Aggregated States via Randomized Least Squares Value Iteration
di: Chen, Yan, et al.
Pubblicazione: (2025)
di: Chen, Yan, et al.
Pubblicazione: (2025)
Intelligent Materials Modelling: Large Language Models Versus Partial Least Squares Regression for Predicting Polysulfone Membrane Mechanical Performance
di: Cao, Dingding, et al.
Pubblicazione: (2026)
di: Cao, Dingding, et al.
Pubblicazione: (2026)
Don't Push the Button! Exploring Data Leakage Risks in Machine Learning and Transfer Learning
di: Apicella, Andrea, et al.
Pubblicazione: (2024)
di: Apicella, Andrea, et al.
Pubblicazione: (2024)
F-GRPO: Don't Let Your Policy Learn the Obvious and Forget the Rare
di: Plyusov, Daniil, et al.
Pubblicazione: (2026)
di: Plyusov, Daniil, et al.
Pubblicazione: (2026)
Transformers Learn Robust In-Context Regression under Distributional Uncertainty
di: Cao, Hoang T. H., et al.
Pubblicazione: (2026)
di: Cao, Hoang T. H., et al.
Pubblicazione: (2026)
Active Slice Discovery in Large Language Models
di: Zhang, Minhui, et al.
Pubblicazione: (2025)
di: Zhang, Minhui, et al.
Pubblicazione: (2025)
DriftXpress: Faster Drifting Models via Projected RKHS Fields
di: Falahati, Ali, et al.
Pubblicazione: (2026)
di: Falahati, Ali, et al.
Pubblicazione: (2026)
KernelSHAP-IQ: Weighted Least-Square Optimization for Shapley Interactions
di: Fumagalli, Fabian, et al.
Pubblicazione: (2024)
di: Fumagalli, Fabian, et al.
Pubblicazione: (2024)
Benchmarking is Broken -- Don't Let AI be its Own Judge
di: Cheng, Zerui, et al.
Pubblicazione: (2025)
di: Cheng, Zerui, et al.
Pubblicazione: (2025)
Don't Forget It! Conditional Sparse Autoencoder Clamping Works for Unlearning
di: Khoriaty, Matthew, et al.
Pubblicazione: (2025)
di: Khoriaty, Matthew, et al.
Pubblicazione: (2025)
Don't Waste Your Time: Early Stopping Cross-Validation
di: Bergman, Edward, et al.
Pubblicazione: (2024)
di: Bergman, Edward, et al.
Pubblicazione: (2024)
Regress, Don't Guess -- A Regression-like Loss on Number Tokens for Language Models
di: Zausinger, Jonas, et al.
Pubblicazione: (2024)
di: Zausinger, Jonas, et al.
Pubblicazione: (2024)
Trust, Don't Trust, or Flip: Robust Preference-Based Reinforcement Learning with Multi-Expert Feedback
di: Hosseini, Seyed Amir, et al.
Pubblicazione: (2026)
di: Hosseini, Seyed Amir, et al.
Pubblicazione: (2026)
Don't Lag, RAG: Training-Free Adversarial Detection Using RAG
di: Kazoom, Roie, et al.
Pubblicazione: (2025)
di: Kazoom, Roie, et al.
Pubblicazione: (2025)
Don't be lazy: CompleteP enables compute-efficient deep transformers
di: Dey, Nolan, et al.
Pubblicazione: (2025)
di: Dey, Nolan, et al.
Pubblicazione: (2025)
Show, Don't Tell: Uncovering Implicit Character Portrayal using LLMs
di: Jaipersaud, Brandon, et al.
Pubblicazione: (2024)
di: Jaipersaud, Brandon, et al.
Pubblicazione: (2024)
xAI-Drop: Don't Use What You Cannot Explain
di: De Luca, Vincenzo Marco, et al.
Pubblicazione: (2024)
di: De Luca, Vincenzo Marco, et al.
Pubblicazione: (2024)
Generative Pre-Trained Transformer for Symbolic Regression Base In-Context Reinforcement Learning
di: Li, Yanjie, et al.
Pubblicazione: (2024)
di: Li, Yanjie, et al.
Pubblicazione: (2024)
Know What You Don't Know: Uncertainty Calibration of Process Reward Models
di: Park, Young-Jin, et al.
Pubblicazione: (2025)
di: Park, Young-Jin, et al.
Pubblicazione: (2025)
Know What You Don't Know: Selective Prediction for Early Exit DNNs
di: Bajpai, Divya Jyoti, et al.
Pubblicazione: (2025)
di: Bajpai, Divya Jyoti, et al.
Pubblicazione: (2025)
Don't throw the baby out with the bathwater: How and why deep learning for ARC
di: Cole, Jack, et al.
Pubblicazione: (2025)
di: Cole, Jack, et al.
Pubblicazione: (2025)
Trust, or Don't Predict: Introducing the CWSA Family for Confidence-Aware Model Evaluation
di: Shahnazari, Kourosh, et al.
Pubblicazione: (2025)
di: Shahnazari, Kourosh, et al.
Pubblicazione: (2025)
Prediction Bottlenecks Don't Discover Causal Structure (But Here's What They Actually Do)
di: Lade, Ankit Hemant, et al.
Pubblicazione: (2026)
di: Lade, Ankit Hemant, et al.
Pubblicazione: (2026)
mHC-lite: You Don't Need 20 Sinkhorn-Knopp Iterations
di: Yang, Yongyi, et al.
Pubblicazione: (2026)
di: Yang, Yongyi, et al.
Pubblicazione: (2026)
Don't stop me now: Rethinking Validation Criteria for Model Parameter Selection
di: Apicella, Andrea, et al.
Pubblicazione: (2026)
di: Apicella, Andrea, et al.
Pubblicazione: (2026)
Reasoning Models Don't Always Say What They Think
di: Chen, Yanda, et al.
Pubblicazione: (2025)
di: Chen, Yanda, et al.
Pubblicazione: (2025)
Don't Blind Your VLA: Aligning Visual Representations for OOD Generalization
di: Kachaev, Nikita, et al.
Pubblicazione: (2025)
di: Kachaev, Nikita, et al.
Pubblicazione: (2025)
Angles Don't Lie: Unlocking Training-Efficient RL Through the Model's Own Signals
di: Wang, Qinsi, et al.
Pubblicazione: (2025)
di: Wang, Qinsi, et al.
Pubblicazione: (2025)
What LLMs Think When You Don't Tell Them What to Think About?
di: Kwon, Yongchan, et al.
Pubblicazione: (2026)
di: Kwon, Yongchan, et al.
Pubblicazione: (2026)
Don't Let Bandit Feedback Pull Continual LLM-Recommender Updates Off Target
di: Kim, Taesan, et al.
Pubblicazione: (2026)
di: Kim, Taesan, et al.
Pubblicazione: (2026)
Don't flatten, tokenize! Unlocking the key to SoftMoE's efficacy in deep RL
di: Sokar, Ghada, et al.
Pubblicazione: (2024)
di: Sokar, Ghada, et al.
Pubblicazione: (2024)
Don't Retrain, Align: Adapting Autoregressive LMs to Diffusion LMs via Representation Alignment
di: Peng, Fred Zhangzhi, et al.
Pubblicazione: (2026)
di: Peng, Fred Zhangzhi, et al.
Pubblicazione: (2026)
You Don't Need Prompt Engineering Anymore: The Prompting Inversion
di: Khan, Imran
Pubblicazione: (2025)
di: Khan, Imran
Pubblicazione: (2025)
Documenti analoghi
-
Ordinary Least Squares is a Special Case of Transformer
di: Tan, Xiaojun, et al.
Pubblicazione: (2026) -
Don't Forget the Critic: Value-Based Data Rehearsal for Multi-Cyclic Continual Reinforcement Learning
di: Poole, Benjamin, et al.
Pubblicazione: (2026) -
Attention Mechanisms Don't Learn Additive Models: Rethinking Feature Importance for Transformers
di: Leemann, Tobias, et al.
Pubblicazione: (2024) -
Do's and Don'ts: Learning Desirable Skills with Instruction Videos
di: Kim, Hyunseung, et al.
Pubblicazione: (2024) -
Out of the Ordinary: Spectrally Adapting Regression for Covariate Shift
di: Eyre, Benjamin, et al.
Pubblicazione: (2023)