Provably Learning from Modern Language Models via Low Logit Rank
Fuente:
arXiv
Salvato in:
| Autori principali: | Golowich, Noah, Liu, Allen, Shetty, Abhishek |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2025
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
Sequences of Logits Reveal the Low Rank Structure of Language Models
di: Golowich, Noah, et al.
Pubblicazione: (2025)
di: Golowich, Noah, et al.
Pubblicazione: (2025)
Model Stealing for Any Low-Rank Language Model
di: Liu, Allen, et al.
Pubblicazione: (2024)
di: Liu, Allen, et al.
Pubblicazione: (2024)
Provably Extracting the Features from a General Superposition
di: Liu, Allen
Pubblicazione: (2025)
di: Liu, Allen
Pubblicazione: (2025)
Smooth Nash Equilibria: Algorithms and Complexity
di: Daskalakis, Constantinos, et al.
Pubblicazione: (2023)
di: Daskalakis, Constantinos, et al.
Pubblicazione: (2023)
Exploration is Harder than Prediction: Cryptographically Separating Reinforcement Learning from Supervised Learning
di: Golowich, Noah, et al.
Pubblicazione: (2024)
di: Golowich, Noah, et al.
Pubblicazione: (2024)
Learning with Monotone Adversarial Corruptions
di: Larsen, Kasper Green, et al.
Pubblicazione: (2026)
di: Larsen, Kasper Green, et al.
Pubblicazione: (2026)
Adversarial Resilience in Sequential Prediction via Abstention
di: Goel, Surbhi, et al.
Pubblicazione: (2023)
di: Goel, Surbhi, et al.
Pubblicazione: (2023)
Tolerant Algorithms for Learning with Arbitrary Covariate Shift
di: Goel, Surbhi, et al.
Pubblicazione: (2024)
di: Goel, Surbhi, et al.
Pubblicazione: (2024)
Breaking the $T^{2/3}$ Barrier for Sequential Calibration
di: Dagan, Yuval, et al.
Pubblicazione: (2024)
di: Dagan, Yuval, et al.
Pubblicazione: (2024)
Self-Supervised Graph Learning via Spectral Bootstrapping and Laplacian-Based Augmentations
di: Bini, Lorenzo, et al.
Pubblicazione: (2025)
di: Bini, Lorenzo, et al.
Pubblicazione: (2025)
An Algorithm for Learning Smaller Representations of Models With Scarce Data
di: de Wynter, Adrian
Pubblicazione: (2020)
di: de Wynter, Adrian
Pubblicazione: (2020)
Robust Second-Order Nonconvex Optimization and Its Application to Low Rank Matrix Sensing
di: Li, Shuyao, et al.
Pubblicazione: (2024)
di: Li, Shuyao, et al.
Pubblicazione: (2024)
Omnipredictors for Regression and the Approximate Rank of Convex Functions
di: Gopalan, Parikshit, et al.
Pubblicazione: (2024)
di: Gopalan, Parikshit, et al.
Pubblicazione: (2024)
On Tradeoffs in Learning-Augmented Algorithms
di: Benomar, Ziyad, et al.
Pubblicazione: (2025)
di: Benomar, Ziyad, et al.
Pubblicazione: (2025)
Anytime-Constrained Reinforcement Learning
di: McMahan, Jeremy, et al.
Pubblicazione: (2023)
di: McMahan, Jeremy, et al.
Pubblicazione: (2023)
Learning-Augmented Priority Queues
di: Benomar, Ziyad, et al.
Pubblicazione: (2024)
di: Benomar, Ziyad, et al.
Pubblicazione: (2024)
Streaming Attention Approximation via Discrepancy Theory
di: Kochetkova, Ekaterina, et al.
Pubblicazione: (2025)
di: Kochetkova, Ekaterina, et al.
Pubblicazione: (2025)
AlgoSelect: Universal Algorithm Selection via the Comb Operator
di: Yao, Jasper
Pubblicazione: (2025)
di: Yao, Jasper
Pubblicazione: (2025)
Learning-Augmented Online Bipartite Fractional Matching
di: Choo, Davin, et al.
Pubblicazione: (2025)
di: Choo, Davin, et al.
Pubblicazione: (2025)
Polynomial-Time Approximability of Constrained Reinforcement Learning
di: McMahan, Jeremy
Pubblicazione: (2025)
di: McMahan, Jeremy
Pubblicazione: (2025)
Learning-Based Algorithms for Graph Searching Problems
di: DePavia, Adela Frances, et al.
Pubblicazione: (2024)
di: DePavia, Adela Frances, et al.
Pubblicazione: (2024)
Efficiently Learning Branching Networks for Multitask Algorithmic Reasoning
di: Li, Dongyue, et al.
Pubblicazione: (2025)
di: Li, Dongyue, et al.
Pubblicazione: (2025)
Online Learning with Probing for Sequential User-Centric Selection
di: Xu, Tianyi, et al.
Pubblicazione: (2025)
di: Xu, Tianyi, et al.
Pubblicazione: (2025)
Approximate Lifted Model Construction
di: Luttermann, Malte, et al.
Pubblicazione: (2025)
di: Luttermann, Malte, et al.
Pubblicazione: (2025)
Learning-augmented smooth integer programs with PAC-learnable oracles
di: He, Hao-Yuan, et al.
Pubblicazione: (2026)
di: He, Hao-Yuan, et al.
Pubblicazione: (2026)
Minimum Weighted Feedback Arc Sets for Ranking from Pairwise Comparisons
di: Vahidi, Soroush, et al.
Pubblicazione: (2024)
di: Vahidi, Soroush, et al.
Pubblicazione: (2024)
Rethinking Flexible Graph Similarity Computation: One-step Alignment with Global Guidance
di: Liu, Zhouyang, et al.
Pubblicazione: (2025)
di: Liu, Zhouyang, et al.
Pubblicazione: (2025)
Differentially Private Kernel Density Estimation
di: Liu, Erzhi, et al.
Pubblicazione: (2024)
di: Liu, Erzhi, et al.
Pubblicazione: (2024)
High-Dimensional Calibration from Swap Regret
di: Fishelson, Maxwell, et al.
Pubblicazione: (2025)
di: Fishelson, Maxwell, et al.
Pubblicazione: (2025)
Constructing Decision Trees from Data Streams
di: Pham, Huy, et al.
Pubblicazione: (2024)
di: Pham, Huy, et al.
Pubblicazione: (2024)
Rethinking Model-based, Policy-based, and Value-based Reinforcement Learning via the Lens of Representation Complexity
di: Feng, Guhao, et al.
Pubblicazione: (2023)
di: Feng, Guhao, et al.
Pubblicazione: (2023)
Parallel Sampling via Counting
di: Anari, Nima, et al.
Pubblicazione: (2024)
di: Anari, Nima, et al.
Pubblicazione: (2024)
Faster Low-Rank Approximation and Kernel Ridge Regression via the Block-Nyström Method
di: Garg, Sachin, et al.
Pubblicazione: (2025)
di: Garg, Sachin, et al.
Pubblicazione: (2025)
Optimizing Text Search: A Novel Pattern Matching Algorithm Based on Ukkonen's Approach
di: Guan, Xinyu, et al.
Pubblicazione: (2025)
di: Guan, Xinyu, et al.
Pubblicazione: (2025)
Demand Selection for VRP with Emission Quota
di: Najar, Farid, et al.
Pubblicazione: (2025)
di: Najar, Farid, et al.
Pubblicazione: (2025)
Fast EXP3 Algorithms
di: Sato, Ryoma, et al.
Pubblicazione: (2025)
di: Sato, Ryoma, et al.
Pubblicazione: (2025)
Uncovering Fairness through Data Complexity as an Early Indicator
di: Ferreira, Juliett Suárez, et al.
Pubblicazione: (2025)
di: Ferreira, Juliett Suárez, et al.
Pubblicazione: (2025)
DiscQuant: A Quantization Method for Neural Networks Inspired by Discrepancy Theory
di: Chee, Jerry, et al.
Pubblicazione: (2025)
di: Chee, Jerry, et al.
Pubblicazione: (2025)
Optimal Classification Trees for Continuous Feature Data Using Dynamic Programming with Branch-and-Bound
di: Brita, Catalin E., et al.
Pubblicazione: (2025)
di: Brita, Catalin E., et al.
Pubblicazione: (2025)
Pareto Optimal Algorithmic Recourse in Multi-cost Function
di: Chen, Wen-Ling, et al.
Pubblicazione: (2025)
di: Chen, Wen-Ling, et al.
Pubblicazione: (2025)
Documenti analoghi
-
Sequences of Logits Reveal the Low Rank Structure of Language Models
di: Golowich, Noah, et al.
Pubblicazione: (2025) -
Model Stealing for Any Low-Rank Language Model
di: Liu, Allen, et al.
Pubblicazione: (2024) -
Provably Extracting the Features from a General Superposition
di: Liu, Allen
Pubblicazione: (2025) -
Smooth Nash Equilibria: Algorithms and Complexity
di: Daskalakis, Constantinos, et al.
Pubblicazione: (2023) -
Exploration is Harder than Prediction: Cryptographically Separating Reinforcement Learning from Supervised Learning
di: Golowich, Noah, et al.
Pubblicazione: (2024)