Supervised Contrastive Representation Learning: Landscape Analysis with Unconstrained Features
Fuente:
arXiv
Salvato in:
| Autori principali: | Behnia, Tina, Thrampoulidis, Christos |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2024
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
Facts in Stats: Impacts of Pretraining Diversity on Language Model Generalization
di: Behnia, Tina, et al.
Pubblicazione: (2025)
di: Behnia, Tina, et al.
Pubblicazione: (2025)
Implicit Geometry of Next-token Prediction: From Language Sparsity Patterns to Model Representations
di: Zhao, Yize, et al.
Pubblicazione: (2024)
di: Zhao, Yize, et al.
Pubblicazione: (2024)
Why Loss Re-weighting Works If You Stop Early: Training Dynamics of Unconstrained Features
di: Zhao, Yize, et al.
Pubblicazione: (2026)
di: Zhao, Yize, et al.
Pubblicazione: (2026)
In-Context Occam's Razor: How Transformers Prefer Simpler Hypotheses on the Fly
di: Deora, Puneesh, et al.
Pubblicazione: (2025)
di: Deora, Puneesh, et al.
Pubblicazione: (2025)
Implicit Optimization Bias of Next-Token Prediction in Linear Models
di: Thrampoulidis, Christos
Pubblicazione: (2024)
di: Thrampoulidis, Christos
Pubblicazione: (2024)
Sharper Guarantees for Learning Neural Network Classifiers with Gradient Methods
di: Taheri, Hossein, et al.
Pubblicazione: (2024)
di: Taheri, Hossein, et al.
Pubblicazione: (2024)
Memorization Capacity of Multi-Head Attention in Transformers
di: Mahdavi, Sadegh, et al.
Pubblicazione: (2023)
di: Mahdavi, Sadegh, et al.
Pubblicazione: (2023)
Thumb on the Scale: Optimal Loss Weighting in Last Layer Retraining
di: Stromberg, Nathan, et al.
Pubblicazione: (2025)
di: Stromberg, Nathan, et al.
Pubblicazione: (2025)
Unlocking the Potential of Prompt-Tuning in Bridging Generalized and Personalized Federated Learning
di: Deng, Wenlong, et al.
Pubblicazione: (2023)
di: Deng, Wenlong, et al.
Pubblicazione: (2023)
Memory capacity of two layer neural networks with smooth activations
di: Madden, Liam, et al.
Pubblicazione: (2023)
di: Madden, Liam, et al.
Pubblicazione: (2023)
Implicit Bias and Fast Convergence Rates for Self-attention
di: Vasudeva, Bhavya, et al.
Pubblicazione: (2024)
di: Vasudeva, Bhavya, et al.
Pubblicazione: (2024)
Advantage Shaping as Surrogate Reward Maximization: Unifying Pass@K Policy Gradients
di: Thrampoulidis, Christos, et al.
Pubblicazione: (2025)
di: Thrampoulidis, Christos, et al.
Pubblicazione: (2025)
Implicit Bias of Spectral Descent and Muon on Multiclass Separable Data
di: Fan, Chen, et al.
Pubblicazione: (2025)
di: Fan, Chen, et al.
Pubblicazione: (2025)
Neural Collapse Beyond the Unconstrained Features Model: Landscape, Dynamics, and Generalization in the Mean-Field Regime
di: Wu, Diyuan, et al.
Pubblicazione: (2025)
di: Wu, Diyuan, et al.
Pubblicazione: (2025)
Geometric Analysis of Unconstrained Feature Models with $d=K$
di: Shen, Yi, et al.
Pubblicazione: (2024)
di: Shen, Yi, et al.
Pubblicazione: (2024)
Diagonalizing the Softmax: Hadamard Initialization for Tractable Cross-Entropy Dynamics
di: Garrod, Connall, et al.
Pubblicazione: (2025)
di: Garrod, Connall, et al.
Pubblicazione: (2025)
Generalization Analysis for Supervised Contrastive Representation Learning under Non-IID Settings
di: Hieu, Nong Minh, et al.
Pubblicazione: (2025)
di: Hieu, Nong Minh, et al.
Pubblicazione: (2025)
On the Properties of Feature Attribution for Supervised Contrastive Learning
di: Arrighi, Leonardo, et al.
Pubblicazione: (2026)
di: Arrighi, Leonardo, et al.
Pubblicazione: (2026)
A Refined Generalization Analysis for Extreme Multi-class Supervised Contrastive Representation Learning
di: Hieu, Nong Minh, et al.
Pubblicazione: (2026)
di: Hieu, Nong Minh, et al.
Pubblicazione: (2026)
On the Optimization and Generalization of Multi-head Attention
di: Deora, Puneesh, et al.
Pubblicazione: (2023)
di: Deora, Puneesh, et al.
Pubblicazione: (2023)
Unconstrained Stochastic CCA: Unifying Multiview and Self-Supervised Learning
di: Chapman, James, et al.
Pubblicazione: (2023)
di: Chapman, James, et al.
Pubblicazione: (2023)
Next-token prediction capacity: general upper bounds and a lower bound for transformers
di: Madden, Liam, et al.
Pubblicazione: (2024)
di: Madden, Liam, et al.
Pubblicazione: (2024)
How Muon's Spectral Design Benefits Generalization: A Study on Imbalanced Data
di: Vasudeva, Bhavya, et al.
Pubblicazione: (2025)
di: Vasudeva, Bhavya, et al.
Pubblicazione: (2025)
Self-Supervised Contrastive Learning is Approximately Supervised Contrastive Learning
di: Luthra, Achleshwar, et al.
Pubblicazione: (2025)
di: Luthra, Achleshwar, et al.
Pubblicazione: (2025)
Neural Collapse in Cumulative Link Models for Ordinal Regression: An Analysis with Unconstrained Feature Model
di: Ma, Chuang, et al.
Pubblicazione: (2025)
di: Ma, Chuang, et al.
Pubblicazione: (2025)
Learning Representations in Video Game Agents with Supervised Contrastive Imitation Learning
di: Celemin, Carlos, et al.
Pubblicazione: (2025)
di: Celemin, Carlos, et al.
Pubblicazione: (2025)
Time Series Representation Learning with Supervised Contrastive Temporal Transformer
di: Liu, Yuansan, et al.
Pubblicazione: (2024)
di: Liu, Yuansan, et al.
Pubblicazione: (2024)
Generalization Analysis for Deep Contrastive Representation Learning
di: Hieu, Nong Minh, et al.
Pubblicazione: (2024)
di: Hieu, Nong Minh, et al.
Pubblicazione: (2024)
Class-attribute Priors: Adapting Optimization to Heterogeneity and Fairness Objective
di: Zhang, Xuechen, et al.
Pubblicazione: (2024)
di: Zhang, Xuechen, et al.
Pubblicazione: (2024)
Understanding Contextual Recall in Transformers: How Finetuning Enables In-Context Reasoning over Pretraining Knowledge
di: Vasudeva, Bhavya, et al.
Pubblicazione: (2026)
di: Vasudeva, Bhavya, et al.
Pubblicazione: (2026)
Supervised Contrastive Frame Aggregation for Video Representation Learning
di: Chowdhury, Shaif, et al.
Pubblicazione: (2025)
di: Chowdhury, Shaif, et al.
Pubblicazione: (2025)
Prototypical Contrastive Learning For Improved Few-Shot Audio Classification
di: Sgouropoulos, Christos, et al.
Pubblicazione: (2025)
di: Sgouropoulos, Christos, et al.
Pubblicazione: (2025)
Neural Collapse for Cross-entropy Class-Imbalanced Learning with Unconstrained ReLU Feature Model
di: Dang, Hien, et al.
Pubblicazione: (2024)
di: Dang, Hien, et al.
Pubblicazione: (2024)
Subgraph Gaussian Embedding Contrast for Self-Supervised Graph Representation Learning
di: Xie, Shifeng, et al.
Pubblicazione: (2025)
di: Xie, Shifeng, et al.
Pubblicazione: (2025)
Neural Multivariate Regression: Qualitative Insights from the Unconstrained Feature Model
di: Andriopoulos, George, et al.
Pubblicazione: (2025)
di: Andriopoulos, George, et al.
Pubblicazione: (2025)
Directional Alignment Mitigates Reward Hacking in Reinforcement Learning for Language Models
di: Deng, Wenlong, et al.
Pubblicazione: (2026)
di: Deng, Wenlong, et al.
Pubblicazione: (2026)
On the Alignment Between Supervised and Self-Supervised Contrastive Learning
di: Luthra, Achleshwar, et al.
Pubblicazione: (2025)
di: Luthra, Achleshwar, et al.
Pubblicazione: (2025)
Fully Unconstrained Online Learning
di: Cutkosky, Ashok, et al.
Pubblicazione: (2024)
di: Cutkosky, Ashok, et al.
Pubblicazione: (2024)
Transformers as Support Vector Machines
di: Tarzanagh, Davoud Ataee, et al.
Pubblicazione: (2023)
di: Tarzanagh, Davoud Ataee, et al.
Pubblicazione: (2023)
Beyond Unconstrained Features: Neural Collapse for Shallow Neural Networks with General Data
di: Hong, Wanli, et al.
Pubblicazione: (2024)
di: Hong, Wanli, et al.
Pubblicazione: (2024)
Documenti analoghi
-
Facts in Stats: Impacts of Pretraining Diversity on Language Model Generalization
di: Behnia, Tina, et al.
Pubblicazione: (2025) -
Implicit Geometry of Next-token Prediction: From Language Sparsity Patterns to Model Representations
di: Zhao, Yize, et al.
Pubblicazione: (2024) -
Why Loss Re-weighting Works If You Stop Early: Training Dynamics of Unconstrained Features
di: Zhao, Yize, et al.
Pubblicazione: (2026) -
In-Context Occam's Razor: How Transformers Prefer Simpler Hypotheses on the Fly
di: Deora, Puneesh, et al.
Pubblicazione: (2025) -
Implicit Optimization Bias of Next-Token Prediction in Linear Models
di: Thrampoulidis, Christos
Pubblicazione: (2024)