LayerMatch: Do Pseudo-labels Benefit All Layers?
Fuente:
arXiv
Saved in:
| Main Authors: | Liang, Chaoqi, Yang, Guanglei, Qiao, Lifeng, Huang, Zitong, Yan, Hongliang, Wei, Yunchao, Zuo, Wangmeng |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
MIRAGE: Assessing Hallucination in Multimodal Reasoning Chains of MLLM
by: Dong, Bowen, et al.
Published: (2025)
by: Dong, Bowen, et al.
Published: (2025)
Toward Understanding BERT-Like Pre-Training for DNA Foundation Models
by: Liang, Chaoqi, et al.
Published: (2023)
by: Liang, Chaoqi, et al.
Published: (2023)
IMWA: Iterative Model Weight Averaging Benefits Class-Imbalanced Learning Tasks
by: Huang, Zitong, et al.
Published: (2024)
by: Huang, Zitong, et al.
Published: (2024)
On the Benefits of Rank in Attention Layers
by: Amsel, Noah, et al.
Published: (2024)
by: Amsel, Noah, et al.
Published: (2024)
Reduction-based Pseudo-label Generation for Instance-dependent Partial Label Learning
by: Qiao, Congyu, et al.
Published: (2024)
by: Qiao, Congyu, et al.
Published: (2024)
Model Decides How to Tokenize: Adaptive DNA Sequence Tokenization with MxDNA
by: Qiao, Lifeng, et al.
Published: (2024)
by: Qiao, Lifeng, et al.
Published: (2024)
MR-GDINO: Efficient Open-World Continual Object Detection
by: Dong, Bowen, et al.
Published: (2024)
by: Dong, Bowen, et al.
Published: (2024)
All Nodes are created Not Equal: Node-Specific Layer Aggregation and Filtration for GNN
by: Wang, Shilong, et al.
Published: (2024)
by: Wang, Shilong, et al.
Published: (2024)
Mechanistic Permutability: Match Features Across Layers
by: Balagansky, Nikita, et al.
Published: (2024)
by: Balagansky, Nikita, et al.
Published: (2024)
Layer-Aware Analysis of Catastrophic Overfitting: Revealing the Pseudo-Robust Shortcut Dependency
by: Lin, Runqi, et al.
Published: (2024)
by: Lin, Runqi, et al.
Published: (2024)
scMRDR: A scalable and flexible framework for unpaired single-cell multi-omics data integration
by: Sun, Jianle, et al.
Published: (2025)
by: Sun, Jianle, et al.
Published: (2025)
CPT: Consistent Proxy Tuning for Black-box Optimization
by: He, Yuanyang, et al.
Published: (2024)
by: He, Yuanyang, et al.
Published: (2024)
Not Only the Last-Layer Features for Spurious Correlations: All Layer Deep Feature Reweighting
by: Hameed, Humza Wajid, et al.
Published: (2024)
by: Hameed, Humza Wajid, et al.
Published: (2024)
Understanding the Benefits of SimCLR Pre-Training in Two-Layer Convolutional Neural Networks
by: Zhang, Han, et al.
Published: (2024)
by: Zhang, Han, et al.
Published: (2024)
Linear Probe Accuracy Scales with Model Size and Benefits from Multi-Layer Ensembling
by: Nordby, Erik, et al.
Published: (2026)
by: Nordby, Erik, et al.
Published: (2026)
Uncertainty-aware Pseudo-label Selection for Positive-Unlabeled Learning
by: Dorigatti, Emilio, et al.
Published: (2022)
by: Dorigatti, Emilio, et al.
Published: (2022)
CGL: Advancing Continual GUI Learning via Reinforcement Fine-Tuning
by: Yao, Zhenquan, et al.
Published: (2026)
by: Yao, Zhenquan, et al.
Published: (2026)
LPT++: Efficient Training on Mixture of Long-tailed Experts
by: Dong, Bowen, et al.
Published: (2024)
by: Dong, Bowen, et al.
Published: (2024)
UPL: Uncertainty-aware Pseudo-labeling for Imbalance Transductive Node Classification
by: Teimuri, Mohammad T., et al.
Published: (2025)
by: Teimuri, Mohammad T., et al.
Published: (2025)
MatchXML: An Efficient Text-label Matching Framework for Extreme Multi-label Text Classification
by: Ye, Hui, et al.
Published: (2023)
by: Ye, Hui, et al.
Published: (2023)
Deciphering the Chaos: Enhancing Jailbreak Attacks via Adversarial Prompt Translation
by: Li, Qizhang, et al.
Published: (2024)
by: Li, Qizhang, et al.
Published: (2024)
Latent Code Augmentation Based on Stable Diffusion for Data-free Substitute Attacks
by: Shao, Mingwen, et al.
Published: (2023)
by: Shao, Mingwen, et al.
Published: (2023)
Not All Layers of LLMs Are Necessary During Inference
by: Fan, Siqi, et al.
Published: (2024)
by: Fan, Siqi, et al.
Published: (2024)
Concise One-Layer Transformers Can Do Function Evaluation (Sometimes)
by: Strobl, Lena, et al.
Published: (2025)
by: Strobl, Lena, et al.
Published: (2025)
The Benefits of Reusing Batches for Gradient Descent in Two-Layer Networks: Breaking the Curse of Information and Leap Exponents
by: Dandi, Yatin, et al.
Published: (2024)
by: Dandi, Yatin, et al.
Published: (2024)
Learnable Multipliers: Freeing the Scale of Language Model Matrix Layers
by: Velikanov, Maksim, et al.
Published: (2026)
by: Velikanov, Maksim, et al.
Published: (2026)
Semi-supervised CAPP Transformer Learning via Pseudo-labeling
by: Gross, Dennis, et al.
Published: (2026)
by: Gross, Dennis, et al.
Published: (2026)
Segmenting Objectiveness and Task-awareness Unknown Region for Autonomous Driving
by: Zheng, Mi, et al.
Published: (2025)
by: Zheng, Mi, et al.
Published: (2025)
Improved Generation of Adversarial Examples Against Safety-aligned LLMs
by: Li, Qizhang, et al.
Published: (2024)
by: Li, Qizhang, et al.
Published: (2024)
Single-View Graph Contrastive Learning with Soft Neighborhood Awareness
by: Sun, Qingqiang, et al.
Published: (2024)
by: Sun, Qingqiang, et al.
Published: (2024)
Where Do Flow Semantics Reside? A Protocol-Native Tabular Pretraining Paradigm for Encrypted Traffic Classification
by: Huang, Sizhe, et al.
Published: (2026)
by: Huang, Sizhe, et al.
Published: (2026)
Similarity is Not All You Need: Endowing Retrieval Augmented Generation with Multi Layered Thoughts
by: Gan, Chunjing, et al.
Published: (2024)
by: Gan, Chunjing, et al.
Published: (2024)
A Single-Layer Model Can Do Language Modeling
by: Wang, Zanmin
Published: (2026)
by: Wang, Zanmin
Published: (2026)
Reliable Pseudo-labeling via Optimal Transport with Attention for Short Text Clustering
by: Yao, Zhihao
Published: (2025)
by: Yao, Zhihao
Published: (2025)
On the Nonlinearity of Layer Normalization
by: Ni, Yunhao, et al.
Published: (2024)
by: Ni, Yunhao, et al.
Published: (2024)
Layer-Aware Influence for Online Data Valuation Estimation
by: Yang, Ziao, et al.
Published: (2025)
by: Yang, Ziao, et al.
Published: (2025)
DiCaP: Distribution-Calibrated Pseudo-labeling for Semi-Supervised Multi-Label Learning
by: Han, Bo, et al.
Published: (2025)
by: Han, Bo, et al.
Published: (2025)
Generalized Category Discovery via Token Manifold Capacity Learning
by: Tang, Luyao, et al.
Published: (2025)
by: Tang, Luyao, et al.
Published: (2025)
Pseudo-label Refinement for Improving Self-Supervised Learning Systems
by: Zia-ur-Rehman, et al.
Published: (2024)
by: Zia-ur-Rehman, et al.
Published: (2024)
Variational Inference on the Final-Layer Output of Neural Networks
by: Wei, Yadi, et al.
Published: (2023)
by: Wei, Yadi, et al.
Published: (2023)
Similar Items
-
MIRAGE: Assessing Hallucination in Multimodal Reasoning Chains of MLLM
by: Dong, Bowen, et al.
Published: (2025) -
Toward Understanding BERT-Like Pre-Training for DNA Foundation Models
by: Liang, Chaoqi, et al.
Published: (2023) -
IMWA: Iterative Model Weight Averaging Benefits Class-Imbalanced Learning Tasks
by: Huang, Zitong, et al.
Published: (2024) -
On the Benefits of Rank in Attention Layers
by: Amsel, Noah, et al.
Published: (2024) -
Reduction-based Pseudo-label Generation for Instance-dependent Partial Label Learning
by: Qiao, Congyu, et al.
Published: (2024)