Using Early Readouts to Mediate Featural Bias in Distillation
Fuente:
arXiv
Saved in:
| Main Authors: | Tiwari, Rishabh, Sivasubramanian, Durga, Mekala, Anmol, Ramakrishnan, Ganesh, Shenoy, Pradeep |
|---|---|
| Format: | Preprint |
| Published: |
2023
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Gradient Coreset for Federated Learning
by: Sivasubramanian, Durga, et al.
Published: (2024)
by: Sivasubramanian, Durga, et al.
Published: (2024)
When and How Long? The Readout-Mediator Angle in Temporal Reasoning
by: Fadnavis, Shreyas, et al.
Published: (2026)
by: Fadnavis, Shreyas, et al.
Published: (2026)
Exploring Gradient Subspaces: Addressing and Overcoming LoRA's Limitations in Federated Fine-Tuning of Large Language Models
by: Mahla, Navyansh, et al.
Published: (2024)
by: Mahla, Navyansh, et al.
Published: (2024)
FairPO: Robust Preference Optimization for Fair Multi-Label Learning
by: Mondal, Soumen Kumar, et al.
Published: (2025)
by: Mondal, Soumen Kumar, et al.
Published: (2025)
Root Causing Prediction Anomalies Using Explainable AI
by: Vishnampet, Ramanathan, et al.
Published: (2024)
by: Vishnampet, Ramanathan, et al.
Published: (2024)
Attention Speaks Volumes: Localizing and Mitigating Bias in Language Models
by: Adiga, Rishabh, et al.
Published: (2024)
by: Adiga, Rishabh, et al.
Published: (2024)
VectorFit : Adaptive Singular & Bias Vector Fine-Tuning of Pre-trained Foundation Models
by: Hegde, Suhas G, et al.
Published: (2025)
by: Hegde, Suhas G, et al.
Published: (2025)
SMART: Submodular Data Mixture Strategy for Instruction Tuning
by: Renduchintala, H S V N S Kowndinya, et al.
Published: (2024)
by: Renduchintala, H S V N S Kowndinya, et al.
Published: (2024)
Inducing Robustness in a 2 Dimensional Direct Preference Optimization Paradigm
by: Shashidhar, Sarvesh, et al.
Published: (2025)
by: Shashidhar, Sarvesh, et al.
Published: (2025)
Consistency Is the Key: Detecting Hallucinations in LLM Generated Text By Checking Inconsistencies About Key Facts
by: Gupta, Raavi, et al.
Published: (2025)
by: Gupta, Raavi, et al.
Published: (2025)
Bandit Guided Submodular Curriculum for Adaptive Subset Selection
by: Chanda, Prateek, et al.
Published: (2025)
by: Chanda, Prateek, et al.
Published: (2025)
The Art of Scaling Reinforcement Learning Compute for LLMs
by: Khatri, Devvrit, et al.
Published: (2025)
by: Khatri, Devvrit, et al.
Published: (2025)
Adaptive Few-Shot Learning (AFSL): Tackling Data Scarcity with Stability, Robustness, and Versatility
by: Agrawal, Rishabh
Published: (2025)
by: Agrawal, Rishabh
Published: (2025)
FlashAttention-3: Fast and Accurate Attention with Asynchrony and Low-precision
by: Shah, Jay, et al.
Published: (2024)
by: Shah, Jay, et al.
Published: (2024)
CoRPO: Adding a Correctness Bias to GRPO Improves Generalization
by: Garg, Anisha, et al.
Published: (2025)
by: Garg, Anisha, et al.
Published: (2025)
Offline Safe Reinforcement Learning Using Trajectory Classification
by: Gong, Ze, et al.
Published: (2024)
by: Gong, Ze, et al.
Published: (2024)
Group Relative Knowledge Distillation: Learning from Teacher's Relational Inductive Bias
by: Li, Chao, et al.
Published: (2025)
by: Li, Chao, et al.
Published: (2025)
Auditing Language Model Unlearning via Information Decomposition
by: Goel, Anmol, et al.
Published: (2026)
by: Goel, Anmol, et al.
Published: (2026)
Text2Insight: Transform natural language text into insights seamlessly using multi-model architecture
by: Sain, Pradeep
Published: (2024)
by: Sain, Pradeep
Published: (2024)
DistillSpec: Improving Speculative Decoding via Knowledge Distillation
by: Zhou, Yongchao, et al.
Published: (2023)
by: Zhou, Yongchao, et al.
Published: (2023)
Feature Distillation is the Better Choice for Model-Heterogeneous Federated Learning
by: Li, Yichen, et al.
Published: (2025)
by: Li, Yichen, et al.
Published: (2025)
Complement Submodular Information Measures for Balanced and Robust Data Selection
by: Iyer, Rishabh
Published: (2026)
by: Iyer, Rishabh
Published: (2026)
Different Horses for Different Courses: Comparing Bias Mitigation Algorithms in ML
by: Ganesh, Prakhar, et al.
Published: (2024)
by: Ganesh, Prakhar, et al.
Published: (2024)
Mitigating Bias in Dataset Distillation
by: Cui, Justin, et al.
Published: (2024)
by: Cui, Justin, et al.
Published: (2024)
Task-Driven Causal Feature Distillation: Towards Trustworthy Risk Prediction
by: Chu, Zhixuan, et al.
Published: (2023)
by: Chu, Zhixuan, et al.
Published: (2023)
Enhanced Pruning Strategy for Multi-Component Neural Architectures Using Component-Aware Graph Analysis
by: Sundaram, Ganesh, et al.
Published: (2025)
by: Sundaram, Ganesh, et al.
Published: (2025)
REMEDI: Relative Feature Enhanced Meta-Learning with Distillation for Imbalanced Prediction
by: Liu, Fei, et al.
Published: (2025)
by: Liu, Fei, et al.
Published: (2025)
Improving Group Fairness in Knowledge Distillation via Laplace Approximation of Early Exits
by: Fasth, Edvin, et al.
Published: (2025)
by: Fasth, Edvin, et al.
Published: (2025)
Early Prediction of Sepsis: Feature-Aligned Transfer Learning
by: Komolafe, Oyindolapo O., et al.
Published: (2025)
by: Komolafe, Oyindolapo O., et al.
Published: (2025)
Two Speeds of Learning: A Representation-Readout Decomposition of Grokking and Double Descent
by: Chou, Chi-Ning, et al.
Published: (2026)
by: Chou, Chi-Ning, et al.
Published: (2026)
Model-Level GNN Explanations via Rule-to-Graph Readout for Logit Reconstruction
by: Lu, Shengyao, et al.
Published: (2025)
by: Lu, Shengyao, et al.
Published: (2025)
LLM-Select: Feature Selection with Large Language Models
by: Jeong, Daniel P., et al.
Published: (2024)
by: Jeong, Daniel P., et al.
Published: (2024)
BPL: Bias-adaptive Preference Distillation Learning for Recommender System
by: Kang, SeongKu, et al.
Published: (2025)
by: Kang, SeongKu, et al.
Published: (2025)
Robust Knowledge Distillation Based on Feature Variance Against Backdoored Teacher Model
by: Chen, Jinyin, et al.
Published: (2024)
by: Chen, Jinyin, et al.
Published: (2024)
Lillama: Large Language Models Compression via Low-Rank Feature Distillation
by: Sy, Yaya, et al.
Published: (2024)
by: Sy, Yaya, et al.
Published: (2024)
Masked Generative Nested Transformers with Decode Time Scaling
by: Goyal, Sahil, et al.
Published: (2025)
by: Goyal, Sahil, et al.
Published: (2025)
Compositional Learning of Visually-Grounded Concepts Using Reinforcement
by: Lin, Zijun, et al.
Published: (2023)
by: Lin, Zijun, et al.
Published: (2023)
Does Your Neural Network Extrapolate? Feature Engineering as Identifiability Bias for OOD Generalization
by: Aguilar, Leonel, et al.
Published: (2026)
by: Aguilar, Leonel, et al.
Published: (2026)
On-Policy Distillation of Language Models: Learning from Self-Generated Mistakes
by: Agarwal, Rishabh, et al.
Published: (2023)
by: Agarwal, Rishabh, et al.
Published: (2023)
TabKD: Tabular Knowledge Distillation through Interaction Diversity of Learned Feature Bins
by: Pereira, Shovon Niverd, et al.
Published: (2026)
by: Pereira, Shovon Niverd, et al.
Published: (2026)
Similar Items
-
Gradient Coreset for Federated Learning
by: Sivasubramanian, Durga, et al.
Published: (2024) -
When and How Long? The Readout-Mediator Angle in Temporal Reasoning
by: Fadnavis, Shreyas, et al.
Published: (2026) -
Exploring Gradient Subspaces: Addressing and Overcoming LoRA's Limitations in Federated Fine-Tuning of Large Language Models
by: Mahla, Navyansh, et al.
Published: (2024) -
FairPO: Robust Preference Optimization for Fair Multi-Label Learning
by: Mondal, Soumen Kumar, et al.
Published: (2025) -
Root Causing Prediction Anomalies Using Explainable AI
by: Vishnampet, Ramanathan, et al.
Published: (2024)