Leveraging Data Symmetries to Select an Optimal Subset of Training Data under Label Noise
Fuente:
arXiv
Saved in:
| Main Authors: | Shubham, Kumar, Karjol, Pavan, K, Kiran M, AP, Prathosh |
|---|---|
| Format: | Preprint |
| Published: |
2026
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Spectral Discovery of Continuous Symmetries via Generalized Fourier Transforms
by: Karjol, Pavan, et al.
Published: (2026)
by: Karjol, Pavan, et al.
Published: (2026)
Interpretable Discovery of One-parameter Subgroups: A Modular Framework for Elliptical, Hyperbolic, and Parabolic Symmetries
by: Karjol, Pavan, et al.
Published: (2025)
by: Karjol, Pavan, et al.
Published: (2025)
Learning Equivariant Functions via Quadratic Forms
by: Karjol, Pavan, et al.
Published: (2025)
by: Karjol, Pavan, et al.
Published: (2025)
FAST: Feature Aware Similarity Thresholding for Weak Unlearning in Black-Box Generative Models
by: Panda, Subhodip, et al.
Published: (2023)
by: Panda, Subhodip, et al.
Published: (2023)
WISER: Weak supervISion and supErvised Representation learning to improve drug response prediction in cancer
by: Shubham, Kumar, et al.
Published: (2024)
by: Shubham, Kumar, et al.
Published: (2024)
You Only Train Once: Differentiable Subset Selection for Omics Data
by: Chopard, Daphné, et al.
Published: (2025)
by: Chopard, Daphné, et al.
Published: (2025)
Class-based Subset Selection for Transfer Learning under Extreme Label Shift
by: Goyal, Akul, et al.
Published: (2024)
by: Goyal, Akul, et al.
Published: (2024)
DataS^3: Dataset Subset Selection for Specialization
by: Hulkund, Neha, et al.
Published: (2025)
by: Hulkund, Neha, et al.
Published: (2025)
Bayesian Pseudo-Coresets via Contrastive Divergence
by: Tiwary, Piyush, et al.
Published: (2023)
by: Tiwary, Piyush, et al.
Published: (2023)
AdaKD: Dynamic Knowledge Distillation of ASR models using Adaptive Loss Weighting
by: Ganguly, Shreyan, et al.
Published: (2024)
by: Ganguly, Shreyan, et al.
Published: (2024)
Selecting Subsets of Source Data for Transfer Learning with Applications in Metal Additive Manufacturing
by: Tang, Yifan, et al.
Published: (2024)
by: Tang, Yifan, et al.
Published: (2024)
f-INE: A Hypothesis Testing Framework for Estimating Influence under Training Randomness
by: Panda, Subhodip, et al.
Published: (2025)
by: Panda, Subhodip, et al.
Published: (2025)
Market-Driven Subset Selection for Budgeted Training
by: Jha, Ashish, et al.
Published: (2025)
by: Jha, Ashish, et al.
Published: (2025)
Learning under Temporal Label Noise
by: Nagaraj, Sujay, et al.
Published: (2024)
by: Nagaraj, Sujay, et al.
Published: (2024)
LoMOE: Localized Multi-Object Editing via Multi-Diffusion
by: Chakrabarty, Goirik, et al.
Published: (2024)
by: Chakrabarty, Goirik, et al.
Published: (2024)
ReMOVE: A Reference-free Metric for Object Erasure
by: Chandrasekar, Aditya, et al.
Published: (2024)
by: Chandrasekar, Aditya, et al.
Published: (2024)
Improve Knowledge Distillation via Label Revision and Data Selection
by: Lan, Weichao, et al.
Published: (2024)
by: Lan, Weichao, et al.
Published: (2024)
Delving into Instance-Dependent Label Noise in Graph Data: A Comprehensive Study and Benchmark
by: Kim, Suyeon, et al.
Published: (2025)
by: Kim, Suyeon, et al.
Published: (2025)
Hybrid Autoencoders for Tabular Data: Leveraging Model-Based Augmentation in Low-Label Settings
by: Naor, Erel, et al.
Published: (2025)
by: Naor, Erel, et al.
Published: (2025)
Impact of Label Noise on Learning Complex Features
by: Vashisht, Rahul, et al.
Published: (2024)
by: Vashisht, Rahul, et al.
Published: (2024)
Leveraging Sub-Optimal Data for Human-in-the-Loop Reinforcement Learning
by: Muslimani, Calarina, et al.
Published: (2024)
by: Muslimani, Calarina, et al.
Published: (2024)
Learning from Noisy Labels for Long-tailed Data via Optimal Transport
by: Li, Mengting, et al.
Published: (2024)
by: Li, Mengting, et al.
Published: (2024)
Relabeling Minimal Training Subset to Flip a Prediction
by: Yang, Jinghan, et al.
Published: (2023)
by: Yang, Jinghan, et al.
Published: (2023)
Bandit Guided Submodular Curriculum for Adaptive Subset Selection
by: Chanda, Prateek, et al.
Published: (2025)
by: Chanda, Prateek, et al.
Published: (2025)
Scales++: Compute Efficient Evaluation Subset Selection with Cognitive Scales Embeddings
by: Bean, Andrew M., et al.
Published: (2025)
by: Bean, Andrew M., et al.
Published: (2025)
Concept Influence: Leveraging Interpretability to Improve Performance and Efficiency in Training Data Attribution
by: Kowal, Matthew, et al.
Published: (2026)
by: Kowal, Matthew, et al.
Published: (2026)
HTM-EAR: Importance-Preserving Tiered Memory with Hybrid Routing under Saturation
by: Singh, Shubham Kumar
Published: (2026)
by: Singh, Shubham Kumar
Published: (2026)
Robust Deep Hawkes Process under Label Noise of Both Event and Occurrence
by: Tan, Xiaoyu, et al.
Published: (2024)
by: Tan, Xiaoyu, et al.
Published: (2024)
Improving Open-world Continual Learning under the Constraints of Scarce Labeled Data
by: Li, Yujie, et al.
Published: (2025)
by: Li, Yujie, et al.
Published: (2025)
Training Data Selection with Gradient Orthogonality for Efficient Domain Adaptation
by: Zhang, Xiyang, et al.
Published: (2026)
by: Zhang, Xiyang, et al.
Published: (2026)
Resurrecting Label Propagation for Graphs with Heterophily and Label Noise
by: Cheng, Yao, et al.
Published: (2023)
by: Cheng, Yao, et al.
Published: (2023)
Controlling Grokking with Nonlinearity and Data Symmetry
by: Salah, Ahmed, et al.
Published: (2024)
by: Salah, Ahmed, et al.
Published: (2024)
TabMDA: Tabular Manifold Data Augmentation for Any Classifier using Transformers with In-context Subsetting
by: Margeloiu, Andrei, et al.
Published: (2024)
by: Margeloiu, Andrei, et al.
Published: (2024)
Downstream Task-Oriented Generative Model Selections on Synthetic Data Training for Fraud Detection Models
by: Cheng, Yinan, et al.
Published: (2024)
by: Cheng, Yinan, et al.
Published: (2024)
The Effect of Data Poisoning on Counterfactual Explanations
by: Artelt, André, et al.
Published: (2024)
by: Artelt, André, et al.
Published: (2024)
Leveraging the Power of Conversations: Optimal Key Term Selection in Conversational Contextual Bandits
by: Liu, Maoli, et al.
Published: (2025)
by: Liu, Maoli, et al.
Published: (2025)
HyperCore: Coreset Selection under Noise via Hypersphere Models
by: Moser, Brian B., et al.
Published: (2025)
by: Moser, Brian B., et al.
Published: (2025)
Gradient Inversion Transcript: Leveraging Robust Generative Priors to Reconstruct Training Data from Gradient Leakage
by: Chen, Xinping, et al.
Published: (2025)
by: Chen, Xinping, et al.
Published: (2025)
Learning-Zone Energy: Online Data Selection for Efficient RL Post-Training
by: Cui, Peng, et al.
Published: (2026)
by: Cui, Peng, et al.
Published: (2026)
Uncertainty-Aware Gradient Signal-to-Noise Data Selection for Instruction Tuning
by: Yuan, Zhihang, et al.
Published: (2026)
by: Yuan, Zhihang, et al.
Published: (2026)
Similar Items
-
Spectral Discovery of Continuous Symmetries via Generalized Fourier Transforms
by: Karjol, Pavan, et al.
Published: (2026) -
Interpretable Discovery of One-parameter Subgroups: A Modular Framework for Elliptical, Hyperbolic, and Parabolic Symmetries
by: Karjol, Pavan, et al.
Published: (2025) -
Learning Equivariant Functions via Quadratic Forms
by: Karjol, Pavan, et al.
Published: (2025) -
FAST: Feature Aware Similarity Thresholding for Weak Unlearning in Black-Box Generative Models
by: Panda, Subhodip, et al.
Published: (2023) -
WISER: Weak supervISion and supErvised Representation learning to improve drug response prediction in cancer
by: Shubham, Kumar, et al.
Published: (2024)