Large-Scale Dataset Pruning in Adversarial Training through Data Importance Extrapolation
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Nieth, Björn, Altstidl, Thomas, Schwinn, Leo, Eskofier, Björn |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2024
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Effective Data Pruning through Score Extrapolation
von: Schmidt, Sebastian, et al.
Veröffentlicht: (2025)
von: Schmidt, Sebastian, et al.
Veröffentlicht: (2025)
How Intermodal Interaction Affects the Performance of Deep Multimodal Fusion for Mixed-Type Time Series
von: Dietz, Simon, et al.
Veröffentlicht: (2024)
von: Dietz, Simon, et al.
Veröffentlicht: (2024)
Sampling-aware Adversarial Attacks Against Large Language Models
von: Beyer, Tim, et al.
Veröffentlicht: (2025)
von: Beyer, Tim, et al.
Veröffentlicht: (2025)
How Human-Like Are Large Language Models? A Register-Aware Linguistic Evaluation Framework
von: Nieth, Björn, et al.
Veröffentlicht: (2026)
von: Nieth, Björn, et al.
Veröffentlicht: (2026)
Efficient Adversarial Training in LLMs with Continuous Attacks
von: Xhonneux, Sophie, et al.
Veröffentlicht: (2024)
von: Xhonneux, Sophie, et al.
Veröffentlicht: (2024)
Adversarial Robustness of Graph Transformers
von: Foth, Philipp, et al.
Veröffentlicht: (2024)
von: Foth, Philipp, et al.
Veröffentlicht: (2024)
Diffusion LLMs are Natural Adversaries for any LLM
von: Lüdke, David, et al.
Veröffentlicht: (2025)
von: Lüdke, David, et al.
Veröffentlicht: (2025)
Closing the Distribution Gap in Adversarial Training for LLMs
von: Hu, Chengzhi, et al.
Veröffentlicht: (2026)
von: Hu, Chengzhi, et al.
Veröffentlicht: (2026)
A Probabilistic Perspective on Unlearning and Alignment for Large Language Models
von: Scholten, Yan, et al.
Veröffentlicht: (2024)
von: Scholten, Yan, et al.
Veröffentlicht: (2024)
Caption-Driven Explorations: Aligning Image and Text Embeddings through Human-Inspired Foveated Vision
von: Zanca, Dario, et al.
Veröffentlicht: (2024)
von: Zanca, Dario, et al.
Veröffentlicht: (2024)
Efficient Time Series Processing for Transformers and State-Space Models through Token Merging
von: Götz, Leon, et al.
Veröffentlicht: (2024)
von: Götz, Leon, et al.
Veröffentlicht: (2024)
Spectral Compact Training: Pre-Training Large Language Models via Permanent Truncated SVD and Stiefel QR Retraction
von: Kohlberger, Björn Roman
Veröffentlicht: (2026)
von: Kohlberger, Björn Roman
Veröffentlicht: (2026)
Soft Prompt Threats: Attacking Safety Alignment and Unlearning in Open-Source LLMs through the Embedding Space
von: Schwinn, Leo, et al.
Veröffentlicht: (2024)
von: Schwinn, Leo, et al.
Veröffentlicht: (2024)
Adversarial Alignment for LLMs Requires Simpler, Reproducible, and More Measurable Objectives
von: Schwinn, Leo, et al.
Veröffentlicht: (2025)
von: Schwinn, Leo, et al.
Veröffentlicht: (2025)
Joint Out-of-Distribution Filtering and Data Discovery Active Learning
von: Schmidt, Sebastian, et al.
Veröffentlicht: (2025)
von: Schmidt, Sebastian, et al.
Veröffentlicht: (2025)
The Bid Picture: Auction-Inspired Multi-player Generative Adversarial Networks Training
von: Shim, Joo Yong, et al.
Veröffentlicht: (2024)
von: Shim, Joo Yong, et al.
Veröffentlicht: (2024)
When to retrain a machine learning model
von: Florence, Regol, et al.
Veröffentlicht: (2025)
von: Florence, Regol, et al.
Veröffentlicht: (2025)
Scaling Adversarial Training via Data Selection
von: Ye, Youran, et al.
Veröffentlicht: (2025)
von: Ye, Youran, et al.
Veröffentlicht: (2025)
Byte Pair Encoding for Efficient Time Series Forecasting
von: Götz, Leon, et al.
Veröffentlicht: (2025)
von: Götz, Leon, et al.
Veröffentlicht: (2025)
Enhancing IMU-Based Online Handwriting Recognition via Contrastive Learning with Zero Inference Overhead
von: Li, Jindong, et al.
Veröffentlicht: (2026)
von: Li, Jindong, et al.
Veröffentlicht: (2026)
Assessing Robustness via Score-Based Adversarial Image Generation
von: Kollovieh, Marcel, et al.
Veröffentlicht: (2023)
von: Kollovieh, Marcel, et al.
Veröffentlicht: (2023)
Joint Relational Database Generation via Graph-Conditional Diffusion Models
von: Ketata, Mohamed Amine, et al.
Veröffentlicht: (2025)
von: Ketata, Mohamed Amine, et al.
Veröffentlicht: (2025)
Model Collapse Is Not a Bug but a Feature in Machine Unlearning for LLMs
von: Scholten, Yan, et al.
Veröffentlicht: (2025)
von: Scholten, Yan, et al.
Veröffentlicht: (2025)
How Class Ontology and Data Scale Affect Audio Transfer Learning
von: Milling, Manuel, et al.
Veröffentlicht: (2026)
von: Milling, Manuel, et al.
Veröffentlicht: (2026)
Is Adversarial Training with Compressed Datasets Effective?
von: Chen, Tong, et al.
Veröffentlicht: (2024)
von: Chen, Tong, et al.
Veröffentlicht: (2024)
Think, Prune, Train, Improve: Scaling Reasoning without Scaling Models
von: Costello, Caia, et al.
Veröffentlicht: (2025)
von: Costello, Caia, et al.
Veröffentlicht: (2025)
Soft Label Pruning and Quantization for Large-Scale Dataset Distillation
von: Lingao, Xiao, et al.
Veröffentlicht: (2026)
von: Lingao, Xiao, et al.
Veröffentlicht: (2026)
Adaptive Pruning for Large Language Models with Structural Importance Awareness
von: Zheng, Haotian, et al.
Veröffentlicht: (2024)
von: Zheng, Haotian, et al.
Veröffentlicht: (2024)
Holistic Adversarially Robust Pruning
von: Zhao, Qi, et al.
Veröffentlicht: (2024)
von: Zhao, Qi, et al.
Veröffentlicht: (2024)
Extracting Unlearned Information from LLMs with Activation Steering
von: Seyitoğlu, Atakan, et al.
Veröffentlicht: (2024)
von: Seyitoğlu, Atakan, et al.
Veröffentlicht: (2024)
AdaDeDup: Adaptive Hybrid Data Pruning for Efficient Large-Scale Object Detection Training
von: Kang, Feiyang, et al.
Veröffentlicht: (2025)
von: Kang, Feiyang, et al.
Veröffentlicht: (2025)
Tokenization vs. Augmentation: A Systematic Study of Writer Variance in IMU-Based Online Handwriting Recognition
von: Li, Jindong, et al.
Veröffentlicht: (2026)
von: Li, Jindong, et al.
Veröffentlicht: (2026)
Training-Free Dataset Pruning for Instance Segmentation
von: Dai, Yalun, et al.
Veröffentlicht: (2025)
von: Dai, Yalun, et al.
Veröffentlicht: (2025)
Towards High Supervised Learning Utility Training Data Generation: Data Pruning and Column Reordering
von: Kwok, Tung Sum Thomas, et al.
Veröffentlicht: (2025)
von: Kwok, Tung Sum Thomas, et al.
Veröffentlicht: (2025)
Measuring Sample Importance in Data Pruning for Language Models based on Information Entropy
von: Kim, Minsang, et al.
Veröffentlicht: (2024)
von: Kim, Minsang, et al.
Veröffentlicht: (2024)
IPPRO: Importance-based Pruning with PRojective Offset for Magnitude-indifferent Structural Pruning
von: Jung, Jaeheun, et al.
Veröffentlicht: (2025)
von: Jung, Jaeheun, et al.
Veröffentlicht: (2025)
Flow Matching with Gaussian Process Priors for Probabilistic Time Series Forecasting
von: Kollovieh, Marcel, et al.
Veröffentlicht: (2024)
von: Kollovieh, Marcel, et al.
Veröffentlicht: (2024)
Bayesian Neural Scaling Law Extrapolation with Prior-Data Fitted Networks
von: Lee, Dongwoo, et al.
Veröffentlicht: (2025)
von: Lee, Dongwoo, et al.
Veröffentlicht: (2025)
Scale Efficient Training for Large Datasets
von: Zhou, Qing, et al.
Veröffentlicht: (2025)
von: Zhou, Qing, et al.
Veröffentlicht: (2025)
Random Aggregate Beamforming for Over-the-Air Federated Learning in Large-Scale Networks
von: Xu, Chunmei, et al.
Veröffentlicht: (2024)
von: Xu, Chunmei, et al.
Veröffentlicht: (2024)
Ähnliche Einträge
-
Effective Data Pruning through Score Extrapolation
von: Schmidt, Sebastian, et al.
Veröffentlicht: (2025) -
How Intermodal Interaction Affects the Performance of Deep Multimodal Fusion for Mixed-Type Time Series
von: Dietz, Simon, et al.
Veröffentlicht: (2024) -
Sampling-aware Adversarial Attacks Against Large Language Models
von: Beyer, Tim, et al.
Veröffentlicht: (2025) -
How Human-Like Are Large Language Models? A Register-Aware Linguistic Evaluation Framework
von: Nieth, Björn, et al.
Veröffentlicht: (2026) -
Efficient Adversarial Training in LLMs with Continuous Attacks
von: Xhonneux, Sophie, et al.
Veröffentlicht: (2024)