Rethinking Large-scale Dataset Compression: Shifting Focus From Labels to Images
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Xiao, Lingao, Liu, Songhua, He, Yang, Wang, Xinchao |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2025
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Soft Label Pruning and Quantization for Large-Scale Dataset Distillation
von: Lingao, Xiao, et al.
Veröffentlicht: (2026)
von: Lingao, Xiao, et al.
Veröffentlicht: (2026)
Are Large-scale Soft Labels Necessary for Large-scale Dataset Distillation?
von: Xiao, Lingao, et al.
Veröffentlicht: (2024)
von: Xiao, Lingao, et al.
Veröffentlicht: (2024)
Understanding Dataset Distillation via Spectral Filtering
von: Bo, Deyu, et al.
Veröffentlicht: (2025)
von: Bo, Deyu, et al.
Veröffentlicht: (2025)
LinFusion: 1 GPU, 1 Minute, 16K Image
von: Liu, Songhua, et al.
Veröffentlicht: (2024)
von: Liu, Songhua, et al.
Veröffentlicht: (2024)
One-shot Federated Learning via Synthetic Distiller-Distillate Communication
von: Zhang, Junyuan, et al.
Veröffentlicht: (2024)
von: Zhang, Junyuan, et al.
Veröffentlicht: (2024)
Training-Free Dataset Pruning for Instance Segmentation
von: Dai, Yalun, et al.
Veröffentlicht: (2025)
von: Dai, Yalun, et al.
Veröffentlicht: (2025)
CoDA: From Text-to-Image Diffusion Models to Training-Free Dataset Distillation
von: Zhou, Letian, et al.
Veröffentlicht: (2025)
von: Zhou, Letian, et al.
Veröffentlicht: (2025)
Distilled Datamodel with Reverse Gradient Matching
von: Ye, Jingwen, et al.
Veröffentlicht: (2024)
von: Ye, Jingwen, et al.
Veröffentlicht: (2024)
StyDeSty: Min-Max Stylization and Destylization for Single Domain Generalization
von: Liu, Songhua, et al.
Veröffentlicht: (2024)
von: Liu, Songhua, et al.
Veröffentlicht: (2024)
Heavy Labels Out! Dataset Distillation with Label Space Lightening
von: Yu, Ruonan, et al.
Veröffentlicht: (2024)
von: Yu, Ruonan, et al.
Veröffentlicht: (2024)
OminiControl: Minimal and Universal Control for Diffusion Transformer
von: Tan, Zhenxiong, et al.
Veröffentlicht: (2024)
von: Tan, Zhenxiong, et al.
Veröffentlicht: (2024)
Teddy: Efficient Large-Scale Dataset Distillation via Taylor-Approximated Matching
von: Yu, Ruonan, et al.
Veröffentlicht: (2024)
von: Yu, Ruonan, et al.
Veröffentlicht: (2024)
Large-scale Dataset Pruning with Dynamic Uncertainty
von: He, Muyang, et al.
Veröffentlicht: (2023)
von: He, Muyang, et al.
Veröffentlicht: (2023)
Rethinking Dataset Distillation: Hard Truths about Soft Labels
von: Dey, Priyam, et al.
Veröffentlicht: (2026)
von: Dey, Priyam, et al.
Veröffentlicht: (2026)
Dataset Color Quantization: A Training-Oriented Framework for Dataset-Level Compression
von: Yu, Chenyue, et al.
Veröffentlicht: (2026)
von: Yu, Chenyue, et al.
Veröffentlicht: (2026)
Introducing Visual Perception Token into Multimodal Large Language Model
von: Yu, Runpeng, et al.
Veröffentlicht: (2025)
von: Yu, Runpeng, et al.
Veröffentlicht: (2025)
Multisize Dataset Condensation
von: He, Yang, et al.
Veröffentlicht: (2024)
von: He, Yang, et al.
Veröffentlicht: (2024)
Pooling Image Datasets With Multiple Covariate Shift and Imbalance
von: Chytas, Sotirios Panagiotis, et al.
Veröffentlicht: (2024)
von: Chytas, Sotirios Panagiotis, et al.
Veröffentlicht: (2024)
Control and Realism: Best of Both Worlds in Layout-to-Image without Training
von: Li, Bonan, et al.
Veröffentlicht: (2025)
von: Li, Bonan, et al.
Veröffentlicht: (2025)
CLEAR: Conv-Like Linearization Revs Pre-Trained Diffusion Transformers Up
von: Liu, Songhua, et al.
Veröffentlicht: (2024)
von: Liu, Songhua, et al.
Veröffentlicht: (2024)
Image Editing As Programs with Diffusion Models
von: Hu, Yujia, et al.
Veröffentlicht: (2025)
von: Hu, Yujia, et al.
Veröffentlicht: (2025)
Top-Down Compression: Revisit Efficient Vision Token Projection for Visual Instruction Tuning
von: li, Bonan, et al.
Veröffentlicht: (2025)
von: li, Bonan, et al.
Veröffentlicht: (2025)
Ungeneralizable Examples
von: Ye, Jingwen, et al.
Veröffentlicht: (2024)
von: Ye, Jingwen, et al.
Veröffentlicht: (2024)
From Local Geometry to Global Pseudo Labeling for Robust Positive Unlabeled Learning under Covariate Shift
von: Gabetni, Firas, et al.
Veröffentlicht: (2026)
von: Gabetni, Firas, et al.
Veröffentlicht: (2026)
Dynamic Base model Shift for Delta Compression
von: Huang, Chenyu, et al.
Veröffentlicht: (2025)
von: Huang, Chenyu, et al.
Veröffentlicht: (2025)
A Label is Worth a Thousand Images in Dataset Distillation
von: Qin, Tian, et al.
Veröffentlicht: (2024)
von: Qin, Tian, et al.
Veröffentlicht: (2024)
Neural Metamorphosis
von: Yang, Xingyi, et al.
Veröffentlicht: (2024)
von: Yang, Xingyi, et al.
Veröffentlicht: (2024)
Flash Sculptor: Modular 3D Worlds from Objects
von: Hu, Yujia, et al.
Veröffentlicht: (2025)
von: Hu, Yujia, et al.
Veröffentlicht: (2025)
Dataset Distillers Are Good Label Denoisers In the Wild
von: Cheng, Lechao, et al.
Veröffentlicht: (2024)
von: Cheng, Lechao, et al.
Veröffentlicht: (2024)
Label Dropout: Improved Deep Learning Echocardiography Segmentation Using Multiple Datasets With Domain Shift and Partial Labelling
von: Islam, Iman, et al.
Veröffentlicht: (2024)
von: Islam, Iman, et al.
Veröffentlicht: (2024)
Learning without Exact Guidance: Updating Large-scale High-resolution Land Cover Maps from Low-resolution Historical Labels
von: Li, Zhuohong, et al.
Veröffentlicht: (2024)
von: Li, Zhuohong, et al.
Veröffentlicht: (2024)
Diversity-Guided MLP Reduction for Efficient Large Vision Transformers
von: Shen, Chengchao, et al.
Veröffentlicht: (2025)
von: Shen, Chengchao, et al.
Veröffentlicht: (2025)
From Pixels to Prose: A Large Dataset of Dense Image Captions
von: Singla, Vasu, et al.
Veröffentlicht: (2024)
von: Singla, Vasu, et al.
Veröffentlicht: (2024)
CRAM: Large-scale Video Continual Learning with Bootstrapped Compression
von: Mall, Shivani, et al.
Veröffentlicht: (2025)
von: Mall, Shivani, et al.
Veröffentlicht: (2025)
CLImage: Human-Annotated Datasets for Complementary-Label Learning
von: Wang, Hsiu-Hsuan, et al.
Veröffentlicht: (2023)
von: Wang, Hsiu-Hsuan, et al.
Veröffentlicht: (2023)
Producing Plankton Classifiers that are Robust to Dataset Shift
von: Chen, Cheng, et al.
Veröffentlicht: (2024)
von: Chen, Cheng, et al.
Veröffentlicht: (2024)
From Label Error Detection to Correction: A Modular Framework and Benchmark for Object Detection Datasets
von: Penquitt, Sarina, et al.
Veröffentlicht: (2025)
von: Penquitt, Sarina, et al.
Veröffentlicht: (2025)
Rethinking Decoders for Transformer-based Semantic Segmentation: A Compression Perspective
von: Wen, Qishuai, et al.
Veröffentlicht: (2024)
von: Wen, Qishuai, et al.
Veröffentlicht: (2024)
Ultra-Resolution Adaptation with Ease
von: Yu, Ruonan, et al.
Veröffentlicht: (2025)
von: Yu, Ruonan, et al.
Veröffentlicht: (2025)
Dataset Distillation as Data Compression: A Rate-Utility Perspective
von: Bao, Youneng, et al.
Veröffentlicht: (2025)
von: Bao, Youneng, et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
Soft Label Pruning and Quantization for Large-Scale Dataset Distillation
von: Lingao, Xiao, et al.
Veröffentlicht: (2026) -
Are Large-scale Soft Labels Necessary for Large-scale Dataset Distillation?
von: Xiao, Lingao, et al.
Veröffentlicht: (2024) -
Understanding Dataset Distillation via Spectral Filtering
von: Bo, Deyu, et al.
Veröffentlicht: (2025) -
LinFusion: 1 GPU, 1 Minute, 16K Image
von: Liu, Songhua, et al.
Veröffentlicht: (2024) -
One-shot Federated Learning via Synthetic Distiller-Distillate Communication
von: Zhang, Junyuan, et al.
Veröffentlicht: (2024)