Are Large-scale Soft Labels Necessary for Large-scale Dataset Distillation?
Fuente:
arXiv
Guardado en:
| Autores principales: | Xiao, Lingao, He, Yang |
|---|---|
| Formato: | Preprint |
| Publicado: |
2024
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
Soft Label Pruning and Quantization for Large-Scale Dataset Distillation
por: Lingao, Xiao, et al.
Publicado: (2026)
por: Lingao, Xiao, et al.
Publicado: (2026)
Rethinking Large-scale Dataset Compression: Shifting Focus From Labels to Images
por: Xiao, Lingao, et al.
Publicado: (2025)
por: Xiao, Lingao, et al.
Publicado: (2025)
Multisize Dataset Condensation
por: He, Yang, et al.
Publicado: (2024)
por: He, Yang, et al.
Publicado: (2024)
Stabilizing, Scaling & Enhancing MeanFlow for Large-scale Diffusion Distillation
por: He, Xiao, et al.
Publicado: (2026)
por: He, Xiao, et al.
Publicado: (2026)
Training-Free Dataset Pruning for Instance Segmentation
por: Dai, Yalun, et al.
Publicado: (2025)
por: Dai, Yalun, et al.
Publicado: (2025)
Large-scale Dataset Pruning with Dynamic Uncertainty
por: He, Muyang, et al.
Publicado: (2023)
por: He, Muyang, et al.
Publicado: (2023)
Vector-Quantized Soft Label Compression for Dataset Distillation
por: Abbasi, Ali, et al.
Publicado: (2026)
por: Abbasi, Ali, et al.
Publicado: (2026)
Dataset Color Quantization: A Training-Oriented Framework for Dataset-Level Compression
por: Yu, Chenyue, et al.
Publicado: (2026)
por: Yu, Chenyue, et al.
Publicado: (2026)
AMD: Automatic Multi-step Distillation of Large-scale Vision Models
por: Han, Cheng, et al.
Publicado: (2024)
por: Han, Cheng, et al.
Publicado: (2024)
Rethinking Dataset Distillation: Hard Truths about Soft Labels
por: Dey, Priyam, et al.
Publicado: (2026)
por: Dey, Priyam, et al.
Publicado: (2026)
InternVid: A Large-scale Video-Text Dataset for Multimodal Understanding and Generation
por: Wang, Yi, et al.
Publicado: (2023)
por: Wang, Yi, et al.
Publicado: (2023)
Long-range Turbulence Mitigation: A Large-scale Dataset and A Coarse-to-fine Framework
por: Xu, Shengqi, et al.
Publicado: (2024)
por: Xu, Shengqi, et al.
Publicado: (2024)
Large-scale Codec Avatars: The Unreasonable Effectiveness of Large-scale Avatar Pretraining
por: Li, Junxuan, et al.
Publicado: (2026)
por: Li, Junxuan, et al.
Publicado: (2026)
Action100M: A Large-scale Video Action Dataset
por: Chen, Delong, et al.
Publicado: (2026)
por: Chen, Delong, et al.
Publicado: (2026)
OLATverse: A Large-scale Real-world Object Dataset with Precise Lighting Control
por: Zhou, Xilong, et al.
Publicado: (2025)
por: Zhou, Xilong, et al.
Publicado: (2025)
SingingHead: A Large-scale 4D Dataset for Singing Head Animation
por: Wu, Sijing, et al.
Publicado: (2023)
por: Wu, Sijing, et al.
Publicado: (2023)
HOIGen-1M: A Large-scale Dataset for Human-Object Interaction Video Generation
por: Liu, Kun, et al.
Publicado: (2025)
por: Liu, Kun, et al.
Publicado: (2025)
OmniHuman: A Large-scale Dataset and Benchmark for Human-Centric Video Generation
por: Zhu, Lei, et al.
Publicado: (2026)
por: Zhu, Lei, et al.
Publicado: (2026)
Video Repurposing from User Generated Content: A Large-scale Dataset and Benchmark
por: Wu, Yongliang, et al.
Publicado: (2024)
por: Wu, Yongliang, et al.
Publicado: (2024)
LMHaze: Intensity-aware Image Dehazing with a Large-scale Multi-intensity Real Haze Dataset
por: Zhang, Ruikun, et al.
Publicado: (2024)
por: Zhang, Ruikun, et al.
Publicado: (2024)
Pool-Select-Refine: Allocation-Aware Generative Dataset Distillation with Soft-Label-Guided Latent Refinement
por: Li, Wenmin, et al.
Publicado: (2026)
por: Li, Wenmin, et al.
Publicado: (2026)
UAV-VisLoc: A Large-scale Dataset for UAV Visual Localization
por: Xu, Wenjia, et al.
Publicado: (2024)
por: Xu, Wenjia, et al.
Publicado: (2024)
FLAIR-HUB: Large-scale Multimodal Dataset for Land Cover and Crop Mapping
por: Garioud, Anatol, et al.
Publicado: (2025)
por: Garioud, Anatol, et al.
Publicado: (2025)
Embody 3D: A Large-scale Multimodal Motion and Behavior Dataset
por: McLean, Claire, et al.
Publicado: (2025)
por: McLean, Claire, et al.
Publicado: (2025)
Parameter-efficient Tuning of Large-scale Multimodal Foundation Model
por: Wang, Haixin, et al.
Publicado: (2023)
por: Wang, Haixin, et al.
Publicado: (2023)
RoScenes: A Large-scale Multi-view 3D Dataset for Roadside Perception
por: Zhu, Xiaosu, et al.
Publicado: (2024)
por: Zhu, Xiaosu, et al.
Publicado: (2024)
TextAtlas5M: A Large-scale Dataset for Dense Text Image Generation
por: Wang, Alex Jinpeng, et al.
Publicado: (2025)
por: Wang, Alex Jinpeng, et al.
Publicado: (2025)
Heavy Labels Out! Dataset Distillation with Label Space Lightening
por: Yu, Ruonan, et al.
Publicado: (2024)
por: Yu, Ruonan, et al.
Publicado: (2024)
ActionHub: A Large-scale Action Video Description Dataset for Zero-shot Action Recognition
por: Zhou, Jiaming, et al.
Publicado: (2024)
por: Zhou, Jiaming, et al.
Publicado: (2024)
WildFake: A Large-scale Challenging Dataset for AI-Generated Images Detection
por: Hong, Yan, et al.
Publicado: (2024)
por: Hong, Yan, et al.
Publicado: (2024)
LED: A Large-scale Real-world Paired Dataset for Event Camera Denoising
por: Duan, Yuxing, et al.
Publicado: (2024)
por: Duan, Yuxing, et al.
Publicado: (2024)
ShapeSplat: A Large-scale Dataset of Gaussian Splats and Their Self-Supervised Pretraining
por: Ma, Qi, et al.
Publicado: (2024)
por: Ma, Qi, et al.
Publicado: (2024)
OpenMaterial: A Large-scale Dataset of Complex Materials for 3D Reconstruction
por: Dang, Zheng, et al.
Publicado: (2024)
por: Dang, Zheng, et al.
Publicado: (2024)
Enhancing Descriptive Image Quality Assessment with A Large-scale Multi-modal Dataset
por: You, Zhiyuan, et al.
Publicado: (2024)
por: You, Zhiyuan, et al.
Publicado: (2024)
Manga109Dialog: A Large-scale Dialogue Dataset for Comics Speaker Detection
por: Li, Yingxuan, et al.
Publicado: (2023)
por: Li, Yingxuan, et al.
Publicado: (2023)
Leader360V: The Large-scale, Real-world 360 Video Dataset for Multi-task Learning in Diverse Environment
por: Zhang, Weiming, et al.
Publicado: (2025)
por: Zhang, Weiming, et al.
Publicado: (2025)
Ev-Layout: A Large-scale Event-based Multi-modal Dataset for Indoor Layout Estimation and Tracking
por: Guo, Xucheng, et al.
Publicado: (2025)
por: Guo, Xucheng, et al.
Publicado: (2025)
Diving into Underwater: Segment Anything Model Guided Underwater Salient Instance Segmentation and A Large-scale Dataset
por: Lian, Shijie, et al.
Publicado: (2024)
por: Lian, Shijie, et al.
Publicado: (2024)
Openstory++: A Large-scale Dataset and Benchmark for Instance-aware Open-domain Visual Storytelling
por: Ye, Zilyu, et al.
Publicado: (2024)
por: Ye, Zilyu, et al.
Publicado: (2024)
MedTrinity-25M: A Large-scale Multimodal Dataset with Multigranular Annotations for Medicine
por: Xie, Yunfei, et al.
Publicado: (2024)
por: Xie, Yunfei, et al.
Publicado: (2024)
Ejemplares similares
-
Soft Label Pruning and Quantization for Large-Scale Dataset Distillation
por: Lingao, Xiao, et al.
Publicado: (2026) -
Rethinking Large-scale Dataset Compression: Shifting Focus From Labels to Images
por: Xiao, Lingao, et al.
Publicado: (2025) -
Multisize Dataset Condensation
por: He, Yang, et al.
Publicado: (2024) -
Stabilizing, Scaling & Enhancing MeanFlow for Large-scale Diffusion Distillation
por: He, Xiao, et al.
Publicado: (2026) -
Training-Free Dataset Pruning for Instance Segmentation
por: Dai, Yalun, et al.
Publicado: (2025)