Distill Video Datasets into Images
Fuente:
arXiv
Saved in:
| Main Authors: | Zhao, Zhenghao, Wang, Haoxuan, Wang, Kai, Shang, Yuzhang, Hong, Yuan, Yan, Yan |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Efficient Multimodal Dataset Distillation via Generative Models
by: Zhao, Zhenghao, et al.
Published: (2025)
by: Zhao, Zhenghao, et al.
Published: (2025)
CaO$_2$: Rectifying Inconsistencies in Diffusion-Based Dataset Distillation
by: Wang, Haoxuan, et al.
Published: (2025)
by: Wang, Haoxuan, et al.
Published: (2025)
Dataset Quantization with Active Learning based Adaptive Sampling
by: Zhao, Zhenghao, et al.
Published: (2024)
by: Zhao, Zhenghao, et al.
Published: (2024)
QuEST: Low-bit Diffusion Model Quantization via Efficient Selective Finetuning
by: Wang, Haoxuan, et al.
Published: (2024)
by: Wang, Haoxuan, et al.
Published: (2024)
PTQ4DiT: Post-training Quantization for Diffusion Transformers
by: Wu, Junyi, et al.
Published: (2024)
by: Wu, Junyi, et al.
Published: (2024)
Supplementing Missing Visions via Dialog for Scene Graph Generations
by: Zhao, Zhenghao, et al.
Published: (2022)
by: Zhao, Zhenghao, et al.
Published: (2022)
VGDFR: Diffusion-based Video Generation with Dynamic Latent Frame Rate
by: Yuan, Zhihang, et al.
Published: (2025)
by: Yuan, Zhihang, et al.
Published: (2025)
FBPT: A Fully Binary Point Transformer
by: Hou, Zhixing, et al.
Published: (2024)
by: Hou, Zhixing, et al.
Published: (2024)
Distilling Long-tailed Datasets
by: Zhao, Zhenghao, et al.
Published: (2024)
by: Zhao, Zhenghao, et al.
Published: (2024)
PackCache: A Training-Free Acceleration Method for Unified Autoregressive Video Generation via Compact KV-Cache
by: Li, Kunyang, et al.
Published: (2026)
by: Li, Kunyang, et al.
Published: (2026)
DKDM: Data-Free Knowledge Distillation for Diffusion Models with Any Architecture
by: Xiang, Qianlong, et al.
Published: (2024)
by: Xiang, Qianlong, et al.
Published: (2024)
Efficient Multitask Dense Predictor via Binarization
by: Shang, Yuzhang, et al.
Published: (2024)
by: Shang, Yuzhang, et al.
Published: (2024)
DD-Ranking: Rethinking the Evaluation of Dataset Distillation
by: Li, Zekai, et al.
Published: (2025)
by: Li, Zekai, et al.
Published: (2025)
DLFR-VAE: Dynamic Latent Frame Rate VAE for Video Generation
by: Yuan, Zhihang, et al.
Published: (2025)
by: Yuan, Zhihang, et al.
Published: (2025)
Attend Locally, Remember Linearly: Linear Attention as Cross-Frame Memory for Autoregressive Video Diffusion
by: Li, Kunyang, et al.
Published: (2026)
by: Li, Kunyang, et al.
Published: (2026)
AdaTooler-V: Adaptive Tool-Use for Images and Videos
by: Wang, Chaoyang, et al.
Published: (2025)
by: Wang, Chaoyang, et al.
Published: (2025)
freePruner: A Training-free Approach for Large Multimodal Model Acceleration
by: Xu, Bingxin, et al.
Published: (2024)
by: Xu, Bingxin, et al.
Published: (2024)
Online Multi-spectral Neuron Tracing
by: Duan, Bin, et al.
Published: (2024)
by: Duan, Bin, et al.
Published: (2024)
Orientation-anchored Hyper-Gaussian for 4D Reconstruction from Casual Videos
by: Wu, Junyi, et al.
Published: (2025)
by: Wu, Junyi, et al.
Published: (2025)
Latent Video Dataset Distillation
by: Li, Ning, et al.
Published: (2025)
by: Li, Ning, et al.
Published: (2025)
Robin3D: Improving 3D Large Language Model via Robust Instruction Tuning
by: Kang, Weitai, et al.
Published: (2024)
by: Kang, Weitai, et al.
Published: (2024)
WildFake: A Large-scale Challenging Dataset for AI-Generated Images Detection
by: Hong, Yan, et al.
Published: (2024)
by: Hong, Yan, et al.
Published: (2024)
E-CAR: Efficient Continuous Autoregressive Image Generation via Multistage Modeling
by: Yuan, Zhihang, et al.
Published: (2024)
by: Yuan, Zhihang, et al.
Published: (2024)
MFSR: MeanFlow Distillation for One Step Real-World Image Super Resolution
by: Wang, Ruiqing, et al.
Published: (2026)
by: Wang, Ruiqing, et al.
Published: (2026)
Temporal Saliency-Guided Distillation: A Scalable Framework for Distilling Video Datasets
by: Gu, Xulin, et al.
Published: (2025)
by: Gu, Xulin, et al.
Published: (2025)
Can World Simulators Reason? Gen-ViRe: A Generative Visual Reasoning Benchmark
by: Liu, Xinxin, et al.
Published: (2025)
by: Liu, Xinxin, et al.
Published: (2025)
LLaVA-PruMerge: Adaptive Token Reduction for Efficient Large Multimodal Models
by: Shang, Yuzhang, et al.
Published: (2024)
by: Shang, Yuzhang, et al.
Published: (2024)
X-Field: A Physically Grounded Representation for 3D X-ray Reconstruction
by: Wang, Feiran, et al.
Published: (2025)
by: Wang, Feiran, et al.
Published: (2025)
Fine-Grained Prototypes Distillation for Few-Shot Object Detection
by: Wang, Zichen, et al.
Published: (2024)
by: Wang, Zichen, et al.
Published: (2024)
Color-Oriented Redundancy Reduction in Dataset Distillation
by: Yuan, Bowen, et al.
Published: (2024)
by: Yuan, Bowen, et al.
Published: (2024)
Alleviating Class Imbalance in Semi-supervised Multi-organ Segmentation via Balanced Subclass Regularization
by: Feng, Zhenghao, et al.
Published: (2024)
by: Feng, Zhenghao, et al.
Published: (2024)
CAD: A General Multimodal Framework for Video Deepfake Detection via Cross-Modal Alignment and Distillation
by: Du, Yuxuan, et al.
Published: (2025)
by: Du, Yuxuan, et al.
Published: (2025)
Towards Patronizing and Condescending Language in Chinese Videos: A Multimodal Dataset and Detector
by: Wang, Hongbo, et al.
Published: (2024)
by: Wang, Hongbo, et al.
Published: (2024)
A Survey of Token Compression for Efficient Multimodal Large Language Models
by: Shao, Kele, et al.
Published: (2025)
by: Shao, Kele, et al.
Published: (2025)
Towards Robust Text-to-Image Person Retrieval: Multi-View Reformulation for Semantic Compensation
by: Yuan, Chao, et al.
Published: (2026)
by: Yuan, Chao, et al.
Published: (2026)
MSVCOD:A Large-Scale Multi-Scene Dataset for Video Camouflage Object Detection
by: Gao, Shuyong, et al.
Published: (2025)
by: Gao, Shuyong, et al.
Published: (2025)
ATOM: Attention Mixer for Efficient Dataset Distillation
by: Khaki, Samir, et al.
Published: (2024)
by: Khaki, Samir, et al.
Published: (2024)
Towards Fast, Memory-based and Data-Efficient Vision-Language Policy
by: Li, Haoxuan, et al.
Published: (2025)
by: Li, Haoxuan, et al.
Published: (2025)
Dataset Distillation for Histopathology Image Classification
by: Cong, Cong, et al.
Published: (2024)
by: Cong, Cong, et al.
Published: (2024)
UniLumos: Fast and Unified Image and Video Relighting with Physics-Plausible Feedback
by: Liu, Ropeway, et al.
Published: (2025)
by: Liu, Ropeway, et al.
Published: (2025)
Similar Items
-
Efficient Multimodal Dataset Distillation via Generative Models
by: Zhao, Zhenghao, et al.
Published: (2025) -
CaO$_2$: Rectifying Inconsistencies in Diffusion-Based Dataset Distillation
by: Wang, Haoxuan, et al.
Published: (2025) -
Dataset Quantization with Active Learning based Adaptive Sampling
by: Zhao, Zhenghao, et al.
Published: (2024) -
QuEST: Low-bit Diffusion Model Quantization via Efficient Selective Finetuning
by: Wang, Haoxuan, et al.
Published: (2024) -
PTQ4DiT: Post-training Quantization for Diffusion Transformers
by: Wu, Junyi, et al.
Published: (2024)