Self-Refining Video Sampling
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Jang, Sangwon, Ki, Taekyung, Jo, Jaehyeong, Xie, Saining, Yoon, Jaehong, Hwang, Sung Ju |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2026
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Avatar Forcing: Real-Time Interactive Head Avatar Generation for Natural Conversation
von: Ki, Taekyung, et al.
Veröffentlicht: (2026)
von: Ki, Taekyung, et al.
Veröffentlicht: (2026)
Frame Guidance: Training-Free Guidance for Frame-Level Control in Video Diffusion Models
von: Jang, Sangwon, et al.
Veröffentlicht: (2025)
von: Jang, Sangwon, et al.
Veröffentlicht: (2025)
EVEREST: Efficient Masked Video Autoencoder by Removing Redundant Spatiotemporal Tokens
von: Hwang, Sunil, et al.
Veröffentlicht: (2022)
von: Hwang, Sunil, et al.
Veröffentlicht: (2022)
BECoTTA: Input-dependent Online Blending of Experts for Continual Test-time Adaptation
von: Lee, Daeun, et al.
Veröffentlicht: (2024)
von: Lee, Daeun, et al.
Veröffentlicht: (2024)
Identity Decoupling for Multi-Subject Personalization of Text-to-Image Models
von: Jang, Sangwon, et al.
Veröffentlicht: (2024)
von: Jang, Sangwon, et al.
Veröffentlicht: (2024)
STELLA: Continual Audio-Video Pre-training with Spatio-Temporal Localized Alignment
von: Lee, Jaewoo, et al.
Veröffentlicht: (2023)
von: Lee, Jaewoo, et al.
Veröffentlicht: (2023)
Continual Learning: Forget-free Winning Subnetworks for Video Representations
von: Kang, Haeyong, et al.
Veröffentlicht: (2023)
von: Kang, Haeyong, et al.
Veröffentlicht: (2023)
WorldMM: Dynamic Multimodal Memory Agent for Long Video Reasoning
von: Yeo, Woongyeong, et al.
Veröffentlicht: (2025)
von: Yeo, Woongyeong, et al.
Veröffentlicht: (2025)
Progressive Fourier Neural Representation for Sequential Video Compilation
von: Kang, Haeyong, et al.
Veröffentlicht: (2023)
von: Kang, Haeyong, et al.
Veröffentlicht: (2023)
StyleLipSync: Style-based Personalized Lip-sync Video Generation
von: Ki, Taekyung, et al.
Veröffentlicht: (2023)
von: Ki, Taekyung, et al.
Veröffentlicht: (2023)
Silent Branding Attack: Trigger-free Data Poisoning Attack on Text-to-Image Diffusion Models
von: Jang, Sangwon, et al.
Veröffentlicht: (2025)
von: Jang, Sangwon, et al.
Veröffentlicht: (2025)
Deconstructing Denoising Diffusion Models for Self-Supervised Learning
von: Chen, Xinlei, et al.
Veröffentlicht: (2024)
von: Chen, Xinlei, et al.
Veröffentlicht: (2024)
ECoFLaP: Efficient Coarse-to-Fine Layer-Wise Pruning for Vision-Language Models
von: Sung, Yi-Lin, et al.
Veröffentlicht: (2023)
von: Sung, Yi-Lin, et al.
Veröffentlicht: (2023)
Visualizing the loss landscape of Self-supervised Vision Transformer
von: Lee, Youngwan, et al.
Veröffentlicht: (2024)
von: Lee, Youngwan, et al.
Veröffentlicht: (2024)
Concept-skill Transferability-based Data Selection for Large Vision-Language Models
von: Lee, Jaewoo, et al.
Veröffentlicht: (2024)
von: Lee, Jaewoo, et al.
Veröffentlicht: (2024)
Multimodal Representation Learning by Alternating Unimodal Adaptation
von: Zhang, Xiaohui, et al.
Veröffentlicht: (2023)
von: Zhang, Xiaohui, et al.
Veröffentlicht: (2023)
Simple yet Effective Semi-supervised Knowledge Distillation from Vision-Language Models via Dual-Head Optimization
von: Kang, Seongjae, et al.
Veröffentlicht: (2025)
von: Kang, Seongjae, et al.
Veröffentlicht: (2025)
Inference-Time Scaling for Flow Models via Stochastic Generation and Rollover Budget Forcing
von: Kim, Jaihoon, et al.
Veröffentlicht: (2025)
von: Kim, Jaihoon, et al.
Veröffentlicht: (2025)
SAFREE: Training-Free and Adaptive Guard for Safe Text-to-Image And Video Generation
von: Yoon, Jaehong, et al.
Veröffentlicht: (2024)
von: Yoon, Jaehong, et al.
Veröffentlicht: (2024)
Representation Alignment for Generation: Training Diffusion Transformers Is Easier Than You Think
von: Yu, Sihyun, et al.
Veröffentlicht: (2024)
von: Yu, Sihyun, et al.
Veröffentlicht: (2024)
Flow Map Distillation Without Data
von: Tong, Shangyuan, et al.
Veröffentlicht: (2025)
von: Tong, Shangyuan, et al.
Veröffentlicht: (2025)
Diffusion Transformers with Representation Autoencoders
von: Zheng, Boyang, et al.
Veröffentlicht: (2025)
von: Zheng, Boyang, et al.
Veröffentlicht: (2025)
Self-Correcting Text-to-Video Generation with Misalignment Detection and Localized Refinement
von: Lee, Daeun, et al.
Veröffentlicht: (2024)
von: Lee, Daeun, et al.
Veröffentlicht: (2024)
Learning from Oblivion: Predicting Knowledge Overflowed Weights via Retrodiction of Forgetting
von: Jang, Jinhyeok, et al.
Veröffentlicht: (2025)
von: Jang, Jinhyeok, et al.
Veröffentlicht: (2025)
SELMA: Learning and Merging Skill-Specific Text-to-Image Experts with Auto-Generated Data
von: Li, Jialu, et al.
Veröffentlicht: (2024)
von: Li, Jialu, et al.
Veröffentlicht: (2024)
Transition Matching Distillation for Fast Video Generation
von: Nie, Weili, et al.
Veröffentlicht: (2026)
von: Nie, Weili, et al.
Veröffentlicht: (2026)
VideoRAG: Retrieval-Augmented Generation over Video Corpus
von: Jeong, Soyeong, et al.
Veröffentlicht: (2025)
von: Jeong, Soyeong, et al.
Veröffentlicht: (2025)
Delta Rectified Flow Sampling for Text-to-Image Editing
von: Beaudouin, Gaspard, et al.
Veröffentlicht: (2025)
von: Beaudouin, Gaspard, et al.
Veröffentlicht: (2025)
FLOAT: Generative Motion Latent Flow Matching for Audio-driven Talking Portrait
von: Ki, Taekyung, et al.
Veröffentlicht: (2024)
von: Ki, Taekyung, et al.
Veröffentlicht: (2024)
REPA-E: Unlocking VAE for End-to-End Tuning with Latent Diffusion Transformers
von: Leng, Xingjian, et al.
Veröffentlicht: (2025)
von: Leng, Xingjian, et al.
Veröffentlicht: (2025)
MCAQ-YOLO: Morphological Complexity-Aware Quantization for Efficient Object Detection with Curriculum Learning
von: Seo, Yoonjae, et al.
Veröffentlicht: (2025)
von: Seo, Yoonjae, et al.
Veröffentlicht: (2025)
Occupancy-Based Dual Contouring
von: Hwang, Jisung, et al.
Veröffentlicht: (2024)
von: Hwang, Jisung, et al.
Veröffentlicht: (2024)
RBF-Solver: A Multistep Sampler for Diffusion Probabilistic Models via Radial Basis Functions
von: Park, Soochul, et al.
Veröffentlicht: (2026)
von: Park, Soochul, et al.
Veröffentlicht: (2026)
Masking meets Supervision: A Strong Learning Alliance
von: Heo, Byeongho, et al.
Veröffentlicht: (2023)
von: Heo, Byeongho, et al.
Veröffentlicht: (2023)
Resolution Chromatography of Diffusion Models
von: Hwang, Juno, et al.
Veröffentlicht: (2023)
von: Hwang, Juno, et al.
Veröffentlicht: (2023)
oculomix: Hierarchical Sampling for Retinal-Based Systemic Disease Prediction
von: Kim, Hyunmin, et al.
Veröffentlicht: (2026)
von: Kim, Hyunmin, et al.
Veröffentlicht: (2026)
SiT: Exploring Flow and Diffusion-based Generative Models with Scalable Interpolant Transformers
von: Ma, Nanye, et al.
Veröffentlicht: (2024)
von: Ma, Nanye, et al.
Veröffentlicht: (2024)
LLM4SGG: Large Language Models for Weakly Supervised Scene Graph Generation
von: Kim, Kibum, et al.
Veröffentlicht: (2023)
von: Kim, Kibum, et al.
Veröffentlicht: (2023)
Psi-Sampler: Initial Particle Sampling for SMC-Based Inference-Time Reward Alignment in Score Models
von: Yoon, Taehoon, et al.
Veröffentlicht: (2025)
von: Yoon, Taehoon, et al.
Veröffentlicht: (2025)
Zigzag Diffusion Sampling: Diffusion Models Can Self-Improve via Self-Reflection
von: Bai, Lichen, et al.
Veröffentlicht: (2024)
von: Bai, Lichen, et al.
Veröffentlicht: (2024)
Ähnliche Einträge
-
Avatar Forcing: Real-Time Interactive Head Avatar Generation for Natural Conversation
von: Ki, Taekyung, et al.
Veröffentlicht: (2026) -
Frame Guidance: Training-Free Guidance for Frame-Level Control in Video Diffusion Models
von: Jang, Sangwon, et al.
Veröffentlicht: (2025) -
EVEREST: Efficient Masked Video Autoencoder by Removing Redundant Spatiotemporal Tokens
von: Hwang, Sunil, et al.
Veröffentlicht: (2022) -
BECoTTA: Input-dependent Online Blending of Experts for Continual Test-time Adaptation
von: Lee, Daeun, et al.
Veröffentlicht: (2024) -
Identity Decoupling for Multi-Subject Personalization of Text-to-Image Models
von: Jang, Sangwon, et al.
Veröffentlicht: (2024)