Reward Guided Latent Consistency Distillation
Fuente:
arXiv
Salvato in:
| Autori principali: | Li, Jiachen, Feng, Weixi, Chen, Wenhu, Wang, William Yang |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2024
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
TC-Bench: Benchmarking Temporal Compositionality in Text-to-Video and Image-to-Video Generation
di: Feng, Weixi, et al.
Pubblicazione: (2024)
di: Feng, Weixi, et al.
Pubblicazione: (2024)
T2V-Turbo: Breaking the Quality Bottleneck of Video Consistency Model with Mixed Reward Feedback
di: Li, Jiachen, et al.
Pubblicazione: (2024)
di: Li, Jiachen, et al.
Pubblicazione: (2024)
T2V-Turbo-v2: Enhancing Video Generation Model Post-Training through Data, Reward, and Conditional Guidance Design
di: Li, Jiachen, et al.
Pubblicazione: (2024)
di: Li, Jiachen, et al.
Pubblicazione: (2024)
EditReward: A Human-Aligned Reward Model for Instruction-Guided Image Editing
di: Wu, Keming, et al.
Pubblicazione: (2025)
di: Wu, Keming, et al.
Pubblicazione: (2025)
Bad Seeing or Bad Thinking? Rewarding Perception for Vision-Language Reasoning
di: Wang, Haozhe, et al.
Pubblicazione: (2026)
di: Wang, Haozhe, et al.
Pubblicazione: (2026)
BlobGEN-Vid: Compositional Text-to-Video Generation with Blob Video Representations
di: Feng, Weixi, et al.
Pubblicazione: (2025)
di: Feng, Weixi, et al.
Pubblicazione: (2025)
Beyond VLM-Based Rewards: Diffusion-Native Latent Reward Modeling
di: Liu, Gongye, et al.
Pubblicazione: (2026)
di: Liu, Gongye, et al.
Pubblicazione: (2026)
SpatialReward: Verifiable Spatial Reward Modeling for Fine-Grained Spatial Consistency in Text-to-Image Generation
di: Zhou, Sashuai, et al.
Pubblicazione: (2026)
di: Zhou, Sashuai, et al.
Pubblicazione: (2026)
Self-Corrected Image Generation with Explainable Latent Rewards
di: Luo, Yinyi, et al.
Pubblicazione: (2026)
di: Luo, Yinyi, et al.
Pubblicazione: (2026)
VisualWebInstruct: Scaling up Multimodal Instruction Data through Web Search
di: Jia, Yiming, et al.
Pubblicazione: (2025)
di: Jia, Yiming, et al.
Pubblicazione: (2025)
LatentEdit: Adaptive Latent Control for Consistent Semantic Editing
di: Liu, Siyi, et al.
Pubblicazione: (2025)
di: Liu, Siyi, et al.
Pubblicazione: (2025)
Efficient Text-driven Motion Generation via Latent Consistency Training
di: Hu, Mengxian, et al.
Pubblicazione: (2024)
di: Hu, Mengxian, et al.
Pubblicazione: (2024)
RL for Consistency Models: Faster Reward Guided Text-to-Image Generation
di: Oertell, Owen, et al.
Pubblicazione: (2024)
di: Oertell, Owen, et al.
Pubblicazione: (2024)
VELMA: Verbalization Embodiment of LLM Agents for Vision and Language Navigation in Street View
di: Schumann, Raphael, et al.
Pubblicazione: (2023)
di: Schumann, Raphael, et al.
Pubblicazione: (2023)
MagicBrush: A Manually Annotated Dataset for Instruction-Guided Image Editing
di: Zhang, Kai, et al.
Pubblicazione: (2023)
di: Zhang, Kai, et al.
Pubblicazione: (2023)
Latent Distillation for Continual Object Detection at the Edge
di: Pasti, Francesco, et al.
Pubblicazione: (2024)
di: Pasti, Francesco, et al.
Pubblicazione: (2024)
MMWorld: Towards Multi-discipline Multi-faceted World Model Evaluation in Videos
di: He, Xuehai, et al.
Pubblicazione: (2024)
di: He, Xuehai, et al.
Pubblicazione: (2024)
Latent Action Control for Reasoning-Guided Unified Image Generation
di: Zhai, Fuxiang, et al.
Pubblicazione: (2026)
di: Zhai, Fuxiang, et al.
Pubblicazione: (2026)
TG-LLaVA: Text Guided LLaVA via Learnable Latent Embeddings
di: Yan, Dawei, et al.
Pubblicazione: (2024)
di: Yan, Dawei, et al.
Pubblicazione: (2024)
Diffusion-Guided Semantic Consistency for Multimodal Heterogeneity
di: Liu, Jing, et al.
Pubblicazione: (2026)
di: Liu, Jing, et al.
Pubblicazione: (2026)
MAGIC: Meta-Ability Guided Interactive Chain-of-Distillation for Effective-and-Efficient Vision-and-Language Navigation
di: Wang, Liuyi, et al.
Pubblicazione: (2024)
di: Wang, Liuyi, et al.
Pubblicazione: (2024)
Consistent Flow Distillation for Text-to-3D Generation
di: Yan, Runjie, et al.
Pubblicazione: (2025)
di: Yan, Runjie, et al.
Pubblicazione: (2025)
CleverDistiller: Simple and Spatially Consistent Cross-modal Distillation
di: Govindarajan, Hariprasath, et al.
Pubblicazione: (2025)
di: Govindarajan, Hariprasath, et al.
Pubblicazione: (2025)
TLCM: Training-efficient Latent Consistency Model for Image Generation with 2-8 Steps
di: Xie, Qingsong, et al.
Pubblicazione: (2024)
di: Xie, Qingsong, et al.
Pubblicazione: (2024)
Visual-Noise Guided In-Context Distillation for Multimodal Large Language Model Unlearning
di: Chen, Junkai, et al.
Pubblicazione: (2026)
di: Chen, Junkai, et al.
Pubblicazione: (2026)
Collaborative Attention and Consistent-Guided Fusion of MRI and PET for Alzheimer's Disease Diagnosis
di: Ma, Delin, et al.
Pubblicazione: (2025)
di: Ma, Delin, et al.
Pubblicazione: (2025)
WildVision: Evaluating Vision-Language Models in the Wild with Human Preferences
di: Lu, Yujie, et al.
Pubblicazione: (2024)
di: Lu, Yujie, et al.
Pubblicazione: (2024)
Saliency-Guided Representation with Consistency Policy Learning for Visual Unsupervised Reinforcement Learning
di: Sun, Jingbo, et al.
Pubblicazione: (2026)
di: Sun, Jingbo, et al.
Pubblicazione: (2026)
Lightweight Remote Sensing Scene Classification on Edge Devices via Knowledge Distillation and Early-exit
di: Zhao, Yang, et al.
Pubblicazione: (2025)
di: Zhao, Yang, et al.
Pubblicazione: (2025)
ALAM: Algebraically Consistent Latent Action Model for Vision-Language-Action Models
di: Tang, Zuojin, et al.
Pubblicazione: (2026)
di: Tang, Zuojin, et al.
Pubblicazione: (2026)
Hyperbolic Distillation: Geometry-Guided Cross-Modal Transfer for Robust 3D Object Detection
di: Ning, Kanglin, et al.
Pubblicazione: (2026)
di: Ning, Kanglin, et al.
Pubblicazione: (2026)
GPI-Net: Gestalt-Guided Parallel Interaction Network via Orthogonal Geometric Consistency for Robust Point Cloud Registration
di: Gu, Weikang, et al.
Pubblicazione: (2025)
di: Gu, Weikang, et al.
Pubblicazione: (2025)
GVD: Guiding Video Diffusion Model for Scalable Video Distillation
di: Li, Kunyang, et al.
Pubblicazione: (2025)
di: Li, Kunyang, et al.
Pubblicazione: (2025)
APLA: Additional Perturbation for Latent Noise with Adversarial Training Enables Consistency
di: Yao, Yupu, et al.
Pubblicazione: (2023)
di: Yao, Yupu, et al.
Pubblicazione: (2023)
Diffusion Reinforcement Learning via Centered Reward Distillation
di: Zhu, Yuanzhi, et al.
Pubblicazione: (2026)
di: Zhu, Yuanzhi, et al.
Pubblicazione: (2026)
RewardMap: Tackling Sparse Rewards in Fine-grained Visual Reasoning via Multi-Stage Reinforcement Learning
di: Feng, Sicheng, et al.
Pubblicazione: (2025)
di: Feng, Sicheng, et al.
Pubblicazione: (2025)
Contrastive Learning Guided Latent Diffusion Model for Image-to-Image Translation
di: Si, Qi, et al.
Pubblicazione: (2025)
di: Si, Qi, et al.
Pubblicazione: (2025)
DGAE: Diffusion-Guided Autoencoder for Efficient Latent Representation Learning
di: Liu, Dongxu, et al.
Pubblicazione: (2025)
di: Liu, Dongxu, et al.
Pubblicazione: (2025)
Calibrating Biased Distribution in VFM-derived Latent Space via Cross-Domain Geometric Consistency
di: Ma, Yanbiao, et al.
Pubblicazione: (2025)
di: Ma, Yanbiao, et al.
Pubblicazione: (2025)
DiffusionTalker: Efficient and Compact Speech-Driven 3D Talking Head via Personalizer-Guided Distillation
di: Chen, Peng, et al.
Pubblicazione: (2025)
di: Chen, Peng, et al.
Pubblicazione: (2025)
Documenti analoghi
-
TC-Bench: Benchmarking Temporal Compositionality in Text-to-Video and Image-to-Video Generation
di: Feng, Weixi, et al.
Pubblicazione: (2024) -
T2V-Turbo: Breaking the Quality Bottleneck of Video Consistency Model with Mixed Reward Feedback
di: Li, Jiachen, et al.
Pubblicazione: (2024) -
T2V-Turbo-v2: Enhancing Video Generation Model Post-Training through Data, Reward, and Conditional Guidance Design
di: Li, Jiachen, et al.
Pubblicazione: (2024) -
EditReward: A Human-Aligned Reward Model for Instruction-Guided Image Editing
di: Wu, Keming, et al.
Pubblicazione: (2025) -
Bad Seeing or Bad Thinking? Rewarding Perception for Vision-Language Reasoning
di: Wang, Haozhe, et al.
Pubblicazione: (2026)