Improving Controllable Generation: Faster Training and Better Performance via $x_0$-Supervision
Fuente:
arXiv
Saved in:
| Main Authors: | Sangare, Amadou S., Maglo, Adrien, Chaouch, Mohamed, Luvison, Bertrand |
|---|---|
| Format: | Preprint |
| Published: |
2026
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
3D-COCO: extension of MS-COCO dataset for image detection and 3D reconstruction modules
by: Bideaux, Maxence, et al.
Published: (2024)
by: Bideaux, Maxence, et al.
Published: (2024)
Benchmarking Adversarial Robustness and Adversarial Training Strategies for Object Detection
by: Winter, Alexis, et al.
Published: (2026)
by: Winter, Alexis, et al.
Published: (2026)
Faster and Better 3D Splatting via Group Training
by: Wang, Chengbo, et al.
Published: (2024)
by: Wang, Chengbo, et al.
Published: (2024)
Fairer Analysis and Demographically Balanced Face Generation for Fairer Face Verification
by: Fournier-Montgieux, Alexandre, et al.
Published: (2024)
by: Fournier-Montgieux, Alexandre, et al.
Published: (2024)
Reliable and Reproducible Demographic Inference for Fairness in Face Analysis
by: Fournier-Montgieux, Alexandre, et al.
Published: (2025)
by: Fournier-Montgieux, Alexandre, et al.
Published: (2025)
Toward Fairer Face Recognition Datasets
by: Fournier-Montgieux, Alexandre, et al.
Published: (2024)
by: Fournier-Montgieux, Alexandre, et al.
Published: (2024)
Improving Supervised Machine Learning Performance in Optical Quality Control via Generative AI for Dataset Expansion
by: Sprute, Dennis, et al.
Published: (2026)
by: Sprute, Dennis, et al.
Published: (2026)
Towards Better & Faster Autoregressive Image Generation: From the Perspective of Entropy
by: Ma, Xiaoxiao, et al.
Published: (2025)
by: Ma, Xiaoxiao, et al.
Published: (2025)
CoRe^2: Collect, Reflect and Refine to Generate Better and Faster
by: Shao, Shitong, et al.
Published: (2025)
by: Shao, Shitong, et al.
Published: (2025)
FlashMesh: Faster and Better Autoregressive Mesh Synthesis via Structured Speculation
by: Shen, Tingrui, et al.
Published: (2025)
by: Shen, Tingrui, et al.
Published: (2025)
Faster Training, Fewer Labels: Self-Supervised Pretraining for Fine-Grained BEV Segmentation
by: Busch, Daniel, et al.
Published: (2026)
by: Busch, Daniel, et al.
Published: (2026)
Valeo Near-Field: a novel dataset for pedestrian intent detection
by: Musabini, Antonyo, et al.
Published: (2025)
by: Musabini, Antonyo, et al.
Published: (2025)
Tracking Meets LoRA: Faster Training, Larger Model, Stronger Performance
by: Lin, Liting, et al.
Published: (2024)
by: Lin, Liting, et al.
Published: (2024)
Schedule On the Fly: Diffusion Time Prediction for Faster and Better Image Generation
by: Ye, Zilyu, et al.
Published: (2024)
by: Ye, Zilyu, et al.
Published: (2024)
FBRT-YOLO: Faster and Better for Real-Time Aerial Image Detection
by: Xiao, Yao, et al.
Published: (2025)
by: Xiao, Yao, et al.
Published: (2025)
Know Your Step: Faster and Better Alignment for Flow Matching Models via Step-aware Advantages
by: Yue, Zhixiong, et al.
Published: (2026)
by: Yue, Zhixiong, et al.
Published: (2026)
FasterDiT: Towards Faster Diffusion Transformers Training without Architecture Modification
by: Yao, Jingfeng, et al.
Published: (2024)
by: Yao, Jingfeng, et al.
Published: (2024)
ReGATE: Learning Faster and Better with Fewer Tokens in MLLMs
by: Li, Chaoyu, et al.
Published: (2025)
by: Li, Chaoyu, et al.
Published: (2025)
RoMa v2: Harder Better Faster Denser Feature Matching
by: Edstedt, Johan, et al.
Published: (2025)
by: Edstedt, Johan, et al.
Published: (2025)
FREE: Faster and Better Data-Free Meta-Learning
by: Wei, Yongxian, et al.
Published: (2024)
by: Wei, Yongxian, et al.
Published: (2024)
Faster and Better: Reinforced Collaborative Distillation and Self-Learning for Infrared-Visible Image Fusion
by: Wang, Yuhao, et al.
Published: (2025)
by: Wang, Yuhao, et al.
Published: (2025)
Semi-Supervised Training to Improve Player and Ball Detection in Soccer
by: Vandeghen, Renaud, et al.
Published: (2022)
by: Vandeghen, Renaud, et al.
Published: (2022)
Better, Stronger, Faster: Tackling the Trilemma in MLLM-based Segmentation with Simultaneous Textual Mask Prediction
by: Liu, Jiazhen, et al.
Published: (2025)
by: Liu, Jiazhen, et al.
Published: (2025)
Bidirectional Sparse Attention for Faster Video Diffusion Training
by: Zhan, Chenlu, et al.
Published: (2025)
by: Zhan, Chenlu, et al.
Published: (2025)
Faster Inference of Flow-Based Generative Models via Improved Data-Noise Coupling
by: Davtyan, Aram, et al.
Published: (2026)
by: Davtyan, Aram, et al.
Published: (2026)
Tracktention: Leveraging Point Tracking to Attend Videos Faster and Better
by: Lai, Zihang, et al.
Published: (2025)
by: Lai, Zihang, et al.
Published: (2025)
The Unmet Promise of Synthetic Training Images: Using Retrieved Real Images Performs Better
by: Geng, Scott, et al.
Published: (2024)
by: Geng, Scott, et al.
Published: (2024)
Can Better Text Semantics in Prompt Tuning Improve VLM Generalization?
by: Kuchibhotla, Hari Chandana, et al.
Published: (2024)
by: Kuchibhotla, Hari Chandana, et al.
Published: (2024)
One View Is Enough! Monocular Training for In-the-Wild Novel View Generation
by: Rahary, Adrien Ramanana, et al.
Published: (2026)
by: Rahary, Adrien Ramanana, et al.
Published: (2026)
CycleCap: Improving VLMs Captioning Performance via Self-Supervised Cycle Consistency Fine-Tuning
by: Krestenitis, Marios, et al.
Published: (2026)
by: Krestenitis, Marios, et al.
Published: (2026)
FlashPortrait: 6x Faster Infinite Portrait Animation with Adaptive Latent Prediction
by: Tu, Shuyuan, et al.
Published: (2025)
by: Tu, Shuyuan, et al.
Published: (2025)
Training-Free Semantic Segmentation via LLM-Supervision
by: Sun, Wenfang, et al.
Published: (2024)
by: Sun, Wenfang, et al.
Published: (2024)
ShareGPT4Video: Improving Video Understanding and Generation with Better Captions
by: Chen, Lin, et al.
Published: (2024)
by: Chen, Lin, et al.
Published: (2024)
Striving for Faster and Better: A One-Layer Architecture with Auto Re-parameterization for Low-Light Image Enhancement
by: An, Nan, et al.
Published: (2025)
by: An, Nan, et al.
Published: (2025)
ISAC: Training-Free Instance-to-Semantic Attention Control for Improving Multi-Instance Generation
by: Jo, Sanghyun, et al.
Published: (2025)
by: Jo, Sanghyun, et al.
Published: (2025)
Learn "No" to Say "Yes" Better: Improving Vision-Language Models via Negations
by: Singh, Jaisidh, et al.
Published: (2024)
by: Singh, Jaisidh, et al.
Published: (2024)
Improving the Generalization of Segmentation Foundation Model under Distribution Shift via Weakly Supervised Adaptation
by: Zhang, Haojie, et al.
Published: (2023)
by: Zhang, Haojie, et al.
Published: (2023)
CoCoCo: Improving Text-Guided Video Inpainting for Better Consistency, Controllability and Compatibility
by: Zi, Bojia, et al.
Published: (2024)
by: Zi, Bojia, et al.
Published: (2024)
Self and Mixed Supervision to Improve Training Labels for Multi-Class Medical Image Segmentation
by: Liu, Jianfei, et al.
Published: (2024)
by: Liu, Jianfei, et al.
Published: (2024)
DExTeR: Weakly Semi-Supervised Object Detection with Class and Instance Experts for Medical Imaging
by: Meyer, Adrien, et al.
Published: (2026)
by: Meyer, Adrien, et al.
Published: (2026)
Similar Items
-
3D-COCO: extension of MS-COCO dataset for image detection and 3D reconstruction modules
by: Bideaux, Maxence, et al.
Published: (2024) -
Benchmarking Adversarial Robustness and Adversarial Training Strategies for Object Detection
by: Winter, Alexis, et al.
Published: (2026) -
Faster and Better 3D Splatting via Group Training
by: Wang, Chengbo, et al.
Published: (2024) -
Fairer Analysis and Demographically Balanced Face Generation for Fairer Face Verification
by: Fournier-Montgieux, Alexandre, et al.
Published: (2024) -
Reliable and Reproducible Demographic Inference for Fairness in Face Analysis
by: Fournier-Montgieux, Alexandre, et al.
Published: (2025)