$Z^2$-Sampling: Zero-Cost Zigzag Trajectories for Semantic Alignment in Diffusion Models
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Li, Haosen, Chen, Wenshuo, Liang, Shaofeng, Wang, Lei, Yuan, Kaishen, Yue, Yutao |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2026
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Oracle Noise: Faster Semantic Spherical Alignment for Interpretable Latent Optimization
von: Li, Haosen, et al.
Veröffentlicht: (2026)
von: Li, Haosen, et al.
Veröffentlicht: (2026)
Guided Path Sampling: Steering Diffusion Models Back on Track with Principled Path Guidance
von: Li, Haosen, et al.
Veröffentlicht: (2025)
von: Li, Haosen, et al.
Veröffentlicht: (2025)
POLARIS: Projection-Orthogonal Least Squares for Robust and Adaptive Inversion in Diffusion Models
von: Chen, Wenshuo, et al.
Veröffentlicht: (2025)
von: Chen, Wenshuo, et al.
Veröffentlicht: (2025)
Delta Score Matters! Spatial Adaptive Multi Guidance in Diffusion Models
von: Li, Haosen, et al.
Veröffentlicht: (2026)
von: Li, Haosen, et al.
Veröffentlicht: (2026)
CoEmoGen: Towards Semantically-Coherent and Scalable Emotional Image Content Generation
von: Yuan, Kaishen, et al.
Veröffentlicht: (2025)
von: Yuan, Kaishen, et al.
Veröffentlicht: (2025)
Learning to Think in Physics: Breaking Shortcut Learning in Scientific Diffusion via Representation Alignment
von: Jia, Haozhe, et al.
Veröffentlicht: (2026)
von: Jia, Haozhe, et al.
Veröffentlicht: (2026)
Ctrl-Z Sampling: Scaling Diffusion Sampling with Controlled Random Zigzag Explorations
von: Mao, Shunqi, et al.
Veröffentlicht: (2025)
von: Mao, Shunqi, et al.
Veröffentlicht: (2025)
ANT: Adaptive Neural Temporal-Aware Text-to-Motion Model
von: Chen, Wenshuo, et al.
Veröffentlicht: (2025)
von: Chen, Wenshuo, et al.
Veröffentlicht: (2025)
LUMA: Low-Dimension Unified Motion Alignment with Dual-Path Anchoring for Text-to-Motion Diffusion Model
von: Jia, Haozhe, et al.
Veröffentlicht: (2025)
von: Jia, Haozhe, et al.
Veröffentlicht: (2025)
Efficient Conditional Diffusion Model with Probability Flow Sampling for Image Super-resolution
von: Yuan, Yutao, et al.
Veröffentlicht: (2024)
von: Yuan, Yutao, et al.
Veröffentlicht: (2024)
Physics-Informed Representation Alignment for Sparse Radio-Map Reconstruction
von: Jia, Haozhe, et al.
Veröffentlicht: (2025)
von: Jia, Haozhe, et al.
Veröffentlicht: (2025)
Before the Body Moves: Learning Anticipatory Joint Intent for Language-Conditioned Humanoid Control
von: Jia, Haozhe, et al.
Veröffentlicht: (2026)
von: Jia, Haozhe, et al.
Veröffentlicht: (2026)
Zigzag Diffusion Sampling: Diffusion Models Can Self-Improve via Self-Reflection
von: Bai, Lichen, et al.
Veröffentlicht: (2024)
von: Bai, Lichen, et al.
Veröffentlicht: (2024)
Consistent Story Generation: Unlocking the Potential of Zigzag Sampling
von: Li, Mingxiao, et al.
Veröffentlicht: (2025)
von: Li, Mingxiao, et al.
Veröffentlicht: (2025)
Free-T2M: Robust Text-to-Motion Generation for Humanoid Robots via Frequency-Domain
von: Chen, Wenshuo, et al.
Veröffentlicht: (2025)
von: Chen, Wenshuo, et al.
Veröffentlicht: (2025)
Unsafe by Reciprocity: How Generation-Understanding Coupling Undermines Safety in Unified Multimodal Models
von: Wang, Kaishen, et al.
Veröffentlicht: (2026)
von: Wang, Kaishen, et al.
Veröffentlicht: (2026)
Guarding the Gate: ConceptGuard Battles Concept-Level Backdoors in Concept Bottleneck Models
von: Lai, Songning, et al.
Veröffentlicht: (2024)
von: Lai, Songning, et al.
Veröffentlicht: (2024)
MedTVT-R1: A Multimodal LLM Empowering Medical Reasoning and Diagnosis
von: Zhang, Yuting, et al.
Veröffentlicht: (2025)
von: Zhang, Yuting, et al.
Veröffentlicht: (2025)
SPARS3R: Semantic Prior Alignment and Regularization for Sparse 3D Reconstruction
von: Tang, Yutao, et al.
Veröffentlicht: (2024)
von: Tang, Yutao, et al.
Veröffentlicht: (2024)
ECHO: Edge-Cloud Humanoid Orchestration for Language-to-Motion Control
von: Jia, Haozhe, et al.
Veröffentlicht: (2026)
von: Jia, Haozhe, et al.
Veröffentlicht: (2026)
SARA: Semantically Adaptive Relational Alignment for Video Diffusion Models
von: Lian, Jiesong, et al.
Veröffentlicht: (2026)
von: Lian, Jiesong, et al.
Veröffentlicht: (2026)
Semantic-Space-Intervened Diffusive Alignment for Visual Classification
von: Li, Zixuan, et al.
Veröffentlicht: (2025)
von: Li, Zixuan, et al.
Veröffentlicht: (2025)
ZigzagPointMamba: Spatial-Semantic Mamba for Point Cloud Understanding
von: Diao, Linshuang, et al.
Veröffentlicht: (2025)
von: Diao, Linshuang, et al.
Veröffentlicht: (2025)
Align-KD: Distilling Cross-Modal Alignment Knowledge for Mobile Vision-Language Model
von: Feng, Qianhan, et al.
Veröffentlicht: (2024)
von: Feng, Qianhan, et al.
Veröffentlicht: (2024)
Neural Electromagnetic Fields for High-Resolution Material Parameter Reconstruction
von: Chen, Zhe, et al.
Veröffentlicht: (2026)
von: Chen, Zhe, et al.
Veröffentlicht: (2026)
Exploring Conditional Multi-Modal Prompts for Zero-shot HOI Detection
von: Lei, Ting, et al.
Veröffentlicht: (2024)
von: Lei, Ting, et al.
Veröffentlicht: (2024)
Attribute Distribution Modeling and Semantic-Visual Alignment for Generative Zero-shot Learning
von: Pu, Haojie, et al.
Veröffentlicht: (2026)
von: Pu, Haojie, et al.
Veröffentlicht: (2026)
ANYPORTAL: Zero-Shot Consistent Video Background Replacement
von: Gao, Wenshuo, et al.
Veröffentlicht: (2025)
von: Gao, Wenshuo, et al.
Veröffentlicht: (2025)
AULLM++: Structural Reasoning with Large Language Models for Micro-Expression Recognition
von: Liu, Zhishu, et al.
Veröffentlicht: (2026)
von: Liu, Zhishu, et al.
Veröffentlicht: (2026)
GDTS: Goal-Guided Diffusion Model with Tree Sampling for Multi-Modal Pedestrian Trajectory Prediction
von: Sun, Ge, et al.
Veröffentlicht: (2023)
von: Sun, Ge, et al.
Veröffentlicht: (2023)
Diffusion-APO: Trajectory-Aware Direct Preference Alignment for Video Diffusion Transformers
von: Zhu, Jingyuan, et al.
Veröffentlicht: (2026)
von: Zhu, Jingyuan, et al.
Veröffentlicht: (2026)
FEALLM: Advancing Facial Emotion Analysis in Multimodal Large Language Models with Emotional Synergy and Reasoning
von: Hu, Zhuozhao, et al.
Veröffentlicht: (2025)
von: Hu, Zhuozhao, et al.
Veröffentlicht: (2025)
Semantic Image Synthesis via Diffusion Models
von: Zhou, Wengang, et al.
Veröffentlicht: (2022)
von: Zhou, Wengang, et al.
Veröffentlicht: (2022)
ConsistNav: Closing the Action Consistency Gap in Zero-Shot Object Navigation with Semantic Executive Control
von: Wang, Haosen, et al.
Veröffentlicht: (2026)
von: Wang, Haosen, et al.
Veröffentlicht: (2026)
Step-level Denoising-time Diffusion Alignment with Multiple Objectives
von: Zhang, Qi, et al.
Veröffentlicht: (2026)
von: Zhang, Qi, et al.
Veröffentlicht: (2026)
Nearly Zero-Cost Protection Against Mimicry by Personalized Diffusion Models
von: Ahn, Namhyuk, et al.
Veröffentlicht: (2024)
von: Ahn, Namhyuk, et al.
Veröffentlicht: (2024)
Unifying Visual and Semantic Feature Spaces with Diffusion Models for Enhanced Cross-Modal Alignment
von: Zheng, Yuze, et al.
Veröffentlicht: (2024)
von: Zheng, Yuze, et al.
Veröffentlicht: (2024)
ELLA: Equip Diffusion Models with LLM for Enhanced Semantic Alignment
von: Hu, Xiwei, et al.
Veröffentlicht: (2024)
von: Hu, Xiwei, et al.
Veröffentlicht: (2024)
Single-Sample Black-Box Membership Inference Attack against Vision-Language Models via Cross-modal Semantic Alignment
von: Li, Jiaqing, et al.
Veröffentlicht: (2026)
von: Li, Jiaqing, et al.
Veröffentlicht: (2026)
Frequency-Enhanced Diffusion Models: Curriculum-Guided Semantic Alignment for Zero-Shot Skeleton Action Recognition
von: Zhou, Yuxi, et al.
Veröffentlicht: (2026)
von: Zhou, Yuxi, et al.
Veröffentlicht: (2026)
Ähnliche Einträge
-
Oracle Noise: Faster Semantic Spherical Alignment for Interpretable Latent Optimization
von: Li, Haosen, et al.
Veröffentlicht: (2026) -
Guided Path Sampling: Steering Diffusion Models Back on Track with Principled Path Guidance
von: Li, Haosen, et al.
Veröffentlicht: (2025) -
POLARIS: Projection-Orthogonal Least Squares for Robust and Adaptive Inversion in Diffusion Models
von: Chen, Wenshuo, et al.
Veröffentlicht: (2025) -
Delta Score Matters! Spatial Adaptive Multi Guidance in Diffusion Models
von: Li, Haosen, et al.
Veröffentlicht: (2026) -
CoEmoGen: Towards Semantically-Coherent and Scalable Emotional Image Content Generation
von: Yuan, Kaishen, et al.
Veröffentlicht: (2025)