TMPO: Trajectory Matching Policy Optimization for Diverse and Efficient Diffusion Alignment
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Li, Jiaming, Zhu, Chenyu, Yi, Nanxi, Bao, Youjun, Sun, Li, Lv, Quanying, Fang, Xiang, Liu, Daizong, Li, Jianjun, He, Kun, Zhou, Bowen, Ma, Zhiyuan |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2026
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Flow Diverse and Efficient: Learning Momentum Flow Matching via Stochastic Velocity Field Sampling
von: Ma, Zhiyuan, et al.
Veröffentlicht: (2025)
von: Ma, Zhiyuan, et al.
Veröffentlicht: (2025)
Multi-Modal Cross-Domain Alignment Network for Video Moment Retrieval
von: Fang, Xiang, et al.
Veröffentlicht: (2022)
von: Fang, Xiang, et al.
Veröffentlicht: (2022)
LPG: Balancing Efficiency and Policy Reasoning in Latent Policy Guardrails
von: Li, Nanxi, et al.
Veröffentlicht: (2026)
von: Li, Nanxi, et al.
Veröffentlicht: (2026)
DCHO: A Decomposition-Composition Framework for Predicting Higher-Order Brain Connectivity to Enhance Diverse Downstream Applications
von: Li, Weibin, et al.
Veröffentlicht: (2025)
von: Li, Weibin, et al.
Veröffentlicht: (2025)
Kinematics-Aware Diffusion Policy with Consistent 3D Observation and Action Space for Whole-Arm Robotic Manipulation
von: Lv, Kangchen, et al.
Veröffentlicht: (2025)
von: Lv, Kangchen, et al.
Veröffentlicht: (2025)
Visualizing and Optimizing Phase Matching in Nonlinear Guided-mode Resonators with the Green's Function Integral Method
von: Liang, Chengkang, et al.
Veröffentlicht: (2025)
von: Liang, Chengkang, et al.
Veröffentlicht: (2025)
I2E: From Image Pixels to Actionable Interactive Environments for Text-Guided Image Editing
von: Yu, Jinghan, et al.
Veröffentlicht: (2026)
von: Yu, Jinghan, et al.
Veröffentlicht: (2026)
MILD: Multi-Layer Diffusion Strategy for Complex and Precise Multi-IP Aware Human Erasing
von: Yu, Jinghan, et al.
Veröffentlicht: (2025)
von: Yu, Jinghan, et al.
Veröffentlicht: (2025)
Variational Trajectory Optimization of Anisotropic Diffusion Schedules
von: Liu, Pengxi, et al.
Veröffentlicht: (2026)
von: Liu, Pengxi, et al.
Veröffentlicht: (2026)
Dual-Stream Decoupled Learning for Temporal Consistency and Speaker Interaction in AVSD
von: Xiao, Junhao, et al.
Veröffentlicht: (2025)
von: Xiao, Junhao, et al.
Veröffentlicht: (2025)
Zero-Shot Cellular Trajectory Map Matching
von: Shi, Weijie, et al.
Veröffentlicht: (2025)
von: Shi, Weijie, et al.
Veröffentlicht: (2025)
Hierarchical Local-Global Transformer for Temporal Sentence Grounding
von: Fang, Xiang, et al.
Veröffentlicht: (2022)
von: Fang, Xiang, et al.
Veröffentlicht: (2022)
Flash-GRPO: Efficient Alignment for Video Diffusion via One-Step Policy Optimization
von: He, Xiaoxuan, et al.
Veröffentlicht: (2026)
von: He, Xiaoxuan, et al.
Veröffentlicht: (2026)
Efficient Diffusion Models: A Comprehensive Survey from Principles to Practices
von: Ma, Zhiyuan, et al.
Veröffentlicht: (2024)
von: Ma, Zhiyuan, et al.
Veröffentlicht: (2024)
Bridging the Modality Gap: Dimension Information Alignment and Sparse Spatial Constraint for Image-Text Matching
von: Ma, Xiang, et al.
Veröffentlicht: (2024)
von: Ma, Xiang, et al.
Veröffentlicht: (2024)
Learning an Efficient Optimizer via Hybrid-Policy Sub-Trajectory Balance
von: Guan, Yunchuan, et al.
Veröffentlicht: (2025)
von: Guan, Yunchuan, et al.
Veröffentlicht: (2025)
On-Policy Self-Evolution via Failure Trajectories for Agentic Safety Alignment
von: Yin, Bo, et al.
Veröffentlicht: (2026)
von: Yin, Bo, et al.
Veröffentlicht: (2026)
Smoothing Meets Perturbation: Unified and Tight Analysis for Nonconvex-Concave Minimax Optimization
von: Li, Jiajin, et al.
Veröffentlicht: (2026)
von: Li, Jiajin, et al.
Veröffentlicht: (2026)
ET-SEED: Efficient Trajectory-Level SE(3) Equivariant Diffusion Policy
von: Tie, Chenrui, et al.
Veröffentlicht: (2024)
von: Tie, Chenrui, et al.
Veröffentlicht: (2024)
PRISM: Robust VLM Alignment with Principled Reasoning for Integrated Safety in Multimodality
von: Li, Nanxi, et al.
Veröffentlicht: (2025)
von: Li, Nanxi, et al.
Veröffentlicht: (2025)
UltraDP: Generalizable Carotid Ultrasound Scanning with Force-Aware Diffusion Policy
von: Chen, Ruoqu, et al.
Veröffentlicht: (2025)
von: Chen, Ruoqu, et al.
Veröffentlicht: (2025)
Visual Decoding and Reconstruction via EEG Embeddings with Guided Diffusion
von: Li, Dongyang, et al.
Veröffentlicht: (2024)
von: Li, Dongyang, et al.
Veröffentlicht: (2024)
Rethinking Video-Language Model from the Language Input Perspective
von: Fang, Xiang, et al.
Veröffentlicht: (2026)
von: Fang, Xiang, et al.
Veröffentlicht: (2026)
Task-Specific Distance Correlation Matching for Few-Shot Action Recognition
von: Long, Fei, et al.
Veröffentlicht: (2025)
von: Long, Fei, et al.
Veröffentlicht: (2025)
Learning Multi-Modal Trajectory Policies for Data-Efficient Robotic Manipulation
von: Chen, Zijia, et al.
Veröffentlicht: (2026)
von: Chen, Zijia, et al.
Veröffentlicht: (2026)
DSSP: Diffusion State Space Policy with Full-History Encoding
von: Guan, Zhiyuan, et al.
Veröffentlicht: (2026)
von: Guan, Zhiyuan, et al.
Veröffentlicht: (2026)
Q-MMR: Off-Policy Evaluation via Recursive Reweighting and Moment Matching
von: Li, Xiang, et al.
Veröffentlicht: (2026)
von: Li, Xiang, et al.
Veröffentlicht: (2026)
AutoTailor: Automatic and Efficient Adaptive Model Deployment for Diverse Edge Devices
von: Liu, Mengyang, et al.
Veröffentlicht: (2025)
von: Liu, Mengyang, et al.
Veröffentlicht: (2025)
Aligning Diffusion Model with Problem Constraints for Trajectory Optimization
von: Li, Anjian, et al.
Veröffentlicht: (2025)
von: Li, Anjian, et al.
Veröffentlicht: (2025)
Temporally Decoupled Diffusion Planning for Autonomous Driving
von: Li, Xiang, et al.
Veröffentlicht: (2026)
von: Li, Xiang, et al.
Veröffentlicht: (2026)
Preference Ranking Optimization for Human Alignment
von: Song, Feifan, et al.
Veröffentlicht: (2023)
von: Song, Feifan, et al.
Veröffentlicht: (2023)
Mamba Policy: Towards Efficient 3D Diffusion Policy with Hybrid Selective State Models
von: Cao, Jiahang, et al.
Veröffentlicht: (2024)
von: Cao, Jiahang, et al.
Veröffentlicht: (2024)
A Hybrid DNN Transformer AE Framework for Corporate Tax Risk Supervision and Risk Level Assessment
von: Song, Zhenzhen, et al.
Veröffentlicht: (2025)
von: Song, Zhenzhen, et al.
Veröffentlicht: (2025)
Reading ability detection using eye-tracking data with LSTM-based few-shot learning
von: Li, Nanxi, et al.
Veröffentlicht: (2024)
von: Li, Nanxi, et al.
Veröffentlicht: (2024)
Beyond Point-Wise Matching: Structural Representation Alignment for Accelerating Diffusion Transformers
von: Xu, Shaodong, et al.
Veröffentlicht: (2026)
von: Xu, Shaodong, et al.
Veröffentlicht: (2026)
Harnesses for Inference-Time Alignment over Execution Trajectories
von: Wang, Boyuan, et al.
Veröffentlicht: (2026)
von: Wang, Boyuan, et al.
Veröffentlicht: (2026)
Online Trajectory Optimization for Arbitrary-Shaped Mobile Robots via Polynomial Separating Hypersurfaces
von: Li, Shuoye, et al.
Veröffentlicht: (2026)
von: Li, Shuoye, et al.
Veröffentlicht: (2026)
Constraint-Aware Diffusion Models for Trajectory Optimization
von: Li, Anjian, et al.
Veröffentlicht: (2024)
von: Li, Anjian, et al.
Veröffentlicht: (2024)
Beyond Static Vision: Scene Dynamic Field Unlocks Intuitive Physics Understanding in Multi-modal Large Language Models
von: Li, Nanxi, et al.
Veröffentlicht: (2026)
von: Li, Nanxi, et al.
Veröffentlicht: (2026)
Gradient-Adaptive Policy Optimization: Towards Multi-Objective Alignment of Large Language Models
von: Li, Chengao, et al.
Veröffentlicht: (2025)
von: Li, Chengao, et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
Flow Diverse and Efficient: Learning Momentum Flow Matching via Stochastic Velocity Field Sampling
von: Ma, Zhiyuan, et al.
Veröffentlicht: (2025) -
Multi-Modal Cross-Domain Alignment Network for Video Moment Retrieval
von: Fang, Xiang, et al.
Veröffentlicht: (2022) -
LPG: Balancing Efficiency and Policy Reasoning in Latent Policy Guardrails
von: Li, Nanxi, et al.
Veröffentlicht: (2026) -
DCHO: A Decomposition-Composition Framework for Predicting Higher-Order Brain Connectivity to Enhance Diverse Downstream Applications
von: Li, Weibin, et al.
Veröffentlicht: (2025) -
Kinematics-Aware Diffusion Policy with Consistent 3D Observation and Action Space for Whole-Arm Robotic Manipulation
von: Lv, Kangchen, et al.
Veröffentlicht: (2025)